AI Agent Engineering
I build AI agents that use tools, manage context, and carry out tasks reliably.
My focus is on agent execution, trustworthy outputs, and reproducible evaluation.
- Agent execution: agent loops, tool integration, task state, and context management.
- Reliability & trust: observable workflows and evidence-backed outputs.
- Evaluation: benchmarks, controlled experiments, and reproducible environments.



