I'm Mert, a Harness Engineer and OSS Maintainer focused on Agentic Systems & Applied AI.
I specialize in agent harnesses, multi-agent orchestration, CodeMode, and agent interoperability. I work with A2A, ACP, MCP, Pydantic AI, LangChain, Codex App Server, and Claude Agent SDK, with a focus on typed APIs, evaluation, and observability.
Website · LinkedIn · PyPI · Email
| Area | Stack |
|---|---|
| Languages | |
| Python & backend | |
| Agent frameworks & runtimes | |
| AI Models | |
| Optimization & evaluation | |
| Testing & observability | |
| Protocols | |
| Development tools | |
| Infrastructure & data |
| Project | What it does |
|---|---|
| VSH | Rust-powered transactional filesystem simulation: execute in a virtual snapshot, inspect changes, and commit through explicit policies. Docs |
| AutoBench | Benchmarking infrastructure for comparing variants, capturing semantic evidence and asset versions, and replaying recorded experiments. Docs |
| ACP Kit | ACP adapters for Pydantic AI and LangChain / LangGraph, with session state, approvals, runtime capability mapping, and remote transport helpers. Docs |
| pydantic-gepa | A typed optimization runtime combining GEPA, Pydantic AI, and Pydantic Evals, with candidate injection, reflection, and checkpoint/resume. Docs |
Interested in collab about agentic systems, harnesses, or evaluation? contact@tomris.dev





