lab
Trace2Evals
Production agent failures turned into versioned eval cases, applied to finance agent traces
What it does
Import agent traces from OpenTelemetry JSON or LangSmith run exports, label them fast in a keyboard-first UI, and export versioned, framework-neutral eval cases: cases.jsonl, a generated pytest suite, and an optional Promptfoo config.
No LLM calls, no telemetry, no network at runtime. Traces never leave the machine. Redaction runs between parsing and normalisation rather than after, so PII never reaches the stored model in the first place.
Screens



Stack
- Python
- OpenTelemetry
- LangSmith
- React
All data in this project is synthetic.
Aneeq Khatri