n8n/packages/@n8n/instance-ai/evaluations/comparison
José Braulio González Valido b93a99945b
feat(ai-builder): Harden the eval harness — crash-recovery journal, parity fixes, extension points (no-changelog) (#34747)
Co-authored-by: Claude Fable 5 <noreply@anthropic.com>
2026-07-23 12:25:55 +00:00
..
bucket-from-evaluation.ts feat(ai-builder): Persist eval expectation verdicts to LangSmith run outputs (no-changelog) (#33788) 2026-07-09 22:21:46 +00:00
compare.ts feat(ai-builder): Persist eval expectation verdicts to LangSmith run outputs (no-changelog) (#33788) 2026-07-09 22:21:46 +00:00
fetch-baseline.ts refactor(ai-builder): Decompose the eval CLI into phase modules with typed seams (no-changelog) (#34691) 2026-07-22 16:25:44 +00:00
format.ts test(ai-builder): Support evals for agent building and for config evals (no-changelog) (#33888) 2026-07-15 12:34:21 +00:00
gate.ts feat(ai-builder): Harden the eval harness — crash-recovery journal, parity fixes, extension points (no-changelog) (#34747) 2026-07-23 12:25:55 +00:00
statistics.ts feat(ai-builder): Add per-PR eval regression detection vs LangSmith baseline (#29456) 2026-05-06 08:15:08 +00:00