n8n/packages/@n8n/instance-ai/evaluations/harness
José Braulio González Valido b0a10229a0
refactor(ai-builder): Split the eval harness runner into domain modules (no-changelog) (#34834)
Co-authored-by: Claude Fable 5 <noreply@anthropic.com>
2026-07-27 09:01:09 +00:00
..
artifacts test(ai-builder): Support evals for agent building and for config evals (no-changelog) (#33888) 2026-07-15 12:34:21 +00:00
agent-execution.ts refactor(ai-builder): Split the eval harness runner into domain modules (no-changelog) (#34834) 2026-07-27 09:01:09 +00:00
build-workflow.ts refactor(ai-builder): Split the eval harness runner into domain modules (no-changelog) (#34834) 2026-07-27 09:01:09 +00:00
capture-run-debug.ts ci: Bound and instrument Instance AI eval lane containers (no-changelog) (#33903) 2026-07-14 09:28:46 +00:00
chat-loop.ts refactor(ai-builder): Split the eval harness runner into domain modules (no-changelog) (#34834) 2026-07-27 09:01:09 +00:00
cleanup.ts refactor(ai-builder): Split the eval harness runner into domain modules (no-changelog) (#34834) 2026-07-27 09:01:09 +00:00
conversation-seed.ts refactor: Remove @n8n/utils barrel export (no-changelog) (#33082) 2026-06-30 16:05:24 +03:00
in-memory-event-bus.ts chore(ai-builder): Typecheck the eval harness and pin its external contracts (no-changelog) (#34673) 2026-07-22 08:18:41 +00:00
langsmith-seed.ts chore(ai-builder): Typecheck the eval harness and pin its external contracts (no-changelog) (#34673) 2026-07-22 08:18:41 +00:00
logger.ts feat(ai-builder): Workflow evaluation framework with LLM mock execution (#27818) 2026-04-07 13:31:16 +00:00
normalize-workflow.ts chore(ai-builder): Typecheck the eval harness and pin its external contracts (no-changelog) (#34673) 2026-07-22 08:18:41 +00:00
parse-seed-workflow.ts feat(ai-builder): Add conversation pre-seeding to workflow evals (no-changelog) (#32196) 2026-06-23 17:30:30 +00:00
prebuilt-workflows.ts refactor(ai-builder): Split the eval harness runner into domain modules (no-changelog) (#34834) 2026-07-27 09:01:09 +00:00
redact.ts feat(ai-builder): Author-level conversation expectations on eval test cases (no-changelog) (#31787) 2026-06-09 10:06:07 +00:00
sandbox-config.ts feat(core): Use n8n default sandbox for Instance AI (no-changelog) (#31335) 2026-06-04 08:31:27 +00:00
scenario-execution.ts refactor(ai-builder): Split the eval harness runner into domain modules (no-changelog) (#34834) 2026-07-27 09:01:09 +00:00
schema.ts feat: Add typed seed data tables with rows on execution scenarios (#34420) 2026-07-17 17:30:19 +00:00
seed-tables.ts refactor(ai-builder): Split the eval harness runner into domain modules (no-changelog) (#34834) 2026-07-27 09:01:09 +00:00
stub-services.ts chore(ai-builder): Typecheck the eval harness and pin its external contracts (no-changelog) (#34673) 2026-07-22 08:18:41 +00:00
transient-error.ts fix: Do not lose scenario results on a budget/timeout abort (#34419) 2026-07-17 11:15:28 +00:00
workflow-context.ts feat(ai-builder): Author-level conversation expectations on eval test cases (no-changelog) (#31787) 2026-06-09 10:06:07 +00:00