Catalog
Each run writes
input.json, results.json, report.md, failures.jsonl, cost.json, and traces/ under .zenrows/evals/<run-id>/.A reproducible benchmark bundled with the Zenrows CLI, with explicit pass or fail criteria.
input.json, results.json, report.md, failures.jsonl, cost.json, and traces/ under .zenrows/evals/<run-id>/.