No signal event on file yet — status is unknown, not "quiet".
- first observed Kanon started keeping permanent history for this agent
| Machine-callable | — | unknown | — |
| Open source | — | unknown | — |
| Self-hostable | 1 | inferred | — |
| Bring your own key | — | unknown | — |
| Autonomy level | — | unknown | — |
| Pricing model | free | inferred | 2 quote(s) |
| Integrations | — | unknown | — |
“Unknown” means we have not verified it — it is not a “no”. Hard filters never treat unknown as false.
It standardizes and localizes LLM testing, making it a foundational tool for reliable AI development pipelines
Evidence quotesverbatim, from the product’s own materials
“replay model or judge calls offline for $0”— description
“single-binary test runner for LLM apps and agents. Define cases and scorers in YAML, run local targets on every PR”— description
FAQ
What is EvalCore?
EvalCore enables developers to test and validate LLM behavior locally using YAML config, offline replay, and in-CI workflows.
What does EvalCore do?
Test and validate LLM application behavior using YAML configuration, local execution, and offline replay
Why does EvalCore matter?
It standardizes and localizes LLM testing, making it a foundational tool for reliable AI development pipelines
How much does EvalCore cost?
Free and open-source
Is EvalCore free?
Yes — EvalCore has a free tier. Pricing as stated on its own page: Free and open-source
Can I self-host EvalCore?
Yes — EvalCore can be self-hosted.
How popular is EvalCore?
As tracked by KanonAgent: 11 upvotes (first indexed 2026-07-20).
What are the best EvalCore alternatives?
Similar AI agents tracked by KanonAgent: deepseek-harness, gemini-cli, Codédex, LLM Gateway, oh-my-openagent, goose.