中文
~/home / Open Source / ProductHunt
ProductHunt · Open Source

EvalCore

EvalCore is a coding AI agent for Test and validate LLM application behavior using YAML configuration, local execution, and offline replay. Pricing: Free and open-source. As of 2026-09-11, KanonAgent records 11 upvotes. First indexed by KanonAgent on 2026-07-20.
EvalCore enables developers to test and validate LLM behavior locally using YAML config, offline replay, and in-CI workflows.
EvalCore — official preview image
11 upvotes
Tracked by Kanon since Jul 20, 2026
no signal yet
status
11/100
momentum · conf 0.48
20d
tracked since 2026-08-22
2
evidence records

No signal event on file yet — status is unknown, not "quiet".

📈 Timelinewhat changed, and when
🔎 Known / Unknown 12 of 23 fields unknown
Machine-callable unknown
Open source unknown
Self-hostable1 inferred
Bring your own key unknown
Autonomy level unknown
Pricing modelfree inferred 2 quote(s)
Integrations unknown

“Unknown” means we have not verified it — it is not a “no”. Hard filters never treat unknown as false.

🤖 Agent teardown · coding
Job to be doneTest and validate LLM application behavior using YAML configuration, local execution, and offline replay
Who it is forAI engineering teams building and refining LLM applications
Prerequisitesself-hostable
PricingFree and open-source
Real cost$0
Traction · why it is rising11 votes signal growing adoption among developers seeking reliable, reproducible testing tools
Why it matters

It standardizes and localizes LLM testing, making it a foundational tool for reliable AI development pipelines

Evidence quotesverbatim, from the product’s own materials

“replay model or judge calls offline for $0”— description
“single-binary test runner for LLM apps and agents. Define cases and scorers in YAML, run local targets on every PR”— description
Open Source
Signal source: ProductHunt
Visit official site →
📛 Official badgefor your site / README
EvalCore badge
Building EvalCore? Pick a style above — the embed code updates live. Deep color control via URL params: bg= / fg= / accent= (hex). It links back to this page.
Share on X
On mobile tap Share for WeChat / RED (Xiaohongshu) / X; on desktop use Copy text and paste into the app.

FAQ

What is EvalCore?

EvalCore enables developers to test and validate LLM behavior locally using YAML config, offline replay, and in-CI workflows.

What does EvalCore do?

Test and validate LLM application behavior using YAML configuration, local execution, and offline replay

Why does EvalCore matter?

It standardizes and localizes LLM testing, making it a foundational tool for reliable AI development pipelines

How much does EvalCore cost?

Free and open-source

Is EvalCore free?

Yes — EvalCore has a free tier. Pricing as stated on its own page: Free and open-source

Can I self-host EvalCore?

Yes — EvalCore can be self-hosted.

How popular is EvalCore?

As tracked by KanonAgent: 11 upvotes (first indexed 2026-07-20).

What are the best EvalCore alternatives?

Similar AI agents tracked by KanonAgent: deepseek-harness, gemini-cli, Codédex, LLM Gateway, oh-my-openagent, goose.

EvalCore alternatives — similar AI agents

deepseek-harnessgemini-cliCodédexLLM Gatewayoh-my-openagentgoose

Where this fits — browse the same shelf

AI Coding AgentsAI agents for Local offline AIFree AI agentsSelf-hosted & open-source AI agents