← Back to Benchmarks
Coding
ClarEval
A benchmark specifically evaluating code agents' ability to ask clarifying questions and seek information when faced with ambiguous programming instructions, not just task completion.
Year2026
This entry was automatically discovered and hasn't been researched yet. Sections below fill in as enrichment completes. Discovered 7/13/2026.
What It Tests
A benchmark specifically evaluating code agents' ability to ask clarifying questions and seek information when faced with ambiguous programming instructions, not just task completion.
Discovery notes
Notes the discovery agent wrote when proposing this benchmark.
Fills practical gap in agent evaluation: ability to handle unclear requirements. Tests a critical real-world capability previously unmeasured in coding benchmarks.