← All projects
jev-coherence-bench
ERA-Fathom/jev-coherence-bench MIT Python Eval & Align
rank#320
Benchmark of a committed-state coherence read against Jev, the typed verifier, on agent runs with injected defects. Standard library only.
- Stars
- 0
- Forks
- 0
- Open issues
- 0
- Last push
- today
Star history starts building from the first daily refresh.
agent-evaluationai-agentsbenchmarkcoding-agentsllm-evaluationpythonverification