← All projects
jev-judge
orq-ai/jev-judge MIT Python Other
rank#378
Judge repeatability study: Jev via Orq classify vs LLM judges on a docs agent. Frozen runs, labels, scripts.
- Stars
- 0
- Forks
- 0
- Open issues
- 0
- Last push
- today
Star history starts building from the first daily refresh.