← All projects
jevbench
dhruvmehra/jevbench MIT Python Eval & Align
rank#293
Benchmark TypeSafe JEV against LLMs, fine-tuned BERT, Laya and zero-shot NLI on text classification: accuracy, calibration, latency, throughput, cost
- Stars
- 1
- Forks
- 0
- Open issues
- 0
- Last push
- today
Star history starts building from the first daily refresh.