Jev Users
← All projects

jevbench

dhruvmehra/jevbench MIT Python Eval & Align

rank#293
What it does

Benchmark TypeSafe JEV against LLMs, fine-tuned BERT, Laya and zero-shot NLI on text classification: accuracy, calibration, latency, throughput, cost

Numbers
Stars
1
Forks
0
Open issues
0
Last push
today

Star history starts building from the first daily refresh.

Open on GitHub →

More in Eval & Align