Launching today

Open LLM Benchmark
Reproducible benchmarks for evaluating AI models
1 follower
Reproducible benchmarks for evaluating AI models
1 follower
An open-source benchmark for comparing AI models across reasoning, coding, instruction following, reliability, speed, and resource requirements. Built to make model evaluation more transparent and reproducible.
Open LLM Benchmark Reviews
Reviews