Launching today

Open LLM Benchmark
Reproducible benchmarks for evaluating AI models
1 follower
Reproducible benchmarks for evaluating AI models
1 follower
An open-source benchmark for comparing AI models across reasoning, coding, instruction following, reliability, speed, and resource requirements. Built to make model evaluation more transparent and reproducible.
No makers yet
It looks like there are no makers for this product.