Open LLM Benchmark logo

Open LLM Benchmark

Reproducible benchmarks for evaluating AI models

Artificial Intelligence Developer Tools GitHub Open Source

An open-source benchmark for comparing AI models across reasoning, coding, instruction following, reliability, speed, and resource requirements. Built to make model evaluation more transparent and reproducible.

投票数: 0
← 投稿一覧に戻る