Hugging Face AI模型与数据集社区,全球最大的开源AI模型托管平台
The AI Leaderboard — independent rankings of GPT, Claude, Gemini, Llama, DeepSeek and 300+ AI models by intelligence, speed and price. Composite LLM Stats Score updated continuously from public ben...
LLM Leaderboard compares 40+ AI models by benchmark score, speed, and API cost. See GPT vs Claude vs Gemini vs Sarvam AI rankings, live model performance, and independent benchmark analysis.
Comparison and analysis of AI models and API hosting providers. Independent benchmarks across key performance metrics including quality, price, output speed & latency.
An automatic evaluator for instruction-following language models. Human-validated, high-quality, cheap, and fast. - tatsu-lab/alpaca_eval