Open-Source LLM and Leaderboard 2026 Collaboration
benchmarks claude deepseek gemini llama open-source reasoning
| Source: Mastodon | Original article
Open-source LLMs ranked in new leaderboard. Top model achieves 82.4% on MMLU-Pro benchmark.
The Open-Source LLM Leaderboard 2026 has been released, providing an independent ranking of top AI models. According to the leaderboard, Kimi K2 achieves impressive scores, including 76.6% on GPQA and 82.4% on MMLU-Pro. Its performance is measured in terms of intelligence points per dollar, with 19.4 points per dollar.
This leaderboard matters as it offers a transparent and continuously updated comparison of over 300 AI models, including GPT, Claude, and Llama. The rankings are based on public benchmarks and live API metrics, allowing developers to make informed decisions when selecting AI models for their projects.
As the AI landscape continues to evolve, it will be interesting to watch how these rankings change over time. With multiple leaderboards available, including those from BenchLM.ai and WhatLLM.org, the competition among AI models is expected to drive innovation and improvement in the field.
Sources
Back to AIPULSEN