We rank open-weight and open-source language models using Chatbot Arena+ ELO, SWE-bench Verified, reasoning benchmarks, and API pricing. Data from OpenLM, Onyx, LLM Stats, and BenchLM leaderboards.
Each model is scored on a weighted 4-criteria matrix. Only open-weight/open-source models (MIT, Apache 2.0, Gemma, Llama licenses).
Sources: Chatbot Arena+ ยท Onyx LLM ยท LLM Stats ยท BenchLM
Our config packs include optimized prompts and system instructions for DeepSeek, Qwen, Gemma, Mistral, and more. Ready to use in any agent.