📊 Qwen3 235B A22B (Reasoning) — the actual numbers GPQA: 70% MMLU-Pro: 82.8% Humanity's Last Exam: 11% Long Context Reas...

📊 Qwen3 235B A22B (Reasoning) — the actual numbers GPQA: 70% MMLU-Pro: 82.8% Humanity's Last Exam: 11% Long Context Reasoning: 0%💰 5.1 intelligence points per dollarMeasured independently, not self-reported →https://olud.ai/leaderboard.html#LLM #Benchmarks #OpenSource #AI

Read Original

Related