📊 Qwen3 VL 8B (Reasoning) — the actual numbers GPQA: 57.9% MMLU-Pro: 74.9% Humanity's Last Exam: 3.3% Long Context Reaso...

📊 Qwen3 VL 8B (Reasoning) — the actual numbers GPQA: 57.9% MMLU-Pro: 74.9% Humanity's Last Exam: 3.3% Long Context Reasoning: 31%⚡ 128.9 tokens/sec💰 16.1 intelligence points per dollarMeasured independently, not self-reported →https://opensourceai.tech/leaderboard.html#LLM #Benchmarks #OpenSource #AI

Read Original

Related