📊 DeepSeek V3.2 (Reasoning) — the actual numbers GPQA: 84% MMLU-Pro: 86.2% Humanity's Last Exam: 22.2% Long Context Reas...

📊 DeepSeek V3.2 (Reasoning) — the actual numbers GPQA: 84% MMLU-Pro: 86.2% Humanity's Last Exam: 22.2% Long Context Reasoning: 65%💰 101.6 intelligence points per dollarMeasured independently, not self-reported →https://opensourceai.tech/leaderboard.html#LLM #Benchmarks #OpenSource #AI

Read Original

Related