📊 K-EXAONE (Reasoning) — the actual numbers GPQA: 78.3% MMLU-Pro: 83.8% Humanity's Last Exam: 13.1% Long Context Reasoni...

📊 K-EXAONE (Reasoning) — the actual numbers GPQA: 78.3% MMLU-Pro: 83.8% Humanity's Last Exam: 13.1% Long Context Reasoning: 55.7%Measured independently, not self-reported →https://opensourceai.tech/leaderboard.html#LLM #Benchmarks #OpenSource #AI

Read Original

Related