📊 K-EXAONE (Reasoning) — the actual numbers GPQA: 78.3% MMLU-Pro: 83.8% Humanity's Last Exam: 13.1% Long Context Reasoning: 55.7%Measured independently, not self-reported →https://opensourceai.tech/leaderboard.html#LLM #Benchmarks #OpenSource #AI
Related
AI NPCs could make games more dynamic... but more dialogue doesn't automatically mean better writing. Players aren't too...
AI NPCs could make games more dynamic... but more dialogue doesn't automatically mean better writing. Players aren't too dumb to notice when human creativity gets replaced by short...
Everything is about to "go dark"Article URL: https://blog.cryptographyengineering.com/2026/08/14/everything-is-about-to-...
Everything is about to "go dark"Article URL: https://blog.cryptographyengineering.com/2026/08/14/everything-is-about-to-go-dark/ Comments URL: https://news.ycombinator.com/item?id=...
A YOLO- and CLIP-based vision-language framework classifies mosquito flight frames of uninfected and Dengue virus seroty...
A YOLO- and CLIP-based vision-language framework classifies mosquito flight frames of uninfected and Dengue virus serotype 2-infected mosquitoes with 98.54% accuracy and 99.91% sen...