📊 Granite 4.0 1B — the actual numbers GPQA: 28.1% MMLU-Pro: 32.5% Humanity's Last Exam: 4.8% Long Context Reasoning: 6%M...

📊 Granite 4.0 1B — the actual numbers GPQA: 28.1% MMLU-Pro: 32.5% Humanity's Last Exam: 4.8% Long Context Reasoning: 6%Measured independently, not self-reported →https://olud.ai/leaderboard.html#LLM #Benchmarks #OpenSource #AI

Read Original

Related

Mastodon discussion 22m ago

https://technews.tw/2026/08/15/watermarks-remover/【開源工具挑戰 AI 溯源,可移除 Anthropic、OpenAI 等模型浮水印】『一項名為 watermarks-remover 的開源...

https://technews.tw/2026/08/15/watermarks-remover/【開源工具挑戰 AI 溯源,可移除 Anthropic、OpenAI 等模型浮水印】『一項名為 watermarks-remover 的開源工具在 GitHub 平台釋出,能從 AI 生成的內容移除不同類型的 AI 溯源辨識標記,包含 Unicode 字元、統...