📊 Nova 2.0 Lite (medium) — the actual numbers GPQA: 76.8% MMLU-Pro: 81.3% Humanity's Last Exam: 8.6% Long Context Reasoning: 58.3%⚡ 217.9 tokens/sec💰 22.4 intelligence points per dollarMeasured independently, not self-reported →https://olud.ai/leaderboard.html#LLM #Benchmarks #OpenSource #AI
Related
Dead text or binding clause? Measuring and restoring constraint influence in black-box LLM dialoguesSource: arXiv cs.AIh...
Dead text or binding clause? Measuring and restoring constraint influence in black-box LLM dialoguesSource: arXiv cs.AIhttps://arxiv.org/abs/2608.12599#MachineLearning
Masowa produkcja treści przez AI drastycznie obniża zarobki realnych autorów na Amazonie. Nowe dane pokazują, że algoryt...
Masowa produkcja treści przez AI drastycznie obniża zarobki realnych autorów na Amazonie. Nowe dane pokazują, że algorytmiczne tytuły przejmują już jedną trzecią list bestsellerów....
RE: https://cosocial.ca/@jfmezei/117097914922561340Good illustration of one of the major things most people don't get: T...
RE: https://cosocial.ca/@jfmezei/117097914922561340Good illustration of one of the major things most people don't get: The AI has no connection between the elements in the image - ...