Alibaba released benchmark scores for Qwen3.8-Max two weeks after claiming superiority without data. The 2.4T-parameter model scores come from internal testing only. Independent verification remains pending. What to watch: third-party results and the unstated license for weights arriving next week. https://www.implicator.ai/alibaba-publishes-the-qwen3-8-max-benchmarks-it-withheld-two-weeks-ago/ #AI #LLMs #benchmarks
Related
Dead text or binding clause? Measuring and restoring constraint influence in black-box LLM dialoguesSource: arXiv cs.AIh...
Dead text or binding clause? Measuring and restoring constraint influence in black-box LLM dialoguesSource: arXiv cs.AIhttps://arxiv.org/abs/2608.12599#MachineLearning
Masowa produkcja treści przez AI drastycznie obniża zarobki realnych autorów na Amazonie. Nowe dane pokazują, że algoryt...
Masowa produkcja treści przez AI drastycznie obniża zarobki realnych autorów na Amazonie. Nowe dane pokazują, że algorytmiczne tytuły przejmują już jedną trzecią list bestsellerów....
RE: https://cosocial.ca/@jfmezei/117097914922561340Good illustration of one of the major things most people don't get: T...
RE: https://cosocial.ca/@jfmezei/117097914922561340Good illustration of one of the major things most people don't get: The AI has no connection between the elements in the image - ...