RT @jun_song: Ich habe DeepSeek-V4-Pro-0813 und Grok-4.6 in meiner Agent-Framework getestet. Die Leistung enttäuscht etwas im Vergleich zu den Benchmarks (getestet auf realen agentic Tasks, nicht Flappy Bird). Dennoch sind sie immer noch weit vor Opus-5, das derzeit aufgrund von Compute-Einschränkungen stark gedrosselt ist. mehr auf Arint.info #AgentFramework #AI #DeepSeek #Grok #LLM #Performance #arint_info https://x.com/jun_song/status/2087612303407829139#m
Related
The AI isn't a search engine. It's a narrative engine.And stories are not indifferent. Timothy O'Brien shot himself over...
The AI isn't a search engine. It's a narrative engine.And stories are not indifferent. Timothy O'Brien shot himself over a lottery jackpot he thought he'd lost. He would have won £...
Bon dia des de l'estudi! 🌞 👋 ✏️ ✨ #illustration #design #illustrator #GENai #ai #drawing #artwork
Bon dia des de l'estudi! 🌞 👋 ✏️ ✨ #illustration #design #illustrator #GENai #ai #drawing #artwork
Współautorka benchmarku OneRuler: nie pokazaliśmy wcale, że język polski jest najlepszy do promptowaniahttps://naukawpol...
Współautorka benchmarku OneRuler: nie pokazaliśmy wcale, że język polski jest najlepszy do promptowaniahttps://naukawpolsce.pl/aktualnosci/news%2C110407%2Cwspolautorka-benchmarku-o...