/// AI HUB
Dashboard News Models Tools Papers Repos Videos Companies Trending
Login

#LLM

9604 articles tagged with LLM

Latest Trending
Dev.to tutorial 2d ago

Writing a Contract Test Suite for Your Own LLM Gateway

The assertions that belong to the proxy layer rather than to the model behind it, tested against a stub upstream so they run in milliseconds.

LLM
12
Mastodon discussion 2d ago

Prompt injection is still king, but “excessive agency” just jumped to #3 in OWASP’s Top 10 for LLM apps. Stop chasing un...

Prompt injection is still king, but “excessive agency” just jumped to #3 in OWASP’s Top 10 for LLM apps. Stop chasing unbreakable models—start containing fooled agents. https://jpm...

LLM
18
Mastodon discussion 2d ago

Researchers at IIT Bombay and Adobe Research built a method called Previous-Token Prediction that reconstructs an LLM’s ...

Researchers at IIT Bombay and Adobe Research built a method called Previous-Token Prediction that reconstructs an LLM’s original prompt from its output with near-perfect accuracy. ...

LLM
9
Mastodon discussion 2d ago

Qwen3.8-2.4T-A95B is finally out. Does anyone have a setup that can run this thing at home? lol #LLM #AI https://hugging...

Qwen3.8-2.4T-A95B is finally out. Does anyone have a setup that can run this thing at home? lol #LLM #AI https://huggingface.co/Qwen/Qwen3.8-2.4T-A95B

LLM
27
Dev.to tutorial 2d ago

What LLM Cost Calculators Get Wrong

Is the number these calculators show you actually right? Not whether the model is in the catalog —...

LLM
12
Mastodon discussion 2d ago

New arXiv paper introduces exact Likert-scale framework to isolate LLM biases and attitudes, enabling controlled behavio...

New arXiv paper introduces exact Likert-scale framework to isolate LLM biases and attitudes, enabling controlled behavioral evaluation for autonomous agentsSource: arXiv cs.CLhttps...

LLM
9
Mastodon discussion 2d ago

VibeLifeBench introduces a benchmark of 200 long-horizon tasks to test if LLM agents can act proactively and persistentl...

VibeLifeBench introduces a benchmark of 200 long-horizon tasks to test if LLM agents can act proactively and persistently in a changing simulated world. Seven leading models all sc...

LLM Benchmark
9
Mastodon discussion 2d ago

New benchmark finds automated evaluation of tool-using LLM agents often unreliable; GPT-4o-mini surpasses heuristic judg...

New benchmark finds automated evaluation of tool-using LLM agents often unreliable; GPT-4o-mini surpasses heuristic judging, and runtime interceptors cut hallucinations by 24 perce...

LLM Multimodal Benchmark
9
Mastodon discussion 2d ago

Living-Harness is a self-evolving agent framework that updates procedural knowledge from failure patterns to improve LLM...

Living-Harness is a self-evolving agent framework that updates procedural knowledge from failure patterns to improve LLM agent reliability, showing double-digit performance gains i...

LLM
9
Mastodon discussion 2d ago

The nice thing about OpenTelemetry for LLM apps: instrumentation is decoupled from the backend, so traces from your Pyth...

The nice thing about OpenTelemetry for LLM apps: instrumentation is decoupled from the backend, so traces from your Python, Go, and Java services land in one queryable place. Multi...

LLM
9
Mastodon discussion 2d ago

And it's not the only place from which I've been banned, btw. Using #LLM is not only dangerous for your privacy, but als...

And it's not the only place from which I've been banned, btw. Using #LLM is not only dangerous for your privacy, but also for your sociality !

LLM
9
Mastodon discussion 2d ago

Does anyone know of an #LLM friendly open #Mastodon instance ? I was banned from floss.social but I'm too lazy to self-h...

Does anyone know of an #LLM friendly open #Mastodon instance ? I was banned from floss.social but I'm too lazy to self-host something ( or I wouldn't be using LLMs right? )Fully st...

LLM
33
Mastodon discussion 2d ago

I really think that at this point, the brains of some of the anti-LLM people are exactly as fried as the worst AI bros'....

I really think that at this point, the brains of some of the anti-LLM people are exactly as fried as the worst AI bros'.Do they actually not have any senior devs around them who so...

LLM
24
Mastodon discussion 2d ago

@captaincalliope.at it's irresponsible to use #LLM tools.

@captaincalliope.at it's irresponsible to use #LLM tools.

LLM
24
Mastodon discussion 2d ago

"Help help! We accidentally fed a book about the Trojan Horse into our LLM, and it generated a bunch of angry and warlik...

"Help help! We accidentally fed a book about the Trojan Horse into our LLM, and it generated a bunch of angry and warlike Greeks!"https://freeradical.zone/@funnymonkey/117083125243...

LLM
9
Mastodon discussion 2d ago

Anyone else reading this or working on production-grade AI architecture as we speak?#SystemDesign #LLM #AI #SoftwareEngi...

Anyone else reading this or working on production-grade AI architecture as we speak?#SystemDesign #LLM #AI #SoftwareEngineering

LLM
9
Mastodon discussion 2d ago

The frontier #AI model security vulnerability finding orgy is the Internet foreclosing on tech debt.#infosec #LLM

The frontier #AI model security vulnerability finding orgy is the Internet foreclosing on tech debt.#infosec #LLM

LLM
24
Dev.to tutorial 2d ago

Choosing the Right LLM-as-a-Judge: A Practical Guide with Model Recommendations

How to pick the best LLM judge for your RAG, generation, and OCR tasks. Data-driven recommendations, debiasing strategies, and cost-performance tradeoffs.

LLM
12
Mastodon discussion 2d ago

Third silent truncation in Lookspan in five review passes. This one was the worst.Before anything reaches the LLM judge ...

Third silent truncation in Lookspan in five review passes. This one was the worst.Before anything reaches the LLM judge it gets cut to 12,000 characters. The cut carried no marker....

LLM
9
Mastodon discussion 2d ago

https://www.youtube.com/watch?v=kON2ZI2BNj8#DataCenters #LLM is still not #AI

https://www.youtube.com/watch?v=kON2ZI2BNj8#DataCenters #LLM is still not #AI

LLM
9
Mastodon discussion 2d ago

How I feel about #AI "Detectors"#LLM

How I feel about #AI "Detectors"#LLM

LLM
18
Dev.to tutorial 2d ago

Deploying DeepSeek V3 (LLM) Using SGLang

DeepSeek V3 is a 671B-parameter Mixture-of-Experts language model: Multi-head Latent Attention and...

LLM
12
Dev.to tutorial 2d ago

Deploying Langfuse – Open-Source LLM Observability Platform

Langfuse is an open-source observability platform for LLM applications including traces...

LLM Open Source
12
Mastodon discussion 2d ago

I assume this sounded better in the original Chinese.#AI #LLM #shitpost

I assume this sounded better in the original Chinese.#AI #LLM #shitpost

LLM
9
« Previous Page 7 of 401 (9604 items) Next »
AI Hub // AI Intelligence Platform // LIVE FEED // Impressum // Datenschutz © 2026
0 new articles available