Classifying free-text queries into a fixed product taxonomy is one of the most common LLM workloads in production, and one of the most quietly expensive ones. A
Why Your LLM Classifier Doesn't Need the Taxonomy: Hypothetical Classification with Embeddings