Mistral has released Mistral OCR 4, advancing document understanding with support for bounding boxes, block classification, and inline confidence scores. Each extracted content block is now localized, classified by type, and accompanied by per-page and per-word confidence metrics, alongside the textual output. The model expands accessibility by supporting 170 languages across 10 language groups, including those that are rare or low-resource, addressing a gap in many existing solutions. Building on these enhancements, Mistral OCR 4 accepts common enterprise document formats: PDF, DOC, PPT, and OpenDocument, broadening its suitability for corporate workflows. For deployment, Mistral OCR 4 runs as a single container and can be fully self-hosted. This design enables organizations to manage cost-sensitive or high-volume operations while maintaining strict data sovereignty by keeping document processing within on-premises infrastructure. These capabilities allow the model to serve not only a...
Related
OpenAI launches a Computer History feature that tracks user activity across apps and sites
OpenAI has introduced Computer History in the ChatGPT desktop app for macOS, a new feature that enables the app to track user activity and interactions across applications and webs...
Google releases Gemini 3.7 Flash with enhanced agentic coding and lower API pricing
Google has introduced Gemini 3.7 Flash, a new general purpose AI model that improves coding, automation, document processing, and agentic performance while cutting API pricing comp...
Google launches Sheets canvas for interactive dashboards with AI prompts
Google has launched Sheets canvas in Google Sheets, allowing users to reformat spreadsheet data into interactive mini-apps using natural language prompts. No coding is needed, and ...