Cohere introduced Parse, a vision-language model for enterprise documents. It processes tables, forms, diagrams, and images into machine-readable output, including Markdown and spatial information, and is available within the Compass retrieval stack.
Context
Document understanding requires more than recognizing words. A value must remain attached to the right heading, unit, or table row. Preserving those relationships helps later search and analysis, while references to the original page make errors easier to inspect. The downstream answer is only as reliable as the structure extracted at the start.
Sources & authors
- Cohere Parse | Enterprise Intelligence at ScaleCohere · August 27, 2026



