Cohere Launches Parse Document Parsing Model and Transcribe Speech Recognition Model
Cohere expanded its enterprise AI model lineup with the introduction of Cohere Parse for document layout parsing and Cohere Transcribe for speech recognition. These additions enable native handling of unstructured multimodal document layouts and audio files directly within Cohere's platform.
Verified State Diff
Impact & Verification Analysis
Enterprise developers, data engineers, and solution architects building document processing, search, and RAG pipelines.
Enables end-to-end data ingestion directly within Cohere's private deployment infrastructure (Model Vault), eliminating third-party dependencies for document parsing and speech transcription.
Full Fact Overview
Cohere has introduced two dedicated models to its product suite: Cohere Parse and Cohere Transcribe. Cohere Parse is designed to parse unstructured, complex enterprise documents such as PDFs and tables into structured data for search and RAG workflows. Cohere Transcribe adds automated speech-to-text capabilities, allowing enterprises to convert audio streams into text that can be processed by Cohere Command, Embed, and Rerank models inside isolated cloud environments like Model Vault.