Cohere’s Parse 5 Promises Efficient Multi-Modal Information Extraction from Complex Documents

InfoQ (AI, ML & Data)
Read full post
Cohere launched Parse 5, a 2.3B-parameter multimodal model that extracts structured data from complex documents like PDFs into Markdown with visual grounding. It uses an 8K-token context window and a custom vision encoder for enterprise-scale workloads, scoring 79.2 on the ParseBench benchmark.

More on this story


More in Business & Enterprise

Apple’s first foldable is the iPhone Duo, and it costs $1,999

Covered by 6 sources

Mistral Seeks to Grow Enterprise Customer Base Through Cloudera Partnership

Covered by 3 sources

Sources: DOJ is investigating whether Nvidia tried to skirt antitrust scrutiny of its 2025 Groq deal, described by Groq as a "nonexclusive licensing agreement" (New York Times)

Covered by 3 sources