Granite 4.0 3B Vision: Compact Multimodal Intelligence for Enterprise Documents
Today we’re excited to announce Granite 4.0 3B Vision, a compact vision-language model (VLM) designed for enterprise document understanding. It’s purpose-built for reliable information extraction from complex documents, forms, and structured visuals. Granite 4.0 3B Vision excels on the following capabilities: Table Extraction: Accurately parsing complex table structures (e.g., multi-row, multi-column, etc.) from document images Chart Understanding: Converting charts and figures into structured machine-readable formats, summaries, or executable code Semantic Key-Value Pair (KVP) Extraction: Identifying and grounding semantically meaningful key-value […]
Read more