Choose a model for the work you need to doMe
← Use cases
Vision & OCR: how to choose a model
Vision models usually return descriptions or structured text. OCR, chart interpretation, and Excel export need separate workflow steps.
Before integrating
- 01Check image format, resolution, and count limits.
- 02Specify whether the output should be prose, a table, or JSON.
- 03Add human or programmatic validation of extracted text.
- 04For Excel, add table validation and file generation.
Understanding images does not imply image editing or direct Excel file output.
Example request: “Extract the table from an image and export it to Excel”Find models for this scenario →
