Skip to main content
Extract structured text from images, scanned documents, receipts, and invoices using vision-language models purpose-built for OCR. No preprocessing, no bounding boxes — send an image and get text back.

Available OCR models


Extract text from an image


Receipt parsing with structured output

Extract specific fields from a receipt photo:

Next steps

  • AI Vision API — analyze images beyond OCR: scene understanding, chart reading, visual Q&A.
  • Extract structured data — combine OCR output with schema-based extraction for production pipelines.
  • Model catalog — browse all available vision and OCR models.