Overview
Sarvam Vision is a 3B-parameter, state-space Vision Language Model (VLM) purpose-built for high-accuracy Document Intelligence. Part of Sarvam's sovereign model series, it treats document understanding as knowledge extraction rather than simple text capture - extracting text, converting complex tables, and preserving layout, reading order, and hierarchy from PDFs and scanned images.
Sarvam Vision delivers world-class accuracy across 23 languages (22 official Indian languages + English), with native support for Indian scripts where most global models fall short. It achieves best-in-class scores on global benchmarks (olmOCR-Bench, OmniDocBench V1.5) for English, and leading accuracy on the Sarvam Indic OCR Bench for Indian languages - outperforming frontier models on Indic document tasks.
Ideal for digitizing scanned archives, Indic OCR at scale, and table-heavy documents such as scientific literature, financial reports, government bulletins, historical manuscripts, textbooks, and newspapers.
Highlights
- Indic-first document intelligence - native, high-accuracy OCR across 22 official Indian languages plus English (23 total).
- Knowledge extraction, not just text - parses complex tables, charts, and multi-column layouts while preserving reading order and document structure; outputs clean HTML or Markdown.
- Efficient & benchmark-leading - a compact 3B state-space VLM that leads on the Sarvam Indic OCR Bench and scores competitively on global benchmarks (olmOCR-Bench, OmniDocBench V1.5).
Introducing multi-product solutions
You can now purchase comprehensive solutions tailored to use cases and industries.
Features and programs
Financing for AWS Marketplace purchases
Pricing
Dimension | Description | Cost/host/hour |
|---|---|---|
ml.g6.xlarge Inference (Batch) Recommended | Model inference on the ml.g6.xlarge instance type, batch mode | $5.00 |
ml.g6e.xlarge Inference (Real-Time) Recommended | Model inference on the ml.g6e.xlarge instance type, real-time mode | $10.00 |
ml.g6.2xlarge Inference (Batch) | Model inference on the ml.g6.2xlarge instance type, batch mode | $5.00 |
ml.g6.4xlarge Inference (Batch) | Model inference on the ml.g6.4xlarge instance type, batch mode | $5.00 |
ml.g6.8xlarge Inference (Batch) | Model inference on the ml.g6.8xlarge instance type, batch mode | $5.00 |
ml.g6.12xlarge Inference (Batch) | Model inference on the ml.g6.12xlarge instance type, batch mode | $20.00 |
ml.g6.16xlarge Inference (Batch) | Model inference on the ml.g6.16xlarge instance type, batch mode | $5.00 |
ml.g6.24xlarge Inference (Batch) | Model inference on the ml.g6.24xlarge instance type, batch mode | $20.00 |
ml.g6.48xlarge Inference (Batch) | Model inference on the ml.g6.48xlarge instance type, batch mode | $40.00 |
ml.g6e.2xlarge Inference (Real-Time) | Model inference on the ml.g6e.2xlarge instance type, real-time mode | $10.00 |
Vendor refund policy
Contact support@sarvam.ai for details
How can we make this page better?
Legal
Vendor terms and conditions
Content disclaimer
Delivery details
Amazon SageMaker model
An Amazon SageMaker model package is a pre-trained machine learning model ready to use without additional training. Use the model package to create a model on Amazon SageMaker for real-time inference or batch processing. Amazon SageMaker is a fully managed platform for building, training, and deploying machine learning models at scale.
Version release notes
Enable Sarvam Vision models to be hosted on all variation of ml.g6 and ml.g6e series
Additional details
Inputs
- Summary
Documents to be digitized — PDFs or scanned page images. Provide a single PDF, individual PNG/JPG page images, or a flat ZIP archive of page images (up to 10). Optionally specify the target language code and output format.
- Limitations for input type
- Upload up to 10 pages for sync endpoint. For higher number of pages use async endpoint
- Input MIME type
- application/pdf
Resources
Vendor resources
Support
Vendor support
AWS infrastructure support
AWS Support is a one-on-one, fast-response support channel that is staffed 24x7x365 with experienced and technical support engineers. The service helps customers of all sizes and technical abilities to successfully utilize the products and features provided by Amazon Web Services.