> For clean Markdown of any page, append `.md` to the page URL.
> For a complete documentation index, see https://docs.sarvam.ai/llms.txt.
> For full documentation content in one file, see https://docs.sarvam.ai/llms-full.txt.
> For AI client integration (Claude Code, Cursor, etc.), connect to the MCP server at https://docs.sarvam.ai/_mcp/server.

# Welcome to Doc Agents

> Document-intelligence workbench for extracting, digitising, and translating documents across 22 Indian languages plus English — no code required.

## One workbench. Every document.

Extract, digitise, and translate any document in Hindi, Bengali, Tamil, Telugu, English, and 18 more Indian languages. No code required.

![Doc Agents Home dashboard with quick-start cards for Extract, Digitise, and Translate, a credit balance, and a template shelf.](https://fdr-prod-docs-files-public.s3.us-east-1.amazonaws.com/sarvam-api-docs.docs.buildwithfern.com/69ad5328ac658caa0060df1f12092180c9f8d7238656cd5df1f4d57f131e3666/sarvam-pages/images/home.png?X-Amz-Algorithm=AWS4-HMAC-SHA256&X-Amz-Content-Sha256=UNSIGNED-PAYLOAD&X-Amz-Credential=AKIA6KXJSKKNFOCF7G4B%2F20260907%2Fus-east-1%2Fs3%2Faws4_request&X-Amz-Date=20260907T124333Z&X-Amz-Expires=604800&X-Amz-Signature=edb49b62bdac84a469feba46ad4f1e06ea220a016541206b85c7b51467ee363f&X-Amz-SignedHeaders=host&x-amz-checksum-mode=ENABLED&x-id=GetObject)

**Doc Agents** is a document-intelligence workbench for turning real-world documents into data you can use. From BFSI paperwork (invoices, KYC, bank and credit-card statements, insurance claims) to government records (land deeds, licensing forms), legal filings, HR resumes, healthcare intake, printed books, and handwritten multilingual manuscripts, the workbench handles it all. Every workflow is one dashboard click at [dashboard.sarvam.ai](https://dashboard.sarvam.ai/akshar/dashboard). No code required.

Under the hood, **Sarvam Vision** powers extraction and digitisation, and **Sarvam Document Translation** powers whole-document translation — all with SOTA accuracy on **22 Indian languages plus English**.

### What you can do

#### Extract fields as structured data

Pull invoice numbers, PAN, dates, amounts, and line-items from any document type as JSON, CSV, or Excel.

#### Digitise whole documents

Convert scans, handwritten forms, land deeds, and historic multilingual manuscripts into structured text, Markdown, HTML, or DOCX.

#### Translate whole documents

Translate PDFs, Word files, and more into Indian languages while preserving layout, then refine in a side-by-side editor.

#### Export as CSV or Excel

Turn table-heavy PDFs such as bank statements and invoices into spreadsheet-ready CSV or Excel output.

#### Build reusable Configs

Save every workflow as a Config. Share with your team. Re-run against thousands of documents with one click.

## Who it's for

Built for anyone whose day-to-day involves documents:

* **BFSI teams** processing invoices, KYC, bank and credit-card statements, GST filings, and insurance claims at scale.
* **Analysts** turning scanned reports and statements into Excel-ready data for pandas, SQL, or dashboards.
* **Developers** wiring document extraction into product flows, back-office pipelines, and RAG systems.
* **Legal, compliance, and government teams** working with land deeds, court records, licensing forms, and notarised documents.
* **HR and healthcare teams** parsing resumes, offer letters, patient intake forms, and prescriptions.
* **Researchers and publishers** digitising historic manuscripts, out-of-print books, and multilingual archives.

## Prerequisites

* A document to test on: **PDF, JPEG, or PNG**, up to **50 MB per file** and **10 pages per project**
* Sufficient credits in the workspace

## Home Page

Open [dashboard.sarvam.ai/akshar/dashboard](https://dashboard.sarvam.ai/akshar/dashboard) and sign in. You land on **Home**, which shows your credit balance, quick-start cards for Extract, Digitise, and Translate, and a shelf of ready-made template configs organised by category. The sidebar has three working groups: **Products** (Extract, Digitise, Translate) and **Workspace** (Projects, Configs).

![Doc Agents Home page with quick-start cards for Extract and Digitise, a credit balance, and a template shelf organised by category.](https://fdr-prod-docs-files-public.s3.us-east-1.amazonaws.com/sarvam-api-docs.docs.buildwithfern.com/d4c6a12b0f4108238f81046e661ca11952f5d810de8b2583adcb6875ed077c56/sarvam-pages/images/home-final.png?X-Amz-Algorithm=AWS4-HMAC-SHA256&X-Amz-Content-Sha256=UNSIGNED-PAYLOAD&X-Amz-Credential=AKIA6KXJSKKNFOCF7G4B%2F20260907%2Fus-east-1%2Fs3%2Faws4_request&X-Amz-Date=20260907T124333Z&X-Amz-Expires=604800&X-Amz-Signature=1f15eeb564ac06d7f7799940c40afb5b9e062b765508182009d72b5e950e7099&X-Amz-SignedHeaders=host&x-amz-checksum-mode=ENABLED&x-id=GetObject)

## Products

**Products** in the sidebar contains the three things the workbench does:

* **Extract** pulls specific fields out of a document (invoice numbers, dates, amounts, names) and returns them as structured JSON, CSV, or Excel. You describe what you want in plain English and the workbench drafts a schema for you.
* **Digitise** converts the whole content of a document into structured text (headlines, paragraphs, tables, images) and exports as HTML, Markdown, DOCX, or Plain Text. Great for scans, handwritten forms, land deeds, and historic multilingual manuscripts.
* **Translate** renders a whole document into one or more languages while preserving layout, then lets you refine each translation in a side-by-side editor with AI assistance and learned rules.

### Extract

**Extract** pulls specific fields out of a document and returns them as structured data: invoice numbers, dates, amounts, PAN, GST IDs, account holders, transaction rows, or whatever fields the document contains.

You describe what you want in plain English (*"rider name, trip date, total amount"*), and the workbench auto-drafts a schema from your prompt. Review and edit the schema (change field types, add nested groups, remove fields you don't need), then approve and run.

The results view puts the source document on the left and a hierarchical **Output** panel on the right. Every field carries a value, a per-field **confidence score**, and a dash placeholder if the field wasn't found. Click any value to correct it inline before export.

![Extract results view with an Uber receipt on the left and a hierarchical Output panel with per-field confidence scores on the right.](https://fdr-prod-docs-files-public.s3.us-east-1.amazonaws.com/sarvam-api-docs.docs.buildwithfern.com/ef5c1e0afc78b498ced40e500614e6cda8ff0af8925efc36247ec6f39aadb3f8/sarvam-pages/images/extract-results.png?X-Amz-Algorithm=AWS4-HMAC-SHA256&X-Amz-Content-Sha256=UNSIGNED-PAYLOAD&X-Amz-Credential=AKIA6KXJSKKNFOCF7G4B%2F20260907%2Fus-east-1%2Fs3%2Faws4_request&X-Amz-Date=20260907T124333Z&X-Amz-Expires=604800&X-Amz-Signature=33ddd5293962a22875de27cd04ef29e97ae833ebb007a97c2389c7d70198092f&X-Amz-SignedHeaders=host&x-amz-checksum-mode=ENABLED&x-id=GetObject)

**Great for:** invoices, receipts, KYC forms, bank and credit-card statements, GST filings, insurance claims, purchase orders, resumes, background verifications, loan applications, prescriptions, ID cards, and any document where you already know the fields you need.

**Download as:** JSON, CSV, or Excel (XLSX).

For the full walkthrough see **[How to Extract structured fields](/docai/how-to/extract-fields-from-a-document/extract-structured-fields)**.

### Digitise

**Digitise** converts the whole content of a document into structured text. Headlines, section titles, paragraphs, tables, and images are all identified and preserved.

The editor puts the source page on the left with **color-coded bounding boxes** on every detected region, and a numbered **Sections** list on the right. Each section has a type dropdown, and table sections open in a dedicated **Table Editor** for cell-by-cell fixes. If a region was missed, click **+ Add boxes** to draw a rectangle and label it from the same list of section types.

The output is clean, machine-readable text with layout preserved, ready to feed into LLMs, RAG pipelines, publishing tools, or archives.

![Digitise editor with a Tamil land deed on the left and the digitised Tamil text on the right, preserving the document's structure.](https://fdr-prod-docs-files-public.s3.us-east-1.amazonaws.com/sarvam-api-docs.docs.buildwithfern.com/e668f46d144c3f10dbce936de074d46b1aabdf983f707b294c67b2bc87e4c224/sarvam-pages/images/digitise-tamil.png?X-Amz-Algorithm=AWS4-HMAC-SHA256&X-Amz-Content-Sha256=UNSIGNED-PAYLOAD&X-Amz-Credential=AKIA6KXJSKKNFOCF7G4B%2F20260907%2Fus-east-1%2Fs3%2Faws4_request&X-Amz-Date=20260907T124333Z&X-Amz-Expires=604800&X-Amz-Signature=647c64a006072f456bda69b79b31b8ef718cdf17fc15e114d2dfb0f43709d2ce&X-Amz-SignedHeaders=host&x-amz-checksum-mode=ENABLED&x-id=GetObject)

Every page is broken into typed **Sections** (headers, paragraphs, tables, images, and more), each with color-coded bounding boxes you can review, reclassify, or edit.

![Digitise editor showing a document with color-coded bounding boxes on the left and a numbered Sections list on the right with types like Header and Image, including AI-generated image descriptions.](https://fdr-prod-docs-files-public.s3.us-east-1.amazonaws.com/sarvam-api-docs.docs.buildwithfern.com/7df1a9d3cd2cba1b6809e5a7f00bfe3aaf47142310a3ab86368211e204d4157a/sarvam-pages/images/digitise-sections.png?X-Amz-Algorithm=AWS4-HMAC-SHA256&X-Amz-Content-Sha256=UNSIGNED-PAYLOAD&X-Amz-Credential=AKIA6KXJSKKNFOCF7G4B%2F20260907%2Fus-east-1%2Fs3%2Faws4_request&X-Amz-Date=20260907T124333Z&X-Amz-Expires=604800&X-Amz-Signature=478086f8b6ae654cced25d46e0e972fa934743be5bfd2a1f5ca414fe5f66a277&X-Amz-SignedHeaders=host&x-amz-checksum-mode=ENABLED&x-id=GetObject)

**Great for:** handwritten forms and notes, land deeds, court records, historic manuscripts, regional-language archives, printed books and journals, contracts, research papers, gazettes, and any document you want as clean text for LLMs, RAG, or publishing.

**Download as:** HTML, Markdown, DOCX, or Plain Text.

For the full walkthrough see **[How to Digitise a document](/docai/how-to/digitise-a-document)**.

### Translate

**Translate** renders a whole document into one or more languages while preserving its structure — headings, paragraphs, images, and layout all stay in place. You then review and refine each translation in a rich side-by-side editor with AI assistance, learned rules, and document-level guidelines.

**Great for:** reports, case studies, books, contracts, marketing collateral, and any document you need in multiple Indian languages without rebuilding the layout.

**Download as:** DOCX, PDF, HTML, EPUB, or TXT.

For the full walkthrough see **[How to Translate a document](/docai/how-to/translate-a-document)**.

### Extract or Digitise: how to choose

#### Use Extract when…

You know which fields you need. You want structured JSON, CSV, or Excel out. You're feeding a database, spreadsheet, or business workflow.

#### Use Digitise when…

You want the whole document. You don't know in advance what you'll query. You're feeding an LLM, publishing, or archiving.

## Workspace

### Configs

A **Config** is a saved recipe: a name, a description of what you want, a language, and an output format. Build once, run against every document of that type. The `Usage` column tells you how many projects have used each Config.

On **Home**, scroll past the Extract and Digitise cards to **Or start with a predefined config**. The template shelf lists **12 ready-made recipes** (9 Extract + 3 Digitise) with filter pills for **All / Extract / Digitise**. Click any card to preview the sample document and the fields it extracts, then start a project in one click.

![Home page showing Or start with a predefined config with All 12, Extract 9, and Digitise 3 filter pills, and template cards for Insurance Claim Form Extraction, Credit Card Statement Extraction, and KYC Document Extraction.](https://fdr-prod-docs-files-public.s3.us-east-1.amazonaws.com/sarvam-api-docs.docs.buildwithfern.com/e0aa72ecb857adbbd6006cd85ec4c1b7ba26e5da4135dbe269e7a1e769e5c943/sarvam-pages/images/way2-predefined-config.png?X-Amz-Algorithm=AWS4-HMAC-SHA256&X-Amz-Content-Sha256=UNSIGNED-PAYLOAD&X-Amz-Credential=AKIA6KXJSKKNFOCF7G4B%2F20260907%2Fus-east-1%2Fs3%2Faws4_request&X-Amz-Date=20260907T124333Z&X-Amz-Expires=604800&X-Amz-Signature=e3d48789f92f91da939d8e28707c554907e8aa0ac0bb9ac3e4bdca3b3a594443&X-Amz-SignedHeaders=host&x-amz-checksum-mode=ENABLED&x-id=GetObject)

Saved Configs you've built also surface on **Home** under **Start with your custom configs**, so your team can kick off a project in one click.

![Home dashboard with Configs highlighted in the sidebar and a Start with your custom configs row showing saved Configs (Receipt Extraction, Invoice extraction, Story extraction).](https://fdr-prod-docs-files-public.s3.us-east-1.amazonaws.com/sarvam-api-docs.docs.buildwithfern.com/a9e8f4451cb2a4b41db49294598b85d50cc94e24ab1f62cc01b4455bf0c4842e/sarvam-pages/images/configs-way1.png?X-Amz-Algorithm=AWS4-HMAC-SHA256&X-Amz-Content-Sha256=UNSIGNED-PAYLOAD&X-Amz-Credential=AKIA6KXJSKKNFOCF7G4B%2F20260907%2Fus-east-1%2Fs3%2Faws4_request&X-Amz-Date=20260907T124333Z&X-Amz-Expires=604800&X-Amz-Signature=f12bafe61c16b5b7f42ddb927d7970964f5b98bfd839195cc8757ad21dfc214e&X-Amz-SignedHeaders=host&x-amz-checksum-mode=ENABLED&x-id=GetObject)

You can also browse the same templates under **Workspace → Configs → Templates**, duplicate any into your own Config, tweak the prompt or language, and share with your team. Or start from scratch with **+ New config**. See **[How to Create a Config](/docai/how-to/extract-fields-from-a-document/create-a-reusable-config)** for the full walkthrough.

![Configs page with Your configs and Templates tabs, All / Extract / Digitise filter pills, and a Usage column.](https://fdr-prod-docs-files-public.s3.us-east-1.amazonaws.com/sarvam-api-docs.docs.buildwithfern.com/3309ddeaf1ba15e24c196353bbb5200c6b49bed01871aeac05b44c0c592501ee/sarvam-pages/images/configs-page.png?X-Amz-Algorithm=AWS4-HMAC-SHA256&X-Amz-Content-Sha256=UNSIGNED-PAYLOAD&X-Amz-Credential=AKIA6KXJSKKNFOCF7G4B%2F20260907%2Fus-east-1%2Fs3%2Faws4_request&X-Amz-Date=20260907T124333Z&X-Amz-Expires=604800&X-Amz-Signature=f8c53d84290c82499c6d30a3c525b8f19e7ac3c14628b70c147454c6f7ca94e5&X-Amz-SignedHeaders=host&x-amz-checksum-mode=ENABLED&x-id=GetObject)

### Projects

A **Project** is a single run of a Config against one document. It has its own status, results view, and downloads. **Workspace → Projects** lists every project across your organisation, sortable by any column and filterable by **Status** or **Type**.

![Projects page with search bar, All statuses and All types filter pills, and a New project button.](https://fdr-prod-docs-files-public.s3.us-east-1.amazonaws.com/sarvam-api-docs.docs.buildwithfern.com/2a3333dc4909d2a41f51be3072970822ac9193b8d646b7dc850ed6349039e4ed/sarvam-pages/images/projects.png?X-Amz-Algorithm=AWS4-HMAC-SHA256&X-Amz-Content-Sha256=UNSIGNED-PAYLOAD&X-Amz-Credential=AKIA6KXJSKKNFOCF7G4B%2F20260907%2Fus-east-1%2Fs3%2Faws4_request&X-Amz-Date=20260907T124333Z&X-Amz-Expires=604800&X-Amz-Signature=94dbfda9165e6246281e09fa0e2e7db2860d75369a32024f553c12b11c1c1ce2&X-Amz-SignedHeaders=host&x-amz-checksum-mode=ENABLED&x-id=GetObject)![Projects pane listing individual project runs with their Config, type, status, and last-updated time.](https://fdr-prod-docs-files-public.s3.us-east-1.amazonaws.com/sarvam-api-docs.docs.buildwithfern.com/8a7bb1b11164eaa15b7cf2b25c2d172fd8f3d225159b5389d1526aa4cfa56bad/sarvam-pages/images/projects-pane.png?X-Amz-Algorithm=AWS4-HMAC-SHA256&X-Amz-Content-Sha256=UNSIGNED-PAYLOAD&X-Amz-Credential=AKIA6KXJSKKNFOCF7G4B%2F20260907%2Fus-east-1%2Fs3%2Faws4_request&X-Amz-Date=20260907T124333Z&X-Amz-Expires=604800&X-Amz-Signature=e38e927357e2c3a760ed7a3d3c019ba92ee3de5a52ab95ab60d482523d58b75c&X-Amz-SignedHeaders=host&x-amz-checksum-mode=ENABLED&x-id=GetObject)

## How the flow works

Every project (Extract or Digitise) follows the same three-step flow.

#### Upload documents

Drag and drop scans, photos, or PDFs. Supported: **PDF, JPEG, PNG**, up to **50 MB per file** and **10 pages per project**.

#### Sarvam Vision processes every page

The model reads each page. For **Extract**, it locates the fields you asked for. For **Digitise**, it converts every page into accurate structured text.

#### Edit and export

Review the output side-by-side with the source. Fix any values by hand. Download in the format that fits your workflow.

## Next steps

#### [Extract structured fields](/docai/how-to/extract-fields-from-a-document/extract-structured-fields)

End-to-end walkthrough of an Extract project, including the AI-drafted schema step.

#### [Digitise a document](/docai/how-to/digitise-a-document)

End-to-end walkthrough of a Digitise project, including bounding boxes, section types, and Add boxes.

#### [Translate a document](/docai/how-to/translate-a-document)

End-to-end walkthrough of a Translate project, from upload through side-by-side editing to export.

#### [Create a Config](/docai/how-to/extract-fields-from-a-document/create-a-reusable-config)

Start from one of the 12 templates or build a custom Config with Name, Pipeline type, Prompt, and Output schema.

#### [Extract as CSV/Excel](/docai/extract-as-csv-excel)

Turn table-heavy PDFs (bank statements, invoices) into spreadsheet-ready output.