# OCR receipts and invoices from a CSV or Google Sheet

**Use case:** 

Reads every receipt or invoice image linked in a CSV, Excel file or Google Sheet, keeps your columns (supplier, note, id) and adds the recognised text, confidence and word count. Swap the file URL for your own sheet shared as 'Anyone with the link can view'.

## Input

```json
{
  "fileUrl": "https://nerolabs-samples.nerolabs.workers.dev/sample-ocr-list.csv",
  "fileFormat": "auto",
  "fileUrls": [
    "https://nerolabs-samples.nerolabs.workers.dev/sample-receipt.png",
    "https://nerolabs-samples.nerolabs.workers.dev/sample-scanned-letter.pdf",
    "https://nerolabs-samples.nerolabs.workers.dev/sample-screenshot.png",
    "https://www.irs.gov/pub/irs-pdf/fw9.pdf"
  ],
  "languages": [
    "eng"
  ],
  "usePdfTextLayer": true,
  "dpi": 200,
  "pageSegmentation": "auto",
  "maxPagesPerFile": 20,
  "minWordConfidence": 0,
  "includePageText": false,
  "keep": "all",
  "keepOriginalFields": true,
  "concurrency": 2,
  "requestTimeoutSecs": 60,
  "maxFileMb": 25,
  "exportFormats": [
    "csv"
  ]
}
```

## Output

```json
{
  "sourceUrl": {
    "label": "File",
    "format": "link"
  },
  "ocrStatus": {
    "label": "Status",
    "format": "text"
  },
  "fileType": {
    "label": "Type",
    "format": "text"
  },
  "pageCount": {
    "label": "Pages",
    "format": "number"
  },
  "ocrPages": {
    "label": "OCR pages",
    "format": "number"
  },
  "textLayerPages": {
    "label": "Text-layer pages",
    "format": "number"
  },
  "wordCount": {
    "label": "Words",
    "format": "number"
  },
  "confidence": {
    "label": "Confidence",
    "format": "number"
  },
  "lowConfidence": {
    "label": "Low confidence",
    "format": "boolean"
  },
  "text": {
    "label": "Text",
    "format": "text"
  },
  "statusDetail": {
    "label": "Detail",
    "format": "text"
  }
}
```

## About this Actor

This example demonstrates how to use [Bulk OCR: Image & Scanned PDF to Text from CSV or Google Sheet](https://apify.com/nerolabs/bulk-ocr-image-pdf-to-text.md) with a specific input configuration. Visit the [Actor detail page](https://apify.com/nerolabs/bulk-ocr-image-pdf-to-text.md) to learn more, explore other use cases, and run it yourself.


## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
This Task's input is already configured above — use it as-is rather than inventing a new one.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For full API examples (JavaScript, Python, CLI, MCP, OpenAPI), see this Task's Actor page: https://apify.com/nerolabs/bulk-ocr-image-pdf-to-text.md

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).
