# Extract page-by-page text from a PDF report

**Use case:** 

Extracts the full text of a single PDF document plus a per-page array, so you can pull the text of any specific page or feed the document section by section into another tool. Metadata (title, author, page count) is included. Replace the URL with your own contract, research paper, or report.

## Input

```json
{
  "urls": [
    "https://css4.pub/2015/textbook/somatosensory.pdf"
  ],
  "includePages": true,
  "convertToMarkdown": false,
  "maxConcurrency": 3,
  "timeoutPerPdfSecs": 60
}
```

## Output

```json
{
  "fileName": {
    "label": "File Name",
    "format": "text"
  },
  "title": {
    "label": "Title",
    "format": "text"
  },
  "author": {
    "label": "Author",
    "format": "text"
  },
  "pageCount": {
    "label": "Pages",
    "format": "number"
  },
  "pdfVersion": {
    "label": "PDF Version",
    "format": "text"
  },
  "fileSizeBytes": {
    "label": "File Size (bytes)",
    "format": "number"
  },
  "url": {
    "label": "URL",
    "format": "link"
  },
  "error": {
    "label": "Error",
    "format": "text"
  }
}
```

## About this Actor

This example demonstrates how to use [PDF Text Extractor](https://apify.com/parsebird/pdf-text-extractor.md) with a specific input configuration. Visit the [Actor detail page](https://apify.com/parsebird/pdf-text-extractor.md) to learn more, explore other use cases, and run it yourself.


## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
This Task's input is already configured above — use it as-is rather than inventing a new one.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For full API examples (JavaScript, Python, CLI, MCP, OpenAPI), see this Task's Actor page: https://apify.com/parsebird/pdf-text-extractor.md

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).
