# Changelog of PDF Text Extractor (`parsebird/pdf-text-extractor`) Actor

- **URL**: https://apify.com/parsebird/pdf-text-extractor/changelog.md
- **Full Actor documentation**: https://apify.com/parsebird/pdf-text-extractor.md

## Changelog

### v1.1 — 2026-08-17

- Fixed a performance issue where Markdown conversion could take over a minute per PDF; Markdown conversion is now roughly 7-8x faster with no change to output quality
- Added size and page-count limits so a single very large or complex PDF can no longer stall a run or use excessive memory — oversized PDFs are skipped with a clear error, and Markdown is skipped (with text and metadata still returned) for very large or very long PDFs
- Added a processing timeout so a PDF that's unusually slow to parse can no longer run indefinitely
- Lowered default and maximum concurrency for more predictable memory usage on batches of large PDFs

### v1.0 — 2026-08-15

- Initial release
- Extract embedded text and metadata from PDF files by URL
- Per-page text breakdown (optional)
- Optional Markdown conversion, preserving headings and lists
- Concurrent downloads with configurable concurrency and per-file timeout
