# Keyword & Keyphrase Extractor from Text (`hipersoft/keyphrase-extractor`) Actor

Pull the most important keywords and multi-word key phrases out of any text — articles, reviews, transcripts or documents. Ranked by relevance with scores, frequency and word count. Great for SEO, tagging, content analysis and RAG. Process one text or many at once. Output JSON, CSV or Excel.

- **URL**: https://apify.com/hipersoft/keyphrase-extractor.md
- **Developed by:** [hiper soft](https://apify.com/hipersoft) (community)
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $0.00015 / keyphrase extracted

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

## Keyword & Keyphrase Extractor from Text

Automatically pull the **most important keywords and multi-word key phrases** out of any text with the **Keyword & Keyphrase Extractor**. Feed it an article, product review, transcript or document and get a ranked list of the phrases that actually matter — with **relevance scores, frequency and word count** — exported as **JSON, CSV or Excel**.

Great for **SEO, content tagging, topic analysis, trend spotting, metadata generation and RAG / LLM pipelines**, this tool surfaces the key ideas in text without any setup, training or code.

### What it does

- 🏷️ **Extract key phrases** — finds meaningful multi-word phrases (not just single words) that summarise the text.
- 📊 **Ranked by relevance** — each phrase comes with a score so you can take the top ones.
- 🔢 **Frequency & length** — see how often each phrase appears and how many words it spans.
- 📚 **Batch many texts** — analyse a whole list of documents in one run, each labelled by document.
- ⚙️ **Tunable** — cap how many phrases to return and set a minimum word length to cut noise.
- ⚡ **Fast & private** — processes text directly with no external calls; nothing to configure.
- 📤 **Export anywhere** — JSON, CSV or Excel, or straight into a spreadsheet, CMS or workflow.

### Example output

```json
{
  "documentIndex": 0,
  "rank": 1,
  "phrase": "minimal generating sets",
  "score": 8.67,
  "frequency": 1,
  "wordCount": 3
}
```

### How to use it

1. Paste your **text** (or add several texts in the list).
2. Choose how many **keyphrases per text** you want and the **minimum word length**.
3. Run, and download the ranked keyphrases as **JSON, CSV or Excel**.

### Input fields

| Field | Description |
|-------|-------------|
| **Text** | The text to analyse. |
| **Multiple texts** | Optional list of separate texts to analyse in one run. |
| **Max keyphrases per text** | How many top-ranked phrases to return per text. |
| **Minimum word length** | Ignore words shorter than this many characters. |
| **Max results** | Maximum total keyphrases across all texts (0 = no limit). |

### Output fields

`documentIndex`, `rank`, `phrase`, `score`, `frequency`, `wordCount`, `collectedAt`.

### Popular use cases

- **SEO & content** — discover the key phrases in an article to guide titles, tags and meta descriptions.
- **Tagging & taxonomy** — auto-generate tags and categories for a content library.
- **Review & feedback analysis** — surface recurring themes across customer reviews or survey answers.
- **Research & summarisation** — pull the core concepts from long documents fast.
- **RAG & LLM pipelines** — attach keyphrase metadata to chunks for better retrieval and filtering.
- **Trend spotting** — extract phrases across many documents to see what's rising.

### FAQ

**Do I need an account or key?**
No. Paste your text and run.

**How does the ranking work?**
Phrases are scored with a well-established keyword-extraction method that rewards distinctive multi-word phrases over common filler words. Higher score = more central to the text.

**What languages are supported?**
It works best on English text. Other languages will still return phrases, but scoring is tuned for English stopwords.

**Can it handle many documents at once?**
Yes — add a list of texts and each is analysed separately, with a document index on every result.

**Can I export to Excel or Google Sheets?**
Yes — results download as JSON, CSV or Excel and integrate with Sheets, CMSs and automation tools.

***

Surface the **keywords and key phrases** that define any text — ranked, scored and export-ready for SEO, tagging and AI pipelines.

# Actor input Schema

## `text` (type: `string`):

The text to analyse. Paste an article, review, transcript or document. For multiple texts, use the list below.

## `texts` (type: `array`):

Optional — a list of separate texts to analyse in one run. Each is processed independently and labelled by document.

## `maxKeyphrases` (type: `integer`):

How many top-ranked keyphrases to return for each text.

## `minChars` (type: `integer`):

Ignore words shorter than this many characters (helps drop noise).

## `maxItems` (type: `integer`):

Maximum total keyphrases to output across all texts (0 = no limit).

## Actor input object example

```json
{
  "texts": [],
  "maxKeyphrases": 20,
  "minChars": 2,
  "maxItems": 0
}
```

# Actor output Schema

## `results` (type: `string`):

The results as dataset items.

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "text": "",
    "texts": []
};

// Run the Actor and wait for it to finish
const run = await client.actor("hipersoft/keyphrase-extractor").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "text": "",
    "texts": [],
}

# Run the Actor and wait for it to finish
run = client.actor("hipersoft/keyphrase-extractor").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "text": "",
  "texts": []
}' |
apify call hipersoft/keyphrase-extractor --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,hipersoft/keyphrase-extractor"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/8vDMBDFGkYarSCNyL/builds/24w1dbUW71Cfbtca6/openapi.json
