# Scrape Embedding Models on Hugging Face for RAG

**Use case:** 

Collect Hugging Face embedding models for RAG and vector search builds: id, author, libraryName, downloads, likes, lastModified. Keyless official Hub API.

## Input

```json
{
  "searchQueries": [
    "llm",
    "text-to-image"
  ],
  "contentTypes": [
    "models"
  ],
  "repoUrls": [
    "meta-llama/Llama-3.1-8B-Instruct"
  ],
  "authors": [
    "meta-llama"
  ],
  "browseHub": true,
  "sortBy": "downloads",
  "pipelineTag": "feature-extraction",
  "libraryName": "sentence-transformers",
  "minDownloads": 0,
  "minLikes": 0,
  "excludeGated": false,
  "includeFullMetadata": true,
  "includeAuthorProfiles": false,
  "enrichContactEmails": false,
  "maxResults": 100,
  "maxResultsPerQuery": 60,
  "monitorMode": false,
  "monitorStoreName": "hugging-face-monitor",
  "maxConcurrency": 5,
  "proxyConfiguration": {
    "useApifyProxy": true
  },
  "urlsFromFile": ""
}
```

## Output

```json
{
  "type": {
    "label": "Type"
  },
  "id": {
    "label": "Repo"
  },
  "author": {
    "label": "Author"
  },
  "pipelineTag": {
    "label": "Task"
  },
  "libraryName": {
    "label": "Library"
  },
  "downloads": {
    "label": "Downloads"
  },
  "likes": {
    "label": "Likes"
  },
  "trendingScore": {
    "label": "Trending"
  },
  "license": {
    "label": "License"
  },
  "parameters": {
    "label": "Params"
  },
  "lastModified": {
    "label": "Updated"
  },
  "popularityScore": {
    "label": "Popularity"
  },
  "url": {
    "label": "URL"
  }
}
```

## About this Actor

This example demonstrates how to use [Hugging Face Scraper - Models, Datasets, Spaces & Leads](https://apify.com/scrapesage/hugging-face-scraper) with a specific input configuration. Visit the [Actor detail page](https://apify.com/scrapesage/hugging-face-scraper) to learn more, explore other use cases, and run it yourself.


## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
This Task's input is already configured above — use it as-is rather than inventing a new one.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For full API examples (JavaScript, Python, CLI, MCP, OpenAPI), see this Task's Actor page: https://apify.com/scrapesage/hugging-face-scraper.md

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).
