Yandex Keyword Suggest Scraper
Pricing
$0.25 / 1,000 keyword founds
Yandex Keyword Suggest Scraper
Turn one seed term into thousands of the real search queries people type on Yandex, ranked in Yandex's own suggestion order, for the city and market you sell into. Covers Russian, Kazakh and Turkish. Gives you the queries themselves — not an estimate of how often they are searched.
Give it a topic and it gives you back the actual search queries people type into Yandex around that topic — thousands of them, in the order Yandex itself suggests them, for the exact city and market you sell into. It is built for SEO and PPC teams working the Russian-speaking market, where the usual Western keyword tools have thin coverage and the official planner is behind a login. Important, and stated up front: this returns the queries people type, not a figure for how many times they type them. No such figure is published anywhere on this source, and this actor will never show you one.
What you can do with it
- Build a long-tail keyword universe from one seed. One topic, expanded with buying words, question words, letters and digits, becomes thousands of distinct real queries in a single run.
- Do local SEO properly. The same seed returns genuinely different lists in Moscow, Yekaterinburg, Kazan and Minsk — different local brands, different phrasing, different intent. Pick the city you actually sell in.
- Find the brands that own a category in each city. Where Yandex treats a query as a shortcut to a particular company, the company's name and web address come back with the keyword — a fast way to see who dominates a niche in a market you do not know yet.
- Separate buying intent from browsing. Turn on the shopping-intent set and the informational noise drops away, leaving the phrasing people use when they mean to purchase.
- Feed a content plan. Question-shaped queries ("how", "what", "where", "which") come back grouped and ready to become articles.
- Cover Kazakh and Turkish markets too, with the same seed workflow and each market's own alphabet.
What you get
One row per distinct keyword, per seed term. An abridged real record:
{"seed": "пицца","keyword": "пицца мия екатеринбург","queryPrefix": "пицца ","position": 3,"expansionLevel": 0,"source": "search","language": "ru","regionId": 54,"regionName": "Yekaterinburg","resolvedRegionCode": 11162,"wordCount": 3,"characterCount": 22,"containsSeed": true,"siteTitle": null,"siteDomain": null,"collectedAt": "2026-08-17T11:57:32.414Z"}
Input reference
| Setting | Type | Default | What it does |
|---|---|---|---|
| Seed terms | list of text | — | The topics you want ideas for, one per line. Up to 50 per run. This is the only setting you must fill in. |
| Market | choice | Russian | Which market's suggestions to collect: Russian, Kazakh or Turkish. Sets both the language you get back and the ordering they arrive in. |
| Region | choice | Moscow, Russia | The city or country whose searchers you want. 17 Russian cities plus Belarus, Kazakhstan, Uzbekistan and Turkey. Local brands and local phrasing differ sharply between cities, and this is what makes the output usable for local work. |
| Expansion depth | whole number, 0–2 | 1 | How far past the seed to go. 0 returns only what the source suggests for the seed as typed. 1 also pairs the seed with buying words, question words, letters and digits — the usual choice, and typically a few thousand ideas per seed. 2 re-expands everything found at level 1 and produces the deepest list. |
| Maximum keywords per seed | whole number, 1–25000 | 2000 | Stops each seed once this many distinct keywords have been collected. This is your main control on how large and how expensive a run is. |
| Also collect shopping-intent and word-completion suggestions | yes / no | no | Adds two further short sets to every seed: shopping-intent queries, which strip informational noise and leave purchase phrasing, and single-word completions, which surface word stems the main list misses. |
| Extra words to combine with each seed | list of text | empty | Your own words to pair with every seed, on top of the built-in buying and question words. Useful for brand names, model numbers or an industry's own vocabulary. Up to 40. |
| Speed | whole number, 1–30 | 10 | How many parts of the job to work on at once. Higher finishes sooner. Leave it alone unless you have a reason not to. |
Output fields
| Field | Description |
|---|---|
seed | The seed term this keyword was found for. |
keyword | The search query itself, exactly as the source publishes it. |
queryPrefix | The starting point that surfaced this keyword. Group by this to see which angle produced which ideas. |
position | Where the source placed this keyword in its own suggestion order for that starting point. 1 is first. This is a rank, not a count, and not a volume. Empty on the one row per completion starting point that the source publishes outside its ranked list, because that row has no rank of its own. |
expansionLevel | 0 for the seed itself, 1 for the first round of combinations, 2 for the deepest round. |
source | Which set the keyword came from: search, shopping or completion. |
language | The market that was collected: ru, kk or tr. |
regionId | The region code you chose. |
regionName | The plain-English name of that region. |
resolvedRegionCode | The region grouping the source actually applied to this answer. 0 means the source did not apply the region you chose for that particular answer — a condition you can see and filter on rather than one hidden from you. Empty where the source publishes no grouping. |
wordCount | How many words the keyword contains. Useful for isolating long-tail phrases. |
characterCount | How many characters the keyword contains. |
containsSeed | Whether the keyword still contains your seed term. Deep expansion drifts, and this lets you keep or drop the drift deliberately. |
siteTitle | Where the source treats the query as a shortcut to a specific company, that company's name. Empty for ordinary keywords — and empty means empty, never a guess. |
siteDomain | The web address of that company, in the same case. Empty for ordinary keywords. |
collectedAt | When the keyword was collected, as a timestamp. |
Pricing
You pay per keyword, and only for keywords that actually reach your results.
| What you pay for | Price |
|---|---|
| Each distinct keyword written to your results | $0.00025 — that is $0.25 per 1,000 keywords |
A keyword is written once per seed term, so you are never billed twice for the same idea inside one seed. There is no second charge of any kind: the expansion is the product, not an upgrade.
Worked example. One seed at the default settings, capped at 2,000 keywords, costs $0.50. A ten-seed research sprint at the same cap costs $5.00. If you push a single seed to the maximum of 25,000 keywords, that run costs $6.25. Set "Maximum keywords per seed" and you have set your bill.
A seed the source publishes nothing for produces no rows, and therefore costs you nothing.
Limits & what this actor cannot do
- It does not tell you how often a keyword is searched. There is no volume, popularity, traffic, competition or difficulty figure anywhere in what this source publishes, so there is none in the output. What you do get is the source's own ordering, which reflects what it suggests first — a useful signal, but a rank, never a number. If you need volume figures, this is not the tool.
- The source publishes a limited number of suggestions for any one starting point. Depth is what turns that into thousands of ideas; a single starting point on its own returns a few dozen at most.
- The company shortcut is rare.
siteTitleandsiteDomainare filled only on the few queries the source itself treats as a shortcut to a particular company, and that is a small minority of any list — measured at 3 rows in 312 across eight commercial topics, roughly 1%. It is a genuine signal about who owns a query, not a column you can expect on most rows, and the rest are empty rather than guessed. - The word-completion set is short by design — the source publishes only a handful of completions for any stem, and no setting increases it.
- The shopping-intent set is shorter still, and it ignores the region you choose. Those rows reflect the source's default market, not your city. The main set does honour your region.
- Some region and market pairings are not applied by the source. When that happens the affected rows carry a resolved region code of
0and the run says so plainly, rather than passing another market's results off as yours. - Results are a snapshot at the moment of collection. The source's suggestions change continuously, and two runs an hour apart will not match exactly.
- The actor reports what the source publishes to the public. It does not sign in, sees nothing behind a login, and cannot reveal anything the source keeps private.
- Deep expansion drifts away from the seed. The large majority of keywords still contain the seed exactly — above 96% on the deep runs we measured — and the rest are related queries the source associates with it.
containsSeedmarks each row so you can keep or drop the drift deliberately. - A seed the source refuses to answer is reported as a failed seed, with a count you can see in the run summary. It is never reported as "no keywords found". If the source stops answering partway through a seed, the run stops that seed, keeps what it already collected and says so, rather than presenting a truncated list as a complete one.
- Turkey can only be collected together with the Turkish market, because that is the only combination the source honours. The run refuses the other combinations rather than quietly returning a different market's results.
- How long a run takes depends on how large the job is and on the source's own response times. No fixed speed is promised.
- The source's terms govern automated access. You are responsible for using the data lawfully and in line with those terms, including any applicable privacy law.
FAQ
Does it tell me how many people search each keyword? No, and nothing that looks like it will. This source does not publish a search-volume figure of any kind, and inventing one would be worse than useless. You get the queries themselves and the order the source ranks them in. Many teams pair this list with their own click or conversion data, which is more reliable than a third-party estimate anyway.
Do I need an account on the source site? No. Nothing is signed in to and nothing is set up.
Does it need my login or password? No. It never asks for one and could not use one.
Can I schedule it? Yes. Run it on any schedule you like and watch how a market's queries shift week to week.
How many keywords will one seed actually give me?
At depth 1, a few thousand for a broad commercial topic. At depth 2 a single broad seed can pass ten thousand distinct queries. A narrow or obscure seed returns far fewer, because the source genuinely has less to suggest for it. Your "Maximum keywords per seed" setting is the ceiling in every case.
Which markets and languages does it cover? Russian, Kazakh and Turkish, each with its own alphabet and its own ordering. Regionally you can choose 17 Russian cities, Minsk and Belarus nationwide, Almaty, Astana and Kazakhstan nationwide, Tashkent, and Turkey nationwide.
Why do the same seeds give different results in different cities? Because people in different cities search for different things, and the source knows it. A pizza query in Yekaterinburg surfaces a chain that only exists there. That is precisely what makes this useful for local work, and it is why the region setting is worth choosing deliberately.