All posts
·9 min readmarket researchinternational

How to research a market in a language you do not speak

Machine translation has quietly made non-English markets researchable. The failure modes are specific, and mostly avoidable if you know where they sit.

The short answer: build the query set from terms that appear inside native sources rather than translating your own, expect to gather twenty to thirty sources instead of fifteen to twenty-five, keep every original product name beside its translation, and treat local prices as ratios rather than numbers. Translation handles what people do and what breaks; it does not handle tone, and it quietly mangles names.

Competition for a product idea is largely competition among people who read the same sources. An English-speaking founder researching an English-speaking market is working from a corpus that several thousand other founders have already read. One language over, the same workflow problem may be discussed at length by practitioners nobody in your competitive set has heard.

What translation handles well, and what it does not

Product research asks unusually translation-friendly questions. “What tool did you use, what went wrong, how long did it take, what did it cost” are concrete claims with little dependence on register or idiom, and machine translation renders them accurately. The parts that degrade are the parts you would not have used anyway.

Research questionSurvives translation?Handling note
What workflow do people followReliablySteps and sequence translate cleanly
What breaks and what it cost themReliablyVerify figures against a second source
Which tools are usedOnly with the original keptNames get translated literally
Sentiment, sarcasm, enthusiasmPoorlyDo not count sentiment from translations
Marketing language and positioningPoorlyNeeds a native reader, not a model

The sentiment row matters more than it looks. A large amount of validation work leans on reading enthusiasm, and enthusiasm is precisely what survives translation worst — irony arrives as praise. Lean instead on the behavioural markers that translate: what was rebuilt, what was abandoned, what was paid for.

Build the query set from the inside out

The most common mistake is translating your English queries and searching with the results. Native practitioners rarely use the direct translation of the English industry phrase; they use a local term, often an abbreviation or a borrowed word with a different spelling, and the translated query returns a completely different corpus.

The reliable approach is bootstrapping. Find two or three genuinely native sources any way you can, read what they call things, and rebuild the query set from their vocabulary. Then iterate: each new source contributes terms the previous ones did not use. This is the standard corpus-building loop run with an extra vocabulary step, and the underlying method is in searching YouTube like a researcher.

Inward-facing versus outward-facing content

A large share of what surfaces for a foreign-market query is content aimed at outsiders: tutorials for expatriates, explainers for people entering the market, comparisons written for an international audience. It is not evidence about local demand. The tell is who the speaker addresses — if they are explaining the market itself, they are not describing it from inside.

Keep every original term next to its translation

Translation systems treat product names as ordinary words often enough to be dangerous. A tool whose name is a common noun comes back as that noun, and your count of which products practitioners actually use silently breaks. The same happens with feature names and with regulatory terms, which are frequently the whole point of the research.

The fix is mechanical: preserve the original string beside the translation for every proper noun, price and figure, and never count from the translated text alone. This is the same evidentiary discipline as timestamping claims, and it decides whether the research can be re-checked later at all — see citing creators and keeping attribution intact.

Translated-first research
  • English queries machine-translated into the target language
  • Tool names counted from translated text
  • Local prices read as absolute numbers
  • Expat tutorials treated as market evidence
Native-vocabulary research
  • Query set rebuilt from terms native sources use
  • Original names preserved beside translations
  • Prices read as ratios to local alternatives
  • Audience of each source identified explicitly

Prices are ratios, not numbers

A monthly price in another market encodes local purchasing power, tax treatment, payment-method norms and the local competitive set. Converting it into your currency produces a figure that looks actionable and is not. What does transfer is the relationship: whether the tool costs a tenth or a third of the labour it replaces, and how it sits against local substitutes.

Payment norms are worth their own note, because they change product scope rather than just pricing. A market where card subscriptions are unusual needs a different billing path, and that is a build decision as much as a commercial one. The broader method for reading price evidence out of practitioner content is in pricing a SaaS from creator content.

Size the corpus larger, not smaller

The instinct is to gather fewer sources because each one costs more effort. That is backwards. A meaningful share of what you collect will turn out to be outward-facing, mistranslated or too old, so the usable remainder shrinks. Planning for twenty to thirty sources leaves room for that attrition; planning for twelve does not.

If the language genuinely has very little practitioner content on your topic, that is itself a finding — sometimes an opportunity, more often a sign the workflow is not yet common there. The judgement between those two readings is the same one in what to do when your niche has almost no videos.

What a foreign-market read should conclude

The output is rarely “launch there”. More often it is one of three things: the same problem exists with a local constraint you would have to build for, the problem does not exist because the surrounding workflow differs, or a local product already owns it and the entry cost is distribution rather than insight.

Each of those points at a different decision. The first frequently turns into a localisation scope question — which parts you build, which you integrate, which you skip — which is the trade-off examined in deciding build versus buy versus integrate from research. The second and third are honest reasons to stay home, and finding them cheaply is the point. Sizing the opportunity before any of it is worth acting on follows the market-size sanity check.

Local constraints are the finding you cannot guess

The most valuable thing a foreign-market read turns up is usually a constraint nobody in your home market has to think about: an invoicing format the tax authority requires, an identity document every account must carry, a payment rail that is dominant locally and unknown elsewhere, a data-residency expectation that is cultural rather than legal.

These never appear in market-sizing material and they routinely decide whether entry is a localisation project or a product rewrite. Practitioners mention them constantly, because they are the parts of the job that generate work, and a single clear account is usually enough to establish that the constraint exists and is load-bearing. Reading them out of content deliberately is the same technique as spotting compliance constraints in practitioner content.

Treat each one as a scope line rather than a red flag. A constraint that takes two weeks to build for is a moat against every competitor who skipped the research; one that requires a local entity or a licence is a genuine stop, and knowing which of the two you are looking at is worth the whole exercise.

What this costs to run

Because the corpus runs larger, a foreign-market read usually sits on the Pro plan rather than Hobby. As of September 2026 Hobby is $19 a month with 25 videos and 2 projects, Pro is $59 with 80 videos and 8 projects, and Studio is $199 with 250 videos, 20 projects and 3 seats. Every plan includes the same pipeline and a 7-day free trial — see the pricing page.

Stop reading. Start shipping.
Research the market nobody in your competitive set can read

Drop in sources in any language and get one synthesis in English, with the original terms and timestamps preserved. 7-day free trial.

Closing thought

The most crowded markets are not the largest ones — they are the ones every founder can read without effort. A language barrier is now mostly a habit barrier, and habits are cheaper to change than markets are.

Frequently asked

Can you really research a market whose language you do not speak?

For demand signals and workflow detail, yes. Machine translation is now reliable enough for the questions product research actually asks — what people do, what breaks, what they pay for — because those are concrete claims rather than nuance-dependent ones. It is not reliable enough for tone, humour or marketing copy.

What breaks first when you rely on translated sources?

Idiom and product names. Translation flattens sarcasm into agreement, and it frequently renders a brand or feature name as its literal meaning, which silently destroys the one thing you were counting. Keeping the original term alongside the translation prevents both.

How do I search in a language I cannot write?

Build the query set from terms that appear inside sources you have already found, rather than translating your English queries. Native search vocabulary rarely matches a translation of the English phrase, and using the wrong terms returns a corpus of tutorials aimed at foreigners rather than at the market itself.

Are auto-generated captions good enough in other languages?

Usually for the general shape of an account, often not for numbers and names, which are exactly the parts you most want. Treat any figure or product name lifted from an auto caption as unverified until a second source repeats it.

Does this work for pricing research?

Only with care. Prices quoted in another market reflect local purchasing power, tax treatment and competitive structure, so they are evidence about that market and almost never transferable to yours. Read them as a ratio to local alternatives rather than as absolute numbers.

How many sources does a foreign-market read need?

Plan for more than a domestic read, not fewer — roughly twenty to thirty rather than fifteen to twenty-five — because a share of what you gather will turn out to be mistranslation, expatriate commentary or tutorials aimed outward rather than inward.

What does a multi-language research pass cost?

As of September 2026, Hobby is $19 a month for 25 videos and 2 projects, Pro is $59 for 80 videos and 8 projects, and Studio is $199 for 250 videos, 20 projects and 3 seats, with a 7-day free trial on every plan. A foreign-market read usually fits Pro because the corpus runs larger than a domestic one.