All posts
·9 min readresearch workflowmarket research

Conference talks and podcasts as research sources: what each one is good for

Tutorials teach you what is settled. Talks and long-form conversation tell you what practitioners are still fighting — which is a year earlier.

The short answer: tutorials describe solved problems, conference talks describe problems practitioners are currently fighting, and podcasts supply the unrehearsed specifics neither of the others includes. A corpus built only from tutorials is systematically about a year behind, because a thing gets taught only once it has stabilised. The most useful part of a talk is usually what the speaker assumes without explaining.

Source selection tends to default to whatever the search box returns first, which is instructional content. That is fine for learning a category and poor for finding an opening in it, since anything with a good tutorial is by definition no longer an open problem.

Where each format sits in the cycle

Formats are not interchangeable containers for the same information. Each one appears at a different moment in a category's life, and knowing which moment you are looking at is most of the skill.

FormatAppears whenBest evidence it providesMain distortion
Conference talkThe problem is live and unsolvedWhat peers treat as settled vs openSpeaker is presenting a best case
Long-form podcastAny time; unstructuredCosts, failures, vendor churn, real numbersSponsored segments; host agreeableness
Tutorial / walkthroughThe answer has stabilisedThe real workflow, step by stepOnly shows paths that worked
Review / comparisonThe category is crowdedWhich alternatives are in the consideration setAffiliate incentives

A corpus that draws from more than one row of that table is much harder to fool, because the distortions do not overlap. Sponsored enthusiasm in a podcast does not survive contact with a walkthrough that shows the tool being abandoned mid-process.

The unexplained assumption is the signal

A conference talk is an argument addressed to peers, which means the speaker skips whatever the room already agrees on. Those skips are a map of the settled ground — and everything the speaker does stop to justify is, by implication, contested.

Read for it deliberately. If four talks in a category all pause to justify the same architectural choice, that choice is unresolved and anyone offering a credible answer has a position. If a step is consistently waved through, that part of the workflow is finished and building there is competing with something that already works.

Reading a talk for its thesis
  • Notes the conclusion the speaker argued for
  • Treats the case study as representative
  • Misses what the room already agreed on
  • Produces a summary you could have read anywhere
Reading a talk for its assumptions
  • Notes what was skipped without explanation
  • Notes what needed defending
  • Treats the Q&A as the honest section
  • Produces a map of settled vs contested ground
The Q&A is worth more than the talk

Questions come from people doing the work, and they are the one part of a conference recording nobody rehearsed. A question that several people ask across different events is a live problem with an audience already assembled around it.

What long-form conversation is uniquely good for

Podcasts are inefficient per minute and unusually rich per hour, and the richness lives in the digressions. Nobody scripts the moment where a founder mentions what the migration actually cost, or why they dropped a vendor after fourteen months — details that would be cut from a produced video and would never appear in a written case study.

Those specifics are exactly what pricing and positioning research needs, because they are the closest public sources get to revealed spending. A single offhand “we were paying about nine hundred a month for that and it wasn't worth it” outweighs a page of speculation, which is the same principle behind pricing research from public content.

The cost is density. A two-hour conversation might contain six usable claims, which makes it a poor use of an afternoon and a good use of a summarisation pass that keeps timestamps so any surprising number can be checked at source — the requirement described in why summaries need timestamps and citations.

Separating the sponsored part

Sponsorship contaminates recommendations, not descriptions. The rule that works is narrow: discard the endorsement, keep the surrounding account of the workflow.

Discarding whole episodes because they carried an ad is over-correction and throws away most of the long-form material in any commercial category. The genuine risk is subtler — a host who books guests from one vendor's orbit produces an entire catalogue that agrees with itself, which reads like consensus and is really one perspective with good distribution. Checking whether agreement traces back to a shared origin is the core discipline in separating hype from signal in creator content.

Finding the talks in the first place

Conference material is poorly served by ordinary search, because talks are titled for a conference programme rather than for anyone looking for them later. Searching the problem you care about will return tutorials almost exclusively.

Two routes work better. The first is to search the event rather than the topic — most established conferences in a field post their full recordings under one channel, so identifying the two or three events practitioners actually attend gets you a browsable archive rather than a search result. The second is to work backwards from people: find someone whose walkthrough proved substantive, then look for whether they have spoken anywhere, since practitioners who explain well tend to get invited to explain again.

The channel-level approach is worth the setup cost because it turns a recurring search problem into a standing source list. Once the three relevant event channels and half a dozen practitioner channels are identified, every future research question in that vertical starts from a known set rather than from a blank search box — the systematic version of which is in analysing channels at scale.

Weighting a mixed corpus

The instinct is to weight by venue prestige. The better axis is proximity: how directly did this speaker experience the thing they are describing?

A hands-on account on a small podcast — someone describing a process they personally ran last month — is stronger evidence than a keynote summarising an industry. Prestige correlates with reach, not with first-hand knowledge, and a corpus weighted by reach will reliably overweight whoever is currently most visible.

Independence still governs everything else. Three talks at the same conference are frequently one conversation with three speakers, and should be counted accordingly when you tally how many sources support a claim.

Dates need the same care, for a reason specific to these formats. A conference recording carries two timestamps — when the talk was given and when it was uploaded — and they can be a year apart. A podcast episode discussing a product decision may be describing something that happened eighteen months before the recording. Anchoring each claim to when the thing described actually happened, rather than when you watched it, prevents a corpus from quietly reporting old conditions as current.

Assembling the mixed corpus

A workable shape for one research question: two or three conference talks to locate the contested ground, three or four podcast episodes for specifics and numbers, and five or six walkthroughs to see the actual workflow. Twelve to fourteen sources, three formats, distortions that do not overlap.

Ordering matters. Talks first tells you which questions are live, so the walkthroughs can be chosen to answer something rather than gathered for general coverage — the same argument for running cheap breadth before expensive depth made in YouTube versus forums for product research.

What it costs

Long-form sources consume the same allowance as short ones, so a podcast-heavy corpus is no more expensive than a tutorial-heavy one — it is simply far more work to read by hand.

As of August 2026 the tiers are $19 for 25 videos and 2 active projects, $59 for 80 videos and 8 projects, and $199 for 250 videos and 20 projects. A fourteen-source mixed corpus fits the entry plan with room for a follow-up; running several categories in parallel is what moves you up. The pricing page carries the detail.

Stop reading. Start shipping.
Mix the formats, keep the sources

Talks, podcast episodes and walkthroughs in one project, summarised with timestamps so any claim can be checked at its origin. 7-day free trial.

Closing thought

A corpus built entirely from tutorials will teach you a category competently and leave you building in the part of it that is already finished. The earlier formats are less polished, harder to process and considerably more useful — because the problems people are still arguing about in public are the only ones still available.

Frequently asked

Are conference talks useful for product research?

Yes, for a specific job: they tell you what practitioners in a field consider a settled problem versus an open one. A talk is a curated argument delivered to peers, so what the speaker assumes without explaining is often more informative than the argument itself.

What do podcasts give you that tutorials do not?

Unrehearsed detail. Long-form conversation produces the specifics people would edit out of a scripted video — what a migration actually cost, why a vendor was dropped, which part of the process nobody enjoys. The signal is in the digressions.

Do sponsored podcast segments distort the research?

They distort recommendations, not descriptions. Treat any tool praised in a sponsored segment as unusable evidence for that tool, while the surrounding conversation about the workflow generally stays reliable. The two need separating rather than discarding together.

Which source type is best for spotting a category shift?

Conference talks, usually a year or so ahead of tutorials. Practitioners present the problem they are currently fighting; tutorials appear once the answer has stabilised enough to teach. The gap between those two moments is the window worth working in.

How should a mixed corpus be weighted?

Not by source type but by independence and by how directly the speaker experienced what they describe. A hands-on account in a small podcast outweighs a keynote summarising other people's work, regardless of the venue's prestige.

Is a two-hour podcast worth including in a corpus?

Only if it is processed properly. Long conversational material has a low density of useful claims but a high density of specifics, so it pays off when it can be summarised with timestamps and searched rather than watched end to end.

What does a mixed-source corpus cost to run?

As of August 2026 the plans are $19, $59 and $199 a month for 25, 80 and 250 videos respectively. Long-form sources consume the same allowance as short ones, so a podcast-heavy corpus of fifteen items sits comfortably inside the entry tier.