•

A good focus group needs three things done well: open-ended questions that build from warm-up to core topics, audio recorded with enough microphone coverage that overlapping speakers stay intelligible, and a transcript that captures speaker labels, pauses, and tone, not just words. Get any one of these wrong and the other two won't save you. This guide walks through all three, whether you're running the session in a conference room or over Zoom, Teams, or Webex.
What to Ask
Open-ended questions do the heavy lifting. Closed questions kill a focus group faster than anything else, because a room full of people nodding "yes" tells you nothing you didn't already assume walking in.
Start with a warm-up. Something low-stakes that gets people talking before the real questions arrive. "Tell us your name and one word for how you feel about this topic" works fine. It's not meant to generate insight. It's meant to get everyone's voice into the room once, so the second time feels normal.
Move into your core questions. This is where the actual research happens. Ask things like "What's your biggest concern about accessing care in your neighborhood?" or "How did you feel the last time you used this product?" These should be specific enough to anchor the conversation but open enough that nobody can answer in one word.
Use follow-up prompts to go deeper. "Can you say more about that?" or "Did anyone else experience it differently?" These two lines do more work than most moderators give them credit for. They're what turns a list of individual comments into an actual group discussion.
Close with room for anything you missed. "Is there anything important we haven't covered?" People sometimes save their most honest comment for the very end, once they've decided the group is a safe place to say it.
How to Record It
Recording sounds simple until six people start talking over each other in a room with bad acoustics. Then it isn't.
In person, a single centered microphone is a starting point, not a solution. Lapel mics or boundary mics placed around the table catch quieter voices and reduce the risk of one dominant speaker drowning out everyone else. Watch the room itself too. A large space with tile floors and high ceilings will produce an echo that no amount of post-processing fully fixes; a smaller room with carpet and soft furniture does the job better. Run a 30-second test before the real session starts. It costs nothing and it tells you immediately if someone's voice isn't reaching the mic.
Online, over Zoom, Teams, or Webex, the fix is mostly in the settings. Record to the cloud rather than a local device, since a dropped connection shouldn't cost you the whole session. If the platform supports separate audio tracks per speaker, turn that on. It makes the resulting transcript dramatically easier to label correctly. Mute participants who aren't speaking if the group is large enough that background noise becomes a problem, and have a co-host manage that so the moderator can stay focused on the conversation.
Either way, get consent before you hit record, and get it again in writing if the session touches health information or falls under IRB review.
What and How to Transcribe
A focus group transcript is not a word-for-word dump of audio into text. Or rather, it should be exactly that, but it needs more structure than a plain dictation would give you.
A useful transcript includes:
Word-for-word dialogue, including filler words like "um" when they carry meaning
Clear speaker labels, such as "Participant 2" or "Moderator"
Notable pauses or reactions: [laughter], [long pause], [group nods]
Honest markers where audio genuinely can't be recovered: [inaudible], [overlapping conversation]
This level of detail matters more than it looks like it should. Imagine a participant answers a sensitive question by lowering their voice almost to a whisper before admitting something they hadn't told anyone else. If the transcript flattens that into plain text, or worse, marks it [inaudible] because it was quiet, you lose the moment that probably mattered most in the entire session. The words are the easy part. The way they were said is the part that gets lost if nobody's paying attention.
This is the kind of detail we build focus group transcription around at Qualtranscribe. Overlapping speakers, accents, and sessions that switch between languages mid-conversation are the normal case, not the exception, and the transcript needs to hold up under that.
After the Transcript Is Done
Once you have a clean transcript, the analysis work is just getting started.
Review it for anything that needs to be removed or anonymized before it's shared outside the research team. Names, employers, and locations have a way of slipping into casual conversation even when nobody meant to disclose them, and it's worth a careful pass before the document goes anywhere. If you need a structured approach to that step, our guide on de-identification, anonymization, and pseudonymization walks through the difference and when each one applies.
From there, most researchers move into coding, whether that's manual or done through software like NVivo, and start pulling out the themes and quotes that will end up in the final report. If you're formatting transcripts for that kind of analysis, our NVivo guide covers how to structure the file so it imports cleanly.
A focus group is a lot of moving parts held together by a moderator and a good set of questions. Get the questions right, record with enough coverage that nothing gets lost, and treat the transcript as data worth handling carefully, and the session earns the effort it took to run.
Need help with your next one? Start here and we'll handle the transcription end to end.
FAQ
How many people should be in a focus group? Most focus groups run with six to ten participants. Fewer than that and you lose the group dynamic; more than that and it gets hard for everyone to get a word in.
Should I record video as well as audio? Video helps if you want to capture body language or note who's speaking when voices overlap, but audio alone is enough for a usable transcript as long as the recording is clear.
How long does a typical focus group transcript take to prepare? That depends on session length, number of speakers, and audio quality. A one-hour session with overlapping speakers takes longer to transcribe accurately than a one-hour interview with a single speaker.
What's the difference between a verbatim and a clean transcript? A verbatim transcript includes every filler word, false start, and pause. A clean transcript smooths out filler words while keeping the meaning intact. Which one you need depends on whether you're coding for linguistic detail or just pulling themes.
Can a focus group transcript include multiple languages? Yes. Sessions that shift between languages mid-conversation, or include participants speaking different native languages, can be transcribed and translated together rather than handled as separate files.
Related Reading
Turn your recordings into analysis-ready transcripts.
Human Transcription
Clean verbatim and full verbatim transcripts, delivered by specialist transcriptionists
AI Transcription
Instant Draft powered by AI, with Smart Insights for analysis-ready output
Translation Services
Accurate translation across 99+ languages for multilingual research workflows
Keep reading
Related articles

The Best Transcription Services for Focus Groups in 2026
Focus groups generate some of the most demanding audio in qualitative research. Six to twelve people talking, sometimes over each other, sometimes in a room with bad acoustics, sometimes over a Zoom call with background noise from a home environment. Getting that audio into a clean, usable transcript is where a lot of research budgets and timelines get tested. Not every transcription service handles this well, and the right one often depends on the kind of focus group you're actually running.
Read article

Portuguese Transcription in Latin American Field Research: A Practical Guide
A research team returns from fieldwork across São Paulo, Recife, and Porto Alegre. Three cities, three clearly distinct accents, two weeks of interviews. Back at the institution, someone books a Portuguese transcriptionist. Nobody specifies which variety of Portuguese they need. The transcripts come back with Nordestino expressions normalized to São Paulo usage, a participant whose name appears in three different spellings, and no timestamps. Technically, the words are mostly right. As research data, the transcripts are close to unusable. This happens because transcription for field research in Brazil requires decisions that general transcription services don't prompt researchers to make.
Read article

How to Turn a 60-Minute Webex Recording into a 2-Minute Read
Picture this: someone drops a Webex link in your inbox with a note saying the answer to your question is "somewhere in the recording." Or you ran a 60-minute KOL interview three days ago, need to quote it accurately in a briefing, and your memory of what was said is already fuzzing at the edges.
Read article
© 2026 Qualtranscribe LLC. Services Provided Globally

