How to Capture Desktop Audio for Demos and Webinars
Your mic only hears the room. When you play a demo, a webinar clip, or shared media, desktop audio capture puts that sound into the transcript too, live and in the same record.

Most transcription tools only listen to your microphone. That works fine when one person is talking into one mic. It breaks down the moment the sound comes from your computer instead of the room: a product demo with narration, a webinar you are co-hosting, a video clip you play for the group, or a remote guest joining through your speakers.
Desktop audio capture fixes this. It records the sound your computer plays and feeds it into the same live transcript as your mic. Nothing gets dropped just because it came out of a speaker instead of a mouth.
Mic audio versus desktop audio
It helps to think in terms of two separate sound sources.
Microphone audio is what your mic picks up: your voice, plus anyone speaking near you in the room. It is the obvious source, and for an in-person conversation it is often all you need.
Desktop audio, sometimes called system audio, is everything your computer plays back. That includes a video, a webinar feed, music, a remote participant heard through your speakers, or the narration track in a recorded demo. Your microphone cannot capture this cleanly. It might pick up a muffled echo through the air, but that does not make a usable transcript.
When you capture desktop audio, Kalima takes the sound straight from your system rather than from the room. The result is clean and in sync, because it is the same signal your computer is sending to the speakers.
When you actually need it
You do not need desktop capture for every session. Reach for it when sound that matters is coming out of your computer, not just out of your mouth.
A few common cases:
- Live demos. You are walking through a product, and a recorded clip has narration or sound effects you want in the record.
- Webinars and co-hosted streams. You want both your commentary and the other speakers in one transcript, even when they join through the app rather than the room.
- Playing shared media. Someone drops a video into the call, and you want what was said in it transcribed, not just your reaction to it.
- Remote guests on speakers. A guest dials in and you hear them through your laptop. Desktop capture grabs their words directly instead of through a tinny mic echo.
- Reviewing recordings together. You replay an earlier session or a reference clip and want that audio searchable too.
If your session is one person speaking into a mic with no other audio in play, mic capture alone is the simpler choice. There is no point capturing silence from your speakers.
Mixed capture: mic and desktop together
The most useful setup for demos and webinars is both at once. You talk, your computer plays the demo, and both land in a single live transcript, side by side, in the order they happened.
This is where doing it live pays off. There is no bot sitting in the meeting as a separate participant, and you are not stitching two recordings together afterward. The mic track and the desktop track are transcribed as the moment unfolds, so the record is already complete by the time you stop. You never have to go back and wonder what was said while you were sharing your screen.
If your participants speak different languages, the same live pipeline applies. The transcript runs as people speak, and side-by-side live translation keeps pace, whether the words came from the room or from the shared media. A clip narrated in one language and your commentary in another both show up readable in the same view.
A quick setup checklist
Before you start a session that involves desktop audio, run through this:
- Pick your sources. Decide whether you need mic only, desktop only, or both. For a demo or webinar, both is usually right.
- Turn on desktop or system audio capture. Enable it before you begin, so the opening of your demo is not missed.
- Test the playback. Play a few seconds of the clip you plan to share and confirm the words appear in the live transcript.
- Watch your levels. If desktop audio is much louder or quieter than your voice, adjust your system volume so both sources read clearly.
- Confirm translation if you need it. When the session is multilingual, check that live translation is on before the first guest speaks.
The goal of capturing desktop audio is simple. The transcript should reflect everything that was said in the session, not only the parts your microphone happened to hear. When a demo, a webinar, or a shared clip is part of the conversation, that audio belongs in the record too. Capturing it live means you finish with a complete transcript, not a partial one you have to rebuild from memory.
Platform guides
Try Kalima on your next call.
Live transcription and translation, no bot in the meeting. Free to start.
Keep reading.

How to Transcribe a Medical Lecture or Conference Talk
The worry with a medical talk is always the terminology, and it is mostly the wrong worry. The model handles clinical vocabulary as a matter of course. What it cannot know is your ward abbreviations, the speaker's surname or an internal study code, and that short list is what context is for.

Google Meet Transcription Button Missing or Transcript Not Generated
If the transcription button is missing in Google Meet, you are most likely on the wrong Workspace edition, on a phone, or in a meeting whose host has not allowed it. This guide sorts the causes and shows a live transcript that runs on your own computer instead.

Teams Transcription Greyed Out or Not Available? Causes and a Workaround
When Teams refuses to transcribe, the reason is almost always a setting you cannot reach: an admin policy, a licence on someone else's tenant, or the fact that you are a guest. This guide walks through the causes and shows a transcript that does not depend on any of them.