The skills
Listening and Typing at the Same Time
Type continuously as you listen instead of buffering. Capture location and nature first, and handle accents, noise and callers who change their story.
Published June 29, 2026 · 7 min read
This is the skill the job is actually built on. A caller talks, you type what matters into fields while they are still talking, and you do it without asking them to slow down or repeat themselves more than necessary. Candidates who can type fast and remember well still struggle here, because doing both at once is a distinct skill rather than the sum of the two.
The buffering trap
Almost everyone starts the same way. You listen to a chunk of what the caller says, hold it, wait for a pause, then type it out. Listen, hold, type. It feels controlled and organized.
It falls apart quickly, for a reason worth understanding. While you are typing the chunk you held, you are not listening, so anything the caller says during that window is lost. Meanwhile the buffer you are carrying decays, so the longer the chunk the more of it degrades before you get it down. And the pauses you are waiting for do not arrive. Distressed people do not speak in tidy sentences with gaps for your convenience.
So you fall behind, and falling behind makes it worse, because now you are holding a longer buffer while trying to catch up on typing. From the outside this looks like a dispatcher asking a caller to repeat themselves three times. From the inside it feels like drowning.
Type continuously instead
The alternative is to type more or less constantly from the moment the caller starts, staying a few words behind them rather than a sentence or two. You are not transcribing, and you are certainly not waiting. You are converting speech into keystrokes on a rolling basis, with the smallest buffer you can manage.
This works because it removes the storage problem. If a word is on the screen a second after it was spoken, it never needs to be held, and holding is where the failures come from. It also keeps your ears free, since typing something you already understood takes far less attention than typing something you are still parsing.
What makes it hard is that it requires typing without looking at your hands, and without looking at what you have typed. If you check your screen mid-sentence to see how you are doing, you have stopped listening. That is why the plain typing work matters first: touch typing is a prerequisite here rather than a nice extra. If you are still hunting for keys, build that foundation before layering audio on top, using our typing module and the guide to improving typing speed and accuracy for dispatch work.
Expect what you type to look rough. Fragments, missing articles, abbreviations, no capitals. That is correct. You are capturing content, not producing prose, and tidying can happen later or not at all.
Location and nature come first
When information arrives faster than you can type it, you need a rule that fires without deliberation, because deliberating costs you the next sentence.
The rule is location first, then nature of the incident. Everything else is negotiable. Location is what allows help to arrive at all, and it is the detail that is hardest to reconstruct if you lose it, since a caller who has become hard to understand or has dropped off can no longer supply it. Nature of the incident is second because it determines what kind of help gets sent.
In practice this means that when a caller opens with a rush of narrative, you are listening past most of it for an address, a cross street, a landmark, a building name, anything locational, and that goes down first. Then what is happening. Then the rest, in whatever order it arrives.
It also means being willing to interrupt. Letting someone talk for forty seconds before you have a location is not politeness, it is a gap in the record. A short, direct question that gets you an address is the more useful choice, and it is what training will teach you to do anyway.
Accents, noise and callers who change their story
Real audio is not clean, and test audio is often deliberately made messy. Three problems come up repeatedly.
Accents and unfamiliar pronunciation
The instinct is to stop and decode the word you did not catch. Do not. Stopping to decode costs you the next several words, and you usually end up losing more than the one item you were chasing. Type what you heard phonetically, keep going, and come back to it if it turns out to matter. Context arriving later often resolves it for free.
Street names are the exception worth interrupting for, because a wrong street is a wrong response. Ask for a spelling, and ask early rather than at the end of the call.
Background noise
Noise does not just make words harder to hear, it consumes attention you needed for typing. The practical countermeasure is to widen your expectations: you will not catch everything, so prioritize hard and let the low-value detail go. Noise is also information in itself. Shouting, traffic, a smoke alarm and a barking dog all tell you something about the scene, and it is worth a few words in the record.
Callers who correct themselves
People give a house number and change it, describe a blue car and then decide it was green, name a street and then say they meant the next one over. This is normal and it is the reason typing continuously beats buffering.
Do not delete the first version. Type the correction next to it. On a test that scores final field contents you will end up leaving the corrected value, but during the call, having both visible means you can ask which is right rather than discovering you overwrote the one that turned out to be accurate. Callers who correct themselves once often correct themselves back.
Working with fields rather than a blank box
Most dispatch entry happens into structured fields, not a free-text window, and that changes the skill. Your hands have to move between fields while your ears stay on the caller, and every tab or click is attention spent on the form rather than the call.
Two habits reduce the cost. Learn the tab order of whatever form you are given, in the first few seconds, before the audio starts. And accept information out of order: if the caller gives you a vehicle before an address, put the vehicle in its field and come back. Fighting the caller for a tidy sequence loses more than it gains.
Our data entry module covers the field-to-field mechanics, and data entry with interruptions adds the competing demand that makes the exercise realistic. A dedicated audio data entry module is on our roadmap and is not built yet, so for now the audio half has to come from elsewhere.
Practicing when you have no audio drills
You do not need purpose-built materials to train this. You need speech you cannot pause, which is easy to find.
- Type along with talk radio, podcasts or news broadcasts. Interview shows work best, because people speak conversationally and reasonably fast. Type continuously and do not rewind, ever. The rule against rewinding is what makes this useful.
- Extract rather than transcribe. Instead of trying to catch every word, capture only names, numbers, places and times. That is much closer to what dispatch entry demands, and it trains the filtering habit alongside the typing.
- Use structured fields. Set up a few labeled fields in a document, then fill them from the audio as the details arrive. This trains the tabbing and out-of-order handling that free-text practice misses.
- Practice with the audio slightly too fast. Nudging playback speed up makes normal speech feel manageable afterward, in the same way that practicing at an uncomfortable typing pace makes your working pace feel easy.
- Add a second demand occasionally. Have someone ask you questions while you type, or run a timer you have to check. This is the condition that separates people who cope from people who do not, and why candidates fail multitasking goes into why.
Twenty minutes of this a few times a week does more than an hour of plain typing, because it trains the specific combination the job asks for rather than one half of it. Pair it with the retention work in memory techniques for retaining caller details, since the small buffer you do carry while typing is exactly the thing those techniques protect.
Common questions
Should I type in full sentences or fragments?
Fragments. Full sentences cost keystrokes you do not have and add nothing. Location, nature, description, direction, in whatever shorthand is unambiguous to you. Fluency is not being scored.
What if I fall behind the caller?
Stop trying to catch up on what you missed and rejoin at the present. Trying to reconstruct the last twenty seconds while the caller keeps talking guarantees you lose the next twenty as well. Get current, then ask a targeted question to fill the specific gap.
Is it acceptable to ask a caller to repeat themselves?
Yes, and it is expected for critical detail, especially addresses, spellings and callback numbers. The concern is the pattern where you ask repeatedly because you are buffering rather than typing continuously. Asking once for an address is professional. Asking three times for everything is a symptom.
How fast do I need to type before adding audio?
There is no fixed threshold, but there is a real prerequisite: you should be able to type without looking at your hands or your screen. Speed matters less than being able to keep your eyes and attention elsewhere while your fingers work.
Do these test sections use real 911 recordings?
Assessments typically use scripted audio recorded for the purpose rather than live call recordings. It is often made deliberately imperfect, with background noise, hesitation and self-correction, because those are the conditions the skill has to survive.