Wikiwand AI
Stenomask

How the Stenomask Made Silent Speech Reporting Possible

Stenomask 9/6/2026

A stenomask is a hand-held microphone built into a padded, soundproof enclosure that fits over the speaker's mouth or nose. Some lightweight versions may be fitted with an elastic neck strap to hold them in place while freeing the user's hands for other tasks. The purpose of a stenomask is to allow a person to speak without being heard by other people, and to keep background noise away from the microphone. It's a microphone with the ability to filter out background noise.

Q1

How does a stenomask prevent others from hearing the operator while still capturing clear speech?

A stenomask places a microphone inside a padded, soundproof enclosure that covers the operator’s mouth, or mouth and nose. The enclosure muffles the voice before it can escape into the room, while also shielding the microphone from surrounding sounds. This lets the operator speak continuously without disrupting settings such as courtrooms or classrooms.

Because the microphone sits very close to the speaker’s mouth in this acoustically isolated space, it receives a strong, relatively clean speech signal. That makes stenomasks effective for speech recognition and real-time transcription even in noisy environments. #History

Court reporter tests his stenomask.

Court reporter tests his stenomask.

Q2

Why can voice writers using stenomasks achieve high transcription speeds and accuracy?

Voice writers can achieve high speeds because they re-speak everything they hear directly into a sound-insulated microphone, rather than first writing shorthand and later transcribing notes. A stenomask blocks surrounding noise and prevents the operator’s voice from disrupting a courtroom or classroom, while the voice writer can also identify speakers and describe gestures, nonverbal responses, and other events as they occur.

When paired with a speech-recognition system trained to that individual’s voice, this method can exceed 180 words per minute with more than 95% accuracy. Voice writers may further improve recognition by adjusting how they pronounce particular words, especially difficult terms. This avoids some limitations of Gregg shorthand and stenotype systems, whose notes can be harder to manage with rapid speech or specialized vocabulary. #History

Q3

How does stenomask reporting compare with stenotype machines and Gregg shorthand?

Stenomask reporting uses a soundproof hand-held microphone enclosure: a trained reporter “re-voices” the proceedings into speech-recognition software, including speaker identifications, gestures, and nonverbal responses. In a noisy courtroom, it keeps both ambient sound out of the microphone and the reporter’s voice from disturbing the room. A trained voice writer using a pre-trained system can exceed 180 words per minute with more than 95% accuracy, and can provide near-real-time plain-text output. See #History.

By contrast, stenotype machines and Gregg shorthand record spoken words as written shorthand. Horace Webb developed the stenomask in the early 1940s because fast speakers and difficult terminology could make shorthand notes hard to manage; before practical speech recognition, shorthand reporters also had to dictate their notes for later typing. The principal drawback of stenomask reporting compared with these conventional methods is its conspicuous appearance: the reporter must visibly speak into a mask-like device.

Q4

What does a voice writer say into the mask beyond the words spoken aloud in a courtroom?

A voice writer re-voices everything heard in the proceeding into the stenomask so a speech-recognition system can produce a real-time transcript. Beyond the spoken courtroom words, they may identify each speaker, note gestures, record unspoken responses, and describe relevant actions as they occur.

These spoken annotations help make the resulting record intelligible and complete, while the mask prevents the voice writer’s dictation from disturbing the courtroom. See #History for the device’s development.

Q5

How did Horace Webb’s early experiments with a cigar box and tomato juice can lead to the first stenomask?

In the early 1940s, Horace Webb wanted a reporting method faster and more reliable than Gregg shorthand, particularly for rapid speech and difficult terminology. He experimented with improvised devices, first using a cigar box and then a tomato juice can, to find a way to speak a complete record aloud without disturbing people nearby.

Those trials led him to place a microphone inside a military aviator’s rubber oxygen mask and pair it with a coffee pot packed with sound-absorbing material. This combination enclosed and muffled the speaker’s voice while preserving it for recording, creating the first stenomask. The United States Navy later judged it the most accurate available system for verbatim reporting and adopted it for court reporting. #History

Q6

Why did the United States Navy adopt the stenomask for court reporting?

The United States Navy adopted the stenomask for court reporting because it judged the device to be the most accurate method of transcription among the known systems of verbatim reporting. Its soundproof mask let reporters repeat proceedings clearly into a microphone while blocking surrounding noise.

Its inventor, Horace Webb, had developed it in the early 1940s to overcome shortcomings of Gregg shorthand, whose notes could become difficult to manage when speakers talked quickly or used complex terminology. By capturing spoken re-voicing directly, the stenomask offered a faster, more reliable route to a complete transcript. #History

Q7

What role did improved speech-recognition software play in modern stenomask transcription?

Improved speech-recognition software made modern stenomask transcription practical as a real-time digital process. Earlier reporters could use the mask to repeat proceedings, but then had to dictate their notes for typing—often requiring about two hours of dictation for each hour transcribed. Speech recognition became accurate enough for routine use in the mid-1990s, removing much of that additional step. See #History.

A trained voice writer can now re-voice what they hear into a stenomask connected to a pre-trained recognition system, producing an immediate text feed in settings such as courtrooms. The mask suppresses the operator’s voice and filters ambient noise, while the software converts the speech to text; trained users can exceed 180 words per minute with more than 95% accuracy and can adjust pronunciations to improve recognition. This makes it possible to distribute proceeding transcripts in plain text immediately and integrate them with legal case-management software.

More Top Questions