Speech Recognition Guide
How to optimize speech in DRE, or troubleshoot any recognition issue
Audio Devices
No microphone signal
- Verify that your microphone is plugged into the audio device
- Make sure the audio device is plugged into the PC and turned on
- In DRE, go to Sound -> Input and verify the Input is correct
- If it is, try switching to another device and back again

Ensure Default Communications Device is how you want
Sometimes, Windows can override the mic input via its “Default Communications Device” like here, where the “Microphone” has taken the communications role:

This may not be intended, and if you’re struggling to get a signal into DRE, try setting Default Communication Device to the same as the Default Device.
Microphone Gain
If DRE doesn’t pick up what you say, but you see the input level meter moving while speaking, try lowering your microphone level gain.
In Windows Start menu, search for “sound” and click the option Change system sounds. Go to the Recording tab and open the microphone selected in DRE. Head to the Levels tab. Lower the slider about 50% of its current level. Click Apply if able, then OK. Then try speaking into DRE.

Alternatively, in DRE you can adjust the microphone input gain.
Go to the side-menu and choose Sound. Then click the Input tab and locate the Input Gain option. Try lowering this by 15% increments and speak into the microphone every time.

Yellow Input Meter
If the Input Level Meter is yellow, you either have selected Push To Talk (PTT) or Hotword in the Sound -> Input -> Input Mode

PTT
- Make sure to set a key or controller button bind to PTT. Map it from Sound -> Input -> Push To Talk (when Input Mode is Push To Talk)
- Push to PTT key or button and see the input color changing to purple:

Hotword
- The Input Meter should already animate when you speak into the microphone but it’ll stay yellow until you say the hotword. By default this is DRE or Ok, DRE
- You can set your own hotwords in Sound -> Input -> Input Mode -> Hotwords

ℹ️ Note
Voice Commands load in when the DRE app starts. This takes a few moments and keeps you from making voice inputs during this period. Wait until the command loaded has been completed, before speaking.
Incorrect Phrases
Make sure whatever you say to DRE has a corresponding grammar in one of the commands.
Search the Voice Commands to be sure.
Note that you can use Intents instead if you want to speak freely

Local Speech Recognizer
To utilize the voice commands in DRE, you need a speech recognizer installed for local Classic recognition. Note you can skip installing local speech recognizers by using Neural Recognition and Intents only.
DRE supports two classic speech recognition systems:
1: Microsoft Speech Platform
Used by DRE’s “Microsoft” classic recognizer mode. This requires:
- Microsoft Speech Platform Runtime 11 – Download the x64 MSI only. On Microsoft’s download page, this may appear as one of two similarly named SpeechPlatformRuntime.msi files.
- A matching Microsoft Speech Platform Runtime Language MSI for your language, e.g. English (MSSpeech_SR_en-US_TELE.msi – Be careful to choose the “SR” and not the “TTS” installer)
Installing Windows “Basic speech recognition” or “Enhanced speech recognition” does not install Microsoft Speech Platform recognizers.
2: Windows Speech Recognition
Used by DRE’s “Windows” classic recognizer mode. This uses the speech recognition components installed with Windows.
DRE Classic Runtime
To select the Classic Runtime in DRE, go to:
Sound → Recognition → Classic Runtime

Microsoft Classic Runtime
Uses Microsoft Speech Platform.
- Recommended for Windows 11 24H2 or newer
- Often works better with noisy audio, poor microphones, and different accents
- Does not require voice training
- Requires separate runtime and language pack installation
- Disables DRE dictation-style commands, including:
AdminChatAdminAdminChatAllAdminChatNotesAddChatStartChatReplyDDChat
Windows Classic Runtime
Uses Windows’ built-in speech recognition.
- Built into Windows
- Can be trained for your voice on older Windows versions
- Allows dictation commands
- Default on Windows versions earlier than Windows 11 24H2
- On Windows 11 24H2 or later, this option may perform worse because classic Windows Speech Recognition is deprecated/replaced by Voice Access
Windows 11 Installation
On Windows 11, install these from:
Settings → Time & language → Language & region → [your language] → Options → Speech recognition
Install:
- Basic speech recognition
- Enhanced speech recognition, if available
On Windows 11 24H2 or newer, DRE recommends the Microsoft Speech Platform runtime unless you have explicitly configured DRE to use the Windows runtime.

Windows 10 Installation
On Windows 10, the wording is usually different. Go to:
Settings → Time & language → Language → [your language] → Options → Speech → Download

Signal:Noise
The Signal-to-Noise ratio (SNR) is the difference between speech volume and background noise levels.
SNR = Signal – Noise
When the noise in your room masks your speech from being too loud, you have a bad SNR.
Good SNR is above 15dB
Check your current Signal-To-Noise Ratio
- Sound -> Recognition -> Speech Recognition Performance Graph
- Speak a couple of commands like Hello DRE, What time is it
- Hover over some of the colored dots
- Inspect the Signal to Noise value in the upper left purple window

Fix a low Signal-To-Noise Ratio
- Remove noise from your environment if possible, or rotate the mic away from the noise sources
- Position the mic closer to your mouth
- Rotate the mic so it’s not picking up your breath
Minimum Confidence
Simultaneously to DRE speech recognizer guessing what you say, it’ll report back how certain it was by producing a confidence value, let’s say 76%.
Minimum Confidence levels in DRE enter the scene to try to restrict how low a confidence a recognition can be before it is rejected. Setting a confidence level of 70% in our example will reject the recognition as it’s 6% below the required.
Shorter sentences often require a higher minimum confidence level, because short noise may be picked up as speech. The longer your input sentence is the lower confidence is required as the recognizer has more data to work with. This is why you will see two sliders to match the Minimum Confidence (Short) and (Long)

Offsets
Use the PTT Offset to adjust all minimum confidence needed when speaking using Push To Talk. When using PTT, DRE is certain you are actually speaking, so a lower confidence level can be enough to validate speech.
Use the Dynamic Grammar Offset to increase the confidence level needed for any commands that have a dynamic tag in the phrase. This can be a driver name, car number, current position, etc.. Since driver names may be semi-obscure from time to time, adding a bit extra required Minimum Confidence level here makes sense.
Check your confidence
- See the Sound -> Recognition -> Speech Recognition Performance Graph
- Inspect the colored dots in the graph and especially their horizontal positioning in relation to the green squared area, which indicates the area of recognized speech
If you find correct speech outside the green area, by hovering over the colored dots, notice if these are left or right of the green area:
Left of the green area
The phrases were rejected by low confidence
Lower the Minimum Confidence sliders to include future speeches like this one

Right of the green area
The phrases were approved but Confidence compared to its Minimum Confidence level was too high
Increase the Minimum Confidence sliders to keep including future speeches like this one, but also make room for rejected other ones

Calibrate with the Wizard
Using the Get Started Wizard in
Settings -> General -> Wizard
can help set the Minimum Confidence values required for your environment.
Speech Pace
Sometimes noise is picked up as speech by the speech recognizer. This is called a false positive (FP) because it was falsely approved.
To counter this from happening, one strategy is to reject all recognitions with slow pacing.
In one situation a slow-paced FP occurred from the command phrase go back.
The input phrase took 1.8 seconds and had an average character duration of about 300ms. Obviously, no one would speak this sentence this slow in reality, so DRE rejects any speech slower than about 150ms per character

Check your Character Durations
The green area represents the approved speech, so we want all phrases you did say to fall inside this square.
If some colored dots you did speak are above the green area, try increasing the Maximum Character Duration sliders, so future speech like those is covered by the green area.
Vice versa, if all your speech is at the lower part of the green area, or if false positives get included in the green area, lower the Maximum Character Duration sliders
Adjust your Maximum Character Duration
Drag the two sliders in Sound -> Recognition -> Character Duration to adjust the limit to just above your normal pace.
Experiment with the levels and make sure to keep (Short) higher than (Long), as short phrases normally are slower-paced than longer ones.
Also, try the Settings -> General -> Wizard to auto-adjust to your pace

