Audio Devices

No microphone signal

  1. Verify that your microphone is plugged into the audio device
  2. Make sure the audio device is plugged into the PC and turned on
  3. In DRE, go to Sound -> Input and verify the Input is correct
  4. If it is, try switching to another device and back again

Choose Input Device

Ensure Default Communications Device is how you want

Sometimes, Windows can override the mic input via its “Default Communications Device” like here, where the “Microphone” has taken the communications role:

DefaultCommsDevice

This may not be intended, and if you’re struggling to get a signal into DRE, try setting Default Communication Device to the same as the Default Device.

Microphone Gain

If DRE doesn’t pick up what you say, but you see the input level meter moving while speaking, try lowering your microphone level gain.

In Windows Start menu, search for “sound” and click the option Change system sounds. Go to the Recording tab and open the microphone selected in DRE. Head to the Levels tab. Lower the slider about 50% of its current level. Click Apply if able, then OK. Then try speaking into DRE.

Windows Sound Recording Mic Gain Input Slider

Alternatively, in DRE you can adjust the microphone input gain.

Go to the side-menu and choose Sound. Then click the Input tab and locate the Input Gain option. Try lowering this by 15% increments and speak into the microphone every time.

DRE App Sound Input Gain Option

Yellow Input Meter

If the Input Level Meter is yellow, you either have selected Push To Talk (PTT) or Hotword in the Sound -> Input -> Input Mode

Yellow input meter

 PTT

  • Make sure to set a key or controller button bind to PTT. Map it from Sound -> Input -> Push To Talk (when Input Mode is Push To Talk)
  • Push to PTT key or button and see the input color changing to purple:

Purple input

 Hotword

  • The Input Meter should already animate when you speak into the microphone but it’ll stay yellow until you say the hotword. By default this is DRE or Ok, DRE
  • You can set your own hotwords in Sound -> Input -> Input Mode -> Hotwords

Hotwords edit

ℹ️ Note

Voice Commands load in when the DRE app starts. This takes a few moments and keeps you from making voice inputs during this period. Wait until the command loaded has been completed, before speaking.

Incorrect Phrases

Make sure whatever you say to DRE has a corresponding grammar in one of the commands.

Search the Voice Commands to be sure.

Note that you can use Intents instead if you want to speak freely

3ztzFob

Local Speech Recognizer

To utilize the voice commands in DRE, you need a speech recognizer installed for local Classic recognition. Note you can skip installing local speech recognizers by using Neural Recognition and Intents only.

DRE supports two classic speech recognition systems:

 

1: Microsoft Speech Platform

Used by DRE’s “Microsoft” classic recognizer mode. This requires:

Installing Windows “Basic speech recognition” or “Enhanced speech recognition” does not install Microsoft Speech Platform recognizers.

2: Windows Speech Recognition

Used by DRE’s “Windows” classic recognizer mode. This uses the speech recognition components installed with Windows.

DRE Classic Runtime

To select the Classic Runtime in DRE, go to:

Sound → Recognition → Classic Runtime

 

Windows 11 Speech Recognition Classic Runtime

Microsoft Classic Runtime

Uses Microsoft Speech Platform.

  • Recommended for Windows 11 24H2 or newer
  • Often works better with noisy audio, poor microphones, and different accents
  • Does not require voice training
  • Requires separate runtime and language pack installation
  • Disables DRE dictation-style commands, including:
    • AdminChatAdmin
    • AdminChatAll
    • AdminChat
    • NotesAdd
    • ChatStart
    • ChatReply
    • DDChat

Windows Classic Runtime

Uses Windows’ built-in speech recognition.

  • Built into Windows
  • Can be trained for your voice on older Windows versions
  • Allows dictation commands
  • Default on Windows versions earlier than Windows 11 24H2
  • On Windows 11 24H2 or later, this option may perform worse because classic Windows Speech Recognition is deprecated/replaced by Voice Access

Windows 11 Installation

On Windows 11, install these from:
Settings → Time & language → Language & region → [your language] → Options → Speech recognition

Install:

  • Basic speech recognition
  • Enhanced speech recognition, if available

On Windows 11 24H2 or newer, DRE recommends the Microsoft Speech Platform runtime unless you have explicitly configured DRE to use the Windows runtime.

Windows 11 Speech Recognition Settings

Windows 10 Installation

On Windows 10, the wording is usually different. Go to:
Settings → Time & language → Language → [your language] → Options → Speech → Download

Ayo8d21

Signal:Noise

The Signal-to-Noise ratio (SNR) is the difference between speech volume and background noise levels.

SNR = Signal – Noise

When the noise in your room masks your speech from being too loud, you have a bad SNR.

Good SNR is above 15dB

Check your current Signal-To-Noise Ratio

  1. Sound -> Recognition -> Speech Recognition Performance Graph
  2. Speak a couple of commands like Hello DREWhat time is it
  3. Hover over some of the colored dots
  4. Inspect the Signal to Noise value in the upper left purple window

Speech Recognition Performance Graph: Signal-To-Noise Ratio

Fix a low Signal-To-Noise Ratio

  • Remove noise from your environment if possible, or rotate the mic away from the noise sources
  • Position the mic closer to your mouth
  • Rotate the mic so it’s not picking up your breath

Minimum Confidence

Simultaneously to DRE speech recognizer guessing what you say, it’ll report back how certain it was by producing a confidence value, let’s say 76%.

Minimum Confidence levels in DRE enter the scene to try to restrict how low a confidence a recognition can be before it is rejected. Setting a confidence level of 70% in our example will reject the recognition as it’s 6% below the required.

Shorter sentences often require a higher minimum confidence level, because short noise may be picked up as speech. The longer your input sentence is the lower confidence is required as the recognizer has more data to work with. This is why you will see two sliders to match the Minimum Confidence (Short) and (Long)

ob1PldU

Offsets

 Use the PTT Offset to adjust all minimum confidence needed when speaking using Push To Talk. When using PTT, DRE is certain you are actually speaking, so a lower confidence level can be enough to validate speech.

Use the Dynamic Grammar Offset to increase the confidence level needed for any commands that have a dynamic tag in the phrase. This can be a driver name, car number, current position, etc.. Since driver names may be semi-obscure from time to time, adding a bit extra required Minimum Confidence level here makes sense.

Check your confidence

  1. See the Sound -> Recognition -> Speech Recognition Performance Graph
  2. Inspect the colored dots in the graph and especially their horizontal positioning in relation to the green squared area, which indicates the area of recognized speech

If you find correct speech outside the green area, by hovering over the colored dots, notice if these are left or right of the green area:

Left of the green area

The phrases were rejected by low confidence

Lower the Minimum Confidence sliders to include future speeches like this one

MZBASqr

Right of the green area

The phrases were approved but Confidence compared to its Minimum Confidence level was too high

Increase the Minimum Confidence sliders to keep including future speeches like this one, but also make room for rejected other ones

YgmArKV

Calibrate with the Wizard

Using the Get Started Wizard in 

Settings -> General -> Wizard

can help set the Minimum Confidence values required for your environment.

Speech Pace

Sometimes noise is picked up as speech by the speech recognizer. This is called a false positive (FP) because it was falsely approved.

To counter this from happening, one strategy is to reject all recognitions with slow pacing.

In one situation a slow-paced FP occurred from the command phrase go back.

The input phrase took 1.8 seconds and had an average character duration of about 300ms. Obviously, no one would speak this sentence this slow in reality, so DRE rejects any speech slower than about 150ms per character

Cj5Gj7M

Check your Character Durations

The green area represents the approved speech, so we want all phrases you did say to fall inside this square.

If some colored dots you did speak are above the green area, try increasing the Maximum Character Duration sliders, so future speech like those is covered by the green area.

Vice versa, if all your speech is at the lower part of the green area, or if false positives get included in the green area, lower the Maximum Character Duration sliders

Adjust your Maximum Character Duration

Drag the two sliders in Sound -> Recognition -> Character Duration to adjust the limit to just above your normal pace.

Experiment with the levels and make sure to keep (Short) higher than (Long), as short phrases normally are slower-paced than longer ones.

Also, try the Settings -> General -> Wizard to auto-adjust to your pace

Maximum Character Duration sliders

More Help on Speech Recognition

See the Speech Recognition Performance Legend in Sound -> Recognition for more tips on how to dial in everything.

Check out the FAQ and Help sections

Get in touch

Email Email icon Discord Discord icon