I have searched for any existing and/or related issues.
I have searched for any existing and/or related discussions.
I have also searched in the CLOSED issues AND CLOSED discussions and found no related items (your issue might already be addressed on the development branch!).
I am using the latest version of Open WebUI.
Installation Method
Docker
Open WebUI Version
0.8.10
Ollama Version (if applicable)
N/A
Operating System
Linux
Browser (if applicable)
Firefox
Confirmation
I have read and followed all instructions in README.md.
I am using the latest version of both Open WebUI and Ollama.
I have included the browser console logs.
I have included the Docker container logs.
I have provided every relevant configuration, setting, and environment variable used in my setup.
I have clearly listed every relevant configuration, custom setting, environment variable, and command-line option that influences my setup (such as Docker Compose overrides, .env values, browser settings, authentication configurations, etc).
I have documented step-by-step reproduction instructions that are precise, sequential, and leave nothing to interpretation. My steps:
Start with the initial platform/version/OS and dependencies used,
Specify exact install/launch/configure commands,
List URLs visited, user input (incl. example values/emails/passwords if needed),
Describe all options and toggles enabled or changed,
Include any files or environmental changes,
Identify the expected and actual result at each stage,
Ensure any reasonably skilled user can follow and hit the same issue.
Expected Behavior
Works
Actual Behavior
Doesn't work. I think my voice isn't being translated (STT) so the assistant doesn't start replying. There are no errors in docker logs.
Steps to Reproduce
Use the docker cuda image.
Start open-webui
Use docker exec open-webui-openWebUI-1 bash -lc "pip install datasets==3.6.0"
Start trying voice mode - doesn't work.
STT works when using the microphone icon "Dictate"
TTS works when using "Read Aloud" on the prompt.
Voice mode (Call) pops up but doesn't respond or do anything.
I have confirmed the GPU is being utilized with nvidia-smi in both instances, so it should be working.
Attached docker-compose and screenshot of the admin panel.
Originally created by @frenzybiscuit on GitHub (Mar 14, 2026).
Original GitHub issue: https://github.com/open-webui/open-webui/issues/22684
### Check Existing Issues
- [x] I have searched for any existing and/or related issues.
- [x] I have searched for any existing and/or related discussions.
- [x] I have also searched in the CLOSED issues AND CLOSED discussions and found no related items (your issue might already be addressed on the development branch!).
- [x] I am using the latest version of Open WebUI.
### Installation Method
Docker
### Open WebUI Version
0.8.10
### Ollama Version (if applicable)
N/A
### Operating System
Linux
### Browser (if applicable)
Firefox
### Confirmation
- [x] I have read and followed all instructions in `README.md`.
- [x] I am using the latest version of **both** Open WebUI and Ollama.
- [x] I have included the browser console logs.
- [x] I have included the Docker container logs.
- [x] I have **provided every relevant configuration, setting, and environment variable used in my setup.**
- [x] I have clearly **listed every relevant configuration, custom setting, environment variable, and command-line option that influences my setup** (such as Docker Compose overrides, .env values, browser settings, authentication configurations, etc).
- [x] I have documented **step-by-step reproduction instructions that are precise, sequential, and leave nothing to interpretation**. My steps:
- Start with the initial platform/version/OS and dependencies used,
- Specify exact install/launch/configure commands,
- List URLs visited, user input (incl. example values/emails/passwords if needed),
- Describe all options and toggles enabled or changed,
- Include any files or environmental changes,
- Identify the expected and actual result at each stage,
- Ensure any reasonably skilled user can follow and hit the same issue.
### Expected Behavior
Works
### Actual Behavior
Doesn't work. I think my voice isn't being translated (STT) so the assistant doesn't start replying. There are no errors in docker logs.
### Steps to Reproduce
Use the docker cuda image.
Start open-webui
Use docker exec open-webui-openWebUI-1 bash -lc "pip install datasets==3.6.0"
Start trying voice mode - doesn't work.
STT works when using the microphone icon "Dictate"
TTS works when using "Read Aloud" on the prompt.
Voice mode (Call) pops up but doesn't respond or do anything.
I have confirmed the GPU is being utilized with nvidia-smi in both instances, so it should be working.
Attached docker-compose and screenshot of the admin panel.
### Logs & Screenshots

[docker-compose.txt](https://github.com/user-attachments/files/26000313/docker-compose.txt)
### Additional Information
.
GiteaMirror
added the bug label 2026-05-21 02:17:39 -05:00
I'm commenting to say that Speech-to-Text, Text-to-Speech, and the Voice mode features all work perfectly fine for me on Ubuntu 24.04.4 LTS using Mozilla Firefox Snap for Ubuntu v148.0.2 (64-bit).
I am not using the cuda tagged image of Open WebUI though. I personally use the dev tagged image.
I'm also not running docker exec open-webui-openWebUI-1 bash -lc "pip install datasets==3.6.0". I just run docker compose up -d.
<!-- gh-comment-id:4064717784 -->
@silentoplayz commented on GitHub (Mar 16, 2026):
I'm commenting to say that Speech-to-Text, Text-to-Speech, and the `Voice mode` features all work perfectly fine for me on Ubuntu 24.04.4 LTS using Mozilla Firefox Snap for Ubuntu v148.0.2 (64-bit).
I am not using the `cuda` tagged image of Open WebUI though. I personally use the `dev` tagged image.
I'm also not running `docker exec open-webui-openWebUI-1 bash -lc "pip install datasets==3.6.0"`. I just run `docker compose up -d`.
<img width="2334" height="663" alt="Image" src="https://github.com/user-attachments/assets/794a254a-d8cd-4d78-9c59-25ce3788868b" />
Blocking a user prevents them from interacting with repositories, such as opening or commenting on pull requests or issues. Learn more about blocking a user.
Originally created by @frenzybiscuit on GitHub (Mar 14, 2026).
Original GitHub issue: https://github.com/open-webui/open-webui/issues/22684
Check Existing Issues
Installation Method
Docker
Open WebUI Version
0.8.10
Ollama Version (if applicable)
N/A
Operating System
Linux
Browser (if applicable)
Firefox
Confirmation
README.md.Expected Behavior
Works
Actual Behavior
Doesn't work. I think my voice isn't being translated (STT) so the assistant doesn't start replying. There are no errors in docker logs.
Steps to Reproduce
Use the docker cuda image.
Start open-webui
Use docker exec open-webui-openWebUI-1 bash -lc "pip install datasets==3.6.0"
Start trying voice mode - doesn't work.
STT works when using the microphone icon "Dictate"
TTS works when using "Read Aloud" on the prompt.
Voice mode (Call) pops up but doesn't respond or do anything.
I have confirmed the GPU is being utilized with nvidia-smi in both instances, so it should be working.
Attached docker-compose and screenshot of the admin panel.
Logs & Screenshots
docker-compose.txt
Additional Information
.
@silentoplayz commented on GitHub (Mar 16, 2026):
I'm commenting to say that Speech-to-Text, Text-to-Speech, and the
Voice modefeatures all work perfectly fine for me on Ubuntu 24.04.4 LTS using Mozilla Firefox Snap for Ubuntu v148.0.2 (64-bit).I am not using the
cudatagged image of Open WebUI though. I personally use thedevtagged image.I'm also not running
docker exec open-webui-openWebUI-1 bash -lc "pip install datasets==3.6.0". I just rundocker compose up -d.@tjbck commented on GitHub (Mar 22, 2026):
Unable to reproduce