Note to first-time contributors: Please open a discussion post in Discussions and describe your changes before submitting a pull request.
Before submitting, make sure you've checked the following:
Target branch: Please verify that the pull request targets the dev branch.
Description: Provide a concise description of the changes made in this pull request.
Changelog: Ensure a changelog entry following the format of Keep a Changelog is added at the bottom of the PR description.
Documentation: Have you updated relevant documentation Open WebUI Docs, or other documentation sources?
Dependencies: Are there any new dependencies? Have you updated the dependency versions in the documentation?
Testing: Have you written and run sufficient tests to validate the changes?
Code review: Have you performed a self-review of your code, addressing any coding standard issues and ensuring adherence to the project's coding standards?
Prefix: To clearly categorize this pull request, prefix the pull request title using one of the following:
BREAKING CHANGE: Significant changes that may affect compatibility
build: Changes that affect the build system or external dependencies
ci: Changes to our continuous integration processes or workflows
chore: Refactor, cleanup, or other non-functional code changes
docs: Documentation update or addition
feat: Introduces a new feature or enhancement to the codebase
fix: Bug fix or error correction
i18n: Internationalization or localization changes
perf: Performance improvement
refactor: Code restructuring for better maintainability, readability, or scalability
style: Changes that do not affect the meaning of the code (white space, formatting, missing semi-colons, etc.)
test: Adding missing tests or correcting existing tests
WIP: Work in progress, a temporary label for incomplete or ongoing work
Changelog Entry
Description
Introduces a global text-to-speech (TTS) queueing system to better manage playback behaviour and eliminate issues with overlapping audio or playback not resuming after cancellation.
Refactoring the TTS management out of the ResponseMessage component for better abstraction and separation of concerns.
Added
TTS class and global TTSManager to handle queuing and cancelling playback state.
Changed
Refactored ResponseMessage to use the new TTSManager.
Extracted TTS logic out of the ResponseMessage component to centralise TTS playback.
Deprecated
None
Removed
None
Fixed
Playback queue not resuming after cancellation (Web API TTS).
Overlapping TTS audio playback (OpenAI (or other TTS engines that produce audio blobs)).
By submitting this pull request, I confirm that I have read and fully agree to the Contributor License Agreement (CLA), and I am providing my contributions under its terms.
🔄 This issue represents a GitHub Pull Request. It cannot be merged through Gitea due to API limitations.
## 📋 Pull Request Information
**Original PR:** https://github.com/open-webui/open-webui/pull/16152
**Author:** [@jcbyte](https://github.com/jcbyte)
**Created:** 7/30/2025
**Status:** ❌ Closed
**Base:** `dev` ← **Head:** `tts-queue`
---
### 📝 Commits (1)
- [`9c7492b`](https://github.com/open-webui/open-webui/commit/9c7492b5c1754a9f2c1571637a676089cce3c6c1) fix: create tts queue
### 📊 Changes
**2 files changed** (+337 additions, -171 deletions)
<details>
<summary>View changed files</summary>
📝 `src/lib/components/chat/Messages/ResponseMessage.svelte` (+41 -171)
➕ `src/lib/utils/tts.ts` (+296 -0)
</details>
### 📄 Description
# Pull Request Checklist
### Note to first-time contributors: Please open a discussion post in [Discussions](https://github.com/open-webui/open-webui/discussions) and describe your changes before submitting a pull request.
**Before submitting, make sure you've checked the following:**
- [x] **Target branch:** Please verify that the pull request targets the `dev` branch.
- [x] **Description:** Provide a concise description of the changes made in this pull request.
- [x] **Changelog:** Ensure a changelog entry following the format of [Keep a Changelog](https://keepachangelog.com/) is added at the bottom of the PR description.
- [x] **Documentation:** Have you updated relevant documentation [Open WebUI Docs](https://github.com/open-webui/docs), or other documentation sources?
- [x] **Dependencies:** Are there any new dependencies? Have you updated the dependency versions in the documentation?
- [x] **Testing:** Have you written and run sufficient tests to validate the changes?
- [x] **Code review:** Have you performed a self-review of your code, addressing any coding standard issues and ensuring adherence to the project's coding standards?
- [x] **Prefix:** To clearly categorize this pull request, prefix the pull request title using one of the following:
- **BREAKING CHANGE**: Significant changes that may affect compatibility
- **build**: Changes that affect the build system or external dependencies
- **ci**: Changes to our continuous integration processes or workflows
- **chore**: Refactor, cleanup, or other non-functional code changes
- **docs**: Documentation update or addition
- **feat**: Introduces a new feature or enhancement to the codebase
- **fix**: Bug fix or error correction
- **i18n**: Internationalization or localization changes
- **perf**: Performance improvement
- **refactor**: Code restructuring for better maintainability, readability, or scalability
- **style**: Changes that do not affect the meaning of the code (white space, formatting, missing semi-colons, etc.)
- **test**: Adding missing tests or correcting existing tests
- **WIP**: Work in progress, a temporary label for incomplete or ongoing work
# Changelog Entry
### Description
- Introduces a global text-to-speech (TTS) queueing system to better manage playback behaviour and eliminate issues with overlapping audio or playback not resuming after cancellation.
- Refactoring the TTS management out of the `ResponseMessage` component for better abstraction and separation of concerns.
### Added
- `TTS` class and global `TTSManager` to handle queuing and cancelling playback state.
### Changed
- Refactored `ResponseMessage` to use the new `TTSManager`.
- Extracted TTS logic out of the `ResponseMessage` component to centralise TTS playback.
### Deprecated
- None
### Removed
- None
### Fixed
- Playback queue not resuming after cancellation (Web API TTS).
- Overlapping TTS audio playback (OpenAI (or other TTS engines that produce audio blobs)).
### Security
- None
### Breaking Changes
- None
---
### Additional Information
- Resolves issue #16150
- Beneficial for #10739
- Improvements made during the development of SSML support for mixed narration blocks.
### Screenshots or Videos
https://github.com/user-attachments/assets/bcfeb60a-9fb3-4631-85e7-ee8814a6f81d
### Contributor License Agreement
By submitting this pull request, I confirm that I have read and fully agree to the [Contributor License Agreement (CLA)](/CONTRIBUTOR_LICENSE_AGREEMENT), and I am providing my contributions under its terms.
---
<sub>🔄 This issue represents a GitHub Pull Request. It cannot be merged through Gitea due to API limitations.</sub>
Blocking a user prevents them from interacting with repositories, such as opening or commenting on pull requests or issues. Learn more about blocking a user.
📋 Pull Request Information
Original PR: https://github.com/open-webui/open-webui/pull/16152
Author: @jcbyte
Created: 7/30/2025
Status: ❌ Closed
Base:
dev← Head:tts-queue📝 Commits (1)
9c7492bfix: create tts queue📊 Changes
2 files changed (+337 additions, -171 deletions)
View changed files
📝
src/lib/components/chat/Messages/ResponseMessage.svelte(+41 -171)➕
src/lib/utils/tts.ts(+296 -0)📄 Description
Pull Request Checklist
Note to first-time contributors: Please open a discussion post in Discussions and describe your changes before submitting a pull request.
Before submitting, make sure you've checked the following:
devbranch.Changelog Entry
Description
ResponseMessagecomponent for better abstraction and separation of concerns.Added
TTSclass and globalTTSManagerto handle queuing and cancelling playback state.Changed
ResponseMessageto use the newTTSManager.ResponseMessagecomponent to centralise TTS playback.Deprecated
Removed
Fixed
Security
Breaking Changes
Additional Information
Screenshots or Videos
https://github.com/user-attachments/assets/bcfeb60a-9fb3-4631-85e7-ee8814a6f81d
Contributor License Agreement
By submitting this pull request, I confirm that I have read and fully agree to the Contributor License Agreement (CLA), and I am providing my contributions under its terms.
🔄 This issue represents a GitHub Pull Request. It cannot be merged through Gitea due to API limitations.