Is your feature request related to a problem? Please describe.
With multimodal LLMs that can process audio now widely available, it would be very helpful to pass an audio file directly to the model, similar to how images can now be sent as part of a message.
Describe the solution you'd like
When uploading an audio file, a dialog popup should ask whether the file should be transcribed to text or passed as context directly.
Originally created by @Simon-Stone on GitHub (Dec 16, 2024).
Original GitHub issue: https://github.com/open-webui/open-webui/issues/7890
**Is your feature request related to a problem? Please describe.**
With multimodal LLMs that can process audio now widely available, it would be very helpful to pass an audio file directly to the model, similar to how images can now be sent as part of a message.
**Describe the solution you'd like**
When uploading an audio file, a dialog popup should ask whether the file should be transcribed to text or passed as context directly.
Blocking a user prevents them from interacting with repositories, such as opening or commenting on pull requests or issues. Learn more about blocking a user.
Originally created by @Simon-Stone on GitHub (Dec 16, 2024).
Original GitHub issue: https://github.com/open-webui/open-webui/issues/7890
Is your feature request related to a problem? Please describe.
With multimodal LLMs that can process audio now widely available, it would be very helpful to pass an audio file directly to the model, similar to how images can now be sent as part of a message.
Describe the solution you'd like
When uploading an audio file, a dialog popup should ask whether the file should be transcribed to text or passed as context directly.