Originally created by @Arche151 on GitHub (May 31, 2024).
I am running the docker image with bundled Ollama (CPU interference) and when I click the stop button during generation, in the UI the text generation stops, but my CPU is still under as much load as before.
And when I start a new chat, the text generations takes forever. Only after restarting the Docker container I get the same speeds as before.
So clicking stop doesn't actually stop the interference.
Originally created by @Arche151 on GitHub (May 31, 2024).
I am running the docker image with bundled Ollama (CPU interference) and when I click the stop button during generation, in the UI the text generation stops, but my CPU is still under as much load as before.
And when I start a new chat, the text generations takes forever. Only after restarting the Docker container I get the same speeds as before.
So clicking stop doesn't actually stop the interference.
Blocking a user prevents them from interacting with repositories, such as opening or commenting on pull requests or issues. Learn more about blocking a user.
Originally created by @Arche151 on GitHub (May 31, 2024).
I am running the docker image with bundled Ollama (CPU interference) and when I click the stop button during generation, in the UI the text generation stops, but my CPU is still under as much load as before.
And when I start a new chat, the text generations takes forever. Only after restarting the Docker container I get the same speeds as before.
So clicking stop doesn't actually stop the interference.