llama-gpt · Issues· 96 open
Open on GitHubLocally synced open issues (discussions stay on GitHub)
- #164
Podman support
Updated Jun 21, 2025 - #129
Unable to build CUDA-enabled image, missing Makefile
Updated Feb 14, 2025 - #163
cude issue : cant get this to work ./run.sh --model 7b --with-cuda
Updated Oct 25, 2024 - #151
Can't run on Linux
Updated Oct 23, 2024 - #162
Some Python related errors resulting in "Host [llama-gpt-api:8000] not yet available..." and exit code of 1
Updated Sep 6, 2024 - #132
nvidia-container-cli error
Updated Aug 8, 2024 - #89
Support file upload
Updated Jun 13, 2024 - #161
Intel CPU support
Updated Jun 4, 2024 - #157
[FEATURE REQUEST] Support for NPU hardware acceleration
Updated Jun 4, 2024 - #159
coroutine error StopIteration if the AI response is empty
Updated May 7, 2024 - #158
Unable to increase max token size
Updated May 7, 2024 - #35
How to use my own models?
Updated Apr 23, 2024 - #121
Error when running model other than 7b
Updated Apr 22, 2024 - #154
Add API Key restriction so the API is not always OPEN
Updated Apr 18, 2024 - #148
Is there the ability to swap the language to answer?
Updated Apr 18, 2024