ipex-llm · Issues· 1484 open
Open on GitHubLocally synced open issues (discussions stay on GitHub)
- #13335
SYCL multi-GPU inference fails with UR_RESULT_ERROR_OUT_OF_DEVICE_MEMORY on Intel Arc Pro (single GPU works)
user issueUpdated Jan 11, 2026 - #13334
Docker image update needed to support Kernel 6.18
user issueUpdated Jan 9, 2026 - #13320
Could not create a primitive descriptor for a matmul primitive during qwen3 32b inference with portable llama.cpp
user issueUpdated Jan 7, 2026 - #12994
llama_load_model_from_file: failed to load model with SYCL/Level Zero on Intel Arc B580 GPU
user issueUpdated Dec 31, 2025 - #13055
Intel UHD Graphics 620 supported?
user issueUpdated Dec 23, 2025 - #12831
Ollama reports model is 100% on CPU when actually running on GPU
user issueUpdated Dec 22, 2025 - #13333
How does ipex_llm overcome the limitation that dynamic input cannot be used on NPUs?
Updated Dec 13, 2025 - #12761
Nonsense output
user issueUpdated Dec 12, 2025 - #13286
The Olama version is outdated and cannot load the model
user issueUpdated Dec 9, 2025 - #13330
Intel, are you OK?
Updated Dec 6, 2025 - #13317
Update Ollama
Updated Dec 5, 2025 - #13328
Model crashes: IPEX-LLM Ollama v2.3.0-nightly: SIGABRT in sdp_xmx_kernel.cpp with Llama 3.1 on Arc B50 Pro
user issueUpdated Dec 2, 2025 - #13331
GPT-OSS 120B ERROR llama.cpp ipex-llm==2.3.0b20251104
Updated Dec 1, 2025 - #13294
ollama version 0.9.3 with latest docker.io/intelanalytics/ipex-llm-inference-cpp-xpu:latest Docker Image
user issueUpdated Dec 1, 2025 - #13301
Run DeepSeek-R1 random output on Ollama .
user issueUpdated Nov 24, 2025