ktransformers · Issues· 510 open
Open on GitHubLocally synced open issues (discussions stay on GitHub)
- #2189
Seems difficult to install
Updated Sep 18, 2026 - #1897
Intel Consumer Level AVX-VNNI Support
Updated Sep 17, 2026 - #2208
A100 not support?
enhancementUpdated Sep 15, 2026 - #2108
GLM5.2 int8: 'AMXMoEWrapper' object has no attribute 'submit_write_weight_scale_to_buffer'
Updated Sep 11, 2026 - #2184
cache_salt/extra_key are computed & carried but never delivered to Req — radix-cache tenant isolation is a no-op (regression vs upstream sglang)
Updated Sep 3, 2026 - #2179
[Feature Request] qwen4_exp support (Qwen3.8-Flash-Next) — CPU/GPU hybrid expert offload via kt-kernel/sglang-kt
enhancementUpdated Aug 28, 2026 - #1608
FAQs | 常见问题
good first issueUpdated Aug 27, 2026 - #2172
[MXFP4 layerwise prefill] SM 8.6 (Ampere) excluded by capability whitelist — DSV4-Flash on dual A10 falls back to serialized full-GPU path
Updated Aug 27, 2026 - #2151
GeneralMoEWrapper.load_weights() cpu_save path reads nonexistent dict keys, guaranteed KeyError
Updated Aug 18, 2026 - #2127
DeepSeek-V4-Flash-0731 MTP failed to start on avx2 + sm89
Updated Aug 8, 2026 - #2150
kt-sft: chunked_prefill_size sized from cutoff_len, never scaled by batch — any per_device_train_batch_size>1 hard-errors
Updated Aug 8, 2026 - #2106
Error with Deepskeep 4 flash with trtllm_fp4_block_scale_moe
Updated Aug 8, 2026 - #2139
When will kvcache‑ai/sglang be synced with sgl‑project/sglang
Updated Aug 5, 2026 - #2113
MiniMax M2.7: AttributeError: 'LlamafileMoEWrapper' object has no attribute 'submit_write_weight_scale_to_buffer'
Updated Aug 5, 2026 - #2118
Support DSpark speculative decoding for DeepSeek-V4-Flash-0731 in sglang-kt
enhancementUpdated Aug 3, 2026