Baike.dev
Tous les outilsTendancesOpen sourceActualitésSoumettre
Connexion
< 返回工具列表
L

LLMKube

> 云原生
开源

Kubernetes operator for self-hosted LLM inference across a heterogeneous GPU fleet: NVIDIA CUDA, AMD Vulkan, and Apple Silicon Metal. Runtimes: llama.cpp, vLLM,

181 stars0 点赞0 次浏览
访问官网GitHub

工具介绍

Kubernetes operator for self-hosted LLM inference across a heterogeneous GPU fleet: NVIDIA CUDA, AMD Vulkan, and Apple Silicon Metal. Runtimes: llama.cpp, vLLM, TGI, mlx-server. Multi-GPU sharding, model caching, OpenAI-compatible endpoints. Apache-2.0, run across homelab and on-prem fleets, actively developed.

> 标签

Goaiapple-siliconautoscalingedge-computing

暂无评论,来聊聊你的看法吧

> 工具信息

发布日期2026年8月1日
最后更新2026年8月1日
分类云原生
定价开源

> 相关工具

K
Kubernetes
容器编排事实标准
H
Helm
Kubernetes 包管理器
S
seaweedfs
SeaweedFS is a distributed storage system for object storage (S3), file systems, and Iceberg tables,
Baike.dev

baike.dev vous aide à découvrir langages, frameworks, bases de données, DevOps et outils cloud native.

Accès rapide

  • Accueil
  • Tous les outils
  • Tendances
  • Open source

À propos

  • À propos
  • Communauté
  • Actualités

Participer

Vous connaissez un excellent outil ? Partagez-le.

Soumettre un outil
© 2026 baike.dev Encyclopédie développeurMis à jour chaque jour · Découvrez de grands outils