#129·llama-gpt

无法构建支持 CUDA 的镜像,缺少 Makefile

作者: VirtualDisk创建于 2023年11月10日更新于 2025年2月14日

Attaching to LLaMA-gpt-LLaMA-gpt-api-cuda-ggml-1, LLaMA-gpt-LLaMA-gpt-ui-1 | [INFO wait] -------------------------------------------------------- LLaMA-gpt-LLaMA-gpt-ui-1 | [INFO wait] Docker-compose-wait 2.12.1 LLaMA-gpt-LLaMA-gpt-ui-1 | [INFO wait] --------------------------- LLaMA-gpt-LLaMA-gpt-ui-1 | [DEBUG wait] Starting with configuration: LLaMA-gpt-LLaMA-gpt-ui-1 | [DEBUG wait] - Hosts to be waiting for: [LLaMA-gpt-api-cuda-ggml:8000] LLaMA-gpt-LLaMA-gpt-ui-1 | [DEBUG wait] - Paths to be waiting for: [] LLaMA-gpt-LLaMA-gpt-ui-1 | [DEBUG wait] - Timeout before failure: 3600 seconds LLaMA-gpt-LLaMA-gpt-ui-1 | [DEBUG wait] - TCP connection timeout before retry: 5 seconds LLaMA-gpt-LLaMA-gpt-ui-1 | [DEBUG wait] - Sleeping time before checking for hosts/paths availability: 0 seconds LLaMA-gpt-LLaMA-gpt-ui-1 | [DEBUG wait] - Sleeping time once all hosts/paths are available: 0 seconds LLaMA-gpt-LLaMA-gpt-ui-1 | [DEBUG wait] - Sleeping time between retries: 1 seconds LLaMA-gpt-LLaMA-gpt-ui-1 | [DEBUG wait] -------------------------------------------------------- LLaMA-gpt-LLaMA-gpt-ui-1 | [INFO wait] Checking availability of host [LLaMA-gpt-api-cuda-ggml:8000] LLaMA-gpt-LLaMA-gpt-api-cuda-ggml-1 | ========== LLaMA-gpt-LLaMA-gpt-api-cuda-ggml-1 | == CUDA == LLaMA-gpt-LLaMA-gpt-api-cuda-ggml-1 | ========== LLaMA-gpt-LLaMA-gpt-api-cuda-ggml-1 | CUDA Version 12.1.1 LLaMA-gpt-LLaMA-gpt-api-cuda-ggml-1 | Container image Copyright (c) 2016-2023, NVIDIA CORPORATION & AFFILIATES. All rights reserved. LLaMA-gpt-LLaMA-gpt-api-cuda-ggml-1 | This container image and its contents are governed by the NVIDIA 深度学习 Container License. LLaMA-gpt-LLaMA-gpt-api-cuda-ggml-1 | By pulling and using the container, you accept the terms and conditions of this license: LLaMA-gpt-LLaMA-gpt-api-cuda-ggml-1 | https://developer.nvidia.com/ngc/nvidia-deep-learning-container-license LLaMA-gpt-LLaMA-gpt-api-cuda-ggml-1 | A copy of this license is made available in this container at /NGC-DL-CONTAINER-LICENSE for your convenience. LLaMA-gpt-LLaMA-gpt-api-cuda-ggml-1 | /models/LLaMA-2-7b-chat.bin model found. LLaMA-gpt-LLaMA-gpt-api-cuda-ggml-1 | make: *** No rule to make target 'build'. Stop. LLaMA-gpt-LLaMA-gpt-api-cuda-ggml-1 | Initializing server with: LLaMA-gpt-LLaMA-gpt-api-cuda-ggml-1 | Batch size: 2096 LLaMA-gpt-LLaMA-gpt-api-cuda-ggml-1 | Number of CPU threads: 12 LLaMA-gpt-LLaMA-gpt-api-cuda-ggml-1 | Number of GPU layers: 10 LLaMA-gpt-LLaMA-gpt-api-cuda-ggml-1 | Context window: 4096 LLaMA-gpt-LLaMA-gpt-ui-1 | [INFO

内容来源: getumbrel/llama-gpt