#143·memvid

[FEATURE] llama.cpp integration

Author: Anico2Created Jan 12, 2026Updated Mar 23, 2026
Labelsenhancement

Feature Description

As stated in the documentation, we can use local running models for privacy reason, e.g. memvid ask knowledge.mv2 \ --question "What are the key findings?" \ --use-model "ollama:qwen2.5:1.5b"

The used endpoint is not compatible with llama.cpp API, e.g. models running under llama-server