Models & Runtime
Run local models without giving up control.
Download GGUF models, save profiles, start llama.cpp, watch logs and metrics, or connect directly to the server when you want full control.
Explore this realm →- local GGUF library and Hugging Face downloads
- profiles, variants, groups and runtime settings
- start/stop control for llama.cpp plus direct server access
- local /v1 serving plus self-hosted and connected providers