Documentation
Guides covering installation, daily use, configuration and troubleshooting.
Getting Started
Install LlamaTray, launch it, and start your first local llama-server.
Usage
Day-to-day LlamaTray workflows: the tray menu, Single Model and Router modes, and the built-in llama.cpp manager.
llama.cpp Manager
The built-in manager that detects your hardware, installs llama.cpp dependencies, and gives you two ways to get a llama-server binary: compile from source (Option A) or download the pre-built release (Option B).
HuggingFace Downloader
Search GGUF repositories, inspect exact file sizes and download models straight into LlamaTray — with live progress, a manual repo/file mode, and automatic model selection after the download.
Configuration
All LlamaTray settings with their defaults: GPU layers, context size, port, sampler presets, mmproj, router options and profiles.
Troubleshooting
Common LlamaTray issues: missing system tray on GNOME, llama-server not found, port conflicts and GPU monitoring gaps.