Reliable model swapping for any local OpenAI/Anthropic compatible server - llama.cpp, vllm, etc
mostlygeek/llama-swap is engineered as a server, cloud service, or developer package. It is primarily deployed via Docker containers or package managers rather than a single desktop executable.