Original caption
Save this for your next build. 1. ollama: run Llama, Mistral, Gemma and hundreds of models with one command 2. llama.cpp: the engine that started local AI, runs on ordinary hardware 3. open-webui: a ChatGPT-style interface for your local models, fully offline 4. vllm: high-throughput serving when it gets serious, built for production 5. jan: an offline ChatGPT alternative as a clean desktop app, no account 6. exo: split big models across your MacBook, iPhone and iPad as one cluster 7. LocalAI: a drop-in OpenAI API replacement on your own hardware, no API bill Which one are you running first? #ai #coding #aitools #claude #claudecode