When you use local LLMs for real workflows, your local AI runners usually end up being in some sort of rotation because no single tool really covers everything you need it to. I usually bounce between LM Studio for casual chat, AnythingLLM for RAG workflows, llama.cpp for cutting-edge features, and a couple more. So you end up with a bunch of different tools for different jobs.