Run LLMs locally in 2026 with Ollama, LM Studio or llama.cpp: hardware needs, best small models, quantization and API se…