Run open-weight LLMs on your own machine — model catalog, quantization guide, and inference setup.
A practical starting point for local LLMs: open-weight model families, GGUF quantization tiers, engine setup for llama.cpp and Ollama, and hardware notes so you can pick a model that fits your machine.
