Run AI models locally — Ollama, LM Studio, vLLM
Ollama v0.32.4 adds MLX Apple GPU support for Laguna and optimizes Qwen3 MoE decoding.
LM Studio 0.4.20 adds support for enterprise internal network endpoints and LM Link model sharing with Bionic.
LocalAI v4.7.1 introduces UI-managed voice cloning profiles, local avatar generation, and improved GPU device controls.
Ollama v0.32.0 launches an interactive agent experience directly in the CLI for chatting, coding, and web searching.
LM Studio 0.4.20 introduces support for enterprise internal network model endpoints and using local models in Bionic over LM Link.
LocalAI v4.7.0 introduces a UI-managed voice cloning library, local avatar generation, and interleaved thinking with tool calls.