ollama stop llama3.2

Category: AI/ML Tooling

ollama stop llama3.2

Unload a model from memory

Forces Ollama to unload llama3.2 from RAM or VRAM immediately, freeing memory that idle models otherwise hold for several minutes. The model stays on disk, so the next request reloads it with a small delay. Use `ollama ps` afterwards to confirm it is gone.
Looking for more? Search all 7,657 commands — works offline, in English or Spanish, and fixes typos.