See llms.txt for all machine-readable content.
This workflow runs manually or every 10 minutes to check which Ollama models are loaded, log state changes to a file, and optionally unload models that have been idle too long to free GPU memory.
/api/ps endpoint to list currently loaded models and their expires_at timestamps.expires_at) and marks each model as keep, pinned-skip, would-unload, or unload while tracking the previous state./api/generate with keep_alive: 0 for each model marked for unload.http://host.docker.internal:11434 when using Docker).logPath to a writable location for the n8n runtime.idleMinutes, assumedKeepAliveMinutes (to match your Ollama keep-alive behavior), and whether pinned models can be unloaded (unloadPinned).dryRun set to true and verify the log output, then set dryRun to false to enable live unloading.