Skip to main content
List all models currently loaded and running on your Ollama server. This function shows active models, when they will expire from memory, and how much VRAM they are using.

Samples

List running models

See which models are currently loaded:
Returns:

Monitor model expiration

Check when models will unload from memory:

Check VRAM usage

See how much video memory models are consuming:

Connect to specific host

Monitor models on a remote Ollama server:

Total resource usage

Calculate total VRAM used by all running models:

Arguments

Returns

TABLE: A table with the following columns: