Skip to main content
Run Large Language Models locally with Ollama Ollama is a fantastic tool for running models locally. Ollama supports multiple open-source models. See the library here. We recommend experimenting to find the best-suited model for your use-case. Here are some general recommendations:
  • llama3.3 models are good for most basic use-cases.
  • qwen models perform specifically well with tool use.
  • deepseek-r1 models have strong reasoning capabilities.
  • phi4 models are powerful, while being really small in size.

Set up a model

Install ollama and run a model using
run model
This gives you an interactive session with the model. Alternatively, to download the model to be used in an Agno agent
pull model

Example

After you have the model locally, use the Ollama model class to access it
View more examples here.

Params

Ollama is a subclass of the Model class and has access to the same params.