/api/chat), not the OpenAI-compatible
/v1 endpoint. Three modes are supported:
For cloud-only setup with the dedicated
ollama-cloud provider id, see
Ollama Cloud. Use ollama-cloud/<model> refs when
you want cloud routing kept separate from a local ollama provider.
The canonical config key is baseUrl. baseURL is also accepted for
OpenAI-SDK-style examples, but new config should use baseUrl.
This page is an index. Ollama is documented on nine pages, one per reader
job. Open the page that matches your task.
Where each section moved
Every anchor the single-page version published still resolves here, so an existing link such as/providers/ollama#node-local-inference keeps working.
Each entry points at the page that now holds the content.
Ollama setup
- Auth rules
- Getting started
- Cloud models through a local host
- Local and LAN hosts
- Remote and Ollama Cloud hosts
- Custom provider ids
- Auth profiles
- Memory embedding scope
- Onboarding (recommended)
- Run onboarding
- Select a model
- Verify
- Manual setup
- Install and start Ollama
- Set a credential
- Select the model
- Common recipes
- Model selection
- Quick verification
- Local model with auto-discovery
- LAN Ollama host with manual models
- Ollama Cloud only
- Cloud plus local through a signed-in daemon
- Multiple Ollama hosts
- Small local model profile
- Advanced configuration
- Legacy OpenAI-compatible mode
- Context windows
- Thinking control
- Reasoning models
- Model costs
- Memory embeddings
- Streaming configuration
- Troubleshooting
- WSL2 crash loop (repeated reboots)
- Ollama not detected
- No models available
- Connection refused
- Remote host works with curl but not OpenClaw
- Model outputs tool JSON as text
- Kimi or GLM returns garbled symbols
- Cold local model times out
- Large-context model is too slow or runs out of memory
Related
Ollama Cloud
Cloud-only setup with the dedicated
ollama-cloud provider.Model providers
Overview of all providers, model refs, and failover behavior.
Model selection
How to choose and configure models.
Ollama Web Search
Full setup and behavior details for Ollama-powered web search.
LM Studio
Another local runner for GGUF or MLX models, as a GUI app or a headless server.
Memory LanceDB
Long-term memory in LanceDB, with local Ollama-compatible embeddings.
Configuration
Full config reference.