Local models
Run Cline with local models using Ollama, LM Studio or Atomic Chat.
Run Cline with local inference on your machine.
Quick Start
- Install a local runtime (Ollama, LM Studio or Atomic Chat)
- Start the local server
- In Cline Settings, select the matching provider
- Select a local model
- Enable Use Compact Prompt in Cline Settings → Features
Hardware Requirements
| RAM | Typical local setup |
|---|---|
| 16-32GB | Small/quantized models |
| 32-64GB | Mid-size coding models |
| 64GB+ | Larger models and bigger context windows |
Runtime Options
Ollama
1) Install
- Download from ollama.com
- Install for your OS
2) Find popular local models
- Browse the Ollama model catalog: ollama.com/search
- Sort/filter by popularity, model size, and latest updates
- Open any model page and copy the
ollama pullcommand
3) Pull and run a model
ollama pull <model-name>
ollama run <model-name>4) Configure Cline
- Open Cline Settings
- Select provider: Ollama
- Base URL:
http://localhost:11434 - Select your model from the dropdown
5) Troubleshooting
- Make sure Ollama is running before sending prompts
- If connection fails, verify
http://localhost:11434 - If model is missing, run
ollama pull <model-name>
LM Studio
1) Install
- Download from lmstudio.ai
- Install and launch the app
2) Find local models
- Browse the LM Studio model catalog: lmstudio.ai/models
- Filter by model family, size, and capabilities
- Pick a model that matches your hardware
3) Download a model
- Open Discover and download a model
4) Start server
- Open Developer tab
- Start server (default:
http://localhost:1234)
5) Configure Cline
- Open Cline Settings
- Select provider: LM Studio
- Keep Base URL as
http://localhost:1234 - Select your model from the dropdown
6) Troubleshooting
- Ensure LM Studio server is running
- Ensure a model is loaded
- If connection fails, verify
http://localhost:1234
Atomic Chat
1) Install
- Download from atomic.chat (macOS Apple Silicon)
- Or build from AtomicBot-ai/Atomic-Chat
2) Load a model
- Open Atomic Chat and download or load a local model from the catalog
3) Start the local API
- Atomic Chat exposes an OpenAI-compatible server at
http://127.0.0.1:1337/v1by default - List loaded models:
curl http://127.0.0.1:1337/v1/models
4) Configure Cline
- Open Cline Settings
- Select provider: Atomic Chat
- Base URL:
http://127.0.0.1:1337/v1(default) - Select your model from the dropdown
5) Troubleshooting
- Make sure Atomic Chat is running before sending prompts
- If connection fails, verify
http://127.0.0.1:1337/v1/models - If the model list is empty, load a model in Atomic Chat first
Recommended Cline Settings for Local Inference
- Enable Use Compact Prompt
- Keep tasks focused (smaller context = faster responses)
- Start a new task when context gets too large