# Local inference

> Run models on your own machine with the local model server in Kilo Desktop.

- 网址：https://funcoding.ai/agents/kilo-code/desktop/features/local-inference/
- 来源：Kilo Code 官方文档原文（英文），MIT 许可，同步于 2026-10-11
- 官方原文：https://kilo.ai/docs/desktop/features/local-inference

---
Run models on your own machine through a managed local model server, so you can work with a model without sending your prompts to a hosted provider.

<div class="callout callout-note">

The local model server requires a [Kilo account](https://funcoding.ai/agents/kilo-code/desktop/settings/ai/#ai-providers). Sign in to use it.

</div>

## Set up a local model

Sign in, turn the server on, and import a model from a GGUF file from the [Local Model Server](https://funcoding.ai/agents/kilo-code/desktop/settings/ai/#local-model-server) settings. That section covers the full set of server and per-model options.

## Pick it in chat

To use an imported model in a chat, open the model picker, then select  **Local Model Server**. Your imported models appear there. Choose one and use it like any other model.
