# Best AI coding agents for Local Models (2026)

> The best AI coding agents platforms for local models: Command Code leads, for developers who want the agent to run against models on their own machine (ollama.

For local models, **Command Code** is our pick: Running models on your own machine is what matters most here, and only Command Code confirms it, listing local model support. Developers who want the agent to run against models on their own machine (Ollama, LM Studio, llama.cpp): no per-token bill, no code leaving the laptop, and offline work. Below is the full ranking and the tradeoffs, or read [how we score](https://www.versusref.com/methodology/).

## What matters for local models

Weight ×5 = decisive, ×1 = relevant.

| Fact | Weight | Command Code | Factory | Goose | Grok Build | Letta Code |
| --- | --- | --- | --- | --- | --- | --- |
| Runs local models | ×5 | ✓  Yes (Oct 2) | ✓  Yes (Oct 2) | ✓  Yes (Oct 2) | ✓  Yes (Oct 2) | ✓  Yes (Oct 2) |
| License | ×3 | Proprietary (npm package 'command-code' declares UNLICENSED) (Oct 2) | n/a | Apache-2.0 (Oct 2) | Apache-2.0 (Oct 2) | Apache-2.0 (Oct 2) |
| Supported model providers | ×3 | n/a | n/a | 45+ providers incl. Anthropic, OpenAI, Gemini +20 more (Oct 2) | xAI (Grok models); custom models via base URL + API key (Oct 2) | Anthropic, OpenAI, Google Gemini, Mistral, Groq +4 more (Oct 2) |
| Bring your own API key | ×2 | Yes (Oct 2) | true (custom models via any OpenAI/Anthropic-compatible... (Oct 2) | n/a | Yes (Oct 2) | Yes (Oct 2) |

- ×5 **Runs local models:** The page exists for agents that can drive a local model; a cloud-only agent cannot win it.
- ×3 **License:** An open-source agent with a local model is the only fully self-contained setup.
- ×3 **Supported model providers:** Broad OpenAI-compatible provider support is usually how local servers plug in.
- ×2 **Bring your own API key:** The same plumbing that accepts your own key usually accepts a local endpoint.

## The ranking, tool by tool

| Rank | Tool | Verdict | Score | Price |
| --- | --- | --- | --- | --- |
| 1 | [Command Code](https://www.versusref.com/coding-agents/tools/command-code/) | Running models on your own machine is what matters most here, and only Command Code confirms it, listing local model support. | 2 of 2 points · 1 matchup | See pricing |
| 2 | [Factory](https://www.versusref.com/coding-agents/tools/factory/) | Running models on your own machine is the deciding factor, and only Factory publishes local model support. | 2 of 2 points · 1 matchup | See pricing |
| 3 | [Goose](https://www.versusref.com/coding-agents/tools/goose/) (OSS) | Running local models is the deciding factor, and only Goose says it does; Claude Code does not publish local model support. | 2 of 2 points · 1 matchup | See pricing |
| 4 | [Grok Build](https://www.versusref.com/coding-agents/tools/grok-build/) (OSS) | Grok Build is built for this. | 2 of 2 points · 1 matchup | See pricing |
| 5 | [Letta Code](https://www.versusref.com/coding-agents/tools/letta-code/) (OSS) | Running against models on your own machine matters most here, and only Letta Code confirms it runs local models; Claude Code does not publish local model support. | 2 of 2 points · 1 matchup | See pricing |
| 6 | [Mistral Vibe](https://www.versusref.com/coding-agents/tools/mistral-vibe/) (OSS) | Running local models is what matters most here, and only Mistral Vibe says it does; Claude Code does not publish local model support. | 2 of 2 points · 1 matchup | See pricing |
| 7 | [OpenHands](https://www.versusref.com/coding-agents/tools/openhands/) (OSS) | Running models on your own machine is what matters most here, and only OpenHands confirms it runs local models; Devin does not publish local model support. | 2 of 2 points · 1 matchup | See pricing |
| 8 | [Pi](https://www.versusref.com/coding-agents/tools/pi/) (OSS) | Running models on your own machine decides this one, and Pi publishes local model support while Claude Code does not. | 2 of 2 points · 1 matchup | See pricing |
| 9 | [Qwen Code](https://www.versusref.com/coding-agents/tools/qwen-code/) (OSS) | Local model support is the deciding factor, and Qwen Code states it runs local models while Claude Code does not publish that. | 2 of 2 points · 1 matchup | See pricing |
| 10 | [Tabby](https://www.versusref.com/coding-agents/tools/tabby/) (OSS) | Both GitHub Copilot and Tabby say they run local models, so the supporting details decide it. | 2 of 2 points · 1 matchup | See pricing |
| 11 | [Zed](https://www.versusref.com/coding-agents/tools/zed/) (OSS) | Local model support is the deciding factor, and Zed states it runs local models while Cursor does not publish that. | 2 of 2 points · 1 matchup | See pricing |
| 12 | [Google Antigravity](https://www.versusref.com/coding-agents/tools/google-antigravity/) | Running against models on your own machine is what matters most here, and only Google Antigravity confirms it runs local models; Claude Code does not publish local model support. | 5 of 6 points · 3 matchups | See pricing |
| 13 | [OpenAI Codex](https://www.versusref.com/coding-agents/tools/openai-codex/) | Running against models on your own machine is what matters most here, and only OpenAI Codex publishes local model support. | 6 of 8 points · 4 matchups | See pricing |
| 14 | [Aider](https://www.versusref.com/coding-agents/tools/aider/) (OSS) | Running models on your own machine matters most here, and Aider supports local models while Claude Code does not publish local model support. | 4 of 6 points · 3 matchups | See pricing |
| 15 | [OpenCode](https://www.versusref.com/coding-agents/tools/opencode/) (OSS) | Running models on your own machine is the deciding factor, and only OpenCode publishes local model support. | 3 of 6 points · 3 matchups | See pricing |
| 16 | [Cline](https://www.versusref.com/coding-agents/tools/cline/) (OSS) | Local model support is the deciding factor, and Cline states it runs local models while Cursor does not publish that. | 2 of 4 points · 2 matchups | See pricing |
| 17 | [Kilo Code](https://www.versusref.com/coding-agents/tools/kilo-code/) (OSS) | Both run local models and both are MIT licensed, so the core requirement is met either way. | 2 of 4 points · 2 matchups | See pricing |
| 18 | [DeepSeek Harness (dsh)](https://www.versusref.com/coding-agents/tools/deepseek-harness/) (OSS) | Neither tool publishes local model support or a list of supported providers, so no one can promise fully offline work here. | 1 of 2 points · 1 matchup | See pricing |
| 19 | [Freebuff](https://www.versusref.com/coding-agents/tools/freebuff/) (OSS) | Neither Claude Code nor Freebuff publishes local model support, so the deciding attribute is missing on both sides. | 1 of 2 points · 1 matchup | See pricing |
| 20 | [JetBrains Junie](https://www.versusref.com/coding-agents/tools/junie/) | Both tools run local models, so the deciding capability is covered either way. | 1 of 2 points · 1 matchup | See pricing |
| 21 | [Kimi Code](https://www.versusref.com/coding-agents/tools/kimi-code/) (OSS) | Neither tool publishes support for running local models, so the deciding factor goes unanswered on both sides. | 1 of 2 points · 1 matchup | See pricing |
| 22 | [GitHub Copilot](https://www.versusref.com/coding-agents/tools/github-copilot/) | Running against models on your own machine is the deciding factor here, and only GitHub Copilot states that it runs local models. | 7.5 of 16 points · 8 matchups | See pricing |
| 23 | [Amp](https://www.versusref.com/coding-agents/tools/amp/) | Frontier multi-model agent + cloud dev environments (orbs); also terminal_agent (Amp CLI, connects to VS Code/Cursor/Windsurf/Zed/Neovim). OWNERSHIP: no longer branded as Sourcegraph - operated by "Amp Frontier Corporation" (privacy policy, modified Sep 10, 2026); about page: "We are an independent agent research lab"; CLI npm package moved from @sourcegraph/amp to @ampcode/cli (May 14, 2026 changelog). Free Hobby tier with BYOK/ChatGPT-sub since Sep 13, 2026. | 0.5 of 2 points · 1 matchup | See pricing |
| 24 | [Augment Code](https://www.versusref.com/coding-agents/tools/augment-code/) | Context-engine coding platform for large codebases; now led by Cosmos cloud agents + Auggie CLI (also ide_extension: VS Code, JetBrains, Vim/Neovim chat). Brief seeded it as an IDE extension; the 2026 pricing page leads with "Try Cosmos". Flat team pricing (up to 50 seats, no per-seat charge) with a pooled $ usage balance. | 0.5 of 2 points · 1 matchup | See pricing |
| 25 | [Cosine](https://www.versusref.com/coding-agents/tools/cosine/) | "The Sovereign AI Lab": coding agent + own post-trained Lumen models, pitched at regulated/air-gapped orgs and niche languages (COBOL, Fortran, Verilog). Formerly branded Genie (Genie CLI introduced Q2 2025). Secondary segment: terminal_agent (cos CLI). | 0.5 of 2 points · 1 matchup | See pricing |
| 26 | [Tabnine](https://www.versusref.com/coding-agents/tools/tabnine/) | Private, governed enterprise coding agent with self-hosted and air-gapped deployment. Now a Tricentis product (acquired 2026-07-30). Secondary segment: terminal_agent (Tabnine Plugin for OpenCode). | 0.5 of 2 points · 1 matchup | See pricing |
| 27 | [Trae](https://www.versusref.com/coding-agents/tools/trae/) | Low-priced AI IDE (Free/Lite/Pro/Pro+/Ultra); product family now split into TraeCode (IDE) and TraeWork (web/desktop/mobile workspace built on SOLO). Owner per brief: ByteDance - the legal entity could not be read from a primary source (privacy policy/ToS are client-rendered); enterprise page cites Douyin as a customer. Chats may be used for training unless Privacy Mode is on. | 0.5 of 2 points · 1 matchup | See pricing |
| 28 | [Verdent](https://www.versusref.com/coding-agents/tools/verdent/) | Plan-Code-Verify multi-model agent; desktop app (primary, most changelog activity), VS Code/JetBrains plugins, Slack/Telegram control; increasingly pitched as an 'AI technical cofounder' with app deployment. | 0.5 of 2 points · 1 matchup | See pricing |
| 29 | [Warp](https://www.versusref.com/coding-agents/tools/warp/) | Agentic terminal / ADE with credit-metered agents and cloud agent orchestration ("Warp Factories"). Secondary segment: cloud_agent. Client source available under AGPL-3.0. | 0.5 of 2 points · 1 matchup | See pricing |
| 30 | [Devin Desktop (formerly Windsurf)](https://www.versusref.com/coding-agents/tools/devin-desktop/) | Agentic AI IDE (ex-Windsurf) inside the Devin family; shares plans with Devin Cloud and Devin CLI. | 1 of 6 points · 3 matchups | See pricing |
| 31 | [Cursor](https://www.versusref.com/coding-agents/tools/cursor/) | Neither tool confirms local model support, which is what matters most here. | 3.5 of 22 points · 11 matchups | See pricing |
| 32 | [Devin](https://www.versusref.com/coding-agents/tools/devin/) | Autonomous cloud coding agent + terminal agent (Devin CLI); plans shared with Devin Desktop. | 1 of 8 points · 4 matchups | See pricing |
| 33 | [Kiro](https://www.versusref.com/coding-agents/tools/kiro/) | Spec-driven AI IDE + CLI (ex-Q Developer CLI) + web/cloud agent; AWS successor to Amazon Q Developer. | 0.5 of 4 points · 2 matchups | See pricing |
| 34 | [Claude Code](https://www.versusref.com/coding-agents/tools/claude-code/) | Terminal-first coding agent included in Claude Pro/Max/Team/Enterprise plans or billed via API; also IDE extension, desktop app and cloud sessions (cloud_agent). | 2 of 38 points · 19 matchups | See pricing |
| 35 | [Crush](https://www.versusref.com/coding-agents/tools/crush/) (OSS) | Source-available (FSL, converts to MIT after 2 years) BYO-model terminal agent; optional Charm-hosted 'Hyper' inference subscription. | 0 of 2 points · 1 matchup | See pricing |
| 36 | [Gemini CLI](https://www.versusref.com/coding-agents/tools/gemini-cli/) (OSS) | Open-source terminal agent (also oss_self_hosted), enterprise/API-key only after consumer users moved to Antigravity CLI. | 0 of 2 points · 1 matchup | See pricing |
| 37 | [GitLab Duo Agent Platform](https://www.versusref.com/coding-agents/tools/gitlab-duo-agent-platform/) | DevSecOps-native agents and flows for GitLab Premium/Ultimate; also IDE extensions (VS Code, JetBrains) and a Duo CLI; self-managed and air-gapped self-hosted-model options. | 0 of 2 points · 1 matchup | See pricing |
| 38 | [gptme](https://www.versusref.com/coding-agents/tools/gptme/) (OSS) | MIT local-first terminal agent; also web UI, desktop app, REST server, MCP/ACP; persistent autonomous-agent template. | 0 of 2 points · 1 matchup | See pricing |

### 1. Command Code

Running models on your own machine is what matters most here, and only Command Code confirms it, listing local model support. [Full Command Code vs Claude Code verdict](https://www.versusref.com/coding-agents/claude-code-vs-command-code/)

### 2. Factory

Running models on your own machine is the deciding factor, and only Factory publishes local model support. [Full Factory vs Devin verdict](https://www.versusref.com/coding-agents/devin-vs-factory/)

### 3. Goose

Running local models is the deciding factor, and only Goose says it does; Claude Code does not publish local model support. [Full Goose vs Claude Code verdict](https://www.versusref.com/coding-agents/claude-code-vs-goose/)

### 4. Grok Build

Grok Build is built for this. [Full Grok Build vs Claude Code verdict](https://www.versusref.com/coding-agents/claude-code-vs-grok-build/)

### 5. Letta Code

Running against models on your own machine matters most here, and only Letta Code confirms it runs local models; Claude Code does not publish local model support. [Full Letta Code vs Claude Code verdict](https://www.versusref.com/coding-agents/claude-code-vs-letta-code/)

### 6. Mistral Vibe

Running local models is what matters most here, and only Mistral Vibe says it does; Claude Code does not publish local model support. [Full Mistral Vibe vs Claude Code verdict](https://www.versusref.com/coding-agents/claude-code-vs-mistral-vibe/)

### 7. OpenHands

Running models on your own machine is what matters most here, and only OpenHands confirms it runs local models; Devin does not publish local model support. [Full OpenHands vs Devin verdict](https://www.versusref.com/coding-agents/devin-vs-openhands/)

### 8. Pi

Running models on your own machine decides this one, and Pi publishes local model support while Claude Code does not. [Full Pi vs Claude Code verdict](https://www.versusref.com/coding-agents/claude-code-vs-pi/)

### 9. Qwen Code

Local model support is the deciding factor, and Qwen Code states it runs local models while Claude Code does not publish that. [Full Qwen Code vs Claude Code verdict](https://www.versusref.com/coding-agents/claude-code-vs-qwen-code/)

### 10. Tabby

Both GitHub Copilot and Tabby say they run local models, so the supporting details decide it. [Full Tabby vs GitHub Copilot verdict](https://www.versusref.com/coding-agents/github-copilot-vs-tabby/)

### 11. Zed

Local model support is the deciding factor, and Zed states it runs local models while Cursor does not publish that. [Full Zed vs Cursor verdict](https://www.versusref.com/coding-agents/cursor-vs-zed/)

### 12. Google Antigravity

Running against models on your own machine is what matters most here, and only Google Antigravity confirms it runs local models; Claude Code does not publish local model support. [Full Google Antigravity vs Claude Code verdict](https://www.versusref.com/coding-agents/claude-code-vs-google-antigravity/)

Running models on your own machine is the deciding factor, and only Google Antigravity publishes local model support. [Full Google Antigravity vs Cursor verdict](https://www.versusref.com/coding-agents/cursor-vs-google-antigravity/)

The deciding factor is running models on your own machine, and only Google Antigravity publishes local model support. [Full Google Antigravity vs Gemini CLI verdict](https://www.versusref.com/coding-agents/gemini-cli-vs-google-antigravity/)

### 13. OpenAI Codex

Running against models on your own machine is what matters most here, and only OpenAI Codex publishes local model support. [Full OpenAI Codex vs Claude Code verdict](https://www.versusref.com/coding-agents/claude-code-vs-openai-codex/)

The deciding factor here is whether the agent can run local models at all. [Full OpenAI Codex vs Cursor verdict](https://www.versusref.com/coding-agents/cursor-vs-openai-codex/)

Both GitHub Copilot and OpenAI Codex say they run local models, so the decision moves to the supporting details. [Full OpenAI Codex vs GitHub Copilot verdict](https://www.versusref.com/coding-agents/github-copilot-vs-openai-codex/)

### 14. Aider

Running models on your own machine matters most here, and Aider supports local models while Claude Code does not publish local model support. [Full Aider vs Claude Code verdict](https://www.versusref.com/coding-agents/aider-vs-claude-code/)

Both run local models and both are Apache-2.0, with OpenAI Codex's license covering the Codex CLI. [Full Aider vs OpenAI Codex verdict](https://www.versusref.com/coding-agents/aider-vs-openai-codex/)

Both Aider and gptme run local models and support bringing your own API key, so the supporting details decide it. [Full Aider vs gptme verdict](https://www.versusref.com/coding-agents/aider-vs-gptme/)

### 15. OpenCode

Running models on your own machine is the deciding factor, and only OpenCode publishes local model support. [Full OpenCode vs Claude Code verdict](https://www.versusref.com/coding-agents/claude-code-vs-opencode/)

Both Crush and OpenCode run local models, so the decision comes down to providers and licensing. [Full OpenCode vs Crush verdict](https://www.versusref.com/coding-agents/crush-vs-opencode/)

### 16. Cline

Local model support is the deciding factor, and Cline states it runs local models while Cursor does not publish that. [Full Cline vs Cursor verdict](https://www.versusref.com/coding-agents/cline-vs-cursor/)

### 17. Kilo Code

Both run local models and both are MIT licensed, so the core requirement is met either way. [Full Kilo Code vs OpenCode verdict](https://www.versusref.com/coding-agents/kilo-code-vs-opencode/)

Both Cline and Kilo Code run local models, so the core requirement is met either way. [Full Kilo Code vs Cline verdict](https://www.versusref.com/coding-agents/cline-vs-kilo-code/)

### 18. DeepSeek Harness (dsh)

Neither tool publishes local model support or a list of supported providers, so no one can promise fully offline work here. [Full DeepSeek Harness (dsh) vs Claude Code verdict](https://www.versusref.com/coding-agents/claude-code-vs-deepseek-harness/)

### 19. Freebuff

Neither Claude Code nor Freebuff publishes local model support, so the deciding attribute is missing on both sides. [Full Freebuff vs Claude Code verdict](https://www.versusref.com/coding-agents/claude-code-vs-freebuff/)

### 20. JetBrains Junie

Both tools run local models, so the deciding capability is covered either way. [Full JetBrains Junie vs GitHub Copilot verdict](https://www.versusref.com/coding-agents/github-copilot-vs-junie/)

### 21. Kimi Code

Neither tool publishes support for running local models, so the deciding factor goes unanswered on both sides. [Full Kimi Code vs Claude Code verdict](https://www.versusref.com/coding-agents/claude-code-vs-kimi-code/)

### 22. GitHub Copilot

Running against models on your own machine is the deciding factor here, and only GitHub Copilot states that it runs local models. [Full GitHub Copilot vs Cursor verdict](https://www.versusref.com/coding-agents/cursor-vs-github-copilot/)

Running models on your own machine is the deciding factor here, and only GitHub Copilot confirms it: it lists local model support, while Claude Code does not publish any. [Full GitHub Copilot vs Claude Code verdict](https://www.versusref.com/coding-agents/claude-code-vs-github-copilot/)

Running models on your own machine is what matters most here, and only GitHub Copilot states that it runs local models. [Full GitHub Copilot vs Devin Desktop (formerly Windsurf) verdict](https://www.versusref.com/coding-agents/devin-desktop-vs-github-copilot/)

Both GitHub Copilot and GitLab Duo Agent Platform state that they run local models, so the deciding capability is matched. [Full GitHub Copilot vs GitLab Duo Agent Platform verdict](https://www.versusref.com/coding-agents/github-copilot-vs-gitlab-duo-agent-platform/)

### 23. Amp

Frontier multi-model agent + cloud dev environments (orbs); also terminal_agent (Amp CLI, connects to VS Code/Cursor/Windsurf/Zed/Neovim). OWNERSHIP: no longer branded as Sourcegraph - operated by "Amp Frontier Corporation" (privacy policy, modified Sep 10, 2026); about page: "We are an independent agent research lab"; CLI npm package moved from @sourcegraph/amp to @ampcode/cli (May 14, 2026 changelog). Free Hobby tier with BYOK/ChatGPT-sub since Sep 13, 2026. No won verdicts for this use case yet; it ranks on ties and near-misses.

### 24. Augment Code

Context-engine coding platform for large codebases; now led by Cosmos cloud agents + Auggie CLI (also ide_extension: VS Code, JetBrains, Vim/Neovim chat). Brief seeded it as an IDE extension; the 2026 pricing page leads with "Try Cosmos". Flat team pricing (up to 50 seats, no per-seat charge) with a pooled $ usage balance. No won verdicts for this use case yet; it ranks on ties and near-misses.

### 25. Cosine

"The Sovereign AI Lab": coding agent + own post-trained Lumen models, pitched at regulated/air-gapped orgs and niche languages (COBOL, Fortran, Verilog). Formerly branded Genie (Genie CLI introduced Q2 2025). Secondary segment: terminal_agent (cos CLI). No won verdicts for this use case yet; it ranks on ties and near-misses.

### 26. Tabnine

Private, governed enterprise coding agent with self-hosted and air-gapped deployment. Now a Tricentis product (acquired 2026-07-30). Secondary segment: terminal_agent (Tabnine Plugin for OpenCode). No won verdicts for this use case yet; it ranks on ties and near-misses.

### 27. Trae

Low-priced AI IDE (Free/Lite/Pro/Pro+/Ultra); product family now split into TraeCode (IDE) and TraeWork (web/desktop/mobile workspace built on SOLO). Owner per brief: ByteDance - the legal entity could not be read from a primary source (privacy policy/ToS are client-rendered); enterprise page cites Douyin as a customer. Chats may be used for training unless Privacy Mode is on. No won verdicts for this use case yet; it ranks on ties and near-misses.

### 28. Verdent

Plan-Code-Verify multi-model agent; desktop app (primary, most changelog activity), VS Code/JetBrains plugins, Slack/Telegram control; increasingly pitched as an 'AI technical cofounder' with app deployment. No won verdicts for this use case yet; it ranks on ties and near-misses.

### 29. Warp

Agentic terminal / ADE with credit-metered agents and cloud agent orchestration ("Warp Factories"). Secondary segment: cloud_agent. Client source available under AGPL-3.0. No won verdicts for this use case yet; it ranks on ties and near-misses.

### 30. Devin Desktop (formerly Windsurf)

Agentic AI IDE (ex-Windsurf) inside the Devin family; shares plans with Devin Cloud and Devin CLI. No won verdicts for this use case yet; it ranks on ties and near-misses.

### 31. Cursor

Neither tool confirms local model support, which is what matters most here. [Full Cursor vs Kiro verdict](https://www.versusref.com/coding-agents/cursor-vs-kiro/)

### 32. Devin

Autonomous cloud coding agent + terminal agent (Devin CLI); plans shared with Devin Desktop. No won verdicts for this use case yet; it ranks on ties and near-misses.

### 33. Kiro

Spec-driven AI IDE + CLI (ex-Q Developer CLI) + web/cloud agent; AWS successor to Amazon Q Developer. No won verdicts for this use case yet; it ranks on ties and near-misses.

### 34. Claude Code

Terminal-first coding agent included in Claude Pro/Max/Team/Enterprise plans or billed via API; also IDE extension, desktop app and cloud sessions (cloud_agent). No won verdicts for this use case yet; it ranks on ties and near-misses.

### 35. Crush

Source-available (FSL, converts to MIT after 2 years) BYO-model terminal agent; optional Charm-hosted 'Hyper' inference subscription. No won verdicts for this use case yet; it ranks on ties and near-misses.

### 36. Gemini CLI

Open-source terminal agent (also oss_self_hosted), enterprise/API-key only after consumer users moved to Antigravity CLI. No won verdicts for this use case yet; it ranks on ties and near-misses.

### 37. GitLab Duo Agent Platform

DevSecOps-native agents and flows for GitLab Premium/Ultimate; also IDE extensions (VS Code, JetBrains) and a Duo CLI; self-managed and air-gapped self-hosted-model options. No won verdicts for this use case yet; it ranks on ties and near-misses.

### 38. gptme

MIT local-first terminal agent; also web UI, desktop app, REST server, MCP/ACP; persistent autonomous-agent template. No won verdicts for this use case yet; it ranks on ties and near-misses.

Source: https://www.versusref.com/coding-agents/best/local-models/
