Twinny

Fully offline

Free, lightweight VS Code copilot that runs entirely on Ollama. Strong on autocomplete.

Editorial verdict: “Best minimal-surface Copilot-replacement that's been Ollama-native since day one.

Coding agent
Free
MIT
4.2 / 5
GitHub ★ 3,500

Compatibility at a glance

Which runtime + OS combos this app works against. Source of truth for "will it run on my setup?"

§ Runtimes supported
ollama
§ OS / platform
macoslinuxwindows
§ Hardware + model hint
Minimum VRAM
8 GB
Recommended starter model
DeepSeek Coder 6.7B Q4_K_M or Qwen 2.5 Coder 7B

What it is

Twinny is for solo Ollama users who want a Copilot replacement without the configuration overhead. It bridges directly to Ollama, delivering autocomplete and inline chat with lower latency than Continue, especially on models like DeepSeek Coder 6.7B Q4_K_M or Qwen 2.5 Coder 7B. The extension is fully offline and MIT-licensed, running on macOS, Linux, or Windows with at least 8 GB VRAM. Its minimal surface area means faster setup and tighter autocomplete focus, but you lose agentic edit mode and JetBrains support. If your workflow is VS Code and you just need local autocomplete that works out of the box, Twinny delivers without the bloat.

✓ Strengths

  • +Tiny config — works out of the box with Ollama running locally
  • +Lower latency than Continue for autocomplete
  • +MIT-licensed, fully open

△ Caveats

  • No JetBrains support
  • Fewer features than Continue (no agentic edit mode)