Skip to content
| Marketplace
Sign in
Visual Studio Code>AI>enginedNew to Visual Studio Code? Get it now.
engined

engined

Rethunk.Tech

|
1 install
| (0) | Free
VS Code language model provider for engined, a local inference door.
Installation
Launch VS Code Quick Open (Ctrl+P), paste the following command, and press enter.
Copied to clipboard
More Info

engined for VS Code

License: MIT VS Code ^1.138.0


Brings engined into VS Code's own Copilot Chat surface: engined's local models sit in the model picker beside Copilot's, its image/vision/audio/search routes attach as ordinary chat tools, and its completions route offers inline ghost text. Everything routes to engined running on the same machine, or reached over VS Code Remote-SSH.

The extension only ever talks to engined.doors; no request goes anywhere else, and the status bar always shows which route answered and whether it stayed local.

Copilot Chat on a local engined model: #enginedImage generates a fantasy kingdom with Chroma, #enginedSpeak narrates it, and the image opens beside the chat

Quick start

systemctl --user start engined && curl http://127.0.0.1:29200/openai/v1/models

Then open Copilot Chat's model picker and choose an engined model. Full runbook (building from source, attaching a tool, verifying, uninstalling): HUMANS.md.

Features

  • Local chat models in Copilot's model picker, streamed through engined
  • Five agent tools: generate/edit images, OCR or describe an image, transcribe or translate audio, synthesize speech, semantic workspace search
  • Inline completions (ghost text), optionally with neighbouring-file context
  • Status bar shows what answered, whether it ran locally, tokens used, and a loading state
  • An Engines view for warming, holding, and stopping engines, with live resource usage
  • A usage report (per-day/per-route requests, tokens, cost, and local/remote split) across every configured door
  • Image attachments work on a text-only engined model that engined bridges to a vision route
  • Copilot's Context Window meter shows the token usage engined reports for each reply

Screenshots

Inline completion: Ornith fills in the haversine formula from the helpers above and the return below
Fill-in-the-middle completions from a local model
Status bar popup: the last chat, background and completion calls side by side, today's usage and the default models
Status popup: what answered, where, and at what cost
Engines view: running engines with RAM and GPU use
Engines view with live resource use
Copilot's Configure Tools listing the four engined tools
engined's tools in Copilot's tool picker
#enginedTranscribe returning the Peter Piper tongue twister from an audio file
Speech-to-text through #enginedTranscribe
Usage report: per-day and per-route requests and tokens
Usage report across routes
Copilot's Context Window meter reading engined's reported token usage
Copilot's Context Window meter with engined's counts

Documentation

Doc Covers
HUMANS.md Build, run, use, and configure the extension
AGENTS.md Contributor map and invariants
CONTRIBUTING.md Commit conventions and test steps
SECURITY.md Vulnerability disclosure
CHANGELOG.md Release notes
docs/tools.md Agent tool parameters
docs/polling.md Model-list polling internals
docs/reasoning-effort.md Reasoning-effort mapping
docs/settings.md Every engined.* setting

License

MIT

  • Contact us
  • Jobs
  • Privacy
  • Manage cookies
  • Terms of use
  • Trademarks
  • Your Privacy Choices
  • Consumer Health Privacy
© 2026 Microsoft