Approval before writes
See the plan and the diff before anything hits disk. Every path resolves against one folder root.
Formerly computer.vodka. Now Subterminal Agents.
A control loop around the model you already run. Plan the turn, run one tool, read stdout or the file diff, then stop or wait for approval. Tools, memory, and the session log live in one workspace folder.
Same harness on Windows, Linux, Android, macOS, iOS, and watchOS.
One harness. Attach the model you already run.
The harness, in a chat
Plan, one tool, observe, stop. Writes, deletes, installs, and outbound sends wait for a yes. After setup, prompts and files stay on your machine.
See the plan and the diff before anything hits disk. Every path resolves against one folder root.
Ollama, llama.cpp, LM Studio, vLLM, or any OpenAI-compatible endpoint on localhost or the LAN.
Notes are .md files you can open or delete. Optional embeddings stay on disk. The transcript is plain text.
Off means the tool is never offered to the model. Tokens would live in a local key store.
First message is password-checked so a random sender cannot act for you.
Connectors as switches
How it works
The same loop on every screen we ship, including a watch glance to confirm a write from the wrist.
Point the harness at a folder and a model. It plans the turn and picks one tool. Nothing writes yet.
Stdout, a file diff, a failed test. The next turn starts only after the harness has seen the result.
Deletes, overwrites, installs, and outbound sends wait. Keep your own backups. No built-in git.
Platforms
Desktop, phone, and wrist. Offline after setup.
Desktop
Folder manager on the desk. Local or LAN model. Works offline after setup.
Phone
Same chat shell. SMS and WhatsApp after a password check on the first message.
Wrist
Glance a task. Approve a write from the wrist. Same gate as the laptop.
Limits
The harness is not a branded model. Thin weights produce thin answers.
Once the binary and model are on the machine, prompts and files do not leave it.
Windows, Linux, Android, macOS, iOS, watchOS. Same loop, plus optional SMS and WhatsApp.
Mail, calendar, docs, chat, drive. Off means the tool is never offered.
SMS or WhatsApp after a password check. Hosted gateway is a subscription extra.
Open a page you name. Grind a folder of PDFs, CSVs, or code inside the workspace root.
Read a repo, patch a file, run the test you approve, paste the failure into the next turn.
Write the thing, show it, wait for send. Connected inboxes are a connector.
Plain markdown notes plus optional local embeddings. Delete a memory by deleting the file.
It can organise clips and write ffmpeg plans you approve.
A 7B or 14B on a laptop is not GPT-class. Hosted extras exist for hard turns.
Useful local models need RAM and, ideally, a GPU.
Network fetch is off until you turn it on, and only for pages you name.
It can mis-scope a folder. Review the diff before you approve.
There is no built-in version control.
The chat on this page is a preview. The loop starts after you install beta 3.8.
Subterminal Agents is the harness, not the model.
Pricing
Install it, point it at your model, work in a folder. No card required. From Q4 2027, hosted extras are metered.
No subscription · ships in beta 3.8 · quality follows your model
Plan, tool, observe, stop. Read, write, search, shell, local fetch inside the folder you choose.
Ollama, llama.cpp, LM Studio, or any OpenAI-compatible endpoint.
Markdown notes and a plain transcript. No account.
Deletes, overwrites, and installs wait for a yes. That gate does not move to the cloud.
No seat, no monthly fee, no phone-home just to keep using one computer.
Early paid build. $69.67 per month with a usage cap. Not unlimited. Per-token metering starts in Q4 2027.
Included allowance each month. Files stay unless you attach them. True per-use metering arrives with the Q4 2027 public beta.
Sync chat and memory notes across a laptop and a phone.
Hosted gateway when the lid is closed.
Gmail, Calendar, Slack, Drive need an OAuth broker we host.
A hosted runner keeps a long task going after the lid closes.
$69.67 is billed every month. Usage is capped. It is not unlimited, and it is not billed per token. Per-token metering starts with the Q4 2027 public beta. This early build is unstable. Crashes, lost files, broken connectors, and unfinished features are expected. By requesting access you accept that Subterminal Agents takes no liability for what happens on your machine, files, accounts, or time. The free local install stays the supported path.
About us
Four people. One product. A wrapped API is not an agent harness.
Works on the loop that actually runs: plan, one tool, read stdout, stop. Wires Ollama, llama.cpp, LM Studio, vLLM, and any OpenAI-style /v1 endpoint. Ships the install on Windows, Linux, Android, and Apple Silicon.
Holds the ship date. Writes folder evals that finish a real task, blocks a release if a gate failed, and keeps hosted extras metered. If a demo only works on one laptop, it does not go out.
Designs the shell you tap: chat, capability switches, approval states, desktop, phone, and the watch glance. Type, motion, and the >_ mark. If a control is loud, it does not ship.
Memory is markdown in the workspace. Every path resolves against one folder root. Deletes, overwrites, installs, and outbound sends wait for a yes. First-message password check on SMS and WhatsApp.
Company
Remote-first across New Zealand, Australia, the United Kingdom, and the United States.
The install is free. Quality follows the model you attach. Hosted extras from Q4 2027 are optional and metered.
Remote-first. Releases ship when folder evals pass, not when a demo looks good on one laptop.
Beta, press, partnerships, and security all go to [email protected]. No form first.
Updates
Public log, newest first. Beta 3.8 is open to request.
Grok, Gemini, OpenAI, Claude, Qwen, DeepSeek, Llama, Mistral, Ollama, any OpenAI-compatible local server. Same install on macOS, iOS, watchOS.
Free local stays the full loop. Hosted extras from public beta are usage-only. No seat fee.
Mail, calendar, docs, and chat sit behind the Apps toggle. Off means the tool is never offered.
SMS and WhatsApp drop a task into the same loop. First message is password-checked.
Ollama, llama.cpp, LM Studio, vLLM, or any OpenAI-compatible endpoint.
Same loop on every desktop and phone we could test. No account after setup.
Deletes, overwrites, installs, and outbound sends wait for a yes.
No hidden vector store. Notes are .md files. Every tool call lands in a plain transcript.
Plan, pick one tool, read the result, stop. Four people, one product. Beta 3.8 first.
Careers
Remote across NZ, AU, UK, and US. Pick a role, then skip or apply. Write to [email protected].
This website does not create an account. The chat is a preview. A beta 3.8 request only leaves the browser if you open your mail app. The installed harness keeps prompts, files, shell output, and embeddings on the machine you run. After setup the local loop does not phone home. Hosted extras are opt-in and metered.
The local install is free. Quality follows the model you attach. You are responsible for what you approve. Keep your own backups. Beta 3.8 is invite-only. Do not use the product for anything you could not stand to put your name on. Security notes: subject “Security”, same address.
Beta 3.8
Public beta remains Q4 2027 until the install path is stable. We will draft the request to [email protected].
Usage is capped. It is not unlimited. It is not billed per token. Metering starts in Q4 2027.
This build is unstable. Crashes, lost files, and broken connectors are expected. We take no liability. The free local install stays the supported path.