Approval before damage
You see the plan and the diff before anything hits disk. Paths resolve against one folder root.
Formerly computer.vodka. Now Subterminal Agents.
The control loop around the model you already run. Plan the turn, dispatch one tool, read stdout or the file diff, then stop or wait on an approval gate. Tools, memory, and the session log live in one workspace folder.
Same harness on Windows, Linux, Android, and the Apple ecosystem: macOS, iOS, and watchOS.
One harness. Point it at the model you already run.
The harness, in a chat
Plan, one tool, observe, stop. Writes, deletes, installs, and outbound sends wait for a yes. After setup, prompts and files do not go to our servers.
You see the plan and the diff before anything hits disk. Paths resolve against one folder root.
Ollama, llama.cpp, LM Studio, vLLM, or any OpenAI-compatible endpoint on localhost or the LAN.
Notes are .md files you can open or delete. Optional local embeddings stay on disk. Plain transcript.
Off means the tool is never offered to the model. Tokens would live in a local key store.
First message is password-checked so a random sender cannot act for you.
Apps as switches
How it works
The same loop on every screen we ship, including a watch glance so you can confirm a write from the wrist.
Point the harness at a folder and a model. It plans the turn and picks one tool. Nothing writes yet.
Stdout, a file diff, a failed test. The next turn only starts after the harness has seen the result.
Deletes, overwrites, installs, and outbound sends wait. Keep your own backups. No built-in git.
Platforms
Desktop
Folder manager on the desk. Local or LAN model. Works offline after setup.
Phone
Same chat shell. SMS and WhatsApp after a password check on the first message.
Wrist
Glance a task. Approve a write from the wrist. Same gate as the laptop.
Limits
The harness is not a branded brain. If the weights are thin, the answers are thin.
Once the binary and model are on the machine, prompts and files do not go to our servers.
Windows, Linux, Android, macOS, iOS, watchOS. Same loop, plus optional SMS and WhatsApp.
Mail, calendar, docs, chat, drive. Off means the tool is never offered.
SMS or WhatsApp after a password check. Hosted gateway is a subscription extra.
Open a page you name. Grind a folder of PDFs, CSVs, or code inside the workspace root.
Read a repo, patch a file, run the test you approve, paste the failure into the next turn.
Write the thing, show it, wait for send. Connected inboxes are a connector.
Plain markdown notes plus optional local embeddings. Delete a memory by deleting the file.
It can organise clips and write ffmpeg plans you approve.
A 7B or 14B on a laptop is not GPT-class. Hosted extras exist for hard turns.
Useful local models need RAM and, ideally, a GPU.
Network fetch is off until you turn it on, and only for pages you name.
It can mis-scope a folder. Review the diff before you approve.
There is no built-in version control.
The chat on this page is a preview. The loop starts after you install beta 3.8.
Subterminal Agents is the harness, not the model.
Pricing
Install it, point it at your model, work in a folder. No card. From Q4 2027, hosted extras are optional and metered.
No subscription · ships in beta 3.8 · quality = your model
Plan, tool, observe, stop. Read, write, search, shell, local fetch inside the folder you choose.
Ollama, llama.cpp, LM Studio, or any OpenAI-compatible endpoint.
Markdown notes and a plain transcript. No account.
Deletes, overwrites, and installs wait for a yes. That gate does not move to the cloud.
No seat, no monthly fee, no phone-home just to keep using one computer.
Early paid build. $69.67 each month with a usage cap. Not unlimited. Per-token metering starts in Q4 2027.
Included allowance each month. Files stay unless you attach them. True per-use metering arrives with the Q4 2027 public beta.
Sync chat and memory notes across a laptop and a phone.
Hosted gateway when the lid is closed.
Gmail, Calendar, Slack, Drive need an OAuth broker we host.
A hosted runner keeps a long task going after the lid closes.
$69.67 is billed every month. Usage is capped. It is not unlimited, and it is not billed per token. Per-token metering starts with the Q4 2027 public beta. This early build is unstable. Crashes, lost files, broken connectors, and unfinished features are expected. By requesting access you accept that Subterminal Agents takes no liability for what happens on your machine, files, accounts, or time. The free local install stays the supported path.
About us
Four people. One product. If a model farm wraps an API and calls it an agent, that is not this.
Owns the loop that actually runs: plan, one tool, read stdout, stop. Wires Ollama, llama.cpp, LM Studio, vLLM, and any OpenAI-style /v1 endpoint. Ships the install on Windows, Linux, Android, and Apple Silicon.
Holds the ship date. Writes folder evals that finish a real task, blocks a release if a gate failed, and keeps hosted extras metered. If a demo only works on one laptop, it does not go out.
Designs the shell you tap: chat, capability switches, approval states, desktop, phone, and the watch glance. Type, motion, and the >_ mark. If a control is loud, it does not ship.
Memory is markdown in the workspace. Every path resolves against one folder root. Deletes, overwrites, installs, and outbound sends wait for a yes. First-message password check on SMS and WhatsApp.
Company
The install is free. Quality follows the model you attach. Hosted extras from Q4 2027 are optional and metered.
Remote-first across New Zealand, Australia, the United Kingdom, and the United States.
Beta, press, partnerships, and security all go to [email protected]. No form first.
Updates
Public log. Newest first. Beta 3.8 is open to request.
Grok, Gemini, OpenAI, Claude, Qwen, DeepSeek, Llama, Mistral, Ollama, any OpenAI-compatible local server. Same install on macOS, iOS, watchOS.
Free local stays the full loop. Hosted extras from public beta are usage-only. No seat fee.
Mail, calendar, docs, and chat sit behind the Apps toggle. Off means the tool is never offered.
SMS and WhatsApp drop a task into the same loop. First message is password-checked.
Ollama, llama.cpp, LM Studio, vLLM, or any OpenAI-compatible endpoint.
Same loop on every desktop and phone we could test. No account after setup.
Deletes, overwrites, installs, and outbound sends wait for a yes.
No hidden vector store. Notes are .md files. Every tool call lands in a plain transcript.
Plan, pick one tool, read the result, stop. Four people, one product. Beta 3.8 first.
Careers
Remote across NZ, AU, UK, and US. Pick a role, then skip or apply. Write to [email protected].
This website does not create an account. The chat is a preview. A beta 3.8 request only leaves the browser if you open your mail app. The installed harness keeps prompts, files, shell output, and embeddings on the machine you run. After setup the local loop does not phone home. Hosted extras are opt-in and metered.
The local install is free. Quality follows the model you attach. You are responsible for what you approve. Keep your own backups. Beta 3.8 is invite-only. Do not use the product for anything you could not stand to put your name on. Security notes: subject “Security”, same address.
Beta 3.8
Public beta stays Q4 2027 until the install path is stable. We’ll draft the request to [email protected].
Usage is capped. It is not unlimited. It is not billed per token. Metering starts in Q4 2027.
This build is unstable. Crashes, lost files, and broken connectors are expected. We take no liability. The free local install stays the supported path.