Install
One install gives you two things: decosa agent, a private coding agent on Decosa's open model, and decosa-mcp, which adds the Decosa tools to Claude Code, Cursor, Codex or any MCP client.
Status: Private beta, 29 Sep 2026. The install line works for beta accounts with repository access; the public package is waiting on release approval. Apache-2.0. Without repository access, ask for beta access.
1. Install
Python 3.10 or newer and uv. It gives you two commands, decosa and decosa-mcp.
# The Decosa agent is in early access: ask for it at https://decosa.ai/contact?topic=self-hostThen save your dk_ key once. It asks for the key without echoing it, stores it with mode 600 and makes one test call, which returns a receipt.
decosa setup2. Run the Decosa agent
- On your own GPUYour vLLM (set up with the code kit at decosa.ai/apps/code), started with --enable-auto-tool-choice --tool-call-parser qwen3_coder. Nothing leaves your machine.
decosa agent --local http://127.0.0.1:8000/v1 - Hosted on Decosa (receipted)Interactive; add -p "<task>" for one task. Don't send proprietary code to the hosted route. Needs Node 22.19+ (it fetches pi, pinned, the first time).
decosa agent - Spend and receipts
decosa usage && decosa verify <receipt-id>
3. Or add Decosa to the agent you already use
- Claude CodeReads the key that decosa setup saved; or add -e DECOSA_API_KEY=dk_... before the --.
claude mcp add decosa -- decosa-mcp - Codex CLI
codex mcp add decosa -- decosa-mcp - Cursor (.cursor/mcp.json)
{ "mcpServers": { "decosa": { "command": "decosa-mcp" } } } - Any MCP client (stdio)
{ "mcpServers": { "decosa": { "type": "stdio", "command": "decosa-mcp", "env": { "DECOSA_API_KEY": "dk_..." } } } } - Any OpenAI-compatible harness (opencode, Continue, Qwen Code)Native tool calling; every response carries a receipt id (x-decosa-receipt).
base URL https://api.decosa.ai/v1 · model qwen3.8-27b · bearer dk_... - Without MCP: the CLI
decosa run grounding --sample outreach-email && decosa verify <receipt-id>
The server's instructions double as a skill file for Claude Code. For a local server without a tool parser, decosa proxy converts tool calls to the model's text format and back.
4. Ask it
- “Check this support answer against these two docs with Decosa, then verify one of the receipts.”
- “Can an open model take over my ticket-classification prompt? Which Decosa tool, and what does a run cost?”
- “Can I self-host Decosa's coding model on an RTX 4090?” (No: the pinned model needs a Blackwell card or 40 GB.)
The tools it adds
- decosa_list_use_cases
- Find the right tool for a job, with its measured cost and time per run
- decosa_use_case
- Its routes, the working example call, samples, limits and request shape
- decosa_run_use_case
- Run it on a sample or your input; saves the full output and returns the receipt ids
- decosa_get_receipt / decosa_verify_receipt
- Fetch a receipt and verify it offline against the pinned signing key
- decosa_verify_record
- Check a signed record (the hash-chained report most tools return)
- decosa_verify_confidential
- Check a confidential run's hardware evidence offline
- decosa_selfhost_plan
- Pinned images and models, whether your GPU fits (by model and architecture), and a smoke test
- decosa_docs
- Search the API contract: keys, errors, limits, receipts
How well it works
- 93%
- of 46 Decosa tasks done right by Qwen3.8 with these tools (100% with thinking on)
- 46%
- the same model as a generic agent with our docs and a raw HTTP tool
- $0.007
- per task at list price, a quarter of the model calls a generic agent makes
Our own eval, 46 tasks written for it, 28 Sep 2026: not a public benchmark. In Claude Code with the MCP server, 4 of 5 hand-run tasks were right the first time. The fifth (can an RTX 4090 run the coding model?) was wrong, and the tool now checks the GPU architecture. With no tools at all, the same model gets 22%.
Cost: the tools are free. Tool runs are billed per run, as on each tool's page. When the agent's own model is Decosa's hosted Qwen, it used about $0.007 per Decosa task at list price in the eval.