MoyaiOpen source cloud agent
HermesWorks with Claude Code and CodexSelf-hosted · 100+ providers through LiteLLM
- [1]The problem
- [2]The results: 79% cheaper
- [3]Why we're open sourcing it
- [4]A cloud agent that keeps working
- [5]Any harness
- [6]Any model, any provider
- [7]Get started
Today we're open sourcing Moyai, the self-hosted cloud coding agent our team runs every day. You give it a task in Slack or the browser, and it opens a pull request while your laptop is closed. It runs Claude Code or Codex on any of the 100+ providers LiteLLM supports
The problem
Our Devin bill hit $101,872 in a single month, and only our own team used it. Our engineers kicked off those sessions and automations on models and routing we couldn't control

We already run a gateway that routes across 100+ providers. We wanted our coding agent on our own models and routing, paying for inference instead of seats
The results: 79% cheaper
Moyai does the same work for about $700 a day. Over those 31 days we'd have paid about $21,700 instead of $101,872, and kept $80,000 of one month's bill
Why we're open sourcing it
Last week we wrote about how we built our own internal Devin in 2 days, and most replies asked to run it themselves. Moyai is the same code we run in production at LiteLLM. You deploy it on your own infrastructure and point it at your own LiteLLM gateway, so your code and credentials stay in your accounts
A cloud agent that keeps working
Each session gets its own cloud workspace with a terminal, a filesystem and a browser. The agent edits code and runs your tests there, then opens a pull request for you to review
You can follow along in the web app and send a correction while it works, or pick the thread back up in Slack the next morning. For large tasks, the agent splits the work across parallel workers on separate machines and collects their results
Any harness
You pick the agent harness for each session from the composer. Moyai runs it in the same isolated workspace with the same tools and permissions
HermesClaude Code, Codex, OpenCode and Deep Agents run through the LiteLLM agent SDK. Adding another harness takes one registry entry
Any model, any provider
Moyai sends model requests through LiteLLM. You can switch from GPT-6 Astra to Claude Opus 5.5 between messages, and LiteLLM attributes each request to the teammate who made it. Provider keys stay on the server, out of the sandbox
LiteLLM supports 100+ providers, and Moyai can use any of them
Get started
You can try the local demo in a couple of minutes without API keys
git clone https://github.com/BerriAI/moyai.git
cd moyai
cp .env.example .env
uv sync --frozen
uv run uvicorn app.main:app --host 127.0.0.1 --port 8787 --workers 1
To run real tasks, set up cloud execution with Modal and your LiteLLM gateway, then connect your apps. Issues and PRs are welcome on GitHub


