Does snapshot/suspend/resume keep processes/RAM alive - or do you need to re-start processes/reload stuff into RAM? How does that work under the hood (CRIU?) and how fast is it?
> People run a pilot agent that scopes work and delegates it to sub-agents, each on its own VM: shape a project with the pilot, and the workers implement it and open PRs. One customer runs hundreds of machines at once, spun up and torn down from the CLI.
Are people spawning VMs for every tool call? If so, would love to understand why so, and why containers are not a good fit?
Hi! No not for every tool call. People are spinning up VMs for tasks that require sustained compute for hours or days. For example, they’ll deploy an agent with tools and a prompt to take an entire feature from spec to PR. Or an auto-research loop to improve the performance of an inference model.
Ohh nice catch! I'll update to the latest version and republish the base images tomorrow. But in the meantime, you can also just rebuild with the flakes: https://github.com/fdmtl/machine0-nixos
Hi! Yes that's right. Sorry if it wasn't clear, but you do pay for storage. Cost is nominal compared to compute ($0.078/GB/month).
The other option is to define your entire environment as code using nix (we have native NixOS support). For example, you can use an agent to author code which declares everything on your machine: packages, libraries, shell, vim config... And then you can take that code and use it to rebuild a new VM on machine0 whenever you like (or somewhere else).
You can totally ask an agent to orchestrate an existing cloud. But their APIs weren't designed for agentic orchestration, so it'll be more expensive in terms of context / turns (machine0 grammar is simple: new, ls, rm...).
The other thing is if you're running large workloads that span many machines (e.g. software factories, model training or RL environments), then over time you'll end up with orphaned artifacts that will need to be maintained (think security groups, volumes, elastic IPs etc).
Ultimately, most of our customers today just want to be able to spin up a powerful & reliable VM without worrying about DevOps or any other kind of maintenance :)
But, I don’t think this is your strongest argument. The APIs for those providers are pretty easy to orchestrate and don’t take that many tokens to use. (Especially if you are hosting on top of one of these providers)
Instead, I think some strengths you could focus on are (a) not being one of those providers, (b) having a better product mix that people want to use, and (c) keeping a minimal design. Clear use-cases, minimal friction, easy to keep the model in your head.
You definitely have a good product here with plenty of reasons to choose you. But your minimal API isn’t a great moat.
OpenRouter is valuable because it has a large catalogue of competing providers for a fungible service. This does the opposite (lock in with a specific vendor).
Hi! You get GPUs, much bigger machines and full control of the VM down to the drivers, kernel etc. It's also a lot cheaper, especially for compute intensive workloads. Also, if you're running agents in the VMs, you get native support for credential and MCP tool injection via profiles. We support NixOS too!
Hi! Sure: (1) running agent fleets for software factories, (2) Model training and RL environments orchestrated by agents and (3) as a backend for agentic products and platforms.
Yes, we're building more tooling around fleets, starting with profiles that let you manage named sets of credentials and MCP tools outside of the VM. We're also looking to support more backends and also BYOC.
I've hopelessly lost track of the "vm for agents, typically with a handy CLI for people also" space. Fly.io sprites. Modal. Blaxel. Morph. Daytona. Runloop. Ascii Box... I'm surely only scratching the surface. Then there's also the incumbent mega clouds for vms like Digital Ocean, and AWS/GCP/Azure compute instances etc. Also Blaxel is a YC company too?
I need an explainer. Each of these products is carving out a particular niche, or competing directly for someone else's niche with better X or Y, and I'd love to see some analysis of the landscape.
The Profiles idea is the interesting part. Injection at creation is the easy half; the hard half is revocation mid-session. If a credential in a profile rotates or gets pulled while a box is up for days, does the running VM keep the old value until restart? For long horizon agents that window is where the risk actually lives.
Hi! OAuth token refresh is handled within the profile, and will automatically get picked up by agents using it. If you actually want to pull or rotate a credential, you can do that too and re-inject.
The pattern that's increasingly common is having a pilot or orchestrator agent sitting on top of the fleet that manages this.
Modal is an ephemeral sandbox, whereas machine0 is a persistent VM you own: root, your own driver/CUDA/kernel, GPU passed straight through, and a fixed GPU per size.
Hi! We're not the cheapest compute on the market. But we are cheaper than most sandbox providers / neoclouds. And customers are happy to pay for agent first DX coupled with the performance and reliability you expect from an established cloud.
Are people spawning VMs for every tool call? If so, would love to understand why so, and why containers are not a good fit?
And resume it later with the full disk ready to go? No billing during the inbetween time?
That’d be huge, but seems wild. How can you economically keep the storage between active sessions?
The other option is to define your entire environment as code using nix (we have native NixOS support). For example, you can use an agent to author code which declares everything on your machine: packages, libraries, shell, vim config... And then you can take that code and use it to rebuild a new VM on machine0 whenever you like (or somewhere else).
Docs here: https://docs.machine0.io/examples/nixos
What are you doing here that my agent couldn't do with: AWS, GCP, Hetzner, DigitialOcean?
Quick read is this is some simple api abstraction? or you're even brokering that compute? Which i would want, why?
The other thing is if you're running large workloads that span many machines (e.g. software factories, model training or RL environments), then over time you'll end up with orphaned artifacts that will need to be maintained (think security groups, volumes, elastic IPs etc).
Ultimately, most of our customers today just want to be able to spin up a powerful & reliable VM without worrying about DevOps or any other kind of maintenance :)
But, I don’t think this is your strongest argument. The APIs for those providers are pretty easy to orchestrate and don’t take that many tokens to use. (Especially if you are hosting on top of one of these providers)
Instead, I think some strengths you could focus on are (a) not being one of those providers, (b) having a better product mix that people want to use, and (c) keeping a minimal design. Clear use-cases, minimal friction, easy to keep the model in your head.
You definitely have a good product here with plenty of reasons to choose you. But your minimal API isn’t a great moat.
[1]: https://news.ycombinator.com/item?id=49323381
I need an explainer. Each of these products is carving out a particular niche, or competing directly for someone else's niche with better X or Y, and I'd love to see some analysis of the landscape.
The pattern that's increasingly common is having a pilot or orchestrator agent sitting on top of the fleet that manages this.
It ranges from cold to orange! From one earth gravity to 32* Kelvin!
If you’re gonna pretend to give pricing give pricing. If you’d rather hide it, don’t throw out $0.013
Makes me mental math how much they actually charge per minute.
This gives you the best of both worlds: agent native, CLI-first DX with the reliability and performance of a traditional cloud.
Well, given DigitalOceans already inflated prices, this certainly won't be cheap.