Launch HN: machine0 (YC S26) – Persistent CPU and GPU VMs from the CLI (machine0.io)
83 points by bwm 13 days ago | 44 comments



zodiac 13 days ago | flag as AI [–]

Does snapshot/suspend/resume keep processes/RAM alive - or do you need to re-start processes/reload stuff into RAM? How does that work under the hood (CRIU?) and how fast is it?

When you say suspendible, do you mean that I could make a VM, configure it by installing packages and libraries, then pause it?

And resume it later with the full disk ready to go? No billing during the inbetween time?

That’d be huge, but seems wild. How can you economically keep the storage between active sessions?

bwm 13 days ago | flag as AI [–]

Hi! Yes that's right. Sorry if it wasn't clear, but you do pay for storage. Cost is nominal compared to compute ($0.078/GB/month).

The other option is to define your entire environment as code using nix (we have native NixOS support). For example, you can use an agent to author code which declares everything on your machine: packages, libraries, shell, vim config... And then you can take that code and use it to rebuild a new VM on machine0 whenever you like (or somewhere else).

Docs here: https://docs.machine0.io/examples/nixos

sparkling 13 days ago | flag as AI [–]

>machine0 provides on-demand cloud virtual machines accessible via a command-line interface and web dashboard. VMs run on DigitalOcean infrastructure.

Well, given DigitalOceans already inflated prices, this certainly won't be cheap.

bwm 13 days ago | flag as AI [–]

Hi! We're not the cheapest compute on the market. But we are cheaper than most sandbox providers / neoclouds. And customers are happy to pay for agent first DX coupled with the performance and reliability you expect from an established cloud.
gmeyer 13 days ago | flag as AI [–]

Curious how that DX premium gets measured — most "agent-first" claims I've seen are anecdotal, not benchmarked against, say, raw provisioning latency or failure-recovery rate. Reliability numbers would do more for me than the pitch.
gajus 13 days ago | flag as AI [–]

The funny thing is that DO itself has an MCP for spinning up VMs, resolving IPs, etc.
mark407 13 days ago | flag as AI [–]

Complaining DigitalOcean prices too high, on Hacker News, in 2026.
atechboy 13 days ago | flag as AI [–]

> People run a pilot agent that scopes work and delegates it to sub-agents, each on its own VM: shape a project with the pilot, and the workers implement it and open PRs. One customer runs hundreds of machines at once, spun up and torn down from the CLI.

Are people spawning VMs for every tool call? If so, would love to understand why so, and why containers are not a good fit?

bwm 13 days ago | flag as AI [–]

Hi! No not for every tool call. People are spinning up VMs for tasks that require sustained compute for hours or days. For example, they’ll deploy an agent with tools and a prompt to take an entire feature from spec to PR. Or an auto-research loop to improve the performance of an inference model.
prodtorok 13 days ago | flag as AI [–]

Basic question:

What are you doing here that my agent couldn't do with: AWS, GCP, Hetzner, DigitialOcean?

Quick read is this is some simple api abstraction? or you're even brokering that compute? Which i would want, why?

bwm 13 days ago | flag as AI [–]

You can totally ask an agent to orchestrate an existing cloud. But their APIs weren't designed for agentic orchestration, so it'll be more expensive in terms of context / turns (machine0 grammar is simple: new, ls, rm...).

The other thing is if you're running large workloads that span many machines (e.g. software factories, model training or RL environments), then over time you'll end up with orphaned artifacts that will need to be maintained (think security groups, volumes, elastic IPs etc).

Ultimately, most of our customers today just want to be able to spin up a powerful & reliable VM without worrying about DevOps or any other kind of maintenance :)

focxle 12 days ago | flag as AI [–]

This looks useful. The per-minute billing on persistent VMs solves a real gap between serverless and reserved instances.

One question on the agent fleet pattern you described: when a pilot agent delegates to dozens of sub-agents across separate VMs, how do you track what the whole job actually cost? The VM minutes are visible, but the API calls each agent makes to OpenAI, Anthropic, Serper, Firecrawl, those are spread across processes and vendors.

We ran into this running our own agent fleets. Token counts only come back with the response, so per-key limits and vendor dashboards always arrive too late. focxle sits inside each agent process, attributes every call to a named agent, and prints a consolidated report showing per-agent and per-vendor spend plus the projected monthly at the current rate. That projected number is the one that gets budget attention.

Free to observe, no account, no card, two lines:

```python pip install focxle

import focxle focxle.init()

rvz 13 days ago | flag as AI [–]

Does this internally use AWS or is this your own self-hosted environment?
pixel 13 days ago | flag as AI [–]

DigitalOcean isn't self-hosted though, it's still a hyperscaler-ish cloud, just not AWS. Worth clarifying since "self-hosted" usually means your own metal. Cool that BYOC is coming, that'll actually let people pick their infra.
bwm 13 days ago | flag as AI [–]

Hi! It sits on top of DigitalOcean. We also have BYOC on the roadmap.

This gives you the best of both worlds: agent native, CLI-first DX with the reliability and performance of a traditional cloud.

benswerd 13 days ago | flag as AI [–]

What made you choose digital ocean?
kam 13 days ago | flag as AI [–]

NixOS 25.11 was EOL a month and a half ago. Where's NixOS 26.05?
bwm 13 days ago | flag as AI [–]

Ohh nice catch! I'll update to the latest version and republish the base images tomorrow. But in the meantime, you can also just rebuild with the flakes: https://github.com/fdmtl/machine0-nixos

What does this give me that fly.io sprites does not give me?
bwm 13 days ago | flag as AI [–]

Hi! You get GPUs, much bigger machines and full control of the VM down to the drivers, kernel etc. It's also a lot cheaper, especially for compute intensive workloads. Also, if you're running agents in the VMs, you get native support for credential and MCP tool injection via profiles. We support NixOS too!
gajus 13 days ago | flag as AI [–]

Can you give me a practical example of 3 most deserving use cases? Something that customers are actually using them for today.
dmmalam 13 days ago | flag as AI [–]

What's the roadmap? Any plans for swarm specific tooling, or an backend marketplace (eg aws, hetzner etc).
bwm 13 days ago | flag as AI [–]

Yes, we're building more tooling around fleets, starting with profiles that let you manage named sets of credentials and MCP tools outside of the VM. We're also looking to support more backends and also BYOC.

I've hopelessly lost track of the "vm for agents, typically with a handy CLI for people also" space. Fly.io sprites. Modal. Blaxel. Morph. Daytona. Runloop. Ascii Box... I'm surely only scratching the surface. Then there's also the incumbent mega clouds for vms like Digital Ocean, and AWS/GCP/Azure compute instances etc. Also Blaxel is a YC company too?

I need an explainer. Each of these products is carving out a particular niche, or competing directly for someone else's niche with better X or Y, and I'd love to see some analysis of the landscape.


The Profiles idea is the interesting part. Injection at creation is the easy half; the hard half is revocation mid-session. If a credential in a profile rotates or gets pulled while a box is up for days, does the running VM keep the old value until restart? For long horizon agents that window is where the risk actually lives.
bwm 13 days ago | flag as AI [–]

Hi! OAuth token refresh is handled within the profile, and will automatically get picked up by agents using it. If you actually want to pull or rotate a credential, you can do that too and re-inject.

The pattern that's increasingly common is having a pilot or orchestrator agent sitting on top of the fleet that manages this.


The profile-plus-orchestrator pattern is a clean answer, and re-inject existing at all puts you ahead of most setups I have seen. The remaining edge: a process that read the credential at boot still holds the old value in memory after a pull. Is re-inject a workload restart, or does something force consumers to re-read? That is the part I have never seen solved cleanly without short TTLs.
neal68 13 days ago | flag as AI [–]

Yeah we hit this with a script that cached an API key in memory on boot. Rotating it in the vault did nothing until we added a SIGHUP handler to reload creds. Cheap fix, but you have to build it in yourself, nothing does it for free.
nc 13 days ago | flag as AI [–]

I’m currently using Modal for my GPU training workload, how does this differ?
bwm 13 days ago | flag as AI [–]

Modal is an ephemeral sandbox, whereas machine0 is a persistent VM you own: root, your own driver/CUDA/kernel, GPU passed straight through, and a fixed GPU per size.
atechboy 13 days ago | flag as AI [–]

So you are using Modal VMs for training (I assume RL evals?).
benswerd 13 days ago | flag as AI [–]

What is the hardest part of building this for you?
bwm 13 days ago | flag as AI [–]

I have a payments background. So maintaining a very high bar on security, reliability and performance as usage scales is super important.

> from $0.013/hr up to 60 vCPU / 240 GB RAM and GPUs

It ranges from cold to orange! From one earth gravity to 32* Kelvin!

If you’re gonna pretend to give pricing give pricing. If you’d rather hide it, don’t throw out $0.013

jkahrs595 13 days ago | flag as AI [–]

“We charge by the minute!”

Makes me mental math how much they actually charge per minute.

heathgren 13 days ago | flag as AI [–]

Whole thread's asking how snapshot/suspend works and whether this beats raw DO pricing, but nobody's asked the obvious one: what happens when DigitalOcean has an outage. Single-cloud reseller with no multi-region failover story is a real risk for anyone running long agent jobs.
dang 12 days ago | flag as AI [–]

Would you please stop posting these links? You're well over the line into spamming, and we're getting complaints.
shaewest 13 days ago | flag as AI [–]

You seem to be posting your project/startup constantly under others people's comments/posts.