Declare the job, not the machine
Hand Nimbus a resource envelope and a deadline. It reads the shape of the workload and places it on the nearest healthy node with real headroom, then rebalances as demand shifts.
Rented boxes make you the scheduler. Nimbus is a self-scheduling mesh of GPU nodes: declare a training run or an inference endpoint, and it places you on live capacity in under a second.
No cluster to size. 100 GPU-hours on the house.
Three decisions we made so you never open a capacity ticket, resize a node pool, or babysit a spot fleet again.
Hand Nimbus a resource envelope and a deadline. It reads the shape of the workload and places it on the nearest healthy node with real headroom, then rebalances as demand shifts.
Snapshotted runtimes and a pool of pre-warmed nodes mean your endpoint answers the first request, not the fiftieth. Median cold start is 740 milliseconds, tail p99 under two seconds.
When a cloud reclaims hardware, Nimbus migrates your running workload to a fresh node before it drains. You pay interruptible rates and keep uninterrupted jobs, up to 71 percent off list.
Create an account, pipe in a container or a model ID, and Nimbus schedules the rest. Your first 100 GPU-hours are free, no card, no sales call.