Racks you can hold by the hour
Neocloud for AI teams. Real H100 racks in us-west, live telemetry, hours you can hold.
- cluster
- C1
- region
- us-west
- gpu
- 8x H100 SXM
- util
- --
- jobs
- --
- uptime 30 d
- --
One rack, one feed. Compute, scheduler, telemetry, documentation. The public cloud is built this way. RACK is, too. Hours you can hold.
util
0 %
jobs
0
gpus
8
- gpu0--
- gpu1--
- gpu2--
- gpu3--
- gpu4--
- gpu5--
- gpu6--
- gpu7--
Preview data, 1 Hz
Run training, inference and batch workloads with the telemetry, scheduling and hours you can hold of a cloud
Cluster C1 ▸
GPU
gpu0
gpu1
gpu2
- 8x H100 SXM 80 GB, one NVLink fabric
- 3.2 Tb/s InfiniBand to shared storage
- Owned rack, power and network
Scheduler ▸
| Job | GPU | State |
|---|---|---|
| job-01k4 | gpu2 | RUNNING |
| job-01k3 | gpu5 | DONE |
| job-01k2 | gpu0 | DONE |
- Reserved hours held for your team
- On demand billed per minute
- Batch fills idle capacity, preempted first
Telemetry ▸
- 1 Hz per GPU: util, memory, temp, power
- 24 h history on every page
- Same feed site, console and reports
Documentation ▸
- ▸ Quickstartguide
- Run a jobguide
- Schedulerarchitecture
- SLAarchitecture
- APIfor-labs
- Quickstart curl and Python, real endpoints
- SLA written before general availability
- Reports capacity, monthly
For years, AI teams had two options for GPU capacity. Both required compromise.
Hyperscaler on demand
Speed, but at a cost
`+--------+`.+--------+~> STATUS ┌─────┐ERROR ──────┘ └CAP-429-WAIT`+-------`+-------`.`. `. | `. |+--------+--------+ || | | | || | | | || | `+ | `+| |`. |`.+--------+--------+--`.`. `. | `. |+--------+--------+ || | | | || | | | || | `+ | `+| |`. |`.+--------+--------+--`.`. `. | `. |+--------+--------+ || | | | || | | | || | `+ | `+
- Waitlists for H100 capacity, with no date
- Opaque telemetry you read a bill, not a rack
- Spot revoked with two minutes notice
- Egress priced to keep you
Owning a rack
Control, but at a cost
`+------------`.`. `. |`. `. |+-------------+ || | || | `+| RACK | `.| |`.+-------------+----`.`. `. |`. `. |+-------------+ `+| | `.| NETWORK |`.+-------------+----`.`. `. |`. `. |+-------------+ `+| | `.| POWER |`.+-------------+----`.`. `. |`. `. |+-------------+ `~> READY| | `.| COLO |`. 18 mo+-------------+
- 18 months from order to first job
- Colo, power, network three contracts, three vendors
- Idle hours paid for whether you train or not
- No exit when the workload moves
For the first time, the speed of a cloud meets hours you can hold.
Real racks
A cluster is a rack we own: eight GPUs, one fabric, one power budget. Soak tested for 72 hours before it is listed.
Predictable prices
Three plans, one number each. No egress, no surprise line items.
API driven from the first hour
One endpoint to submit a job, one to read its telemetry. The same calls the console makes.
Two business days to an answer
Access is by invitation during preview. Tell us what you run, we reply within two business days.
Pricing. Per GPU hour, in USD.
On demand ▸
Billed per minute with a one minute minimum. No commitment, capacity as it is free on C1.
Job
train-1
Job
eval-1
Reserved ▸
One month or more. Hours are held for your team and nobody else runs on them.
- MONTH_1720h
- MONTH_2720h
- MONTH_3720h
- SPOT0h
Batch ▸
Preemptible. Runs when reserved capacity is idle, and is the first to give way.
Job
batch-embed
plan = "batch" preempt = true price = 1.19 fills = "idle reserved"
Compared to list prices at the four largest providers, updated monthly.
Regions. us-west online, eu-central next.
Drag to rotate
Regions
- online
us-west
Hillsboro, Oregon
- next
eu-central
Frankfurt
Median round trip to us-west
- San Francisco14 ms
- New York68 ms
- Frankfurt146 ms
Capacity by region
| region | clusters | gpus | util |
|---|---|---|---|
| us-west | 1 | 8 | -- |
| eu-central | 0 | 0 | planned |
Request access. Write the request, we reply within two business days.
cluster = "c1" region = "us-west" email = "" team = "" size = "" workload = "" hours = "" plan = ""
Pick a workload and hours
- 01Review we read the request and check capacity on C1
- 02Invite console access and an API key
- 03First job a reserved window, same afternoon