steading.sh
Run serious AI models, privately,
for cents an hour.
You need AI that doesn't send your data to anyone — for research,
for sensitive documents, for work that can't leave your hands. Your laptop can't run the good
models, and cloud AI platforms charge monthly whether you use them or not.
Steading fixes this. It rents the right GPU from the cheapest market,
installs the fastest serving stack for that exact card, connects the model to your tools —
and shuts everything down the moment you're done. You pay for hours, not months.
What that looks like
One command. Cost shown before anything is rented. Auto-shutoff so a forgotten
session can never become a bill.
$ steading up qwen3.5-27b
host RTX 6000Ada · 49GB — cheapest card that fits, of 64 offers
cost $0.52/hr · hard cap $2.07 · auto-off after 30 min idle
ready ~9 min
● ready your private endpoint: http://127.0.0.1:57778/v1
$ steading down
machine wiped and returned — total: $0.08
What we do for you
Pick the hardwareWe shop the GPU spot market live and rank
offers by real cost — including download time you'd otherwise pay for.
Tune the enginevLLM, SGLang, llama.cpp, MLX — we choose and
configure the fastest one for your model and card, from measurements, not guesswork.
Guard your walletCost preview before renting, a hard spend cap,
idle auto-shutoff, and verified teardown. Surprise bills are structurally impossible.
Keep it privateThe model runs on a machine only you rent,
reached through an encrypted tunnel only you hold. Prompts are never logged, stored, or
shared — there's no one to share them with.
The economics
A 27-billion-parameter model: $0.52/hour.
An evening of experiments: about $2.
A month of daily research use: $15–40 — only the hours you actually run.
Managed AI clouds charge 2.5–4× for the same GPU. Subscriptions charge
you every month either way. We add 10% to the market price of the card, and that's it.
Who it's for
Researchers & academicsSensitive data, IRB constraints,
grant budgets. Run open models on hardware you control, spend exactly what the work
needs.
Anyone whose laptop hit the wallYou tried running models
locally and met the 8GB reality. Rent a 49GB card for the evening instead of buying a
$5,000 one.
Pricing
Do it yourself
$0
open-source tool, your own GPU-market account
The full tool, free forever. You pay the GPU market directly; Steading adds nothing.
Get notified at launch
Managed
market +10%
founding members: $20/mo, locked
No accounts to create, no market to learn. We carry the hardware relationships and the
tuning; you run one command and pay one bill.
Become a founding member
Rented market GPUs are right for open models and shareable-risk data. For data
that can't touch anyone's hardware but yours, we ship the same tuned stack as inspectable
packages for machines you own — talk to us.