NEW

GPU instances are live. Use code LOCAL10 for 10% off your first month.

Deploy now
manifesto

Your models.
Your machine.
Your rules.

Something strange happened to computing. We used to own our machines. Then we started renting our own thoughts back, by the token.

Every prompt you send to a hosted API is a small act of trust: trust that it will not be logged, not be trained on, not be subpoenaed, not be repriced next quarter. Trust is fine. Ownership is better.

Intelligence is becoming a utility. Utilities should run on hardware that answers to you.

The open model movement already did the hard part. Llama, DeepSeek, Qwen, Flux: weights you can download, inspect and keep forever. The only missing piece was a machine worth running them on. That is the piece we sell. Nothing more.

We do not host an AI. We do not operate models, inspect workloads or sit between you and your GPU. We rack good hardware, wire it to fast networks, hand you root, and get out of the way.

A flat price is a moral position.

Per-token billing puts a taxi meter on curiosity. It makes you hesitate before the tenth question, batch your thoughts, ration your own tools. A machine with a flat price does the opposite: the more you use it, the smarter the deal gets. We think that is the right direction for the incentive to point.

So this is the whole company: dedicated servers, honest specs, one number a month, and the strong opinion that the most important technology of this decade should be something you can own.

Rent the machine. Own everything else.
Locally · locally.host