LLM Hangar / Private LLM hosting

Private LLM hosting in your own AWS, Nebius, RunPod or Verda account

Hosting and deployment infrastructure for open models. Not the "Private LLM" iOS and Mac app.

LLM Hangar deploys an open model on a GPU instance in your cloud account and chosen region. You control the resources and provider relationship; we handle deployment through the access you grant and record the actions in your audit log.

What "private" means here

With the default node-local gateway, requests travel directly between your client and the endpoint on your instance over HTTPS. A private edge also runs in your account. If you choose the hosted edge, prompts and responses pass through a gateway we operate. See gateway placement and access and the sub-processor list.

The model weights and serving process run in your cloud account. Your own client, logging and external tools also affect where data is sent or retained.

Each deployment supports named keys with individual limits. Keys are shown once when created. Key changes reach an edge gateway within seconds; on the default gateway, they take effect at the next start.

The security page explains cloud access, credential storage and teardown checks.

What you get

What it costs

Two bills. LLM Hangar is a flat $39 a month on the Lab plan (€34 in EUR), with a 7-day free trial that requires no credit card. Your cloud provider bills you directly for the GPU hours, at your own provider's rates, with your own credits and reserved capacity if you have them. We never see that invoice.

Two reference points for the GPU side, both cited rather than estimated:

The wizard shows the exact hourly and monthly estimate for the shape and region you pick before you confirm anything, and a cost calculator lets you compare that against a per-token API at your own volume.

Who it is for

Limits to consider

Frequently asked questions

Is this the Private LLM app?

No. Private LLM is an iOS and Mac app that runs small models on your device. LLM Hangar is hosting infrastructure: it deploys open models onto GPU instances inside your own AWS, Nebius or RunPod account and gives your applications a private, key-authenticated, OpenAI-compatible endpoint.

Do my prompts ever reach LLM Hangar?

With the default gateway, requests go directly to your instance. The optional hosted edge proxies requests through infrastructure we operate. Gateway placement and any audit-capture settings determine the request path and retention; review these before sending sensitive data.

Can I keep the infrastructure if I cancel?

Yes. The instance, disk, network and endpoint are resources in your own account, tagged and owned by you. They stay after you cancel; you can manage them in your own console. Access for LLM Hangar is granted through a role or key you can revoke at any time.

Which providers and regions?

AWS, Nebius, RunPod and Verda. Regions come from your linked provider. The EU-only option pins every resource of a deployment to EU member-state regions, for example eu-central-1, eu-west-1, eu-west-3, eu-north-1, eu-south-1 and eu-south-2 on AWS.

Start a 7-day free trial