Need it air-gapped? Qadenza can run entirely inside your own network. See how →
closed beta

Private doesn't mean out of reach.

Run AI on hardware you actually own, and reach it as easily as any cloud app. No token limits. No one else's servers. Just your machine, wherever you are.

private by default no token limits macOS · Linux · Windows
You're on the list. We'll email you when your invite is ready.
Enter a valid email to join the list.
Connection paths 3 devices · 2 online
studio
M2 Ultra · 192 GB
direct · 12ms
rtx-box
2× RTX 4090
relay
relayed · 41ms
homelab
RX 7900 XTX
down · 6m ago

Feels like a cloud app. Runs on your machine.

Everything you'd expect from a hosted assistant, without handing your data to one.

01

Chat from anywhere

Open a conversation with whichever model is running on your hardware, from your laptop, your phone, or someone else's computer.

Qadenza finds the fastest path to your machine automatically, so it feels like a hosted app instead of a home network project.

Qadenza chat: a model running on a Mac explains how to build and run a program, with copyable code blocks
02

Models that fit, no token limits

Browse models and see whether each one fits your hardware before you download it. Install with one click, on Ollama or llama.cpp.

Got more than one GPU? Split a model across them. There's no meter running: use it as much as your hardware can handle.

Qadenza models page: memory in use, the loaded model, and whether each installed model fits right now
03

Use it where you already work

The VS Code extension turns your own models into a coding agent: it reads and searches your code, proposes changes, and shows every edit as a diff for you to approve.

Qadenza also speaks MCP and the OpenAI API, so Claude Code, Cursor and anything with a custom base URL can use the models on your machine.

$ code --install-extension Monolyth.qadenza Sign in → approve the code in your dashboard → pick a model

Your files are read and edited locally. Only the conversation goes to your model.

04

Everything else, in one place

Pull models, watch temps and load, generate images, and see what's connected across every machine you own. Agents update themselves.

One dashboard for all of it, whether you're checking in from your desk or from across the country.

Qadenza hardware page: CPU, RAM, disk, GPU temperature and power, with an hour of usage history for four GPUs
air-gapped

No internet? No problem.

For teams that can't let data leave the building, we deploy the whole Qadenza stack inside your network: agents, dashboard and connection layer. Models load from your own storage, and nothing depends on reaching us.

Talk to us about air-gapped

qadenzabeta@mlyth.org

  • Runs on isolated networks with no outbound connection
  • Models and updates delivered offline, on your schedule
  • Same chat, model manager and VS Code extension your team gets online
  • Your hardware, your network, your data. Nothing leaves.

One line to connect

Run one command on any machine you want to reach later. It sets up the service, offers to install the engines, and pairs the machine to your account. Beta invite required.

$ curl -fsSL https://gitlab.com/api/v4/projects/86433000/packages/generic/qadenza-agent/latest/install-macos.sh | sh

Join the closed beta

Leave your email and we'll send an invite when a spot opens.

You're on the list. We'll email you when your invite is ready.
Enter a valid email to join the list.

Request a beta invite

We invite people in small batches and email you when it's your turn.

Which machines would run Qadenza?