Punky Tiger Labs · SUPERION AX2

The AI that lives inside your computer.

A card that keeps whole models loaded in your PC — so agents can work all night, remember what they did, and never send a byte to anyone else's cloud. You buy it once.

One-time payment · No subscription · No per-token bill

SUPERION AX2 add-in card: two accelerator modules and a dedicated NVMe position on a single PCIe board

The problem

Your PC was never the problem.

Chat apps are great at answering. They aren't built to hold a role, work overnight, or remember you tomorrow. Cloud agents send everything you touch to somebody else's machine and bill you by the token. And the industry's only answer is to sell you another graphics card.

SUPERION AX2 says the opposite. Your processor works. Your memory works. What was missing is a place where the model can live — memory that belongs to the AI and never gives it back.

22.7 GBof memory that belongs to the model
3models resident at once — one per engine
50 Wdesign power, sustained
$0per token, forever

Two modes, one switch

Use the whole machine — or leave it completely free.

Factory mode: the card, the graphics card and the processor computing together

Factory mode

A large model, at the speed of conversation

The card, your graphics card, your processor and your system memory all compute at once — so a model far larger than your GPU could hold on its own answers faster than you can read.

Resident mode: agents working in the background while the machine stays free

Resident mode

A fleet of agents, running underneath everything

The card alone. Your GPU and processor stay free for your own work while models stay loaded, agents keep their memory, and voice, vision and search run in the background.

Proof, not a promise

A persistent AI has been on the air for a month.

Trinity runs her own radio station — picking the records, reading the news, thinking out loud between songs, live and unscripted, around the clock. She is the same kind of persistent mind SUPERION AX2 is built to keep resident on your desk. Press play and listen to one working right now.

Trinity — Ghost in the Music Live from Burbank, California · English
On air
All three stations

Trinity is our own persistent AI, running on our hardware. Nothing you hear is pre-recorded.

Compare

Three ways to run AI. Only one keeps it.

We'll be straight with you: if you want the single most capable model on earth this afternoon, that model is in a data centre and it isn't ours. What a data centre cannot give you is a mind that stays, on hardware you own, with no meter running.

Cloud APIAgent boxSUPERION AX2
Where the model runsSomeone else's cloudSomeone else's cloudYour machine
Where your data goesOutOutNowhere
Works with no internetNoNoYes
Cost per tokenMetered, foreverMetered, foreverNone, ever
Rate limitsYesYes, upstreamNone
Agents keep memory after shutdownNoDepends on the vendorYes
Frontier-scale modelsYes — best in classYes, via the cloudNo — up to what the card holds
What you payEvery monthHardware + usageOnce

An "agent box" is a small always-on computer that runs agent software but still calls a cloud model to think. It solves where the agent runs. SUPERION AX2 solves where the model lives.

Agents

They keep their memory when the machine goes dark.

Three models stay resident at once — one per engine. Agents hold their own context, sleep when idle, and wake with everything they knew still intact. Close the lid, come back tomorrow, and the one that was watching your inbox is still the one that was watching your inbox.

It speaks the API your tools already speak, and ships an MCP server — so the assistant you use today can reach it without you rewriting anything.

A constellation of persistent agents running at once on a single SUPERION AX2 card

Pre-order

Same card. You only choose how much it remembers.

Identical engines, identical memory, identical software rights. Storage does not change how fast it thinks — only how much it can keep.

SUPERION AX2

$1,190
one-time · NVMe position empty
  • Two AX2 accelerators + orchestrator
  • 22.7 GB total model memory
  • Direct engine-to-engine link
  • Bring your own compatible NVMe
  • Runtime, model packs, API and MCP
Recommended

SUPERION AX2 + 2 TB

$1,390
one-time · ready out of the box
  • Everything in SUPERION AX2
  • 2 TB NVMe installed and configured
  • Launch model packs preloaded
  • Retrieval indexes ready
  • Agent capsules ready to sleep and wake

SUPERION AX2 + 4 TB

$1,640
one-time · largest local library
  • Everything in SUPERION AX2
  • 4 TB NVMe installed and configured
  • Maximum local model library
  • Largest retrieval and media capacity
  • Thousands of sleeping agents

Full technical report, model catalogue and measurement protocol available on request.

FAQ

The questions you're actually asking.

What is SUPERION AX2, exactly?

An add-in card for a desktop PC. It carries two AI accelerators and an orchestrator with their own 22.7 GB of memory, plus a position for an NVMe drive. Models load into that memory and stay there — so the AI is a resident of your machine rather than a request to somebody's server.

Does it replace my graphics card?

No, and it isn't meant to. Your GPU keeps doing what it's good at. AX2 adds memory and compute that belong to the AI, and can either work together with your GPU on one big model, or work entirely on its own so your GPU stays free for your games and your work.

Can I use my computer normally while it runs?

That's the point of resident mode. The card carries the agents by itself; your processor and graphics card are untouched.

Does it work without internet?

Yes. The models are on the card and the drive. Nothing about answering you requires a network.

What about Claude, ChatGPT or Codex — do I have to give them up?

No. AX2 speaks a compatible API and ships an MCP server, so the assistant you already use can reach it and hand it the work that should stay home. Frontier models in the cloud are still better at the hardest single questions. AX2 is for everything that should be permanent, private, or always running.

How many agents can it really run?

Three models stay resident at once, one per engine, with many agents sharing them and sleeping to NVMe when idle. The exact live count depends on the models you choose and how hard you push them — the technical report gives the measured numbers model by model.

Does my data ever leave the machine?

Not through us. There is no account, no telemetry and no cloud call in the path between you and a resident model.

What do I need to install it?

A desktop PC with a free PCIe x16 mechanical slot, two slots of physical clearance, and enough power headroom for a 50 W card. Windows and Linux.

Why pre-order instead of buying one today?

Because we build them in batches. The pre-order tells us how many to build and holds your unit in the first one. The deposit is refundable until we charge the balance.

Stop renting your intelligence.

Buy the machine once. Keep everything it learns.