// PERSONAL AGENT INFERENCE

Flat-rate inference for the personal agent you already run.

Connect your agent to LunaRoute and give it a private GLM-5.2 inference lane for $19.99/mo. Same API, no token meter, zero data retention.

inference for your existing agent — not another assistant app
fixed monthly price — no metering, no token anxiety
private by default — zero data retention
one request at a time — priced for patient background work; see how it works
LEFT CLAW
$19.99 FLAT. FOREVER.
RIGHT CLAW
ZERO DATA RETENTION
WORKS GREAT WITH THE AGENTS YOU ALREADY RUN —
OPENCLAW
NANOCLAW
HERMES
// 01 — THE PLAN

One inference plan for one personal agent.

Your agent brings the tasks. LunaRoute supplies the private GLM-5.2 inference. No tiers, no token calculator, no surprise bill — just one lane for one agent at a fixed price.

PERSONAL AGENT INFERENCE PLAN
$19.99/month
billed monthly · cancel anytime · this price is not "introductory"
flat-rate GLM-5.2 inference
bring your own agent or client
one lane — one request at a time
unlimited requests at the fixed price
zero data retention, contractually
replies may wait for quiet GPU capacity
A lane means one inference request can run at a time. Perfect for personal agents doing background work; multi-agent and household plans are coming later.
// GET NOTIFIED AT LAUNCH

Personal Agent Inference isn't open yet. Drop your email and we'll ping you the second it goes live — flat price, zero data retention, no surprises.

// 02 — HOW IT WORKS

Why personal-agent inference can be flat-rate.

Your agent calls LunaRoute like any OpenAI-compatible inference endpoint. The difference: Personal Agent requests are handled one at a time, and our scheduler slots each one into quiet moments on the GPU fleet, between realtime traffic.

That tradeoff makes the price simple. Your agent can summarize mail, plan your week, draft replies and clean up notes without watching a token meter. It may wait a few extra seconds to start — but the bill never moves.

your agent --> request
                     |
                     v
            +-------------------------+
realtime -->| LUNAROUTE SCHEDULER     |
 traffic    | metered programs        |
            | professional plans      |
            | personal plans .....>   |
            +-------------------------+

wait: usually seconds    extra cost: $0.00, always
// 03 — ZERO DATA RETENTION

It's your life in there.
We treat it that way.

Your agent may send deeply personal context — messages, calendar entries, notes, drafts, half-written apologies. That inference traffic runs under LunaRoute's private inference policy: zero data retention, processed in memory, not stored, not trained on and contractually enforced at every hop.

what we keep: timestamps and queue status — enough to run the scheduler
what we don't: bodies, prompts, outputs. anything that's you
no training, no analytics, no “anonymized” anything
// 04 — FAQ

Fair questions.

Give your agent a private inference lane.

$19.99/mo or $199.99/year. Flat-rate GLM-5.2 inference for the personal agent you already run. Zero retention. Cancel anytime.

GET THE PLAN →
LUNAROUTE © 2026 · personal agents welcome
no lobsters were logged in the making of this page