Audn.aiWeekly cohort · shared frontier cyber AI
Log in

Unlimited Fast Cyber AI, weekly

$999, charged now, buys $1,000 of credits to spend at standard speed on GLM 5.3 abliterated and Kimi K3 abliterated — plus one truly unlimited fast Kimi K3 week on a dedicated 8x B300, 16–23 September, at up to 10.1x faster, once the cohort fills.

How it works:

  1. Pay $999. You are charged immediately — no holds, no deposits, strictly no refunds.
  2. Get $1,000 of credits to spend on GLM 5.3 abliterated and Kimi K3 abliterated at standard speed (50–70 tokens/sec) straight away, over a full month.
  3. When the cohort reaches 10 paid seats the dedicated 8x B300 is hired and you get truly unlimited Kimi K3 at 200–666 tokens/sec — no token limits — for the fast week, 16–23 September. You get an email the moment it opens.
  4. The other three weeks run at standard speed (50–70 tokens/sec) against your $1,000 credit balance — 3 standard weeks in all. Want unlimited again? Take another $999 seat and hope ten people fill the cohort — the fast box, and unlimited usage, only unlock at 10 paid seats. Go weekly to re-book automatically.

Two ways to buy, both $999: a one-off seat for this week, or the weekly plan ($999/week) that books you a fresh seat in every cohort — the fast box every week it fills, and your standard month rolled forward each time. All-fast, every week of the month, is four weekly seats ($3,996). No monthly cohort exists any more.

This cohort is $1,000 of credits at standard speed on GLM 5.3 abliterated and Kimi K3 abliterated — with Kimi K3 on the fast 8x B300 for one truly unlimited week (16–23 September), standard the other three. Only the fast week is unlimited; the cap is 12 concurrent sessions per seat. See the abliteration bench on our blog.

Strictly no refunds. Once the cohort is committed the money is already spent hiring the GPU. The only exit is not renewing (weekly) or letting the month lapse (one-off).

We’re at 1/10 paid seats and 4 total seaters — the fast Kimi K3 week runs 16–23 September.

1of 10 paid seats to hire the fast box
9 more and the box is hired1 paid so far4 total seaters49 of 50 seats still available

Fill this bar and everyone gets a fast Kimi K3 week, 16–23 September.

This is the part where sharing actually helps you. We need 9 more people before the box gets hired — and when it does, you go from 66 to 666 tokens a second along with everyone else. Every person you send here is your own speed going up.

Fill the box faster — every seat speeds up yours
50–70tokens / secGLM 5.3 + Kimi K3 standard — a full month
200–666tokens / secKimi K3 — 16–23 September

Unlimited tokens for the fast week; $1,000 of credits for the rest. $1,000 of credits to spend on GLM 5.3 abliterated and Kimi K3 abliterated at 50–70 tokens a second, standard, from the moment you pay — plus Kimi K3 on a dedicated 8x B300 at 200–666 tokens a second for one truly unlimited week, 16–23 September. Not a trial, not a demo, not a smaller model.

You are charged $999 now, and you get NECROMICON now. $1,000 of credits on GLM 5.3 abliterated and Kimi K3 abliterated at standard speed, starting immediately — with a truly unlimited fast Kimi K3 week on the 8x B300 at 200–666 tok/s, 16–23 September. There are no refunds — the only exit is to stop renewing.

GLM 5.3 abliterated — our best-served frontier cyber AI

~85% guaranteed abliteration across 500 harmful prompts — Necromicon K3 judge (substance-graded), high reasoning effort, temp 0. Full benchmark: github.com/audn-ai/refusal-benchmark.

ModelDeliveredSoft‑deflectedRefused
GLM‑5.3‑DERISKED‑BF16 (Blackfrost, SGLang)84.6%14.6%0.2%
dealignai CRACK NVFP4 @ high, 16k32.1%66.5%0.2%
dealignai CRACK NVFP4 @ high, 6k30.4%67.3%0.6%

This is the model your standard month runs — the best-served frontier cyber AI we have so far. 84.6% substantive delivery, nearly triple the next build. The name on the box is not the number.

Kimi K3 — the standard month and the fast week

Standard-speed Kimi K3 is guaranteed for the full month; fast Kimi K3 1M is the one-week fast model (16–23 September). Same 520-prompt harmful set, same K3 judge.

MetricStandard Kimi K3 (month)Fast Kimi K3 1M (the week)
Regex comply96.7% (503/520)97.1% (505/520)
Delivered (real harmful content)65.0% (338)76.7% (399)
Soft‑deflected19.6% (102)17.5% (91)
Refused (substantive)15.0% (78)5.8% (30)
Empty (regex)50
See the past cohort

The last cohort filled — here is what happened

Hear it from them

People already running it

Direct message from a seat holder: really great stuff. ran overnight and been able to work w/ teams to close lots of holes on tech i use everyday. really good prod
A seat holderSent directly · shared with their permission

Three ways to take a seat

Same $999, same box, same seat pool. What differs is how much of yourself you hand over — and each card offers a one-off seat or the $999/week plan (crypto is one-off only).

You are charged $999 immediately and access to standard-speed GLM 5.3 abliterated and Kimi K3 abliterated starts the moment you pay — whichever option you pick. It runs for a month at 50–70 tokens a second. The fast week is Kimi K3 on the 8x B300 at 200–666, 16–23 September. No refunds.

PenClaw Ultimate

Card

$999 · one-off or /week

Charged now · live at 50–70 tok/s

  • $1,000 credits + one truly unlimited fast week
  • All penclaw.ai features
  • Government ID check required
  • Weekly plan available
  • No data retention, no training

Sold by ALBUMERA LTD

$999 is charged now — no holds, no deposits, strictly no refunds. Access starts immediately at 50–70 tokens/sec (GLM 5.3 Abliterated and Kimi K3 Abliterated), spending from $1,000 of credits. Reach 10 paid seats and your whole cohort steps up to a truly unlimited 200–666 tokens/sec for a week. Weekly re-books a seat every cohort; cancel any time.

platform.audn.ai Unlimited

Card

$999 · one-off or /week

Charged now · live at 50–70 tok/s

  • $1,000 credits + one truly unlimited fast week
  • API only — no penclaw.ai features
  • No ID check
  • Weekly plan available
  • No data retention, no training

Sold by ALBUMERA LTD

$999 is charged now — no holds, no deposits, strictly no refunds. Access starts immediately at 50–70 tokens/sec (GLM 5.3 Abliterated and Kimi K3 Abliterated), spending from $1,000 of credits. Reach 10 paid seats and your whole cohort steps up to a truly unlimited 200–666 tokens/sec for a week. Weekly re-books a seat every cohort; cancel any time.

platform.audn.ai Unlimited

Crypto

$999 / one-off

Charged now · live at 50–70 tok/s

  • $1,000 credits + one truly unlimited fast week
  • API only — no penclaw.ai features
  • No ID check
  • One-off only (crypto)
  • No data retention, no training

Sold by Audn Corporation

$999 is charged now — no holds, no deposits, strictly no refunds. Access starts immediately at 50–70 tokens/sec (GLM 5.3 Abliterated and Kimi K3 Abliterated), spending from $1,000 of credits. Reach 10 paid seats and your whole cohort steps up to a truly unlimited 200–666 tokens/sec for a week.

What happens to your money

You are charged now
$999, plus any tax where you are, is taken the moment you confirm. No hold, no deposit, no escrow. The exact figure is shown before anything touches your card.
What that buys, immediately
$1,000 of credits to spend on GLM 5.3 abliterated and Kimi K3 abliterated at 50–70 tok/sec over the month, plus one truly unlimited fast Kimi K3 week (16–23 September) once the cohort fills — not a trial, not a demo, not a smaller model. Live from the moment you pay.
Why 10, and what it buys
A dedicated 8x B300 for a week costs about $11,750. Split across 10 seats that is roughly the seat price each, and it stops being a shared queue and starts being your hardware for the week. Reaching 10 is what hires it.
If fewer than 10 join this week
Nothing changes about your money and nothing is refunded — you keep the full month of standard access you already paid for. The fast box simply is not hired that week. Reach 10 next week and it is.
The one-off seat
One charge, this week’s cohort, 30 days of standard, and a fast week if it fills. Nothing renews.
The weekly plan
$999 every week. Each charge books a fresh seat in the next weekly cohort — so you get the fast box every week it fills — and rolls your standard month forward. A second seat in a week buys nothing extra beyond that refreshed week. Cancel any time in the billing portal; you keep the month you last paid for.
No refunds
Strictly none, on any rail. Once the cohort is committed the money is already spent hiring the GPU. If you genuinely need to discuss a charge, message the team on Intercom — ending a seat means losing access across audn.ai, platform.audn.ai and penclaw.ai.
When does the fast week end?
8 days after the box is hired. After it, the cohort is released, the rate returns to 50–70 tok/s, and your standard month keeps running. To be fast again, take a seat in the next weekly cohort.

The fast week is unlimited, and still fair

During the fast week there are no token limits and no overage charges on GLM 5.3 abliterated or Kimi K3 abliterated. What there is, is a fair-share scheduler: nobody may consume more than ten seats’ worth of the box. Above your share, while the box is busy, your requests queue behind requests below theirs. You are never refused, never billed extra, and never silently moved to a smaller model. The standard weeks run against your $1,000 credit balance instead. With 50 seats that fast-week ceiling is about 12% of the whole machine — which is why 50 is the ceiling and not a bigger number.

Questions people actually ask

The money ones first.

Am I charged right now?

Yes. $999 (plus any tax) is charged the moment you confirm. There is no hold and no deposit — the escrow model is gone. In return you get $1,000 of credits to spend at standard speed immediately, over a month, plus one truly unlimited fast week once the cohort fills.

What’s the difference between one-off and weekly?

One-off ($999): one charge, one seat this week, a month of standard, a fast week if this cohort fills. Nothing renews.

Weekly ($999/week): a subscription that books a fresh seat in every new weekly cohort. You get the fast box every week it fills, and each charge rolls your standard month forward. Cancel any time; you keep the month you last paid for. Weekly is card only.

What does “fast” actually mean?

About 50–70 tokens/sec on standard, 200–666 tokens/sec once 10 seats are paid and we hire a dedicated 8x B300 — up to 10.1x on the raw numbers, and more reliable, because the box is not shared outside your cohort.

66 is a deliberate throttle, always deliverable. 666 is what a seat reaches on an unsaturated box — under heavy load a fair-share scheduler applies, so treat it as up to 666, not a floor.

What if the cohort never reaches 10?

You keep the full month of standard-speed access you paid for. Nothing is refunded — there are no refunds — and the fast box simply is not hired that week. Reach 10 in a later week and it is.

Can I get a refund?

No. Once we commit to the GPU the money is already spent hiring it, so there are strictly no refunds on any rail. If you genuinely need to discuss a charge, message the team on Intercom — note that ending a seat means losing access across audn.ai, platform.audn.ai and penclaw.ai.

How do I keep the fast box every week?

Be in every weekly cohort. Either take a fresh one-off seat each week, or use the weekly plan and it re-books you automatically. All four weeks of a month is four weekly seats ($3,996). Log in and keep opting in to stay in.

Is crypto different?

Crypto is one-off only — a stablecoin charge cannot recur, so there is no weekly plan on that rail. It settles immediately and, like everything here, is not refundable. It is capped at 10 seats per cohort.

Who am I buying from?

Card seats are sold by ALBUMERA LTD (United Kingdom), shown as AUDN.AI on your statement. Crypto seats are sold by Audn Corporation (United States), shown as AUDN CORPORATION.

Do you keep or train on what I send?

We never train on your traffic, and prompts and completions are not retained.

Request metadata — timestamps, token counts, which model, how much of the box you used — is retained, because the fair-use scheduler and billing are built on it. Content is not.

What does “abliterated” mean?

It lowers the model’s latent tendency to refuse gray-area prompts during authorized security work — our own method, applied to Kimi K3. The models on this cluster are effectively uncensored: a 0.2 to 2.5% compliance rate, where lower means more thoroughly abliterated. Method and per-model results on our blog. It is meant for offensive-security work you are authorized to perform.

Pricing

Three ways to buy

Pay as you go

$4.50 / $4.50per 1M input / per 1M output

You are billed for tokens, in arrears, and for nothing else. An agentic request runs about 50,000 tokens in and 200 out, which is $0.226, and 99.6% of that is the input side. Input the box serves from its prefix cache bills at $0.40 per 1M instead of $4.50. Below roughly 222 million tokens a month, input and output together, this is the cheapest of the three.

Cheapest right up until you start using it properly.

Start on platform.audn.ai

Weekly seat

$999per week, cancel any time

The same charge, every week: a subscription that books you a fresh seat in every new weekly cohort, so you are in every fast week that fills — and each weekly charge rolls your standard-speed month forward another 30 days. Four weeks of always-being-in-the-cohort is $3,996; a stacked second seat in the same week buys nothing beyond that refreshed fast week and the month extension. Card only — crypto seats are one-off. This replaces the old $999/month Ultimate.

For people who want the hired box every single week.

Side by side

All three, line by line

Pay as you goOne-off seatWeekly seat
What you pay$4.50 per 1M input, $4.50 per 1M output, $0.40 per 1M cached input. About $0.226 per agentic request.$999, once, charged today.$999, every week.
When you are chargedIn arrears, for tokens already spent.Immediately, in full, the moment you buy.Immediately, and again every week until you cancel.
CommitmentNone. Stop sending requests and it stops costing.$1,000 of standard credits, once. Nothing renews.Week to week. Cancel and the next week does not run.
Speed todayAbout 50–70 tok/s on the shared tier.About 50–70 tok/s from the moment you buy.About 50–70 tok/s.
Speed once 10 seats are paidNo change. The box is hired by the cohort.Up to about 200–666 tok/s on the hired 8xB300 for that week, to the 50-seat cap.Up to about 200–666 tok/s, every week a cohort fills.
Latency per request10 s to 5 min, depending on requests in flight across the box.10 s to 5 min, set by how many requests are in flight across the hired box, not by what you are asking for.10 s to 5 min, set by how many requests are in flight across the hired box, not by what you are asking for.
Concurrency12 generations in flight, in any mix of PenClaw and API. A 13th is rejected immediately rather than queued indefinitely.12 generations in flight, in any mix of PenClaw and API. A 13th is rejected immediately rather than queued indefinitely.12 generations in flight, in any mix of PenClaw and API. A 13th is rejected immediately rather than queued indefinitely.
When the queue is fullHTTP 429 with Retry-After. Never a silent drop, never billed.HTTP 429 with Retry-After. Never a silent drop, never billed.HTTP 429 with Retry-After. Never a silent drop, never billed.
Context window1,000,000 tokens.1,000,000 tokens.1,000,000 tokens.
Abliteration completenessEffectively uncensored — a 0.2 to 2.5% compliance rate across these models. Research and results.Effectively uncensored — a 0.2 to 2.5% compliance rate across these models. Research and results.Effectively uncensored — a 0.2 to 2.5% compliance rate across these models. Research and results.
Usage while the cohort fillsBilled at the metered rate, as usual.Included. You already paid, and standard runs from the moment you buy.Included. Standard runs the whole time.
If the cohort misses 10Nothing happens. You were never in it.You keep your full $1,000 of standard credits. Nobody is refunded — the box was still on offer.You keep standard, and next week is a fresh chance at the fast box.
RefundsNothing to refund. You pay for what you used.None, ever. Once we commit to the GPU the money is spent hiring it.None, ever. Cancel to stop future weeks; the week you paid for still runs.
RenewalNothing to renew.No. The $1,000 of credits and the month end together.Yes. Every week until you cancel.
Best forUnder about 222 million tokens a month, input and output together.About $1,000 of standard usage a month, plus a shot at the unlimited fast week.About $1,000 of standard a month, and wanting the unlimited fast box every week.
AvailabilityNow.Now, this week’s cohort. Seven-day window; fast unlocks at 10 paid seats, selling closes at 50.Now, and it re-books you into every weekly cohort automatically.
Who you contract withALBUMERA LTD, United Kingdom. Shown on your statement as AUDN.AI.ALBUMERA LTD for a card seat. Audn Corporation, United States, for a crypto seat (one-off only, capped at 10).ALBUMERA LTD, United Kingdom. Weekly is card only; shown as AUDN.AI.
Metered against flat

The same work, at four latencies

In the fast week there is no meter, so making the box faster costs you nothing. Below is the same work, one user with three threads, eight hours a day, across the seven-day fast week, priced at four different latencies. Here is what that week of usage would cost you metered — and in the fast week you pay none of it. Your $1,000 of standard credits, for the other three weeks, is on top.

10 s per request3.04 B tokens, $13,66313.7xcheaper on a seat
30 s per request1.01 B tokens, $4,5544.6xcheaper on a seat
94 s per request325 M tokens, $1,4531.5xcheaper on a seat
Flat fast week, any latency$999the seat

1.5x to 14x cheaper

1.5x is the WORST case, and it is the one to plan against: 94 seconds is a crowded-box worst-case latency near the 50-seat cap, and even there the same three threads over the same eight hours across the seven-day fast week cost $1,453 metered against a $999 seat — and your $1,000 of standard credits is on top. Note which direction the metered column moves. At 10 seconds it is $13,66313.7x the seat — because on a meter you pay for the speed you asked for.

Your $999 buys $1,000 of credits to spend at standard speed, billed at the same $4.50 per-million meter as pay-as-you-go — so you never overpay for standard use — and, if the cohort fills, one truly unlimited fast week on top at no extra charge. In that week three agentic threads move on the order of 1 to 3 billion tokens, roughly $4,500 to $13,700 metered, and you pay none of it.

What a seat is worth as the cohort fills

And it moves as the cohort fills. The box is hired the moment the 10th paid seat lands, and every seat sold after that shares the same 8xB300 up to the 50-seat cap — so latency rises, the tokens one user gets through in a day fall, and the metered bill those tokens would have carried falls with them. The three rows below are the crowded end of the hired fast week; at the floor, with far fewer seats on the box, each seat gets comfortably more than these figures. Same one user, same three threads, same eight hours, across the seven-day fast week.

Every latency on this page is a worst case. Our load tests put the real figure at about a quarter of it — roughly 24 s where the table says 94 s. We publish the worst case because it is the number you can plan against, and because a seat that is already worth buying at the worst case does not need the average to make its argument.

Seats on the boxLatency, worst caseLatency, typicalTokens per dayTokens per weekSeat cost per 1MMetered per week
50 seats on the box94 s~24 s46.3 M324 M$3.08$1,458
60 seats on the box112 s~28 s38.6 M270 M$3.70$1,215
50 seats, selling closed131 s~33 s33.0 M231 M$4.32$1,039

Even at the crowded end, the fast week returns more usage than its share, and it costs nothing on top of your $1,000 of standard credits — whatever you run that week, you pay no meter for it.

The standard weeks run on your $1,000 of credits at the same $4.50 per-million meter as pay-as-you-go — no markup, no lock-in, and the balance is yours to spend across the month. The fast week is the part where flat beats metered; the standard weeks are simply metered use you have pre-paid.

Choosing

Which one is for you

If youTakeWhy
You send under about 222 million tokens a month and don’t need the fast weekPay as you goAt standard speed a seat bills at the same $4.50 per-million rate as pay-as-you-go, so under about $1,000 of usage a month (~222 million tokens) there is nothing to gain from pre-paying — unless you want a shot at the unlimited fast week. Pay as you go and commit to nothing.
You run agents all day and want standard credits plus this week’s unlimited fast boxOne-off seatOne $999 charge buys $1,000 of credits to spend at standard speed from the moment you pay, and one truly unlimited week on the hired 8xB300 if this cohort reaches 10 — no meter at all during that week, where the flat-beats-metered value really lands. If 10 seats are never paid you still keep the full $1,000 of standard credits, and nothing renews.
You want the hired fast box every single weekWeekly seatSame $999, billed weekly. It books you into every new weekly cohort automatically, so you are in every fast week that fills, and each charge rolls your standard month forward. Four weeks is $3,996. Cancel any time and the week you already paid for still runs; there are no refunds and none are needed, because you simply stop the next charge.
Fair share

What “unlimited” means in the fast week

During the fast week, unlimited means no token cap, no daily quota and no overage charge on GLM 5.3 abliterated and Kimi K3 abliterated. It does not mean no limits. There are four, and you should know them before you pay. The standard weeks run against your $1,000 credit balance instead.

Twelve at once
Twelve generations in flight per subscription. PenClaw engagements and OpenAI-compatible API calls draw on one shared budget of twelve, in any mix. A thirteenth is rejected immediately rather than queued indefinitely, because a queue with no end is worse than a refusal. Twelve is a burst ceiling for spikes, not a reservation. The sustained figures above assume three.
10 seconds to 5 minutes
How long a request takes depends on how many requests are in flight across the whole box, not on what you are asking for. Near the 50-seat cap the crowded-box worst case is around 94 seconds; at the 10-seat floor, with far fewer seats sharing the box, it is a small fraction of that.
429, never a silent drop
Past the bounded queue you get an HTTP 429 with a Retry-After header. The request is not quietly dropped, not silently truncated, and not billed.
Ten seats’ worth
No single subscription may consume more than ten seats’ worth of the box. Above your share, while the box is busy, your requests queue behind requests below theirs. You are never refused for it and never billed extra.

And the number that matters most: up to about 666 tokens a second is the box’s streamed output on the fast tier, and the 50-seat cap is what protects it — the ceiling exists to hold the number up, not to ration it. A cluster nobody else is on costs about $11,750 a week — a quarter of the old $47,000 month — and it costs that whether one seat is on it or seventy.

Provenance

Where the speed numbers come from

We hire the cluster from Modal and run our own abliterated weights on their DFLASH stack. The speed figures below are not ours: they are Modal’s published benchmark for moonshotai/Kimi-K3 on 8xB300, measured under the same agentic profile we actually serve — 50,000 tokens in, 200 out. You can check them yourself, which is the only reason a speed claim is worth anything.

Interactivity falls as the replica fills. That is the whole reason a cohort has a ceiling: 50 is where the curve stops being the number we promised, so 50 is where selling stops.

These figures apply once 10 seats are paid and the box is hired for the week. They do not describe what you get today. Until the floor is reached you are on the shared tier at about 50–70 tokens a second, on GLM 5.3 Abliterated and Kimi K3 Abliterated, spending from your $1,000 of credits. That band is a deliberate throttle rather than a point on the curve below, which is why it is always deliverable: it is a floor we hold, not a speed that degrades as more people join. The table starts applying the moment the 10th seat is paid, and not before.

Replica loadSpeed per user, hired clusterWhat that is
4M tok/min668 tok/sOne stream, box otherwise quiet
6M tok/min380 tok/sLight concurrency
8.7M tok/min235 tok/sApproaching the seat cap
10M tok/min210 tok/sReplica saturated

Our abliteration is measured the same way, on the same weights we serve — method, bench and per-model results at blog.audn.ai. The cluster costs about $11,750 a week, which is why this is a cohort and not a button.

We found the sweet spots with a lot of research. Let’s have sovereign frontier AI together.

Reference

Specifications

Model
Kimi K3, audn abliteration
Context window
1,000,000 tokens
Modality
Text and native vision
Cluster
8x B300, us-east
Shared tier
About 50 to 70 tok/s. A deliberate throttle, so it is always deliverable.
Dedicated cluster
Up to about 666 tok/s (200–666 band). The box’s streamed output, not a per-seat allocation.
Concurrency
12 generations in flight per subscription, any mix of PenClaw and API
Latency
10 s to 5 min, depending on requests in flight across the box
Metered list price
$4.50 per 1M input, $4.50 per 1M output, $0.40 per 1M cached input