Stealth drop · appeared Aug 20, 2026 · free for about a week

Ox Alpha: The Mystery AI Model Just Dropped

Nobody knows who built it. The internet found out anyway.

Ox Alpha, the mystery AI model, just dropped — and nobody knows who built it.

A frontier model appeared out of nowhere two days ago, wearing a mask.

It holds a million tokens at once, it can watch video, and right now it costs nothing.

Independent testers ran it against Claude and GPT, and the scores turned heads.

Then internet detectives found fingerprints hidden inside the model itself.

By the end of this page you'll know the whole story before the mask comes off.

THE MYSTERY Who built it? provider: "stealth" What it is1M · video · free The testsvs Claude + GPT Whodunitthe fingerprints 99% sure A detective story in three stops Someone spent hundreds of millions — then refused to say who they are
The actual sources ↓
§1Tested live · my Agent OS · Free AI Coder

I plugged the mystery model into my own machine. Here's it running.

These are real recordings from the Free AI Coder section of my Agent OS — Ox Alpha wired in as an engine, the same day it appeared.

What you're watching: the Free AI Coder engine dropdown switching from OmniRoute to Ox Alpha · stealth 1M — the status chip flips to "Gateway live" the moment OpenRouter confirms the model is up.

What you're watching: I ask it straight — "Who built you?" It answers: "I was built by an undisclosed organization, and I'm the model known as ox-alpha." Even the model keeps the mask on.

What you're watching: one build prompt, and Ox Alpha writes a full animated page that renders live in the preview pane — then gets saved to my workspace. It even added the caption "IDENTITY WITHHELD" on its own.

§2Here's what happened

A model appeared with no name attached.

On August 20th a listing called Ox Alpha showed up on OpenRouter and OpenCode — no company, no logo, no announcement, just the word "stealth" as the provider.

Ox Alpha provider: "stealth" company: — none — announcement: — none — ? ? ? ? August 20, 2026 — the mask goes on
Tweet 1 · the drop

OpenCode announces the free week

This is the post that started it. OpenCode lists the whole offer: 1M context, multi-modal, zero data retention, "generous rate limits, near unlimited usage" — and the line everyone quoted: "We have capacity for 100T tokens per day, lets see what you can do." It's sitting at 5.8 million views.

§3Why this is strange

Someone spent a fortune — then gave it away.

These four facts don't normally go together.

1,048,576token context
3inputs · text image video
$0for about a week
0names attached
Hundreds of millionsto build a frontier model Given awayfree · about one week Maker: hiddenrefuses to say who This combination should not exist Why would anyone do that? Hold the question.
§4Where it lives · OpenRouter in one picture

OpenRouter is a train station for AI models.

Connect once, and hundreds of models pull up to the same platform — that's how a masked model can appear overnight and reach everyone.

You one connection OpenRouter the station hundreds of models Claude GPT GLM 🎭 Ox Alpha A new train pulled in this week — wearing a mask
§5What the listing says

Built for long jobs.

The listing calls it a reasoning model for coding, sustained agentic work, and production workloads — plain English: an agent that works for hours without losing the plot.

Tweet 2 · the listing

OpenRouter's own description

Here's the official wording from OpenRouter: "a frontier model built for efficient coding, sustained agentic work, and real-world production use" — with the 1M token window and text, image and video input spelled out. Their follow-up note adds the part that matters: it's free, and this provider does not train on your prompts.

"Sustained agentic work" = jobs that run for hours hour 1hour 2hour 3 …loses the plot typical model Ox Alpha
§6Spec 1 · the context window

A million tokens is a very big desk.

The context window is the model's desk — a million tokens means months of notes, reports and plans on that desk at once, all visible while it works.

a normal desk ~128k tokens notes reports plans codebase videos Ox Alpha's desk 1,048,576 tokens Put months of your business on the desk — it sees all of it at once
§7Spec 2 · the inputs

It can watch video. Most models can't.

Text, images, and video in — the video part is rare, and it becomes the biggest clue in the detective story later, so remember it.

📄 Text 🖼 Images 🎬 Video 🎭 Ox Alpha remember this — it matters later 🔍 Every lab chops video into tokens differently — like handwriting
§8Spec 3 + 4 · the price and the wild claim

Free — and 100 trillion tokens a day of capacity.

Very few companies on Earth have that kind of computing power sitting around, and that single number cuts the list of suspects way down.

Who can even serve this? a busy startup's daily tokens Ox Alpha's claimed capacity 100T / day The suspect list collapses small labs startups universities a handful of giants with idle superclusters Hold that thought — the detectives use it later
§9Is it actually good? · the DeepSWE run

An anonymous free model outscored the giants.

Researcher Ben Davis ran it through ten real software engineering tasks on DeepSWE — here's how the run landed.

DeepSWE · 10-task community run (Ben Davis) 🎭 Ox Alpha 80% Claude Fable 5 65% GLM 5.3 62% GPT 5.6 52%
Tweet 3 · the numbers spread

The score card that confused everyone

KC posted the comparison that went around: "gpt-5.6-sol → 52% · Fable → 65% · Ox Alpha → 80%+. People are starting to get very confused." It quotes Ben Davis's original run — his own words: "I am very confused." Read those names again: Claude, GPT — beaten by a model with no name on it.

§10The honest caveat

Ten tasks is a small sample.

There are no official benchmarks and these community numbers aren't independently verified — treat them as a strong early signal, not proof.

the full benchmark — hundreds of tasks, not yet run 1 2 3 4 5 6 7 8 9 10 strong signalnot proof I don't do hype — ten tasks means variance. Keep watching.
§11The second signal · production on day one

Real teams shipped it on day one.

Nous Research plugged Ox Alpha into their Hermes Agent project, and the code editor Zed integrated it — on its first day online. That tells you something benchmarks can't.

Tweet 4 · production adoption

Nous Research turns the taps on

This matters to me personally, because Hermes is one of the main agents inside my own Agent OS. Nous didn't just test the anonymous model — they put it in their portal and wrote: "We have capacity for 1 quadrillion tokens per day. Let the tokens flow." When a team whose tools I run daily routes real work through a masked model, I pay attention.

🎭 Ox Alpha Hermes AgentNous Research · day one Zed editorintegrated · day one "Yes, we'll route real work through this" — that's the real vote
§12Same day · in my Agent OS

New engine drops → swap it in. That's the whole point.

I plugged Ox Alpha into my Agent OS the same day through OpenRouter, right alongside Hermes and Claude — and here's the page it built for me, running full screen.

What you're watching: the page Ox Alpha built inside my Free AI Coder, opened full screen — twinkling starfield, pulsing golden orb, and its own "IDENTITY WITHHELD" caption. The mystery model has a sense of humour.

Your Agent OS — never changes workflows · memory · goals · one dashboard Claude Hermes OpenClaw 🎭 Ox Alpha A new model drops → you swap it in and test on real work in minutes
THINKING IT? "Doesn't running an Agent OS burn a fortune in tokens?"

No — that's the biggest myth about it. This whole guide runs on a free model, the everyday 90% runs on free local models and free APIs, and for frontier work it drives the CLIs you already pay for — your Claude subscription already includes the Claude CLI.

It's a layer on top of what you already own, not a new meter. And inside the AI Profit Boardroom there are full token-optimisation tutorials, so usage never becomes a worry.

§13Skip the setup Skip the setup

Get the Agent OS built for you.

You just watched Ox Alpha running inside mine. This week specifically, we're testing Ox Alpha inside the Agent OS — so when a stealth model drops, you test it in your own system the same day instead of watching from the sidelines.

The full Agent OS zip — your Claude, Hermes and OpenClaw in one dashboard, sharing one memory
The OpenRouter engine-swap setup — plug in Ox Alpha and every new model like it, from a dropdown
A 30-day roadmap — set it up step by step, with a video tutorial walking you through it
Four coaching calls every week — this week they're going deep on routing new models into your workflows
Daily step-by-step tutorials — agents that bring in leads and win you customers
A prompt library and a member map — find people near you running these exact setups
4,000+ business owners inside — plenty started with zero AI experience, someone is always online
Get the Agent OS → Inside the AI Profit Boardroom · skool.com/ai-profit-lab
Set up in an afternoon · used in 38 countries · new engines added the week they ship
"Okay. Now the detective story. Because this is where it gets really fun."
§14The detective story begins

Every model has a fingerprint.

Nobody claimed Ox Alpha, so the community went hunting for the tiny habits a maker can't easily hide — and Ben Davis found three of them.

🎭 Ox Alpha 🔍 Clue 1 · the video encoderhow it chops video into tokens 🔍 Clue 2 · the tokenizerthe dictionary it breaks text with 🔍 Clue 3 · the way it says nohow it rejects audio input Not a literal fingerprint — a technical one. Like handwriting.
§15Clue 1 · the video encoder

Four test videos. Token counts matched GLM exactly.

Davis fed Ox Alpha four controlled videos and counted the tokens — the counts matched Zhipu's GLM-5V-Turbo token for token, while every other candidate had a clearly different signature.

Same 4 videos → count the tokens each model spends 🎭 Ox Alpha GLM-5V-Turbo MATCH Xiaomi MiMo ✗ different Qwen ✗ different Every lab chops video differently — these two chop it identically
§16Clue 2 · the tokenizer

25 prompts. Same dictionary — plus a 75-token wrapper.

Across 25 prompts, Ox Alpha's token counts matched GLM-5.3 exactly apart from a constant 75-token hidden wrapper — matching that precisely requires an identical vocabulary.

25 prompts · token counts, model vs model gap = exactly 75 tokens, every time — Ox Alpha — GLM-5.3 Identical vocabulary + a constant hidden system wrapper
§17Clue 3 · the way it says no

Even its "no" gave it away.

Ox Alpha rejects audio input with the same behavior as GLM-5V, while Xiaomi's MiMo — the leading rival theory — happily accepts audio. They even checked the emoji rate: it matched Zhipu's models.

🔊 🎭 Ox Alpha"no" — rejection style A GLM-5V"no" — rejection style A Xiaomi MiMo"sure!" — accepts audio same "no" = same family 🧬 MiMo theory: eliminated ✗ The way a model refuses something became evidence
§18The verdict

99% certain it's Zhipu's GLM 5 series.

Stack the clues and Davis puts it at 99% a Zhipu GLM 5 model, with other analysis around 90% that this is the next-generation multimodal model — possibly what people will call GLM 5.5. Officially, it is still unconfirmed.

video encoder → GLM ✓ tokenizer → GLM ✓ audio "no" + emoji → GLM ✓ 99% → Zhipu · GLM 5 series ~90% → the next-gen multimodal · maybe "GLM 5.5" OFFICIALLY UNCONFIRMED
§19Why the theory fits · the pattern

This is the fifth masked model in six months.

The previous four were all eventually claimed by Chinese labs — anonymous debut, a burst of free traffic, then the company steps forward.

🎭Pony α→ GLM-5Zhipu 🎭stealth→ MiMo-V2-ProXiaomi 🎭stealth→ Ling-2.6-flashAnt Group 🎭stealth→ LongCat-2.0Meituan 🎭Ox Alpha→ ? Six months · five masks · four already came off Anonymous debut → free traffic → the company steps forward
§20Why labs do this · the playbook

The free week isn't charity. It's paying for feedback with compute.

Launch masked: if it wins, take the mask off and enjoy the applause — if it stumbles, fix it quietly and nobody connects it to you.

🎭 Releasemask on · free week Millions test itreal work · real breakage Lab learnsevery chat teaches 😎 Wins → mask offtake the applause 😶 Stumbles → silencefix it, nobody knows Benchmarks are easy to game. Strangers hammering it aren't. Millions of real conversations, bought with a week of compute
§21One more layer · the timing

Six days after GLM 5.3, the missing piece appeared.

Zhipu shipped GLM 5.3 on August 14th with no image or video input — the number one community ask was a vision version, and six days later an anonymous model with GLM's exact fingerprints shows up seeing both.

Aug 14GLM 5.3 shipstext-only · no eyes the ask"give us a vision version"the #1 community request Aug 20 · 6 days later🎭 Ox Alpha appearssees images AND video ✓ The timing fits like a glove The community asked for eyes — a masked model with GLM's fingerprints brought them
§22The size clue · decode speed

It even runs at GLM speed.

Testers measured how fast it generates text: within about 6 percent of GLM-5V-Turbo — the size class where 100 trillion free tokens a day is a real stress test, not a fantasy.

🎭 Ox Alpha decode speed GLM-5V-Turbo decode speed ~6%apart Similar speed → similar architecture → similar size class The only size where "100T free tokens a day" makes business sense
"So what does all this mean for you? Three things actually matter here."
§23Thing 1 · the frontier is now sometimes free

The gap between them and you keeps shrinking.

A year or two ago this capability was locked behind enterprise contracts — now it shows up anonymously on a public platform, and anyone with an OpenRouter account can point their agents at it.

Access to frontier AI, over time huge companies you 🎭 this week: $0 That shrinking gap is the real story under every one of these drops
§24Thing 2 · the desk + the eyes together

A million tokens plus video changes what you can hand off.

I've been feeding mine long recordings of my own screen work and having it write up what I did — that used to take hours by hand.

🎬 2-hour screen recordingyour actual work session 📚 months of contextnotes · plans · reports 🎭 watches + reads 📝 finished write-up hours of your time, back Most people type short questions — this holds your whole month and watches your screen
Tweet 5 · what the big desk unlocks

"We started handing it the whole repo"

Ziwen wired Ox Alpha into their coding setup and described exactly this shift: "we stopped picking the files we think matter and started handing it the whole repo… this week is for the jobs that are usually too big." No API key, no billing, a million tokens of context — until the window shuts.

§25Thing 3 · the honest warning

Free plus anonymous means you test smart.

The listing says prompts and completions are retained by the provider — just not used for training — and the provider is anonymous, so keep client data on models with a name.

your data tests · drafts clients · numbers 🚦 the rule 🎭 Ox Alphaonly what you'd post publicly Named modelsclient data · private details "zero retention" headline ≠ the listing's fine print — retained, not trained on You don't know whose servers your words sit on. Test hard — with public-safe work.
"Now, I know what some of you are thinking. Let me deal with it head on."
§26Belief 1 · "this moves too fast"

Wrong: "A new model every week — I can't keep up, so why bother starting?"

Right: The speed is exactly why people who build a system win. My Agent OS stays the same, my workflows stay the same — only the engine underneath swaps out. Ox Alpha appeared on a Thursday and ran my tasks by that evening.

Your system — built once engine: Claude engine: Hermes engine: GLM engine: 🎭 Ox Alpha workflows unchanged ✓ no system = drowning in drops system = enjoying the pace Build the frame once — then every new engine makes it stronger
§27Belief 2 · "I'm not technical enough"

Wrong: "I'm not technical enough for something called a stealth reasoning model."

Right: You just followed a story about tokenizer fingerprints and understood every bit of it. Choosing a model on OpenRouter is picking from a dropdown menu — you watched me do it in the first video on this page.

engine: OmniRoute ▾ OmniRoute · free pool 9Router · 573 models 🎭 Ox Alpha · stealth 1M ✓ That's the whole "technical" part — a dropdown
§28Belief 3 · "I'll wait for the dust to settle"

Wrong: "I'll wait until there's one obvious winner."

Right: The dust is not going to settle — five stealth models in six months is the new normal. You need one working system and the habit of testing new engines when they appear. That habit takes an afternoon to build.

🎭 🎭 🎭 🎭 🎭 ? ZhipuXiaomiAntMeituanOx Alphanext month Waiting for calm means waiting forever Other business owners get faster every month while the dust keeps not settling
Don't take my word for it

Members post their wins every day — agency owners, ecom founders, course creators, solo operators across 38 countries. Real businesses, real numbers, in their own words.

Read the 158-page wins doc →
§29What happens next

The mask comes off around August 27.

The free window is expected to close around the 27th, the reveal usually lands after the preview ends, and Zhipu's GLM 5.3 weights unlock after a safety review about two weeks out — the timelines line up for a very interesting end of August.

todayfree · test it hard ~Aug 27free window closes then…🎭 → 😮 the revealGLM 5.3 weights ~2 weeks out too A very interesting end of August If the fingerprints are right, you knew the whole story before the headline
§30Your move Use it, don't just watch it

Join us while we test Ox Alpha inside the Agent OS.

This week specifically, the AI Profit Boardroom is testing Ox Alpha inside the Agent OS — the system where your Claude, your Hermes and your OpenClaw share one memory and work together. Readers watch the mystery from the sidelines. Operators run the mystery model on their own work the same day it drops.

The Agent OS zip file — download it, unzip it, follow the video tutorial
The 30-day roadmap — from nothing installed to agents running your day
Daily updates — we ship new versions as new engines like Ox Alpha appear
Four coaching calls a week — going deep right now on routing new models through OpenRouter
Daily tutorials — agents that bring in leads and win you customers, step by step
A prompt library + member map — compare notes with people near you running these setups
4,000+ business owners — lots started with zero AI experience, someone is always online to help
Get the Agent OS → Inside the AI Profit Boardroom · skool.com/ai-profit-lab
The mask comes off around August 27 — and when it does, you'll already know the whole story
"See you in the next one."