Models·3 min read
By BitsMindsSource: TechCrunch

Ox Alpha Is Free, Frontier-Class, and Unclaimed

A reasoning model with a million-token context window appeared on OpenRouter on August 20 under the provider name "stealth". Five days later, no lab has admitted building it.

LISTED ON OPENROUTER Ox Alpha BUILT BY 1M CONTEXT FREE TEXT IMAGE VIDEO No lab has claimed it since the August 20 preview BITSMINDS.COM
Share:

On August 20 a model called Ox Alpha appeared on OpenRouter and in the OpenCode client, listed under the provider name "stealth". Its card describes a reasoning model built for coding, sustained agentic work and production workloads. It carries a 1,048,576-token context window, accepts text, images and video, and will return up to roughly 131,000 tokens in a single completion. It costs nothing to use. Nobody has said who made it.

That combination is why the listing spread as fast as it did. A million-token context and video input put Ox Alpha in the same bracket as the frontier systems that cost real money per million tokens, and developers who ran it against their own agent harnesses came away describing it as genuinely competitive rather than a curiosity. Stripe chief executive Patrick Collison called the model "very impressive" — a notable endorsement, though one worth reading alongside the fact that Stripe acquired OpenRouter for $7 billion earlier this month.

OpenRouter has been precise about what it is and is not. The company says it routes requests to Ox Alpha but is not the model's developer, owner or provider, and that the third party behind it has chosen to remain anonymous for the duration of the preview. The data terms differ by route: OpenCode says its paths retain nothing and train on nothing, while OpenRouter says the unnamed provider does retain request data but does not train on it. Access on OpenRouter requires an account and an API key; OpenCode serves it through its free Zen gateway tier and its paid Go plan. OpenCode has advertised aggregate capacity of 100 trillion tokens a day, a service-wide figure rather than a per-user allowance.

The guessing has mostly pointed at China. Tokenizer probing and behavioural fingerprinting led several developers toward Z.ai's GLM family, which shipped GLM-5.3 earlier in August and has the coding pedigree to match. The theory is not airtight: Ox Alpha takes image and video input, and GLM-5.3 is text-only. Competing threads have floated an unreleased Microsoft MAI checkpoint, a Phi derivative, and an IBM Granite lineage, with roughly the confidence you would expect from evidence assembled out of sampling quirks and refusal styles. No lab has stepped forward.

Anonymous previews have become a recognisable pre-launch tactic — a way to collect unbiased evaluations and real agentic traffic before a name and a price tag anchor expectations. What makes this one unusual is the scale of the giveaway. Serving a million-token context free to an open gateway is expensive, and whoever is paying that bill is buying something specific: comparison data gathered before anyone knew whose model they were grading.

The preview is reportedly a roughly one-week window, which would put the cutoff around August 27. Until it closes, the most interesting artefact here is not the benchmark score but the experiment itself — a frontier-tier model evaluated entirely on its output, with the brand stripped off.

Want AI news before everyone else?

The morning's most important AI stories, straight to your inbox. No fluff.

Related Articles

Claude Haiku 5.5 vs GPT-6 Luna: the BitsMinds Lab build-off A terracotta disc carved into a labyrinth, with the Claude mark inlaid at its centre, faces a brushed silver disc with the OpenAI mark inlaid over the shading of a crescent moon. A VS badge sits between them, the model names are painted below, and the line above reads Same briefs, same price. BITSMINDS LAB · SAME BRIEFS, SAME PRICE VS CLAUDE HAIKU 5.5 GPT-6 LUNA BITSMINDS.COM
Models

Claude Haiku 5.5 vs GPT-6 Luna: Lost in Thought

Mistral Large 4, le Chonk The pixel-block Mistral logo cast as a thick, heavy slab with deep extruded sides, its face banded yellow to orange to red, standing on a dark warm floor under a single light, above a label reading Large 4, 1.05T parameters, 52B active. MISTRAL LARGE 4 1.05T PARAMS · 52B ACTIVE BITSMINDS.COM
Models

Mistral Large 4: Europe’s 1-Trillion-Parameter Open Model

Claude Haiku 5.5: Anthropic's fastest model, at a tenth of Haiku 4.5's price A terracotta and ivory stopwatch, tipped to show its depth, has the official Claude asterisk inlaid in clay as the hub of its green sweep hand. A paper tag tied to the crown reads minus 90 percent against Haiku 4.5. Beside it, the title names Claude Haiku 5.5 and quotes its API prices for prompts up to 100,000 tokens: 10 US cents per million input tokens and 50 cents per million output tokens. The stopwatch is an editorial metaphor for Anthropic's claim that Haiku 5.5 is its fastest model at standard speed, not an Anthropic product; the 90 percent figure applies to prompts up to 100,000 tokens, and longer prompts cost five times as much. 51015202530354045505560 CLAUDE HAIKU 5.5 PER-TOKEN PRICE −90% VS HAIKU 4.5 PROMPTS UP TO 100K CLAUDE Haiku 5.5 $0.10 INPUT $0.50 OUTPUT PER MILLION TOKENS BITSMINDS.COM
Models

Claude Haiku 5.5 Is Out at a Tenth of Haiku 4.5's Price