Ox Alpha appeared on OpenRouter on August 20, 2026 as a free preview reasoning model with an unnamed developer. Community fingerprinting points hard at Z.ai's GLM family, but nothing is officially confirmed.
2026-08-24 · by Moe Ameen
On August 20, 2026, a model named "stealth/ox-alpha" appeared on OpenRouter with no named developer, listed as a reasoning model built for coding, sustained agentic work, and production workloads, and offered free during a limited preview. OpenRouter's listing describes it as a stealth model whose developer has "chosen to remain anonymous during this preview." A few days later TechCrunch ran it under the plain question everyone was asking: who's behind Ox Alpha? Stripe CEO Patrick Collison — Stripe recently agreed to acquire OpenRouter — called it "very impressive."
The specs on the listing are what drew developers in: a context window of 1,048,576 tokens (about 1M), output in the low hundreds of thousands of tokens, text, image, and video input, and tool calling — all at a preview price of zero. Within days it was among the most-used models on developer platforms that carried it, riding the free window.
On origin, everything is inference, not confirmation. Community "fingerprinting" tools that probe a model's tokenizer and serving infrastructure reported close matches to Z.ai's GLM-5.3, and an independent tester found Ox Alpha's token counts lined up with GLM-5.3 across a batch of prompts, off by only a small constant. Other observers floated an unreleased Microsoft model. No lab has claimed it. The viral "it tops GPT-5.6 and Claude on coding" numbers came from small, unaudited user tests, so treat them as preliminary. Stealth releases like this are a recurring pattern now — a lab drops an unbranded model to gather real-world usage and reactions before attaching its name.
The practical move on a launch like this is to resist standardizing on the model at all. Ox Alpha is exciting and it might be excellent, but it is anonymous, free-for-now, and possibly gone in a week — the opposite of something to wire your workflow around. The lesson the stealth-model churn keeps teaching is that the model is the interchangeable part; the pipeline that turns a model's output into finished content is what you actually own. [Kompozy](/) is built to be that pipeline. It runs managed [Claude and OpenAI models](/glossary/credit-based-pricing) for copy under a [Persona Brief](/glossary/persona-brief) — a known, governed drafting layer for brand work you don't want routed through an unidentified lab — and then does the part no raw model does: generates the media, holds it on-brand, and ships it.
So a creator watching Ox Alpha go viral doesn't need to switch anything. Bring an idea into Kompozy and it fans into [18 formats](/glossary/output-buckets) — [Persona Shorts](/glossary/persona-shorts), [Persona Frames](/glossary/persona-frames) video, carousels, quote cards, blogs, newsletters — each rendered brand-exact and then scheduled and published across the eight social platforms plus blog and email on [Autopilot](/glossary/autopilot), through one review gate. If a stealth model turns out to be the real deal and gets a name and an API, great; the production-and-publishing layer around it never had to change. That is the point: chase models and you rebuild every month, own the shipping layer and you don't.
Ox Alpha is an anonymous "stealth" reasoning model that appeared on OpenRouter on August 20, 2026, built for coding and agentic work and free during a preview. Officially no one has claimed it; the strongest public evidence — tokenizer fingerprinting and matching token counts — points at Z.ai's GLM-5.3, with some speculation about an unreleased Microsoft model.
It was free on OpenRouter during its preview window, described as roughly a week from the August 20 launch. Stealth-preview pricing is temporary by design, so verify current access and cost on OpenRouter before relying on it.
Not directly. It is an anonymous, free-for-now raw model that generates no video, images, or posts and may not persist past its preview. The safer approach is a model-agnostic content engine like Kompozy that uses known governed models for copy and owns the media generation and publishing, so a churning model layer never breaks your pipeline.