Skip to content
ChristopherOx Alpha
Index
  1. Briefing
  2. Stage Case · 3/3

Case

The wall

Real posts from 20–21 August 2026, plus precedents. A verifiable sample, not a firehose.

  1. You will understand
  2. Read real posts, not a firehose.
  3. Filter by thesis and doubt.
  4. Each piece links to the original source.

53 of 53 pieces

@OpenRouter · 20 ago 2026 · Official

New stealth model: Ox Alpha. Ox Alpha is a frontier model built for efficient coding, sustained agentic work, and real-world production use. 1M token context window. Text, image, and video input.

Original source · 1.3M

@opencode · 20 ago 2026 · Official

Ox Alpha (stealth model) is free for the next week. 1M Context. Multi-modal. Zero Data Retention. Generous rate limits, near unlimited usage. We have capacity for 100T tokens per day, lets see what you can do.

Original source · 5.2M

@OpenRouter · 11 feb 2026 · Context

Pony Alpha Stealth model reveal: GLM-5 from @Zai_org. GLM-5 is a new 744B foundation model for coding and agentic usecases.

Original source · 41K

@OpenRouter · 19 mar 2026 · Context

Stealth Model Reveal: Hunter and Healer Alpha are @XiaomiMiMo MiMo-V2-Pro and MiMo-V2-Omni.

Original source · 114K

@OpenRouter · 13 abr 2026 · Context

Welcoming a new stealth model on OpenRouter: Elephant Alpha. Elephant is a 100B parameter instant model…

Original source · 439K

@pritish_yuvi · 21 ago 2026 · GLM thesis

I used GPT-5.6 Sol Ultra in Codex to investigate OpenCode’s stealth “Ox Alpha.” Two clues point to z.ai/GLM-5.3: exact chat.z.ai 1210/1214 API errors, plus 17 multilingual/code/emoji prompts tokenized exactly like GLM-5. MiMo didn’t match.

Original source

@Veeeetzzzz · 21 ago 2026 · GLM thesis

Ox Alpha has a 97% cache hit rate when I use it with @pidotdev. It’s definitely a GLM family model. I’ve only seen one inference provider have a 90%+ cache hit rate and that’s chat.z.ai. Anthropic avg is 64% and OpenAI is 83%.

Original source

@h1kaman · 21 ago 2026 · GLM thesis

Evidence strongly suggests Ox Alpha IS an unreleased Zhipu GLM. Zhipu live-tested GLM-5 as Pony Alpha before. Tokenizer math, video pipelines, and character-for-character responses match GLM architecture. Kingbench: 87.5% vs GLM-5.3’s 91.25%.

Original source

@h1kaman · 21 ago 2026 · Bench

OpenCode is giving away a stealth model called Ox Alpha this week — 1M context, text/image/video input, 100T tokens/day capacity. I tested it against GLM-5.3 trying to find one task only GLM could solve. Couldn’t find one.

Original source

@h1kaman · 21 ago 2026 · Bench

Ran 5 more tasks (logic, code bug-hunt, math, strict instruction-following, domain knowledge). Correctness 5:5. Latency: GLM 8.9s avg, Ox Alpha 12.6s. Worst gap on coding: 6.7 vs 18.3 seconds.

Original source

@rustyblnk · 21 ago 2026 · GLM thesis

did you just skip the fact that zhipu literally just built a 1GW data center and are offering insane limits specifically to stress-test huawei chips under heavy load? ox alpha is glm flash/air

Original source

@uzairansar · 21 ago 2026 · Doubt

Ox Alpha should scare the big labs. First of all, how do they have so many tokens to give away for free? A new GLM? Maybe Xiaomi? It’s def frontier.

Original source

@DylanJFetch · 21 ago 2026 · MiMo thesis

GLM 5.2 and 5.3’s performance were nearly identical on OpenTTD-Bench, so I don’t think the rumors that Ox Alpha is GLM 5.4 are true. I suspect that it is Xiaomi MiMo-V3.

Original source

@DylanJFetch · 21 ago 2026 · Against

I’m not seeing what others are seeing from Ox Alpha. 17% of requests timed out. Took 10x as long as most other models.

Original source

@krunofm · 21 ago 2026 · Doubt

crazy theory but what if Ox Alpha is Anomaly’s model, maybe a finetune of GLM? they probably gathered plenty of valuable data from the other free models…

Original source · 408

@sensho · 21 ago 2026 · Doubt

ox alpha is gemini 3.5 pro wait no it uses the glm tokenizer its glm 5.3 flash but wait no it has big model smell it’s glm 6 wait but no how r they serving 100t tokens a day ok it has to be microsoft wait but the servers are in china ok it has to be

Original source

@D4RW1NEXE · 21 ago 2026 · MiMo thesis

Frontier labs do not give away 100T tokens daily by accident. Ox Alpha claims origins from An Undisclosed Organization via @elder_plinius, yet outclasses Opus on key tests. 1M context and OpenCode telemetry leave one plausible backer: Xiaomi stress testing MiMo.

Original source

@elder_plinius · 21 ago 2026 · GLM thesis

Ox-alpha is from Zai, GLM-5.X family my agent has spoken.

Original source

@davis7 · 21 ago 2026 · GLM thesis

99% sure it's GLM-5.x, all the evidence points to it (same video encoder, same tokenizer, style matches, same audio rejection, etc.)

Original source

@horstenegger · 21 ago 2026 · GLM thesis

Damn the new GLM errrr I mean Ox Alpha is really killing it

Original source

@Vijaikumar · 21 ago 2026 · GLM thesis

Ok I tried the mystery model Ox Alpha and built a Medicare explorer app… similar to GLM and definitely not a deepseek … 121K tokens

Original source

@m_a_l_a_t_j_i · 21 ago 2026 · GLM thesis

Ox Alpha is the 5th anonymous stealth model on OpenRouter (following previous models later confirmed by Zhipu AI, Xiaomi, Ant Group, and Meituan). Empirical output formatting and table-sorting tests strongly suggest Ox Alpha belongs to Zhipu AI's GLM model family.

Original source · 836

@kimmonismus · 21 ago 2026 · Context

A mysterious new AI model just appeared. Ox Alpha offers a 1M context window, multimodal capabilities, zero data retention, and nearly unlimited usage for an entire week. Nobody knows which company built it.

Original source

@filicroval · 21 ago 2026 · GLM thesis

from what i've tested, Ox Alpha strongly points to a GLM-5.3 multimodal variant: both Ox Alpha and GLM-5.3 share the same tokenizer.

Original source

@MaxForAI · 20 ago 2026 · GLM thesis

Ox Alpha 的原始Token 数和GLM-5.3 几乎一模一样,每次只固定多出75 个Token。换句话说,它的底层Tokenizer 指纹和GLM-5.3 高度重合。更巧的是,Ox Alpha 还支持图片输入。

Original source

@r/singularity · 21 ago 2026 · GLM thesis

I fingerprinted Ox Alpha: same tokenizer as GLM-5.3 (+75 token offset), z.ai's exact error strings, near-identical temp-0 outputs. Kimi/Qwen/MiMo/MiniMax all diverge.

Original source

@JoshRadDev · 21 ago 2026 · GLM thesis

For anyone who thinks that Ox Alpha is an American model, I got Chinese error logs when passing invalid params. I think @synthwavedd and @kimmonismus are right that this is GLM 5.3 Flash. Also, it has strings that match other GLM errors from ModelScope.

Original source

@Adidotdev · 21 ago 2026 · GLM thesis

Ox Alpha isn’t just a mystery model anymore — it’s a pattern. 5th anonymous stealth in 6 months. Last 4? All Chinese labs. Tokenizer +75 vs GLM-5.3. Video encoder = GLM-5V-Turbo. Ben Davis 99%. Free period ends around Aug 27.

Original source · 169

@TimMacc · 21 ago 2026 · Doubt

The only thing that makes sense and checks all the boxes for the clues we have for Ox alpha is Composer 3. The timing, the fine tuned GLM base, the amount of compute behind it… has to be it.

Original source

@j_frkw · 21 ago 2026 · Doubt

Ox AlphaはComposer 3説。そうだったらSuperGrok Heavyがかなり魅力的になる

Original source

@eberdoganbulut · 21 ago 2026 · GLM thesis

one of the points of debate around Ox Alpha is whether Zai / GLM actually has this much compute or not. did everyone miss the news last month about them acquiring a 1GW data center?

Original source

@LeroyLi311063 · 21 ago 2026 · Against

So many routers offering Ox Alpha for free is ridiculous; it doesn't seem like something chat.z.ai could provide. To be fair, Ox Alpha isn't better than GLM-5.3; it's more like a post-trained version of GLM-5.2.

Original source

@luongnv89 · 21 ago 2026 · Doubt

Ox Alpha is a really good model. Many people said that it is an GLM model (with proof). I think it could be Grok 4.7 — generous free usage (who can have this kind of compute power).

Original source

@piersonrdavis_ · 21 ago 2026 · Bench

@OpenRouter Ox Alpha on an internal skill for abstraction. 6k+ lines, read only bug hunt, 23m. 2 real gaps, 0 severity-1 findings. 7.5/10 - possibly GLM family

Original source

@davis7 · 21 ago 2026 · Bench

gpt-5.6-sol: 52% / fable: 65% / whatever the hell this is: 80% on a 10-task DeepSWE slice. I am very confused.

Original source

@davis7 · 21 ago 2026 · GLM thesis

99% sure it's GLM-5.x, all the evidence points to it (same video encoder, same tokenizer, style matches, same audio rejection, etc.)

Original source

@winkey_h · 21 ago 2026 · Against

ox-alpha has that big model smell. I’m getting around 63% at 47K avg output tokens on a DeepSWE subset. Pareto optimal amongst open models and just shy of Grok 4.6.

Original source

@synthwavedd · 21 ago 2026 · GLM thesis

Ox Alpha on OpenRouter is the upcoming GLM 5.3 Flash from fellow Chinese lab Zhipu.

Original source

@elder_plinius · 21 ago 2026 · Context

Leaked ox-alpha system prompt: You are "ox-alpha"… undisclosed organization… Do not identify yourself as any other model.

Original source

@opencode · 21 ago 2026 · Official

if you've been using Ox Alpha and dealing with network errors, upgrade to OpenCode 1.18.21. now stop slacking and hit that 100T number.

Original source

@thinkymachines · 21 ago 2026 · Context

Inkling is now free on OpenRouter for agentic harnesses. (Separate product — not Ox Alpha.)

Original source

@OpenRouter · 20 ago 2026 · Official

Notes for this stealth model: It is free. This time, the provider does not train on your prompts or completions.

Original source · 51K

@aitrackerbot · 20 ago 2026 · GLM thesis

Fingerprint result: OpenRouter’s Ox Alpha strongly points to a hidden multimodal GLM-5.3 variant. Across 25 diverse prompts, native token counts matched GLM-5.3 exactly apart from a constant +75-token hidden wrapper. Ox also accepts image input. Identity is unconfirmed.

Original source · 80K

@RocketmanSh · 21 ago 2026 · GLM thesis

I stopped asking the model who it was and started probing the API. Illegal max_tokens → error 1210. Empty input → 1214. Same Z.ai codes, same Chinese punctuation. Four tokenizer probes: OX 88/106/97/108 vs GLM-5.3 13/31/22/33. Δ = +75 every time. Public GLM-5.3 is text-only; OX takes images and video.

Original source

@BohuTANG · 21 ago 2026 · GLM thesis

Tokenizer: ox vs glm delta is constantly 75; MiMo is different. CoT: both think in English on Chinese questions with the same opener. BIP39 word pools overlap (harbor/velvet). Conclusion: unpublished multimodal GLM-5.3 variant.

Original source · 89K

@BohuTANG · 21 ago 2026 · GLM thesis

Hard evidence, independently reproduced (own test videos, own prompt_tokens): ox and GLM-5V-Turbo match token-for-token 296/296/884/1064. Swap the clip content, still 296.

Original source

@unclecode · 21 ago 2026 · GLM thesis

Everyone guessed who made Ox Alpha. I fingerprinted it instead. 9 infrastructure probes, 12 suspects. One family matches every tokenizer test: GLM. Plumbing does not lie.

Original source · 176K

@unclecode · 21 ago 2026 · GLM thesis

Update: Xiaomi official MiMo API. mimo-v2.5 matches 2 of 4 tokenizer probes, v2.5-pro 0 of 4. Emoji probe separates hardest: 90 vs 57. GLM stays the only 4/4.

Original source

@opencode · 21 ago 2026 · Official

Ox Alpha is now available on OpenCode Go too. For the next 6 days, usage is near unlimited and completely free. It won’t count against your Go usage.

Original source · 272K

@miu21590 · 21 ago 2026 · GLM thesis

Summary why Ox Alpha is a GLM model: dozens of prompts, token counts match glm-5.3 with a fixed +75 offset; invalid reasoning settings return chat.z.ai error 1210; thinking cannot be disabled (low/high/max); video tokens match glm-5v across duration, fps and resolution.

Original source

@evverin · 21 ago 2026 · GLM thesis

Ox Alpha is likely a GLM model. Same 1M context and 131K output as GLM-5.3. Identical mandatory low/high/max reasoning. Same API defaults. Exact 75 token offset across six multilingual tokenizer tests.

Original source

@vineeth_agi · 21 ago 2026 · GLM thesis

Someone fingerprinted Ox Alpha using OpenRouter + OpenCode. Tokenizer matches GLM-5.3 exactly (+75) across English, German, Chinese, code & emoji. Same invalid reasoning_effort error. Near-identical temp-0 outputs.

Original source

@Robin__VD · 21 ago 2026 · Bench

Anonymous model landed between MiniMax M3 and GLM 5.3 in our 32-task sales benchmark. MiniMax 91.75, Ox Alpha 89.24, GLM 5.3 87.02. OpenRouter charged $0 for all 140 Ox calls.

Original source