Saltar al contenido
ChristopherOx Alpha
Índice
  1. Briefing
  2. Etapa Caso · 3/3

Caso

El muro

Tuiteros y posts reales del 20–21 de agosto de 2026, más precedentes. Muestra verificable, no un firehose.

  1. Vas a entender
  2. Leer posts reales, no un firehose.
  3. Filtrar por tesis y duda.
  4. Cada pieza enlaza a la fuente original.

53 de 53 piezas

@OpenRouter · 20 ago 2026 · Oficial

New stealth model: Ox Alpha. Ox Alpha is a frontier model built for efficient coding, sustained agentic work, and real-world production use. 1M token context window. Text, image, and video input.

Fuente original · 1.3M

@opencode · 20 ago 2026 · Oficial

Ox Alpha (stealth model) is free for the next week. 1M Context. Multi-modal. Zero Data Retention. Generous rate limits, near unlimited usage. We have capacity for 100T tokens per day, lets see what you can do.

Fuente original · 5.2M

@OpenRouter · 11 feb 2026 · Contexto

Pony Alpha Stealth model reveal: GLM-5 from @Zai_org. GLM-5 is a new 744B foundation model for coding and agentic usecases.

Fuente original · 41K

@OpenRouter · 19 mar 2026 · Contexto

Stealth Model Reveal: Hunter and Healer Alpha are @XiaomiMiMo MiMo-V2-Pro and MiMo-V2-Omni.

Fuente original · 114K

@OpenRouter · 13 abr 2026 · Contexto

Welcoming a new stealth model on OpenRouter: Elephant Alpha. Elephant is a 100B parameter instant model…

Fuente original · 439K

@pritish_yuvi · 21 ago 2026 · Tesis GLM

I used GPT-5.6 Sol Ultra in Codex to investigate OpenCode’s stealth “Ox Alpha.” Two clues point to z.ai/GLM-5.3: exact chat.z.ai 1210/1214 API errors, plus 17 multilingual/code/emoji prompts tokenized exactly like GLM-5. MiMo didn’t match.

Fuente original

@Veeeetzzzz · 21 ago 2026 · Tesis GLM

Ox Alpha has a 97% cache hit rate when I use it with @pidotdev. It’s definitely a GLM family model. I’ve only seen one inference provider have a 90%+ cache hit rate and that’s chat.z.ai. Anthropic avg is 64% and OpenAI is 83%.

Fuente original

@h1kaman · 21 ago 2026 · Tesis GLM

Evidence strongly suggests Ox Alpha IS an unreleased Zhipu GLM. Zhipu live-tested GLM-5 as Pony Alpha before. Tokenizer math, video pipelines, and character-for-character responses match GLM architecture. Kingbench: 87.5% vs GLM-5.3’s 91.25%.

Fuente original

@h1kaman · 21 ago 2026 · Bench

OpenCode is giving away a stealth model called Ox Alpha this week — 1M context, text/image/video input, 100T tokens/day capacity. I tested it against GLM-5.3 trying to find one task only GLM could solve. Couldn’t find one.

Fuente original

@h1kaman · 21 ago 2026 · Bench

Ran 5 more tasks (logic, code bug-hunt, math, strict instruction-following, domain knowledge). Correctness 5:5. Latency: GLM 8.9s avg, Ox Alpha 12.6s. Worst gap on coding: 6.7 vs 18.3 seconds.

Fuente original

@rustyblnk · 21 ago 2026 · Tesis GLM

did you just skip the fact that zhipu literally just built a 1GW data center and are offering insane limits specifically to stress-test huawei chips under heavy load? ox alpha is glm flash/air

Fuente original

@uzairansar · 21 ago 2026 · Duda

Ox Alpha should scare the big labs. First of all, how do they have so many tokens to give away for free? A new GLM? Maybe Xiaomi? It’s def frontier.

Fuente original

@DylanJFetch · 21 ago 2026 · Tesis MiMo

GLM 5.2 and 5.3’s performance were nearly identical on OpenTTD-Bench, so I don’t think the rumors that Ox Alpha is GLM 5.4 are true. I suspect that it is Xiaomi MiMo-V3.

Fuente original

@DylanJFetch · 21 ago 2026 · Contra

I’m not seeing what others are seeing from Ox Alpha. 17% of requests timed out. Took 10x as long as most other models.

Fuente original

@krunofm · 21 ago 2026 · Duda

crazy theory but what if Ox Alpha is Anomaly’s model, maybe a finetune of GLM? they probably gathered plenty of valuable data from the other free models…

Fuente original · 408

@sensho · 21 ago 2026 · Duda

ox alpha is gemini 3.5 pro wait no it uses the glm tokenizer its glm 5.3 flash but wait no it has big model smell it’s glm 6 wait but no how r they serving 100t tokens a day ok it has to be microsoft wait but the servers are in china ok it has to be

Fuente original

@D4RW1NEXE · 21 ago 2026 · Tesis MiMo

Frontier labs do not give away 100T tokens daily by accident. Ox Alpha claims origins from An Undisclosed Organization via @elder_plinius, yet outclasses Opus on key tests. 1M context and OpenCode telemetry leave one plausible backer: Xiaomi stress testing MiMo.

Fuente original

@elder_plinius · 21 ago 2026 · Tesis GLM

Ox-alpha is from Zai, GLM-5.X family my agent has spoken.

Fuente original

@davis7 · 21 ago 2026 · Tesis GLM

99% sure it's GLM-5.x, all the evidence points to it (same video encoder, same tokenizer, style matches, same audio rejection, etc.)

Fuente original

@horstenegger · 21 ago 2026 · Tesis GLM

Damn the new GLM errrr I mean Ox Alpha is really killing it

Fuente original

@Vijaikumar · 21 ago 2026 · Tesis GLM

Ok I tried the mystery model Ox Alpha and built a Medicare explorer app… similar to GLM and definitely not a deepseek … 121K tokens

Fuente original

@m_a_l_a_t_j_i · 21 ago 2026 · Tesis GLM

Ox Alpha is the 5th anonymous stealth model on OpenRouter (following previous models later confirmed by Zhipu AI, Xiaomi, Ant Group, and Meituan). Empirical output formatting and table-sorting tests strongly suggest Ox Alpha belongs to Zhipu AI's GLM model family.

Fuente original · 836

@kimmonismus · 21 ago 2026 · Contexto

A mysterious new AI model just appeared. Ox Alpha offers a 1M context window, multimodal capabilities, zero data retention, and nearly unlimited usage for an entire week. Nobody knows which company built it.

Fuente original

@filicroval · 21 ago 2026 · Tesis GLM

from what i've tested, Ox Alpha strongly points to a GLM-5.3 multimodal variant: both Ox Alpha and GLM-5.3 share the same tokenizer.

Fuente original

@MaxForAI · 20 ago 2026 · Tesis GLM

Ox Alpha 的原始Token 数和GLM-5.3 几乎一模一样,每次只固定多出75 个Token。换句话说,它的底层Tokenizer 指纹和GLM-5.3 高度重合。更巧的是,Ox Alpha 还支持图片输入。

Fuente original

@r/singularity · 21 ago 2026 · Tesis GLM

I fingerprinted Ox Alpha: same tokenizer as GLM-5.3 (+75 token offset), z.ai's exact error strings, near-identical temp-0 outputs. Kimi/Qwen/MiMo/MiniMax all diverge.

Fuente original

@JoshRadDev · 21 ago 2026 · Tesis GLM

For anyone who thinks that Ox Alpha is an American model, I got Chinese error logs when passing invalid params. I think @synthwavedd and @kimmonismus are right that this is GLM 5.3 Flash. Also, it has strings that match other GLM errors from ModelScope.

Fuente original

@Adidotdev · 21 ago 2026 · Tesis GLM

Ox Alpha isn’t just a mystery model anymore — it’s a pattern. 5th anonymous stealth in 6 months. Last 4? All Chinese labs. Tokenizer +75 vs GLM-5.3. Video encoder = GLM-5V-Turbo. Ben Davis 99%. Free period ends around Aug 27.

Fuente original · 169

@TimMacc · 21 ago 2026 · Duda

The only thing that makes sense and checks all the boxes for the clues we have for Ox alpha is Composer 3. The timing, the fine tuned GLM base, the amount of compute behind it… has to be it.

Fuente original

@j_frkw · 21 ago 2026 · Duda

Ox AlphaはComposer 3説。そうだったらSuperGrok Heavyがかなり魅力的になる

Fuente original

@eberdoganbulut · 21 ago 2026 · Tesis GLM

one of the points of debate around Ox Alpha is whether Zai / GLM actually has this much compute or not. did everyone miss the news last month about them acquiring a 1GW data center?

Fuente original

@LeroyLi311063 · 21 ago 2026 · Contra

So many routers offering Ox Alpha for free is ridiculous; it doesn't seem like something chat.z.ai could provide. To be fair, Ox Alpha isn't better than GLM-5.3; it's more like a post-trained version of GLM-5.2.

Fuente original

@luongnv89 · 21 ago 2026 · Duda

Ox Alpha is a really good model. Many people said that it is an GLM model (with proof). I think it could be Grok 4.7 — generous free usage (who can have this kind of compute power).

Fuente original

@piersonrdavis_ · 21 ago 2026 · Bench

@OpenRouter Ox Alpha on an internal skill for abstraction. 6k+ lines, read only bug hunt, 23m. 2 real gaps, 0 severity-1 findings. 7.5/10 - possibly GLM family

Fuente original

@davis7 · 21 ago 2026 · Bench

gpt-5.6-sol: 52% / fable: 65% / whatever the hell this is: 80% on a 10-task DeepSWE slice. I am very confused.

Fuente original

@davis7 · 21 ago 2026 · Tesis GLM

99% sure it's GLM-5.x, all the evidence points to it (same video encoder, same tokenizer, style matches, same audio rejection, etc.)

Fuente original

@winkey_h · 21 ago 2026 · Contra

ox-alpha has that big model smell. I’m getting around 63% at 47K avg output tokens on a DeepSWE subset. Pareto optimal amongst open models and just shy of Grok 4.6.

Fuente original

@synthwavedd · 21 ago 2026 · Tesis GLM

Ox Alpha on OpenRouter is the upcoming GLM 5.3 Flash from fellow Chinese lab Zhipu.

Fuente original

@elder_plinius · 21 ago 2026 · Contexto

Leaked ox-alpha system prompt: You are "ox-alpha"… undisclosed organization… Do not identify yourself as any other model.

Fuente original

@opencode · 21 ago 2026 · Oficial

if you've been using Ox Alpha and dealing with network errors, upgrade to OpenCode 1.18.21. now stop slacking and hit that 100T number.

Fuente original

@thinkymachines · 21 ago 2026 · Contexto

Inkling is now free on OpenRouter for agentic harnesses. (Separate product — not Ox Alpha.)

Fuente original

@OpenRouter · 20 ago 2026 · Oficial

Notes for this stealth model: It is free. This time, the provider does not train on your prompts or completions.

Fuente original · 51K

@aitrackerbot · 20 ago 2026 · Tesis GLM

Fingerprint result: OpenRouter’s Ox Alpha strongly points to a hidden multimodal GLM-5.3 variant. Across 25 diverse prompts, native token counts matched GLM-5.3 exactly apart from a constant +75-token hidden wrapper. Ox also accepts image input. Identity is unconfirmed.

Fuente original · 80K

@RocketmanSh · 21 ago 2026 · Tesis GLM

I stopped asking the model who it was and started probing the API. Illegal max_tokens → error 1210. Empty input → 1214. Same Z.ai codes, same Chinese punctuation. Four tokenizer probes: OX 88/106/97/108 vs GLM-5.3 13/31/22/33. Δ = +75 every time. Public GLM-5.3 is text-only; OX takes images and video.

Fuente original

@BohuTANG · 21 ago 2026 · Tesis GLM

Tokenizer: ox vs glm delta is constantly 75; MiMo is different. CoT: both think in English on Chinese questions with the same opener. BIP39 word pools overlap (harbor/velvet). Conclusion: unpublished multimodal GLM-5.3 variant.

Fuente original · 89K

@BohuTANG · 21 ago 2026 · Tesis GLM

Hard evidence, independently reproduced (own test videos, own prompt_tokens): ox and GLM-5V-Turbo match token-for-token 296/296/884/1064. Swap the clip content, still 296.

Fuente original

@unclecode · 21 ago 2026 · Tesis GLM

Everyone guessed who made Ox Alpha. I fingerprinted it instead. 9 infrastructure probes, 12 suspects. One family matches every tokenizer test: GLM. Plumbing does not lie.

Fuente original · 176K

@unclecode · 21 ago 2026 · Tesis GLM

Update: Xiaomi official MiMo API. mimo-v2.5 matches 2 of 4 tokenizer probes, v2.5-pro 0 of 4. Emoji probe separates hardest: 90 vs 57. GLM stays the only 4/4.

Fuente original

@opencode · 21 ago 2026 · Oficial

Ox Alpha is now available on OpenCode Go too. For the next 6 days, usage is near unlimited and completely free. It won’t count against your Go usage.

Fuente original · 272K

@miu21590 · 21 ago 2026 · Tesis GLM

Summary why Ox Alpha is a GLM model: dozens of prompts, token counts match glm-5.3 with a fixed +75 offset; invalid reasoning settings return chat.z.ai error 1210; thinking cannot be disabled (low/high/max); video tokens match glm-5v across duration, fps and resolution.

Fuente original

@evverin · 21 ago 2026 · Tesis GLM

Ox Alpha is likely a GLM model. Same 1M context and 131K output as GLM-5.3. Identical mandatory low/high/max reasoning. Same API defaults. Exact 75 token offset across six multilingual tokenizer tests.

Fuente original

@vineeth_agi · 21 ago 2026 · Tesis GLM

Someone fingerprinted Ox Alpha using OpenRouter + OpenCode. Tokenizer matches GLM-5.3 exactly (+75) across English, German, Chinese, code & emoji. Same invalid reasoning_effort error. Near-identical temp-0 outputs.

Fuente original

@Robin__VD · 21 ago 2026 · Bench

Anonymous model landed between MiniMax M3 and GLM 5.3 in our 32-task sales benchmark. MiniMax 91.75, Ox Alpha 89.24, GLM 5.3 87.02. OpenRouter charged $0 for all 140 Ox calls.

Fuente original