- Briefing
- Etapa Caso · 3/3
Caso
El muro
Tuiteros y posts reales del 20–21 de agosto de 2026, más precedentes. Muestra verificable, no un firehose.
- Vas a entender
- Leer posts reales, no un firehose.
- Filtrar por tesis y duda.
- Cada pieza enlaza a la fuente original.
53 de 53 piezas
@OpenRouter · 20 ago 2026 · Oficial
New stealth model: Ox Alpha. Ox Alpha is a frontier model built for efficient coding, sustained agentic work, and real-world production use. 1M token context window. Text, image, and video input.
Fuente original · 1.3M
@opencode · 20 ago 2026 · Oficial
Ox Alpha (stealth model) is free for the next week. 1M Context. Multi-modal. Zero Data Retention. Generous rate limits, near unlimited usage. We have capacity for 100T tokens per day, lets see what you can do.
Fuente original · 5.2M
@OpenRouter · 11 feb 2026 · Contexto
Pony Alpha Stealth model reveal: GLM-5 from @Zai_org. GLM-5 is a new 744B foundation model for coding and agentic usecases.
Fuente original · 41K
@OpenRouter · 19 mar 2026 · Contexto
Stealth Model Reveal: Hunter and Healer Alpha are @XiaomiMiMo MiMo-V2-Pro and MiMo-V2-Omni.
Fuente original · 114K
@OpenRouter · 13 abr 2026 · Contexto
Welcoming a new stealth model on OpenRouter: Elephant Alpha. Elephant is a 100B parameter instant model…
Fuente original · 439K
@pritish_yuvi · 21 ago 2026 · Tesis GLM
I used GPT-5.6 Sol Ultra in Codex to investigate OpenCode’s stealth “Ox Alpha.” Two clues point to z.ai/GLM-5.3: exact chat.z.ai 1210/1214 API errors, plus 17 multilingual/code/emoji prompts tokenized exactly like GLM-5. MiMo didn’t match.
@Veeeetzzzz · 21 ago 2026 · Tesis GLM
Ox Alpha has a 97% cache hit rate when I use it with @pidotdev. It’s definitely a GLM family model. I’ve only seen one inference provider have a 90%+ cache hit rate and that’s chat.z.ai. Anthropic avg is 64% and OpenAI is 83%.
@h1kaman · 21 ago 2026 · Tesis GLM
Evidence strongly suggests Ox Alpha IS an unreleased Zhipu GLM. Zhipu live-tested GLM-5 as Pony Alpha before. Tokenizer math, video pipelines, and character-for-character responses match GLM architecture. Kingbench: 87.5% vs GLM-5.3’s 91.25%.
@h1kaman · 21 ago 2026 · Bench
OpenCode is giving away a stealth model called Ox Alpha this week — 1M context, text/image/video input, 100T tokens/day capacity. I tested it against GLM-5.3 trying to find one task only GLM could solve. Couldn’t find one.
@h1kaman · 21 ago 2026 · Bench
Ran 5 more tasks (logic, code bug-hunt, math, strict instruction-following, domain knowledge). Correctness 5:5. Latency: GLM 8.9s avg, Ox Alpha 12.6s. Worst gap on coding: 6.7 vs 18.3 seconds.
@rustyblnk · 21 ago 2026 · Tesis GLM
did you just skip the fact that zhipu literally just built a 1GW data center and are offering insane limits specifically to stress-test huawei chips under heavy load? ox alpha is glm flash/air
@uzairansar · 21 ago 2026 · Duda
Ox Alpha should scare the big labs. First of all, how do they have so many tokens to give away for free? A new GLM? Maybe Xiaomi? It’s def frontier.
@DylanJFetch · 21 ago 2026 · Tesis MiMo
GLM 5.2 and 5.3’s performance were nearly identical on OpenTTD-Bench, so I don’t think the rumors that Ox Alpha is GLM 5.4 are true. I suspect that it is Xiaomi MiMo-V3.
@DylanJFetch · 21 ago 2026 · Contra
I’m not seeing what others are seeing from Ox Alpha. 17% of requests timed out. Took 10x as long as most other models.
@krunofm · 21 ago 2026 · Duda
crazy theory but what if Ox Alpha is Anomaly’s model, maybe a finetune of GLM? they probably gathered plenty of valuable data from the other free models…
Fuente original · 408
@sensho · 21 ago 2026 · Duda
ox alpha is gemini 3.5 pro wait no it uses the glm tokenizer its glm 5.3 flash but wait no it has big model smell it’s glm 6 wait but no how r they serving 100t tokens a day ok it has to be microsoft wait but the servers are in china ok it has to be
@D4RW1NEXE · 21 ago 2026 · Tesis MiMo
Frontier labs do not give away 100T tokens daily by accident. Ox Alpha claims origins from An Undisclosed Organization via @elder_plinius, yet outclasses Opus on key tests. 1M context and OpenCode telemetry leave one plausible backer: Xiaomi stress testing MiMo.
@elder_plinius · 21 ago 2026 · Tesis GLM
Ox-alpha is from Zai, GLM-5.X family my agent has spoken.
@davis7 · 21 ago 2026 · Tesis GLM
99% sure it's GLM-5.x, all the evidence points to it (same video encoder, same tokenizer, style matches, same audio rejection, etc.)
@horstenegger · 21 ago 2026 · Tesis GLM
Damn the new GLM errrr I mean Ox Alpha is really killing it
@Vijaikumar · 21 ago 2026 · Tesis GLM
Ok I tried the mystery model Ox Alpha and built a Medicare explorer app… similar to GLM and definitely not a deepseek … 121K tokens
@m_a_l_a_t_j_i · 21 ago 2026 · Tesis GLM
Ox Alpha is the 5th anonymous stealth model on OpenRouter (following previous models later confirmed by Zhipu AI, Xiaomi, Ant Group, and Meituan). Empirical output formatting and table-sorting tests strongly suggest Ox Alpha belongs to Zhipu AI's GLM model family.
Fuente original · 836
@kimmonismus · 21 ago 2026 · Contexto
A mysterious new AI model just appeared. Ox Alpha offers a 1M context window, multimodal capabilities, zero data retention, and nearly unlimited usage for an entire week. Nobody knows which company built it.
@filicroval · 21 ago 2026 · Tesis GLM
from what i've tested, Ox Alpha strongly points to a GLM-5.3 multimodal variant: both Ox Alpha and GLM-5.3 share the same tokenizer.
@MaxForAI · 20 ago 2026 · Tesis GLM
Ox Alpha 的原始Token 数和GLM-5.3 几乎一模一样,每次只固定多出75 个Token。换句话说,它的底层Tokenizer 指纹和GLM-5.3 高度重合。更巧的是,Ox Alpha 还支持图片输入。
@r/singularity · 21 ago 2026 · Tesis GLM
I fingerprinted Ox Alpha: same tokenizer as GLM-5.3 (+75 token offset), z.ai's exact error strings, near-identical temp-0 outputs. Kimi/Qwen/MiMo/MiniMax all diverge.
@JoshRadDev · 21 ago 2026 · Tesis GLM
For anyone who thinks that Ox Alpha is an American model, I got Chinese error logs when passing invalid params. I think @synthwavedd and @kimmonismus are right that this is GLM 5.3 Flash. Also, it has strings that match other GLM errors from ModelScope.
@Adidotdev · 21 ago 2026 · Tesis GLM
Ox Alpha isn’t just a mystery model anymore — it’s a pattern. 5th anonymous stealth in 6 months. Last 4? All Chinese labs. Tokenizer +75 vs GLM-5.3. Video encoder = GLM-5V-Turbo. Ben Davis 99%. Free period ends around Aug 27.
Fuente original · 169
@TimMacc · 21 ago 2026 · Duda
The only thing that makes sense and checks all the boxes for the clues we have for Ox alpha is Composer 3. The timing, the fine tuned GLM base, the amount of compute behind it… has to be it.
@j_frkw · 21 ago 2026 · Duda
Ox AlphaはComposer 3説。そうだったらSuperGrok Heavyがかなり魅力的になる
@eberdoganbulut · 21 ago 2026 · Tesis GLM
one of the points of debate around Ox Alpha is whether Zai / GLM actually has this much compute or not. did everyone miss the news last month about them acquiring a 1GW data center?
@LeroyLi311063 · 21 ago 2026 · Contra
So many routers offering Ox Alpha for free is ridiculous; it doesn't seem like something chat.z.ai could provide. To be fair, Ox Alpha isn't better than GLM-5.3; it's more like a post-trained version of GLM-5.2.
@luongnv89 · 21 ago 2026 · Duda
Ox Alpha is a really good model. Many people said that it is an GLM model (with proof). I think it could be Grok 4.7 — generous free usage (who can have this kind of compute power).
@piersonrdavis_ · 21 ago 2026 · Bench
@OpenRouter Ox Alpha on an internal skill for abstraction. 6k+ lines, read only bug hunt, 23m. 2 real gaps, 0 severity-1 findings. 7.5/10 - possibly GLM family
@davis7 · 21 ago 2026 · Bench
gpt-5.6-sol: 52% / fable: 65% / whatever the hell this is: 80% on a 10-task DeepSWE slice. I am very confused.
@davis7 · 21 ago 2026 · Tesis GLM
99% sure it's GLM-5.x, all the evidence points to it (same video encoder, same tokenizer, style matches, same audio rejection, etc.)
@winkey_h · 21 ago 2026 · Contra
ox-alpha has that big model smell. I’m getting around 63% at 47K avg output tokens on a DeepSWE subset. Pareto optimal amongst open models and just shy of Grok 4.6.
@synthwavedd · 21 ago 2026 · Tesis GLM
Ox Alpha on OpenRouter is the upcoming GLM 5.3 Flash from fellow Chinese lab Zhipu.
@elder_plinius · 21 ago 2026 · Contexto
Leaked ox-alpha system prompt: You are "ox-alpha"… undisclosed organization… Do not identify yourself as any other model.
@opencode · 21 ago 2026 · Oficial
if you've been using Ox Alpha and dealing with network errors, upgrade to OpenCode 1.18.21. now stop slacking and hit that 100T number.
@thinkymachines · 21 ago 2026 · Contexto
Inkling is now free on OpenRouter for agentic harnesses. (Separate product — not Ox Alpha.)
@OpenRouter · 20 ago 2026 · Oficial
Notes for this stealth model: It is free. This time, the provider does not train on your prompts or completions.
Fuente original · 51K
@aitrackerbot · 20 ago 2026 · Tesis GLM
Fingerprint result: OpenRouter’s Ox Alpha strongly points to a hidden multimodal GLM-5.3 variant. Across 25 diverse prompts, native token counts matched GLM-5.3 exactly apart from a constant +75-token hidden wrapper. Ox also accepts image input. Identity is unconfirmed.
Fuente original · 80K
@RocketmanSh · 21 ago 2026 · Tesis GLM
I stopped asking the model who it was and started probing the API. Illegal max_tokens → error 1210. Empty input → 1214. Same Z.ai codes, same Chinese punctuation. Four tokenizer probes: OX 88/106/97/108 vs GLM-5.3 13/31/22/33. Δ = +75 every time. Public GLM-5.3 is text-only; OX takes images and video.
@BohuTANG · 21 ago 2026 · Tesis GLM
Tokenizer: ox vs glm delta is constantly 75; MiMo is different. CoT: both think in English on Chinese questions with the same opener. BIP39 word pools overlap (harbor/velvet). Conclusion: unpublished multimodal GLM-5.3 variant.
Fuente original · 89K
@BohuTANG · 21 ago 2026 · Tesis GLM
Hard evidence, independently reproduced (own test videos, own prompt_tokens): ox and GLM-5V-Turbo match token-for-token 296/296/884/1064. Swap the clip content, still 296.
@unclecode · 21 ago 2026 · Tesis GLM
Everyone guessed who made Ox Alpha. I fingerprinted it instead. 9 infrastructure probes, 12 suspects. One family matches every tokenizer test: GLM. Plumbing does not lie.
Fuente original · 176K
@unclecode · 21 ago 2026 · Tesis GLM
Update: Xiaomi official MiMo API. mimo-v2.5 matches 2 of 4 tokenizer probes, v2.5-pro 0 of 4. Emoji probe separates hardest: 90 vs 57. GLM stays the only 4/4.
@opencode · 21 ago 2026 · Oficial
Ox Alpha is now available on OpenCode Go too. For the next 6 days, usage is near unlimited and completely free. It won’t count against your Go usage.
Fuente original · 272K
@miu21590 · 21 ago 2026 · Tesis GLM
Summary why Ox Alpha is a GLM model: dozens of prompts, token counts match glm-5.3 with a fixed +75 offset; invalid reasoning settings return chat.z.ai error 1210; thinking cannot be disabled (low/high/max); video tokens match glm-5v across duration, fps and resolution.
@evverin · 21 ago 2026 · Tesis GLM
Ox Alpha is likely a GLM model. Same 1M context and 131K output as GLM-5.3. Identical mandatory low/high/max reasoning. Same API defaults. Exact 75 token offset across six multilingual tokenizer tests.
@vineeth_agi · 21 ago 2026 · Tesis GLM
Someone fingerprinted Ox Alpha using OpenRouter + OpenCode. Tokenizer matches GLM-5.3 exactly (+75) across English, German, Chinese, code & emoji. Same invalid reasoning_effort error. Near-identical temp-0 outputs.
@Robin__VD · 21 ago 2026 · Bench
Anonymous model landed between MiniMax M3 and GLM 5.3 in our 32-task sales benchmark. MiniMax 91.75, Ox Alpha 89.24, GLM 5.3 87.02. OpenRouter charged $0 for all 140 Ox calls.