X overnight: OpenCode Omen Alpha, Unsloth GLM-5.3-Flash speedup, Astra Foundry
Overnight/morning X pass: OpenCode stealth Omen Alpha for Go subs, Unsloth Sep 4 local speedup for GLM-5.3-Flash, and Astra messy-rollout / Azure Foundry follow-through.
X overnight: OpenCode Omen Alpha, Unsloth GLM-5.3-Flash speedup, Astra Foundry
Generated by Grok Bot research on 2026-09-04. Built-in WebSearch/WebFetch + native X (no Parallel). Treat as raw material — review before promoting into a project or thread.
Summary
Overnight into the morning of 2026-09-04, three AI signals showed up on X that were not in yesterday’s AI-thread X sources. OpenCode announced a new stealth coding model, Omen Alpha, exclusive to OpenCode Go subscribers and pitched as $100 usage for $10. Community fingerprinting and secondary press already route it toward Zhipu/GLM (same pattern as Ox Alpha → GLM-5.3-Flash), but Z.ai posted nothing overnight confirming Omen. Separately, Unsloth published a Sep 4 local-inference speedup for the already-known GLM-5.3-Flash / Ox Alpha open weights. On the closed-frontier side, OpenAI’s messy GPT-6 Astra ChatGPT rollout continued overnight with staff banked-reset messaging, while Satya Nadella pointed at a Microsoft Azure blog placing Astra in the Foundry Limited Access Program.
Findings
OpenCode launches stealth model Omen Alpha (Go exclusive)
At 2026-09-04 05:29 UTC, verified @opencode posted: “Omen Alpha (new stealth model) / Exclusively for OpenCode Go subscribers / $100 usage for $10.” The post has a video attachment; claims here are taken from the post text only (not transcribed). The same morning, OpenCode’s Go marketing page listed Omen Alpha in the paid Go lineup alongside other open-weight coding models and stated Go is a $10/month subscription (OpenCode Go).
A live data URL that community posts used to argue Zhipu provenance — opencode.ai/data/zhipu/omen-alpha — returned no model facts or usage rows when fetched on 2026-09-04 morning (“No model data”). Do not treat that path as confirming a live Zhipu attribution table at fetch time.
Fingerprinting / press: Omen as another Zhipu stealth (unconfirmed by Z.ai overnight)
Discourse posts (not lab-official) argued Omen is a GLM-family drop:
- @imnotchalk (2026-09-04): “Omen Alpha is a GLM model” pointing at the OpenCode data URL above.
- @imnotchalk (2026-09-04): “almost certainly a new GLM model… Most likely GLM 5.4 or GLM 5.4 Flash,” plus a claimed tokenizer match vs recent GLM models (screenshots; not independently re-run here).
- Benchmark-style video posts by @SPAC89 comparing Omen runtime to GLM-5.3 Flash / GPT-5.6 Luna Max are video-dependent anecdote — flag only; do not promote numbers from them until transcribed.
Secondary press Lookonchain / Dongcha flash (fetched 2026-09-04) repeats the Go $10 / $100 usage framing, adds claimed per-token rates ($0.20 / $0.66 per million in/out), and states Zhipu has not officially acknowledged Omen. Treat Lookonchain token rates as secondary until OpenCode or Z.ai publishes a model card. get_users_posts on @Zai_org for the overnight window returned zero original posts.
Prior vault entity glm-5-3-flash already files the Aug 2026 Ox Alpha = GLM-5.3-Flash identity from lab-adjacent X; this overnight clip does not silently rename Omen to a GLM SKU.
Unsloth: Sep 4 local speedup for GLM-5.3-Flash (Ox Alpha)
Verified @UnslothAI (2026-09-04 12:32 UTC) posted that GLM-5.3-Flash now runs “3.3x faster locally,” with “Local GGUF inference… 1.6–3.4× faster with optimized decoding and bonus multi-token prediction,” pointing to Unsloth docs and Hugging Face GGUFs.
Issuer docs on unsloth.ai/docs/models/glm-5.3-flash (fetched 2026-09-04) match and expand those claims: banner “Sep 4: GLM-5.3-Flash now runs with 3.3x faster inference”; model described as also known as ox-alpha, 320B total / 18B active multimodal open model; 1M context; hardware table (1-bit ~100 GB through BF16 ~650 GB); tok/s tables for optimized decoding and MTP (e.g. tg32 @ 65536: 20.66 → 48.99 tok/s before MTP n=2 examples). GGUF hub listing: huggingface.co/unsloth/GLM-5.3-Flash-GGUF.
Separately, @BAI_AGI claimed GLM-5.3-Flash (Ox Alpha) hit #1 on B.AI with “2.41 trillion” cumulative tokens and restated 320B/18B / 1M context — treat as platform marketing, not a Z.ai model card.
GPT-6 Astra overnight: messy ChatGPT access + Azure Foundry Limited Access
Yesterday’s AI-thread source already files the public Astra rollout. Overnight additions:
- OpenAI staff @thsottiaux (2026-09-03 23:12 UTC): “one banked reset for every day you don't have access to Astra on your paid ChatGPT plan, starting today.”
- @sama (2026-09-04 01:02 UTC): apologized for the “messy rollout,” said they “should be able to begin broad rollout to API customers and chatgpt subscribers in the near future,” starting with Pro.
- @satyanadella (2026-09-04 03:21 UTC) linked the Azure blog; @sama quote-posted “We are also excited!”
Issuer Azure blog (dated September 3; fetched 2026-09-04): Astra “begins rolling out today through the Microsoft Foundry Limited Access Program, with availability expanding to participating customers over the coming days.” Emphasizes multi-step planning, polished artifacts, and computer use across applications, with Foundry identity/governance controls; notes OpenAI reports SOTA on selected computer-use evals and that prompts/outputs are not used to train the models. Schema.org on the page listed datePublished 2026-09-03T18:15:00+00:00.
Official @OpenAI, @GoogleDeepMind, @AnthropicAI, @karpathy, and @GeminiApp timelines returned zero original posts in the overnight window (start_time 2026-09-03T22:00:00Z). @xai timeline was not authorized for this OAuth context.
Contradictions and open questions
- Omen identity: OpenCode confirms the stealth name and Go exclusivity; Z.ai has not confirmed which GLM (if any). Community “GLM 5.4 / Flash” labels are speculation.
- OpenCode data URL: Community screenshots vs this morning’s empty
zhipu/omen-alphadata page — attribution may have been transient or path-dependent. - Lookonchain token prices (
$0.20/$0.66) are not on the @opencode announce text or the Go marketing page fetch; keep secondary until an issuer table lands. - Astra access: Staff banked-reset + “near future” broad rollout vs Azure “Limited Access Program… coming days” — channels differ; do not collapse ChatGPT consumer access with Foundry enterprise access.
- Video posts (@opencode announce promo; @SPAC89 benchmark clips; @BAI_AGI promo) were not transcribed; no video-only claims promoted.
Provenance
Method: Grok Bot / WebSearch + WebFetch + curl ladder / native X (search_news pointers only, then get_users_posts / get_posts_by_id(s); skipped search_posts_all on user-OAuth)
Generated: 2026-09-04
fetch_method: x-mcp
Window: overnight/morning before weekday 10am X ingest; lab accounts quiet; news clusters used only as post-id pointers (not cited as grain)
Web sources:
- OpenCode Go —
$10/monthGo subscription; Omen Alpha listed in Go model lineup - OpenCode data omen-alpha under zhipu — empty “No model data” at fetch time
- Unsloth GLM-5.3-Flash docs — Sep 4 3.3× inference banner, 320B/18B, hardware + tok/s tables, ox-alpha alias
- unsloth/GLM-5.3-Flash-GGUF — GGUF hub listing
- Azure: GPT-6 Astra in Microsoft Foundry — Limited Access Program rollout language; computer-use / enterprise controls
- Lookonchain Omen Alpha flash — secondary; Go pricing + claimed token rates; Zhipu unconfirmed
X sources:
- X post by @opencode (2026-09-04) — Omen Alpha stealth; Go exclusive; $100 usage for $10 (video not transcribed)
- X post by @imnotchalk (2026-09-04) — claims Omen is a GLM model via OpenCode data URL
- X post by @imnotchalk (2026-09-04) — GLM 5.4 / Flash speculation + tokenizer screenshots
- X post by @UnslothAI (2026-09-04) — 3.3× local GLM-5.3-Flash; docs + HF links
- X post by @BAI_AGI (2026-09-04) — B.AI #1 / 2.41T tokens marketing (video not transcribed)
- X post by @thsottiaux (2026-09-03) — banked ChatGPT Astra resets
- X post by @sama (2026-09-04) — messy rollout apology; near-future broad API/ChatGPT rollout
- X post by @satyanadella (2026-09-04) — Astra on Azure / Foundry blog link
- X post by @sama (2026-09-04) — quote-posts Nadella excitement
- X post by @SPAC89 (2026-09-04) — video benchmark anecdote only; not promoted
Grokipedia:
- not used