Model comparison · lineage and evidence

Omen Alpha vs Ox Alpha

Compare the two stealth AI coding models side by side: Omen Alpha (likely 2.0 successor) and Ox Alpha (1.0 / GLM-5.3-Flash). Pricing, access, context, benchmark evidence, and which one fits your coding or agentic workflow.

Evidence standard: public + community-reportedReviewed: September 5, 202615 comparison points
01 / Quick answer

Omen Alpha is the likely 2.0 successor; test before switching.

Omen Alpha is not literally the same model as Ox Alpha. The community position is that Omen Alpha is the newer, separately distributed stealth model — a likely 2.0 successor — with different observed pricing, behavior, context, and benchmark results. Ox Alpha was later identified as Z.ai’s GLM-5.3-Flash; Omen Alpha remains unconfirmed, though evidence points toward a GLM-5.x lineage.

Decision rule: if you already use Ox Alpha, keep your current workflow and evaluate Omen Alpha in parallel on representative prompts. If you are choosing a new model, test both active routes. This page narrows the choice; it does not replace a controlled test on your codebase.
02 / The relationship

Two stealth models, one likely lineage.

Ox Alpha appeared first as an anonymous coding model with a reported 1M-token context and distribution through OpenRouter and OpenCode. It was later identified as Z.ai GLM-5.3-Flash. See the Wavect account for background.

Omen Alpha appeared around September 4, 2026 — a separate model with different pricing, a reported ~500K-token context, and the OpenCode vendor path zhipu/omen-alpha. Behavioral fingerprints support a GLM-family hypothesis, but no official Z.ai statement confirms Omen Alpha’s identity.

The evidence supports this shorthand: Ox Alpha = earlier model (GLM-5.3-Flash); Omen Alpha = likely 2.0 successor (possible GLM-5.x variant, unconfirmed).

03 / Side-by-side

Omen Alpha vs Ox Alpha: full comparison.

The following values combine public listings and community reports. Items marked as community-reported are not vendor specifications.

Omen Alpha vs Ox Alpha — evidence-scoped comparison
AttributeOmen Alpha (likely 2.0)Ox Alpha (earlier model)
StatusActive stealth previewEarlier stealth model; identified as GLM-5.3-Flash
Likely lineagePossible Z.ai / GLM-5.x variant; unconfirmedZ.ai GLM-5.3-Flash
Input price$0.20 / 1M tokensTokenra: $0.08 · 60% off$0.075 / 1M tokens
Output price$0.66 / 1M tokensTokenra: $0.26 · 60% off$0.25 / 1M tokens
Cached input$0.04 / 1M tokens$0.015 / 1M tokens
Context window~500K tokens (community/platform-reported)1M tokens (public model listings)
Max output128K tokens (platform-reported)128K tokens
Input modalitiesText + image (community/platform-reported)Text + image + video (public listings)
ReasoningSupported (platform-reported)Supported
API formatOpenAI-compatible Chat CompletionsOpenAI-compatible Chat Completions
Model IDomen-alphastealth/ox-alpha
Primary accessOpenCode Go · tokenra.ioOpenRouter · OpenCode · oxalpha.io
Subscription / offerOpenCode Go: reported $10/mo with $100 Omen Alpha creditPay per use; oxalpha.io lists pricing at roughly 40% of reference rates
Benchmark23.14/40 in the dated OpenCode snapshot →No directly comparable same-run benchmark available
PrivacyNot used for training · stated 0-day retentionCheck the active provider’s terms
Vendor confirmed?No — no company has publicly claimed itYes — identified as Z.ai GLM-5.3-Flash
Evidence standard reminder. Community-reported values are not vendor-confirmed specifications. Verify current pricing, limits, privacy terms, and route availability before committing to production.
04 / Pricing

What each model costs right now.

Omen Alpha lists at $0.20 input / $0.66 output per million tokens, with Tokenra offering a 60% discount at $0.08 / $0.26. OpenCode Go is reported at $10 per month with $100 in monthly Omen Alpha credit. See full pricing details →

Ox Alpha lists at $0.075 input, $0.25 output, and $0.015 cached input per million tokens. The oxalpha.io offer is described as roughly 40% of reference pricing. Check the active route before budgeting.

Pricing is provider-controlled. Promotions, credits, and route availability can change independently of the model. Always confirm live terms.
05 / Benchmark evidence

What the available numbers say — and don’t say.

In the dated OpenCode snapshot, Omen Alpha scored 23.14 / 40, with an average cost of $0.03 per prompt and average time of 01:51.

A direct numerical comparison with Ox Alpha is not currently defensible:

  • The available Ox Alpha results use different task sets, dates, or evaluators.
  • Labels such as ox-alpha-max on other leaderboards belong to separate measurement contexts.
  • A fair comparison requires the same prompts, evaluator, settings, and date.

Read the full Omen Alpha benchmark methodology →

06 / Decision guide

When should you choose each model?

START WITH OMEN ALPHA

For the active release.

Choose it when you want the newer stealth model, current OpenCode Go access, or Tokenra’s discounted route.

START WITH OX ALPHA

For known lineage.

Choose it when the confirmed GLM-5.3-Flash identity, reported 1M context, or lower listed rate matters more.

Practical starting points
If you prioritizeStart withWhy
Newest stealth releaseOmen AlphaIt is the active newer release with current OpenCode Go and Tokenra access.
Larger reported contextOx AlphaIts public listings report 1M versus Omen Alpha’s reported ~500K.
Video inputOx AlphaPublic listings describe text, image, and video input.
Lowest listed token ratesCompare live routesOx Alpha has lower listed rates; Tokenra’s Omen Alpha discount is close on input/output.
Vendor transparencyOx AlphaIt has been identified as GLM-5.3-Flash; Omen Alpha remains unconfirmed.
Production coding reliabilityRun bothYour repository, prompts, tools, retries, and failure policy determine the result.

To test Omen Alpha, follow the setup guide or use the API reference.

07 / FAQ

Omen Alpha vs Ox Alpha questions.

Is Omen Alpha better than Ox Alpha?

There is no universal answer. Omen Alpha’s dated OpenCode result is 23.14/40, but a fair comparison requires the same prompts, evaluator, and date. It is a newer likely successor, not automatically better in every scenario.

Is Omen Alpha the same as Ox Alpha?

No. Ox Alpha was the earlier stealth model later identified as GLM-5.3-Flash. Omen Alpha is a newer, separately distributed model generally treated as its likely 2.0 successor.

Which model costs less?

Ox Alpha’s listed rates are $0.075 input, $0.25 output, and $0.015 cached input per million tokens. Omen Alpha’s official rates are $0.20/$0.66, while Tokenra advertises $0.08/$0.26. Confirm live pricing and promotions.

Can I still use Ox Alpha?

Availability depends on the provider route. Check OpenRouter, OpenCode, or the applicable provider for current status. Omen Alpha is the newer active model available through OpenCode Go and tokenra.io.

Why does Ox Alpha report 1M context while Omen Alpha reports ~500K?

Omen Alpha’s ~500K figure is community/platform-reported, not a vendor model card. A provider can also apply operational limits below a model’s theoretical maximum. Test the active route.

Compare them on your own codebase.

Get an API key, run representative prompts, and let controlled results — not assumptions — decide.

Get API access