Reasoning and coding
Use it for complex reasoning, repository-level software engineering, code generation, review, and debugging.
Build with the Omen Alpha API for coding, sustained agentic work, and complex reasoning. Transparent usage-based pricing starts at $0.20 per 1M input tokens. Follow the quick guide below to get an API key and send your first request.
Omen Alpha (Ox Alpha 2.0) is built for coding, complex reasoning, and agentic production use with a 128K-token context window.
Use it for complex reasoning, repository-level software engineering, code generation, review, and debugging.
Evaluate sustained agentic work with representative tasks, clear stopping conditions, and application-level validation.
Omen Alpha corresponds to GLM-5.3-Flash, a native multimodal model developed by Zhipu AI. It supports text, image, and video input on the enabled integration.
Omen Alpha is intended for workloads where reasoning depth, sustained execution, and production-oriented evaluation matter.
Register for TokenRa, open the dashboard, create an API key, confirm that Omen Alpha is enabled for your account, and then send an OpenAI-compatible request from your server. Keep the key private and verify the live model identifier before production use.
Register, then open the TokenRa dashboard.
Open the API key or token section, create a key, and store it server-side.
Confirm the enabled model ID, test a representative workload, and add production controls.
curl https://tokenra.io/zen/go/v1/chat/completions \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"ox-alpha","messages":[{"role":"user","content":"Explain the Omen Alpha architecture"}]}'Omen Alpha is a paid model on TokenRa. No credit card is required to get started. Access is subject to account eligibility, provider availability, and applicable rate limits; verify the live TokenRa console before production use.
Standard input pricing for Omen Alpha.
Generated output pricing for Omen Alpha.
Cached input read pricing through TokenRa.
Omen Alpha (Ox Alpha 2.0) is developed by Zhipu AI as GLM-5.3-Flash. TokenRa provides the API gateway and is not the model developer or owner.
Omen Alpha corresponds to GLM-5.3-Flash, the official Zhipu AI name following the Ox Alpha preview.
| Detail | Ox Alpha | Omen Alpha / GLM-5.3-Flash |
|---|---|---|
| Status | Preview code name | Current model name |
| Input price | See live console | $0.20 / 1M tokens |
| Output price | See live console | $0.66 / 1M tokens |
| Context | See live console | 128K tokens |
| Provider | Third-party model | Zhipu AI |
See the Ox Alpha model guide for historical preview context.
Prompts and completions are retained by the provider and are not used for training. Review the applicable provider and TokenRa terms before sending sensitive or regulated information.
The Omen Alpha API provides OpenAI-compatible access to Omen Alpha, a reasoning model for coding, sustained agentic work, complex reasoning, and production workloads.
Register for TokenRa, open the dashboard, create an API key, confirm that Omen Alpha is enabled, and use the key in a server-side integration.
Omen Alpha costs $0.20 per 1M input tokens, $0.66 per 1M output tokens, and $0.04 per 1M cached read tokens through TokenRa. Check the console for current rates and limits.
Omen Alpha (Ox Alpha 2.0) is developed by Zhipu AI as GLM-5.3-Flash. TokenRa provides the API gateway and is not the model developer or owner.
Omen Alpha corresponds to GLM-5.3-Flash, a native multimodal model developed by Zhipu AI. It supports text, image, and video input on the enabled integration.
They are retained by the provider and are not used for training.
Omen Alpha is associated with the Ox Alpha 2.0 naming used in this guide. TokenRa is an independent API gateway and is not the model developer or owner.