Reasoning and coding
Use it for complex reasoning, repository-level software engineering, code generation, review, and debugging.
Build with the Omen Alpha API for coding, sustained agentic work, and complex reasoning. Transparent usage-based pricing starts at $0.13 per 1M input tokens. Follow the quick guide below to get an API key and send your first request.
Omen Alpha is built for coding, complex reasoning, and agentic production use with a 1M-token context window.
Use it for complex reasoning, repository-level software engineering, code generation, review, and debugging.
Evaluate sustained agentic work with representative tasks, clear stopping conditions, and application-level validation.
Omen Alpha supports text and image input on the enabled integration, with video behaviour closely matching. Confirm current capabilities in the TokenRa console.
Omen Alpha is intended for workloads where reasoning depth, sustained execution, and production-oriented evaluation matter.
Register for TokenRa, open the dashboard, create an API key, confirm that Omen Alpha is enabled for your account, and then send an OpenAI-compatible request from your server. Keep the key private and verify the live model identifier before production use.
Register, then open the TokenRa dashboard.
Open the API key or token section, create a key, and store it server-side.
Confirm the enabled model ID, test a representative workload, and add production controls.
curl https://tokenra.io/zen/go/v1/chat/completions \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"omen-alpha","messages":[{"role":"user","content":"Explain the Omen Alpha architecture"}]}'Omen Alpha is a paid model on TokenRa. No credit card is required to get started. Access is subject to account eligibility, provider availability, and applicable rate limits; verify the live TokenRa console before production use.
Standard input pricing for Omen Alpha.
Generated output pricing for Omen Alpha.
Cached input read pricing through TokenRa.
Omen Alpha's developer has not been officially disclosed; speculation points to a Zhipu AI model (glm-5.3-highspeed), but this is unconfirmed. TokenRa provides the API gateway and is not the model developer or owner.
Omen Alpha and GLM-5.3-Flash are separate models. The table below compares what is publicly documented for each; items that are not officially confirmed are marked as such.
| Detail | GLM-5.3-Flash | Omen Alpha |
|---|---|---|
| Developer | Zhipu AI (Z.ai), publicly confirmed | Not disclosed. OpenCode data paths point to Zhipu and the model sometimes identifies itself as GLM; some users report it has been confirmed not to be Zhipu. |
| Parameters | 320B total / 18B active (MoE) | Not disclosed |
| Architecture | Native multimodal MoE with hybrid sparse and linear attention | Closely matched (tokenizer and vision/video behaviour are nearly identical) |
| Context length | 1M tokens | Not officially disclosed; commonly cited as 500K, measured at 969K or more (close to 1M) |
| Max output | About 128K tokens | Community reports about 128K |
| Input modalities | Text, image, and video | Text and image (video behaviour matches) |
| Reasoning | Supported (effort: low / high / max) | Reasoning supported |
| Open source | Yes (MIT licence, downloadable on Hugging Face) | No (API only) |
| Availability | Broad (Z.ai, Together, self-hosting) | Available through TokenRa |
See the GLM-5.3-Flash model guide for the confirmed specification.
Prompts and completions are retained by the provider and are not used for training. Review the applicable provider and TokenRa terms before sending sensitive or regulated information.
The Omen Alpha API provides OpenAI-compatible access to Omen Alpha, a reasoning model for coding, sustained agentic work, complex reasoning, and production workloads.
Register for TokenRa, open the dashboard, create an API key, confirm that Omen Alpha is enabled, and use the key in a server-side integration.
Omen Alpha costs $0.13 per 1M input tokens, $0.5 per 1M output tokens, and $0.03 per 1M cached read tokens through TokenRa. Check the console for current rates and limits.
Omen Alpha's developer has not been officially disclosed; speculation points to a Zhipu AI model (glm-5.3-highspeed), but this is unconfirmed. TokenRa provides the API gateway and is not the model developer or owner.
Omen Alpha supports text and image input on the enabled integration, with video behaviour closely matching. Confirm current capabilities in the TokenRa console.
They are retained by the provider and are not used for training.
Omen Alpha's developer has not been officially disclosed. Names mentioned belong to their respective rights holders. TokenRa is an independent API gateway and is not the model developer or owner.