Reasoning and coding
Use it for complex reasoning, repository-level software engineering, code generation, review, and debugging.
Build with the Ox Alpha API for coding, sustained agentic work, and complex reasoning. Usage-based pricing: $0.14/M input and $0.52/M output. Follow the quick guide below to get an API key and send your first request.
Ox Alpha is built for long-horizon software engineering, complex reasoning, and production use. It provides a 1M-token context window and accepts text, image, and video input.
Use it for complex reasoning, repository-level software engineering, code generation, review, and debugging.
Handle sustained agentic work with clear stopping conditions and application-level validation.
Ox Alpha supports text, image, and video input on the enabled integration.
Ox Alpha delivers reasoning for workloads that require sustained execution and production-oriented evaluation.
Register for TokenRa, open the dashboard, create an API key, confirm that Ox Alpha is enabled for your account, and then send an OpenAI-compatible request from your server. Keep the key private and verify the live model identifier before production use.
Register, then open the TokenRa dashboard.
Open the API key or token section, create a key, and store it server-side.
Confirm the enabled model ID, test a representative workload, and add production controls.
curl https://tokenra.io/v1/chat/completions \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"ox-alpha","messages":[{"role":"user","content":"YOUR_TEST_PROMPT"}]}'Ox Alpha uses transparent usage-based pricing through TokenRa: $0.14 per 1M input tokens, $0.52 per 1M output tokens, and $0.04 per 1M cached read tokens. No credit card is required to get started. Access is subject to account eligibility, provider availability, and applicable rate limits; verify the live TokenRa console before production use.
Standard input pricing for Ox Alpha.
Generated output pricing for Ox Alpha.
Check the TokenRa console for current RPM and quota limits.
Ox Alpha is the preview code name of GLM-5.3-Flash (official name), developed and operated by Z.ai (Zhipu AI). TokenRa provides API access to it as a routing gateway, and is not the model developer or owner.
Prompts and completions are retained by the provider and are not used for training. Review the applicable provider and TokenRa terms before sending sensitive or regulated information.
The Ox Alpha API provides OpenAI-compatible access to Ox Alpha, a reasoning model for coding, sustained agentic work, complex reasoning, and production workloads.
Register for TokenRa, open the dashboard, create an API key, confirm that Ox Alpha is enabled, and use the key in a server-side integration.
Ox Alpha costs $0.14 per 1M input tokens, $0.52 per 1M output tokens, and $0.04 per 1M cached read tokens. No credit card is required to get started. Confirm current availability, rate limits, and account restrictions in the TokenRa console.
It is developed and operated by Z.ai (Zhipu AI); Ox Alpha is the preview code name of GLM-5.3-Flash. TokenRa provides API access to it as a routing gateway, and is not the model developer or owner.
The stated capabilities include text, image, and video input, subject to the current enabled integration.
They are retained by the provider and are not used for training.
Ox Alpha is the preview code name of GLM-5.3-Flash, developed by Z.ai (Zhipu AI). Provider and related names belong to their respective rights holders. TokenRa provides API access as a routing gateway and is not the model developer or owner.