TokenRaAI Model Catalog
Ox Alpha API ยท Usage-based pricing via TokenRa

Ox Alpha API: $0.14/M Input, $0.52/M Output

Build with the Ox Alpha API for coding, sustained agentic work, and complex reasoning. Usage-based pricing: $0.14/M input and $0.52/M output. Follow the quick guide below to get an API key and send your first request.

$0.14 inputPer 1M input tokens.
$0.52 outputPer 1M output tokens.
1M contextConfirm current account limits.

Built for coding and sustained agentic work

Ox Alpha is built for long-horizon software engineering, complex reasoning, and production use. It provides a 1M-token context window and accepts text, image, and video input.

01

Reasoning and coding

Use it for complex reasoning, repository-level software engineering, code generation, review, and debugging.

02

Agentic workflows

Handle sustained agentic work with clear stopping conditions and application-level validation.

03

Visual context

Ox Alpha supports text, image, and video input on the enabled integration.

Where Ox Alpha fits

Ox Alpha delivers reasoning for workloads that require sustained execution and production-oriented evaluation.

  • Long-horizon software engineering and coding tasks.
  • Complex reasoning that benefits from structured evaluation.
  • Agentic workflows that combine planning, tools, and application checks.
  • Production workloads requiring representative quality, latency, error, and usage testing.

How to get an Ox Alpha API key

Register for TokenRa, open the dashboard, create an API key, confirm that Ox Alpha is enabled for your account, and then send an OpenAI-compatible request from your server. Keep the key private and verify the live model identifier before production use.

1

Create a TokenRa account

Register, then open the TokenRa dashboard.

2

Create your API key

Open the API key or token section, create a key, and store it server-side.

3

Send a test request

Confirm the enabled model ID, test a representative workload, and add production controls.

curl https://tokenra.io/v1/chat/completions \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"ox-alpha","messages":[{"role":"user","content":"YOUR_TEST_PROMPT"}]}'

Ox Alpha pricing: $0.14 input, $0.52 output

Ox Alpha uses transparent usage-based pricing through TokenRa: $0.14 per 1M input tokens, $0.52 per 1M output tokens, and $0.04 per 1M cached read tokens. No credit card is required to get started. Access is subject to account eligibility, provider availability, and applicable rate limits; verify the live TokenRa console before production use.

Input

$0.14 / 1M tokens

Standard input pricing for Ox Alpha.

Output

$0.52 / 1M tokens

Generated output pricing for Ox Alpha.

Rate limits

Account-based

Check the TokenRa console for current RPM and quota limits.

Z.ai (Zhipu AI)

Ox Alpha is the preview code name of GLM-5.3-Flash (official name), developed and operated by Z.ai (Zhipu AI). TokenRa provides API access to it as a routing gateway, and is not the model developer or owner.

Data handling stated for this preview

Prompts and completions are retained by the provider and are not used for training. Review the applicable provider and TokenRa terms before sending sensitive or regulated information.

Production checklist

  • Keep API keys on the server and separate credentials by environment.
  • Confirm the 1M-token context and supported input modalities against the live integration before relying on them.
  • Run representative coding, agentic, reasoning, and visual-context workloads.
  • Set explicit timeouts, bounded retries, tracing, and application-side budgets.
  • Review data handling, provider terms, privacy requirements, and output rights for your use case.

Ox Alpha API FAQ

What is the Ox Alpha API?

The Ox Alpha API provides OpenAI-compatible access to Ox Alpha, a reasoning model for coding, sustained agentic work, complex reasoning, and production workloads.

How do I get an Ox Alpha API key?

Register for TokenRa, open the dashboard, create an API key, confirm that Ox Alpha is enabled, and use the key in a server-side integration.

How much does Ox Alpha cost, and do I need a credit card?

Ox Alpha costs $0.14 per 1M input tokens, $0.52 per 1M output tokens, and $0.04 per 1M cached read tokens. No credit card is required to get started. Confirm current availability, rate limits, and account restrictions in the TokenRa console.

Who develops Ox Alpha?

It is developed and operated by Z.ai (Zhipu AI); Ox Alpha is the preview code name of GLM-5.3-Flash. TokenRa provides API access to it as a routing gateway, and is not the model developer or owner.

What input does Ox Alpha support?

The stated capabilities include text, image, and video input, subject to the current enabled integration.

Are prompts and completions used for training?

They are retained by the provider and are not used for training.

Ox Alpha is the preview code name of GLM-5.3-Flash, developed by Z.ai (Zhipu AI). Provider and related names belong to their respective rights holders. TokenRa provides API access as a routing gateway and is not the model developer or owner.