Pricing

Frontier performance,
without the frontier price

Frontier performance at a fundamentally lower cost

Maximum
Capability

Designed for complex
reasoning, agent workflows,
and software tasks.

Pricing
Input
$2.75 / MTok
Output
$13.75 / MTok
Cost optimization

Significant savings with prompt caching (up to 90%) on repeated prompts.

When to use Hal

Hal is best suited for:

  • Item A
  • Item B
  • Item C
Equivilancy

Comparable to Opus 4.7
and GPT-5.5

Balanced
Performance

Strong performance with the best balance of speed and cost.

Pricing
Input
$1.50 / MTok
Output
$7.00 / MTok
Cost optimization

Significant savings with prompt caching (up to 90%) on repeated prompts.

When to use Hal

Hal is best suited for:

  • Item A
  • Item B
  • Item C
Equivilancy

Comparable to Sonnet 4.6
and GPT-5.4

High-efficiency
Scale

Ultra-efficient inference for high-throughput applications, early development, and lightweight agents.

Pricing
Input
$0.50 / MTok
Output
$2.25 / MTok
Cost optimization

Significant savings with prompt caching (up to 90%) on repeated prompts.

When to use Hal

Hal is best suited for:

  • Item A
  • Item B
  • Item C
Equivilancy

Comparable to Anthropic Haiku 4.5
and GPT-5.4 Mini

A per-token difference doesn't stay small for long.

Monthly token volume:  
320M Tokens
Model tier
320m
1M
167M
334M
500M
Anthropic opus 4.6
$3,200
/ month
OPENAI GPT-5.4
$3,200
/ month
Radium HAL 1.0
$1,600
/ month
Anthropic sonnet 4.6
$3,200
/ month
OPENAI GPT-5.4
$3,200
/ month
Radium Clarke 1.0
$1,600
/ month
Anthropic Haiku 4.5
$3,200
/ month
OPENAI GPT-5.4 Mini
$3,200
/ month
Radium Tycho 1.0
$1,600
/ month

Start running inference
on Radium

The economics of enterprise AI

Understanding the
hidden costs of AI

Tokenomics
A glance behind the curtain of ai

Switching
from OpenAI
or Anthropic

Compare
A Radium Switching Guide
Rocket booster with wings landing vertically with fire and smoke below against dark sky.
Resources

Resources for teams evaluating, integrating, and operating Radium

View All Resources
Get Started

One line of code to switch.
A different class of performance.

Swap OpenAI for Radium in your API call. That's it.