FrontierRL
0,001

Tokens / second

Frontier Inference

Decoding throughput, tokens per second.

4X faster than the frontier lead

Star outputs measured on workstation GPUs, not datacenter-grade.

At Budget-Level Prices

Calculated per 1M output tokens, at each model’s published price.

How far does $100 go?

~75× more tokens per dollar

Watch Star run.

Why Star is different

Not another LLM.
A new architecture, built for agents from day one.

It understands live video, audio, and text, then acts across real software—fast enough for the moment, efficient enough to run continuously.

Programs

TNT MIT TNT Research Together.ai Together.ai Accelerator

Stop rationing.