Decoding throughput, tokens per second.
4X faster than the frontier lead
Star outputs measured on workstation GPUs, not datacenter-grade.
Calculated per 1M output tokens, at each model’s published price.
How far does $100 go?
~75× more tokens per dollar
Why Star is different
Not another LLM.
A new architecture, built for agents from day one.
It understands live video, audio, and text, then acts across real software—fast enough for the moment, efficient enough to run continuously.
Programs
MIT TNT Research
Together.ai Accelerator