Decoding throughput, tokens per second.
~60× faster than commercial models
Star outputs measured on workstation GPUs, not datacenter-grade.
Calculated per 1M output tokens, at each model’s published price.
How far does $100 go?
~75× more tokens per dollar
Why Star is different
Star is our own model.
Not a wrapper, not a reseller.
Everyone else is busy making their big models faster.
We started from the other end, then trained in a totally different way.
Built by engineers from



Programs
MIT TNT Research
Together.ai Accelerator