Decoding throughput, tokens per second.
4X faster than the frontier lead
Star outputs measured on workstation GPUs, not datacenter-grade.
Calculated per 1M output tokens, at each model’s published price.
How far does $100 go?
~75× more tokens per dollar
Why Star is different
Not another LLM.
A new architecture, built for agents
from day one.
Actually parse live video, audio, text, data and act in real time.
Programs
MIT TNT Research
Together.ai Accelerator