FrontierRL

Watch Star run.

Our LLM is uniquely good at high-throughput data processing: video, audio, and live text streams, all understood in real time. Whether gaming, using software, or conducting research, our model always runs at competitive speed.

Video

Video games are the ultimate stress test for real time video processing. Here Star reads the screen and reacts in real time, no game-specific scripting. See more at Star plays games and Star for robotics.

Reacting in real timeSmash 64
Beat Pokémon for $6Pokémon
Multiple AgentsStarCraft II

Software

See how Star uses real software GUIs to get the job done.

Rebuilding the MIT Racecar
Prompt to part change
Designing Presentations

Relevant benchmarks to show comparison:

Robotics

Star can parse a live feed at inhuman speeds, noticing change, reacting fast enough that the answer still applies by the time the joint moves. See more at Star for robotics.

Research

Sourcing, competitive analysis, and long web workflows done at inhuman speed, for pennies.

Sourcing LinkedIn candidates
Competitive biotech analysis (slowed down)
Multi-window research

Benchmarks

The full comparison, relative to published or in-house data. Full set at the benchmarks page.

EvalFrontierRL StarClaude Fable 5Claude Fable 5 (Xhigh)Claude Opus 4.8Claude Opus 5Claude Sonnet 5DeepSeek v4 FlashDeepSeek v4 ProGemini 3 ProGemini 3.1 ProGemini 3.1 Pro PreviewGemini 3.6 FlashGPT‑5GPT‑5.5GPT‑5.6 LunaGPT‑5.6 SolGPT‑5.6 Sol (High)GPT‑5.6 Sol (Max)GPT‑5.6 TerraKimi K3Qwen3.7 MaxSource for other model's metrics
BenchCAD (python tool)81.75%51.8%55.8%73.9%83.4%78.2%openai.com/index/gpt-5-6
IFEval97.97%91.6%93.4%96.64%94.2%In-house tested by FrontierRL
OSWorld Verified85.36%85%83.4%76.2%78.7%anthropic.com/claude/mythos
Terminal-Bench 2.175.28%80.52%84.64%74.53%70.79%73.78%76.40%85.77%80.90%openai.com/index/gpt-5-6
Video-MME v283%77%66.1%44.7%67%69%In-house tested by FrontierRL

Stop rationing.