Fable-tier, 100 tok/s. #1 on Design Arena.
Opus-tier. 744B MoE, 1M context.
27B dense, low latency.
230B MoE for agentic workflows.
1M-context model, served fast.
Classify every agent trace for any behavior that matters, in under 90ms.
Merge AI-generated code edits instantly.
AI search subagent with sub-6s searches.
Verbatim context compaction for long-running agents.
Auto-route each prompt to the best model.
Engineering deep dives and product updates.
Up to $5K in API credits for startups.
Talk to the team about your use case.
How we train and deploy models.
Join a small team shipping daily.
WarpGrep v2 is an RL-trained parallel search subagent that lifts every major coding model to #1 on SWE-Bench Pro. 15.6% cheaper, 28% faster, and now handling multi-repo, package, and log search.