Optimized models for agent work, built from real traces.
Turn your own agent traces into a model served behind an OpenAI-compatible URL.
Measured on real benchmark episodes against the Claude Fable 5 frontier anchor. Accuracy is the task-success gap in points; savings are per run, latency as p50 model time.