Fireworks said partner LangChain generates billions of agent-trace tokens a day, making frontier closed judges too costly. Fine-tuning a Qwen base on Fireworks matched GPT-5.5 / Opus-level judging at up to about 100x lower cost, and Fireworks is pushing its Training API for similar workflows.
Key Takeaways
- ✓LangChain produces billions of agent-trace tokens daily; frontier closed judges are too expensive.
- ✓Qwen fine-tuned on Fireworks matched GPT-5.5 / Opus-level judging at up to ~100x lower cost.
- ✓Training API is positioned for builders to train similar low-cost judge/agent models.