Research —> Product :)
very excited to start rolling out our fine-tuned Trace Judge model built earlier this month with the great
@FireworksAI_HQ team
there’s a mountain of Agent Improvement gold sitting in everyone’s agent traces, the cheaper & faster you can understand that data, the better feedback you can collect to improve your agents
fine-tuning helps us serve + help you build custom, ultra-cheap models to read every single Trace your agents produce to look for Errors, Product Feedback, or really anything you or your company care about. we find that we can do this while matching frontier performance after fine-tuning
from there let the experiments begin! ex: tagging data for finetuning, proposing harness engineering experiments
Trace understanding applied at scale with models purpose built for your most important tasks
if this is interesting please reach out! we’d love to have you try it and talk learn about how we can help you build for your use cases