In the process of migrating the repo I built at Braintrust for
@HamelHusain and
@sh_reya's AI evals course ( to an all-
@pydantic stack: PydanticAI + Logfire + pydantic-evals.
Been living in these tools for a while now and the progress on observability and evals is impressive.
I'll be blogging and posting findings here as I go, including where I think models like
@typesafeai's Jev fit into the end-to-end evals pipeline.
Stay tuned ...