You don't need Fable for the most complex tasks, from training models for protein prediction to optimising compilers
Our open source Zenith harness takes base models to the top of FrontierSWE via adaptive self improvement
GLM 5.2 next 👀
Agent work getting stuck in sessions is vexing.
We're open-sourcing CommonGround Kernel so you and your agents can coordinate, hand off, and build on each other's work.
Blog:
GitHub:
so we built psql_bm25s.
exact BM25 retrieval. native Postgres access method. ~23x faster than pg_search on the standard benchmark.
retrieval stops being a budget item. the harness stops rationing. the agent gets to look things up like it should have the whole time.
New research: long-running agents often fail by stopping too early, not because the model can't make progress.
We tested 5 harness designs across 8 long-horizon coding tasks.
Our new orchestration harness, Zenith, wins 5/8 at 43% the cost of the strongest baseline.
We are happy to share early results from Logos, our novel first-principles augmented intelligence system, that has enabled insightful results across domains.
We start with series of results in physics
Today's is a lovely result hiding in Special Relativity for 121 years.