Register and share your invite link to earn from video plays and referrals.

Ali Ghodsi
@alighodsi
Databricks CEO & Co-founder, UC Berkeley Faculty
Joined January 2010
272 Following    213K Followers
An extremely important functionality for agents is to simply extract fields out of PDFs. This turns out to be harder than people think because LLMs are primarily trained on predicting the next tokens. This leads them to "autocorrect" things that they shouldn't autocorrect. We launched an AI Extract capability that just excels at doing just this task with very high accuracy (95% vs 87% for others) and extremely low cost. Check out this blog on how we did it. The function can of course be called directly from SQL and be used throughout the platform.
Show more
0
75
1.4K
156
Forward to community