Register and share your invite link to earn from video plays and referrals.

Francis deSouza
@fdesouza
630 Following    12.1K Followers
An honor to meet with His Highness the Crown Prince of Kuwait to discuss how @Scale_AI can support Kuwait’s digital transformation and develop Kuwaiti AI talent.
The nation that can measure the AI frontier will be the one that steers it. Our future should not be steered by speculative warnings. We have to be driven by evidence. That was the clearest takeaway from this week's AI conversations at the UN General Assembly. Governments need to get serious about testing frontier AI, and nowhere is that more urgent than in Washington. Getting serious means three things: → Clear responsibility for testing across agencies → Access to models before they're deployed → Funding for the experts and tools to do the work Scale AI's cyber research shows what testing can reveal. We've seen AI agents carry out an attacker's instructions while still finishing the user's task, so the user never knew anything went wrong. We've also seen models refuse to help people defend their own systems. Both are failures. Real testing has to measure the harm AI can cause and the help it fails to give. Government shouldn't have to rely solely on AI developers to explain these failures. It needs its own testing capacity, with the expertise to ask hard questions and the tools to answer them. Scale has worked with the U.S. government for years to test the risks and capabilities of AI models, and we're accelerating that work in the months ahead. Here's what we think America should do next:
Show more
Congrats to @AIatMeta on Muse's success! It's been our pleasure to partner with them on their mission to bring personal superintelligence to everyone.
This week I joined @RoyalFamily, govt ministers and leaders to discuss how AI can best serve the public good. AI can expand opportunity and solve some of our toughest challenges. It’s essential that we pursue that potential responsibly, in a way that benefits people broadly. Rigorous, independent testing - like we do at @Scale_AI - is an essential step to earning public trust. Governments and independent evaluators need to work together to deeply and continually assess AI's capabilities, opportunities, and risks. That allows us to make sound decisions about its development and use. Conversations like this one are a step in the right direction.
Show more
We’re moving to a multi-model world and that’s a good thing. Enterprises will use multiple models, optimized for different scenarios. Thanks for having me on @andrewrsorkin and @BeckyQuick
How do you know if an agent is enterprise-ready? NEWS: Today Scale AI is introducing READY, our Reliable Enterprise Agent Deployment benchmarks, to answer that question. READY measures agents on reliability, the human oversight required to hit a defined reliability target, and the resulting cost, together.
Show more
Big News: I’m joining @scale_AI as CEO, starting August 10. Scale sits at a rare intersection, working with the top AI labs to push the frontier while helping enterprises and governments actually deploy AI they can verify and trust. I’m excited to lead a company with a mission to develop reliable AI systems for the most important decisions. More to come.
Show more