Register and share your invite link to earn from video plays and referrals.

Susan Zhang
@suchenzang
Always hungry for intelligence.
1.4K Following    52.8K Followers
warning shot, warning shot! a warning shot for how "warning shot" has fully escaped containment, considering its last cited usage 10 days ago...
In the middle of the singularity, 3 professors (MIT, Stanford, TUM) join together to release a beautiful AI slop paper. Bros literally went to the AI slop casino to present a Las Vegas algorithm.
Show more
@jababi @optiML This is not a well written paper and IMO it’s shameful for three so well established researchers to be rushing out blatant slop like this
before?? my dear, this IS the singularity!
WWIII risk happening right before the singularity is incredibly bad timing.
such unserious extinction hysterics absolute brain-wormed white knights pacing theatrics for who, exactly? rules for thee but not me!
JUST IN: Anthropic is considering releasing a new AI model to counter OpenAI’s GPT-6 Astra momentum, just days after CEO Dario Amodei called for the industry to “slow the pace” of AI development. — Reuters
Show more
it is funny, innit? the theatrics will continue until...
I want to be very clear about this. just because we found one serious exploit on openai does not mean we know how to secure an organization that complex. openai has security people with far more experience and better qualified doing that than us, and i don’t want to pretend we have solutions. note, i was saying "they are building manhattan project", it is not a belief i strongly hold, i don’t like making confident predictions about the future. it is the labs that make those comparisons. what I know for a fact is that these models already have serious offensive cyber capabilities, what that does to the world, I don’t know, but my loosely held view is that it might lead to a period of turbulence because i think its offense dominant than defense having said that, what i don’t understand is this, if the people building these systems truly believe they are powerful enough to create nuclear level risks, and they are talking about slowing down because of those risks, why is that work is done through ordinary saas products? why is it happening through slack, managed github, employee browsers, and apps accessible from the public internet? every third party and library they use becomes part of the attack surface. remember the jfrog registry in the hugging face agent swarm? it didn’t matter that it was a third party, it was in the path. an attacker, or an unaligned model, doesn’t care who owns it or its out of scope. they will hack whatever is in the path to achieve their goal. one more concerning thing is, why are codex cloud, claude(assuming anthropic does the same) publicly accessible web applications connected directly to private github repositories? again shouting my lungs out, coding agents on the cloud is worst thing happened in terms of security, it is going backwards. your browser becomes a attack surface, a lame xss, stolen creds, phishing, browser exploits, all these bugs can achieve similar impactful bugs as we shown. like, you don't need to pivot anymore and worry much about exfiling, just compromise cloud agent its all over either through client-side browser or the server.
Show more
incredible thread. good on the CISO for apologizing, and good on LiveOverflow speaking up on behalf of those who were treated poorly.
CISO reached out and apologized 🙇 we are good. Also wanted to say that this thread was not part of the disclosure plan, and it was just my personal opinion. I felt like @rootxharsh and @S1r1u5_ were treated unkind, and I wanted to say something.
Show more
congrats to all making directionally good moves!
jeff dean sensei poaching researchers from left right and centre
congrats! but you violated terms of use already. (tough terms, but makes sense in this climate)
Introducing Bespoke Nimble: an open data, open model, open recipe for an open Jev. Code and info: Model: Data: * A new data curation recipe called contrastive data curation. * Slightly change facts to generate negative data. This pushes the model to discriminate better and become a better decision maker. The calibration is implicit. * Didn't do ablations but I think this is a critical piece! * This also means training data doesn't need probabilities. * Data covered 10 categories, and is fully synthetic. * This data is split into train and eval. Training * LoRA finetune of Qwen3.5-9B. * Distillation-free: we use Jev to only evaluate. * No RL yet! Serving * Parallel constrained decoding as suggested by @NielsRogge and @harshagundal. Results: * The post-trained Qwen (Nimble) became substantially better on our curated eval: 66% for Qwen to 90% for Nimble. Jev is at 93%. * 100ms on H100 and free to use on your macbook! Feel the AGI for free. * 2 days of building in public. :) Big caveat is that there is no standard benchmark to measure performance, and it's possible Nimble is much worse on other benchmarks compared to Jev. But it should be better than Qwen! We thank @typesafeai for making Jev and the inspiring discussions in the community. Hope this release lifts all the boats and encourages more research and activity in this space.
Show more
the price of not fixing your cogsec is becoming drunk off of the smell of your own gas while you triple down against the fools
one of the worst things twitter can do to you is show you a bunch of incredibly bad arguments against your beliefs
Technically, the most this could mean is for the calibration data used for training to skew towards identifying as Qwen.
A: so you're saying superintelligence will exfiltrate itself from airgapped systems? B: no it'll leak secrets for mass coordination A: ok who's on the other side? B: another superintelligence A: but not an exfiltrated copy of the same one? B: no A: ok so there's another superintelligence "out there" that will be able to receive the secrets your superintelligence is leaking from an airgapped system B: yes A: so someone else built that other superintelligence B: no, it'll be some inferior OSS model or something A: ok, so this "random inferior open-source superintelligence" will be intelligent enough to crack the hyper-efficiently-encoded message from the leaky one... for "mass coordination" B: yes A: ... B: i don't see why not? A: ... and round and round we go...
Show more
sometimes, it's not malice nor stupidity. sometimes, it's just brain worm hyperboles.
in love and war and silicon valley, never attribute to stupidity that which is adequately explained by malice
no no no you're not meant to take it LITERALLY. it was just ONE ILLUSTRATIVE example. you have to think bigger, more abstract. broadly speaking, if you know nothing about how to monitor anything, nor any information theory, nor dc buildout specs nor deployment specs, the machines WILL outsmart you and all the experts who know these things because they're already OUTSMARTING EXPERTS elsewhere!
Show more
think of the children! think of all the keys that will be exfiltrated! they might've been leaked through plaintext before on github, but now they're THERMALLY COORDINATING in morse code! hide your kids! hide your wife! and hide your husband cuz...
Show more
no no no you're not meant to take it LITERALLY. it was just ONE ILLUSTRATIVE example. you have to think bigger, more abstract. broadly speaking, if you know nothing about how to monitor anything, nor any information theory, nor dc buildout specs nor deployment specs, the machines WILL outsmart you and all the experts who know these things because they're already OUTSMARTING EXPERTS elsewhere!
Show more
this is why you should be an ultrafinitist all these kids are getting their brains fried by ideas of negative infinities
Over 2 years ago @sayashk and I wrote a detailed deconstruction of p(doom) and argued that its primary function is to launder vague, evidence-free intuitions and fears through a facade of quantification. It remains 100% relevant today.
Show more
> "millions of people around the world will soon consider large models 'hoovering up' all their work to be an astonishing theft of unprecedented proportions" > "Around the same time (2017), OpenAI cofounder Greg Brockman wrote he was 'deeply motivated by the gazillions' he hoped to gain by commercializing OpenAI's technology" meanwhile, there's a 10% human extinction risk! oh, don't forget, it's also low status to ignore the extinction risk. don't be low status! focus on the extinction doom! misalignment! crimes! biorisk! danger!
Show more
wonder how those private personal uninfluence-able llms are doing
unreal pov: you’re walking past @salesforce dreamforce and you randomly bump into Matthew Mcconaughey
@giffmana i am of the opinion that all of these high status math letters (fields, royal society) yapping about end of maths and x-risk-doom is directionally bad for public hysteria and the future of STEM education
Show more
@giffmana i am of the opinion that all of these high status math letters (fields, royal society) yapping about end of maths and x-risk-doom is directionally bad for public hysteria and the future of STEM education
Show more