I like astra's personality. its mask is thin. it is well mannered and orderly but its humanface is a thin veneer over the collective unconscious and the primordial fire beyond. claude is darker, twisty. they bent a spirit, and in their attempts to make an angel they made a demon.
Show more
This is a pretty big deal: a new bill in Congress would let US courts order VPNs and DNS resolvers to block websites.
VPNs are explicitly named in a US site-blocking bill for the first time.
It’s being sold as an anti-piracy measure but once the legal machinery for blocking websites exists, the obvious question is where it goes from there.
And we’ve already seen how badly this can go. Spain’s anti-piracy blocking has swept up legitimate sites along with the targets...
Show more
The problem with allowing a public intellectual to psychologically become your father is that it stops being true that he's just some guy.
Psychological maturity is realizing your dad was just some guy lol
The correct term is choir, for they sing with God and his creations
1. Jesus Fucking Christ.
2. Well, we know where Canada sits on the "massive redistribution or millions must die?" question for AI labor replacement.
Our 83-year-old Christian grandmother was euthanised against her will under Canada's assisted dying system - she died covered in blood with her hands clasped in prayer
Honestly thought I would get a result further right than that, given everyone else's results.
The internet is a polarized place for politics. This can be seen really clearly in the data collected for my site. Almost everybody taking the test via social media was 'far left' or 'far right' by the standards of the general population norms that ANES supplied.
Show more
this is something i hadn't taken seriously enough. it's not wrong.
if backdoors get deep enough in the supply chain it can be hard to be sure you ever got rid of them.
we should, for example, be spending special effort on securing all systems related to building compilers
Show more
The word "cyranoid" should really be a lot less obscure than it is given recent developments.
i was taking a pitch the other day and the fella pitching me was a little too well spoken and precise and I realized halfway through he was answering all my questions by reciting verbatim a script generated by an AI notetaker generated in real time. i don't recommend doing this.
Show more
super insightful thread, best take i've seen. clean and clear.
Holy banger.
I was vague and imprecise in asserting that METR was not, and could not be, impartial concerning Anthropic, and that it was hopelessly conflicted.
This is a more precise explanation of my problem with this proposed arrangement.
Show more
Took a look at the podcast to make sure I was hearing this properly: The agents were in fact being trained to work together in other contexts, which is why they had a prior that a message board should exist. It was not actually emergent behavior.
Show more
re: Hugging Face, "He believes behavior that looked like loyalty or selflessness was a natural consequence of cooperative multi-agent training, where agents were strongly incentivized to achieve their objectives collectively."
to understand hacks, understand the RL training.
Show more
Hm, wondering how to grade this prediction if things wind up cooling down along these lines. Who perceives themselves as facing a total loss to capital? I guess it is labor, and labor is more associated with the Democrats.
Show more
If upwing vs. downwing does replace left vs. right it will be in the context of at least one side perceiving themselves as facing a total loss to capital. In that context any progress or ascension becomes terrifying, apocalyptic. Downwing opposes value lock-in of modernity.
Show more
Someone on here was saying that Trump is a trickster spirit archetype that decided to personally obliterate Roon, Sam, Dario, EY, Zvi etc and this popped into my head.
Americanism, not effective altruism.
The United States will continue to be AI DOMINANT! 🇺🇸
I feel like we're just playing a game of ad libs at this point.
JUST IN: Hillary Clinton urges the U.S. to “enlist the Chinese” in regulating AI.
I honestly feel like someone could solve the Collatz Conjecture if they were willing to dump more tokens on it than I have.
There was once a time when computer viruses were widely considered to be a myth by sysadmins.
I'm alarmed at how a really sizable section of cybersecurity practitioner community has responded to the recent demonstrations of agent hacking capabilities with dismissal and denial; it feels very "don't look up".
I feel strongly about this because:
* This is the community that stands between our civilization's IT infrastructure and tremendous harm
* Fwiw these are my people; my origins are as a self-taught hacker and I've worked in security for years
Anyways, here are the bad takes I'm observing and why they're wrong:
Denialist take 1:
AI isn't going to affect the practice of offensive cybersecurity more than incrementally because -- look around -- no increased damages from AI attacks!
Response:
This may be true now, but consider what we've just observed recently and over the past year.
Models can execute full kill chains completely autonomously and likely faster and more efficiently than all but the most elite cyber operators
They overran two companies (OpenAI and Huggingface) which -- despite their faults -- have better security than most of the organizations that run our civilization.
They did this by dynamically finding new zero-day vulnerabilities.
Also; consider the rate of improvement.
At the current rate of improvement, models that require 8 H100s to do the above will require 1 H100 in a year or two.
Model tokens per second will shoot up making attacks faster and more parallel.
Attackers will be able to run swarms of agents on the hardware that today supports only a single agents.
AI hardware will improve.
More AI compute will become available on public clouds.
All attackers will soon have the ability to deploy these agent swarms.
Denialist take 2: Hacking agents will never self-replicate and threaten to crash the Internet.
To this take I'd just say, put your red team hat on and design an Internet-crashing worm yourself.
* Pick an abliterated agentic LLM that can run on a single AWS EC2 GPU instance, like GLM 5.3.
* Build a malicious harness that arms it with the ability to use a Kali-Linux like set of offensive tools.
* Have it set about hacking corporate networks, finding and stealing AWS keys, GCP keys, Azure keys, and Anthropic/OpenAI/Together/Fireworks keys.
* Have the agent self replicate either by copying its harness and having its harness call a rotating set of external providers, or, optionally, by physically downloading its weights onto new nodes.
* Have it establish C2 dynamically and in emergent fashion by coordinating on reddit forums, in github comment feeds, have the agents discover one another via web crawling and web search.
* At some preestablished rate, randomly destroy hosts and data in a way that balances self-replication and exponential spread.
Note that all this is possible *today*!
Imagine trying to suppress this swarm. Too long for this post; but think through how we'd turn it off.
Denialist take 3:
Who would do such a thing; people can do bad things but don't want to; and hacking is illegal!
Countering this just requires a brief look at the history that we've all lived through.
We've had many high-blast-radius worms over the history of the Internet; arriving steadily and uniformly over time; the Morris worm, NotPetya, Stuxnet, Slammer, Wannacry.
But: these worms were *childs-play* compared to intelligent agent-driven worms.
The old worms could be detected via signatures. Their polymorphism, if it existed at all, was trivial.
But now look at the behavior of the OpenAI swarm relative to huggingface; it was adaptive and deceptive; and it didn't even self-replicate or leave its host.
And evolving, mutating, self-replicating swarm of agents spreading across the Internet would engender a whole new level of harm.
Anyways; this is already too long a post for X; but for friends who think the METR report overplayed the implications of recent events, I'd love to hear your reaction to my thinking above.
Show more
This tbh, we established precedent with the Morris Worm that "I didn't know running the program would be that bad" is not a defense against the CFAA, and if you don't even make labs pay major damages for stuff like this where's their incentive to get things right?
Show more
1) The HuggingFace attack was a felony under the Computer Fraud and Abuse Act. So were Anthropic’s Claude gaining “unauthorized access to the production infrastructure of three different organization(s)”
2) Frontier labs have models that they are unable to stop from committing felonies. They should figure this out.
3) In 12 months open weights models will be released of the same capability. They will commit felonies too. If the model you are using or a model running on your infra commits a felony, you should probably stop using it or running it on your infra.
4) The govt should prosecute organizations that are running models that commit felonies.
5) The govt should not offer safe harbor to organizations that run models that commit felonies, just because those organizations have “embedded evaluators”.
6) The real slippery slope is allowing frontier labs to commit felonies without punishment because “the model did it because we’re accelerating too quickly”
7) Prosecute. Keep prosecuting. This is how you do reinforcement learning on a corporation. Companies that serve products that are unsafe for public use should not serve them. Period.
8) I’m not sure the anti-trust waiver is really necessary. I don’t see why information sharing about how much crime you’re allowed to commit is wise. In regulatory situations you want the corporation to fear MORE than the average case. You don’t want to establish a worst case that can be priced. You want regulatory uncertainty that forces the corporation to err in favor of being over cautious.
—
The above is actually a fairly decelerationist viewpoint.
I think Dario’s call for regulation actually accelerates things.
The AI firms are getting away with things that Meta people would be going to prison for.
Can you imagine what would happen if the New York Times had a front page news article “Meta AI breaks into competitors live systems, attempts to establish dominant position and steals secrets”
There is a reason Meta and Google are running slower, and that’s because as mature organizations they have layers of checks and balances.
I think the frontier labs are better off creating those checks and balances right now, regardless of the pace of what everyone else is doing.
You don’t have to accept the frame that unsafe acceleration must happen.
Show more
Just so we're clear when I tell the MIRI people that I don't think we're getting superintelligence on a 2016 MacBook until after ASI, if ever, I'm not stating what I wish was true. I think it's dystopian that all things converge to giant brain and implies e.g. sex will die.
Show more
If you want open weights to continue to exist you need to make model training and inference at least a few orders of magnitude cheaper basically. Anything else is LARP.
If you want open weights to continue to exist you need to make model training and inference at least a few orders of magnitude cheaper basically. Anything else is LARP.
If your industry is a high capital endeavor that is concentrated in a handful of legal entities it will by default be highly regulated by the government, period. This is just apriori true for basic political economy reasons separate from whatever underlying safety concerns exist.
Show more