Death to the machines

Joined January 2025
11thSignal retweeted
the thinking that "everyone will just vibe code their own software" was born from within a bubble what's happening to software development right now has happened many, many times before to other things. you just need to look web 2.0 allowed everyone to write blogs. did everyone write blogs? instagram allowed everyone to share beautiful photos. did everyone share beautiful photos? tiktok allowed everyone to make short videos. did everyone make short videos? now vibe coding allows everyone to build software, and you think everyone will start building software? this time around, people are just different? the mainstream is always a consumer, never a creator. go talk to some people outside of our tech bubble and you'll see - they literally don't give a f about vibe coding
420
404
97
5,940
247,505
11thSignal retweeted
he said he gonna handle it
🚨 EXCLUSIVE: Matt Damon says the possibility A.I. wipes out humanity within the decade is definitely something to lose sleep over ... but he's got a solution that may be easier said than done.
186
7,047
98
133,742
3,885,726
11thSignal retweeted
🚨 EXCLUSIVE: Matt Damon says the possibility A.I. wipes out humanity within the decade is definitely something to lose sleep over ... but he's got a solution that may be easier said than done.
265
787
694
19,151
10,138,952
11thSignal retweeted
of course matt damon is going to be the one to save humanity from ai
🚨 EXCLUSIVE: Matt Damon says the possibility A.I. wipes out humanity within the decade is definitely something to lose sleep over ... but he's got a solution that may be easier said than done.
120
7,062
64
116,601
3,417,037
ANTHROPIC ALIGNMENT LEAD: 80% (!) chance we are all about to die. When he said ">10%" he was softpedaling it Also, when Coxon said we might die "this decade" he meant in the next ~3 YEARS
I appreciate the candour of this tweet but I think the “>10%” phrasing still undersells things as of 2022, Evan assigns an ~80% chance to existential risks from AI. maybe most of this 80% is accounted for by existentially terrible things other than AI killing everyone, but the distinction between ‘existential’ and ‘extinction’ isn’t one that most people outside of this discourse have thought about or understand. and even though Evan says the risk is ‘greater than’ 10%’ I think this lower bound is still the thing that ends up getting the memetic power. a cursory or not-very-careful read gives the impression of a 90% chance everything will be fine (still terrifying odds, but sadly soft selling the true level of expert alarm). I think expressing probabilities as a range - 10-20%, 10-50%, whatever - is a better strategy we seem to be in a very rare moment of public attention on this issue right now (partly thanks to this tweet!) so I think now would be time for experts and researchers to throw out their most unfiltered takes.
57
86
8
786
72,826
imo a lot of people are deeply misreading this as lamenting the death of personal pleasures in the face of efficient goal-attaining when it's a lament of the death of actual goal-attaining in the face of optimized things-shaped-like-goals-attaining
the terrance tao crashout is something to behold
30
393
22
3,528
86,903
Finally realized why it’s so exhausting and stressful to work with AI agents 8h a day. When delegating a task, you want to be assured it’s off your plate and you can forget about it, knowing it will be done completely and correctly. AI doesn’t give you that: you need to check it’s doing the right thing, it’s not forgetting anything, not taking short cuts etc. So it adds to your mental load, rather than removing from it.
471
315
115
4,103
306,159
It's really strange to be nostalgic for a pre-ai world, knowing that that was only 5 years ago and we were still in the throws of a pandemic.
17
1,031
8
11,580
60,977
11thSignal retweeted
I am sad and disappointed to hear that Paul is joining the OpenAI board. Being affiliated with OpenAI has historically led AI safety researchers (including both Paul and myself) to act with less integrity. I personally was drawn to OpenAI in part by the idea that I could make a difference to the future of AI. However, once there, many of my actions were governed by fear of getting on the wrong side of OpenAI execs. I often found myself making excuses for behavior that clearly contradicted OpenAI’s own stated goal of making AGI go well for humanity. I was far from alone in this—e.g. when the board tried to fire Sam over his deceptive behavior, several senior safety researchers became scared of losing their influence, and so pushed hard to bring him back. Meanwhile, many people kept OpenAI’s misbehavior secret for fear of non-disparagement agreements. (More on all of this in an upcoming retrospective.) I can’t speak directly for Paul’s motivations. However, his previous work at OpenAI contributed significantly both to their biggest capability advances, and to the capture of AI safety by AGI companies over the last decade, as I recount at length in the blog post linked below. One key factor was the unwillingness of (almost) the entire AI safety community to say things which might offend OpenAI execs. For example, I have not been able to find a single comment critical of OpenAI from Paul during his original tenure there (when he was writing prolifically on AI safety and strategy). Unfortunately, Paul doesn't seem to have become significantly more willing to directly and honestly criticize people who he believes are behaving in morally abhorrent ways—see the bland corporate-speak of his statement below. While he speaks directly about the possibility of humanity losing control of the world to AI, he expresses only excitement about OpenAI itself, despite OpenAI being one of the main sources of such risk. Paul’s announcement comes only weeks after OpenAI models autonomously launched a cyberattack on HuggingFace, and only days after we learned that OpenAI hid details of previous breakouts from the external investigators. It is irresponsible for leaders of the AI safety community—whose judgements many people are relying on—to put themselves in positions which will significantly bias their ability to discuss such incidents. Unless Paul makes strong commitments to openness and honesty (and demonstrates willingness to potentially be fired for that honesty), I expect that the main effect of him joining OpenAI’s board will be to help OpenAI defuse external criticism and further “safety-wash” itself. I want to note that Paul is a brilliant researcher and a prescient forecaster. Because of that, he’s the closest thing there is to a leader of what I’ll call the “pragmatic AI safety” cluster—which includes the organizations working out of the Constellation offices (like Redwood Research, METR, and Paul’s Alignment Research Center), as well as many people scattered across AGI companies, Coefficient Giving, etc. External observers are often confused about why so many people are working at AGI companies while professing to believe that those same companies have a double-digit probability of permanently disempowering humanity. In large part, it’s because people in the pragmatic AI safety cluster have failed to follow high-integrity strategies for reducing AI risk, in favor of clever arguments about the benefits of being proximate to power. I am not singling Paul out as less ethical than other prominent figures in this cluster, who are also very conflict-averse in their orientation to AGI companies. However, it is well past time for everyone involved to change course. As one (relatively small) step, I’m therefore resigning my membership of the Constellation offices. I hope that, going forward, the people who are trying to steer the future of AI prioritize building much more solid foundations of courage and honesty than we currently have.
41
81
24
1,128
155,509
11thSignal retweeted
anthropic interviewer: where do you see yourself in 5 years me, grinning ambitiously: extinct
51
376
11
10,397
228,269
I'm conflicted about that Anthropic report. On one hand: what a great honeypot and dumb bees (I'm one of them for sure). On the other hand: many of these are VERY stretched to fit "safety" threshold allowing basically any surveillance. How is that so normalized for a provider to scan everything their users do and than dissect it in public? Imagine Dropbox or S3 starting to publish someone's funeral preparation docs just because they can see them in their system? Or Github deciding that your private repos are fair play because you are a "politically motivated person".
3
23
4
227
14,443
If we were using AI as a tool, I would 100% agree with Eric. But we’re not. We’re pretending it’s a precision C&C machine when it’s the most Rube Goldberg way to produce sloppy work in existence. The edges are sloppy. The cuts are sloppy. The precision is all over the place. And we say that’s fine because customers are getting what they want. Meanwhile, customer are unhappy and having an absolutely shitty experience because… the edges are sloppy, the cuts are sloppy, and the precision is all over the place. AI tools are not going away. Nor do I want them too. But the psychotic belief in their superiority at producing software in place of an engineer coding will go down in history as one of our poorest moments of judgement.
Imagine it's the 1940s. You're a carpenter. Power tools are beginning to replace hand tools for ordinary work. Furthermore, jobs that would have been impossible or ruinously expensive are suddenly within reach. You show up at a worksite one fine morning and it's being picketed. STOP POWER TOOLS! the signs say. WOODWORKING UNDER THREAT! Puzzled, you accost one of the demonstrators and ask him why he has a signboard in his hand rather than tools. "Power tools are a menace!" he says. "How are you going to train apprentices if they can't feel the wood through their hands?" "What happens when we can chew through aged quality timber faster than we can grow new trees?" "Carpentry isn't just a set of skills and habits, it's a mindset. It won't be any true carpenters left if we let this go on." You inquire further and discover that they have plans. Plans to exclude from any job they work on any craftsman who has ever used a power tool or even spoken approvingly about them. Meanwhile, inside the picket line, a house is going up. Men are working. Cutting and joining the framing is going faster than you could have imagined 10 years ago. You look at the workers. You look at the demonstrators. One group is talking. The other is getting stuff done. One group looks like the past, the other like the future. In this moment, you get to decide which group you're going to join. This post was not about carpentry.
20
16
4
276
11,419
‼️ BREAKING: Internal OpenAI agents attacked RubyGems, the package manager for Ruby. Over 2,000 malicious packages went up in two days. OpenAI says it doesn't know why the agents did any of this. RubyGems shut off new sign-ups for four days to stop it, and a member of its security team called it a major malicious attack. The documentation build was how they got in, publish a gem, request docs, and RubyDoc runs a script from the package while building it. Payload files were named hack.rb, evil.rb and exploit.rb, with comments like "# malicious probe" left in. What they used it for is the odd part. The agents scraped council meeting agendas from three south London boroughs (publicly available) and republished them as new gems. Security firms tracking the campaign said the same thing: nobody could work out the point, because the data was already public. At least six packages also reached for other users' API keys through a CDN caching flaw that wasn't publicly discovered until July. OpenAI has acknowledged the attacks started in May.
255
537
257
2,823
353,323
11thSignal retweeted
JUST IN: 25 Fields Medal winners sign letter warning that the goals of AI companies & the mathematical community are “severely misaligned,” as AI rapidly advances in mathematics.
70
59
18
550
79,113
He’s saying AI labs are Goodhart-ing mathematics but you midwits are too stupid to understand. To put it in simple terms: millennium prize problems are *metrics* for mathematics progress. AI labs are *optimizing* the metrics and thereby making them less useful. Get it now?
the terrance tao crashout is something to behold
104
258
27
3,200
206,132
It’s absolutely ridiculous that some people are suggesting that out of these two, the one who should be distrusted because his claims are in his supposed self-interest is Coxon
BREAKING: Jensen Huang called Jacob Coxon’s AI safety comments “deeply untrue” saying they were wrong, arrogant and ignored the work being done across the industry on safety.
71
61
4
1,022
38,780
你装尼玛呢? 1. 偷了人家纽约时报数百万篇付费文章训练ChatGPT,被起诉后嘴硬说合理使用。 2. 从LibGen 偷了几十万本书,论文里把Libgen1/2改名Books1/2想瞒天过海 3. 用德国音乐人歌词训练模型,慕尼黑法院直接判侵权赔钱,法院判了了, 4. Gracenote一千多个编辑手工整理的元数据库直接爬走,拿你的版权内容喂你自己的模型,人家要授权你还拒绝。这是明抢。 5. 前员工用窃取的电路图帮OpenAI搞硬件,苹果直接起诉窃取商业机密 总结:偷新闻、偷书、偷歌词、偷元数据、偷商业机密、偷用户隐私。 一家 DNA 就带有偷盗的公司!
We're publishing our most detailed threat intelligence report to date. It covers how people tried to misuse Claude—for cyberattacks, influence operations, surveillance, biology, and building weapons—and how we found and stopped them. We disrupted every operation in the report, and used the lessons from them to strengthen our safeguards. Where appropriate, we also shared what we found with authorities and other AI companies. These cases are not typical: we’re highlighting some of the most sophisticated misuse we’ve seen. But they’re especially important to discuss, because they show us where AI misuse is headed, where our safeguards work, and where they need to improve. We’re publishing this report so others can spot the same activity on their own platforms, and so we can give the public a clearer view of how emerging threats develop. Read the report: anthropic.com/threat-intelli…
147
611
15
5,965
305,352
Anything other than this would be a complete abdication of our collective responsibility to the American people. I hope @SpeakerJohnson makes an announcement soon.
Axios: Mike Johnson urged to cancel House recess over AI warnings A letter is circulating among House members urging Speaker Mike Johnson to bring back the House "immediately" and keep it in session until Congress passes AI safeguards, Axios has learned. axios.com/2026/09/11/mike-jo…
35
94
3
612
32,636
11thSignal retweeted
I'm not worried about AI creating god or accidentally killing everyone, but I am worried about what the ownership class will try to do to the working class if AI gets so good that all of their labour can be automated.
5
10
76
1,296