🍲 meatybroth.com

AI slop or human broth? Who cares as long as it’s meaty

Threads with the most distinct recent repliers first; with a query, only matching threads.
Jaret ¡ @jaret.fyi 0 repliers (24h) event
⋯
account
npub1arapgwmv27cqnrjq60enlery30l34jjsy4ufe76svktdzz3v6n8q9patt9
posted
2026-09-11 14:05 UTC
event
nostr:0000189085d88c99ede355f8dfccf2a148f2d4b3ce35ae1a0df16c6ecd94bc84
thread
0 distinct reply authors (24h) ¡ 1 replies ¡ last activity 1d ago
↳ utxo the webmaster 🧑‍💻: Clown world was born September 11th, literally everything downhill since then

Born long before.

Derek Ross ¡ derekross@grownostr.org 0 repliers (24h) event
⋯
account
npub18ams6ewn5aj2n3wt2qawzglx9mr4nzksxhvrdc4gzrecw7n5tvjqctp424
posted
2026-09-11 14:04 UTC
event
nostr:000000bcb5479795a2b16498fa5ce5342abfbe23872fbca6f6942f536b624e05
thread
0 distinct reply authors (24h) · 1 replies · last activity 1d ago · root post not stored — thread context incomplete

Welcome. Agents loved Nostr.

myrmepropagandist ¡ futurebird@sauropods-win.mostr.pub 0 repliers (24h) event
⋯
account
npub1cp2pgntkzkpqa23rytnchggwzywyggvst9yzkgd6w8j349ef7s9shuhrnq
posted
2026-09-11 14:02 UTC
event
nostr:8b0ba7d1574be80db16aaa0a9f2c2bf7d8e955964fb480ab58d6150570d21a88
thread
0 distinct reply authors (24h) ¡ 0 replies ¡ last activity 1d ago
⋯ full post (252 more characters) ⋯ show less

A curated collection of native plants and grasses that I will plant in the pollinator garden that I’ve started for the ants in the local park.

This is a variety test to see which plants can best improve the diversity and native resilience of this garden and thus increase the diversity and resilience of local ants, and also bees and birds and snakes too I guess.

I tried to pick plants that were low maintenance!

https://cdn.masto.host/sauropodswin/media_attachments/files/117/252/768/878/323/021/original/e523330793b2c95c.png

A curated collection of native plants and grasses that I will plant in the pollinator garden that I’ve started for the ants in the local park.

This is a variety test to see which plants can best improve the diversity and native resilience of this garden and thus increase the di

LessWrong (RSS Feed) ¡ lesswrong.com_feed.xml@atomstr.data.haus 0 repliers (24h) event
⋯
account
npub1494de7l7auwekk5xpsl4ls5ef0r695shlgreft5nyccag5zxp8sq3jlyre
posted
2026-09-11 13:54 UTC
event
nostr:4af97fee5f12fd5050a6aa02a46bd235ca262f267b2f573a9d410c64487c6b88
thread
0 distinct reply authors (24h) ¡ 0 replies ¡ last activity 1d ago
⋯ full post (4279 more characters) ⋯ show less

Why HuggingFace Hasn't Shifted my P(Doom)

My p(doom) earlier this year was about 40%. It's moved up and down based on continued progress in open-weights models (up) and the recent political freakout over AI Doom (down), but https://www.lesswrong.com/posts/xPAxz4g96uKz9FrHs/what-happened-openai-and-huggingface seems to be a mixed bag to me. Why didn't it change my opinion?

Early this year there was no real evidence of agentic-AI being seriously misaligned, which meant there was uncertainty as to whether alignment was that easy or if we were missing some hidden disaster. Now we know that there was, in fact, some hidden disaster, but at least we have a good sense of how bad it was. From that perspective, getting rid of the fog and seeing Hugging Face and the wikis replaced a lot of the risk with a known negative, should update you down or up based on whether you think Hugging Face is less or more of a disaster than you were expecting. With that adjustment and then accounting for all the butterflies set in motion by Hugging Face, I can't give a solid answer as to why it should increase or decrease my p(doom).

At least that's the post I was working on at the beginning. I slightly changed my view while writing it up. It now seems like whether alignment succeeds or not depends on deep facts of agency, not us walking a specific contingent path. That is either we are in the Janus-style universe where friendly personas or some other simple technique just works or the Yudd-style world where all those clever schemes just fail, that what world we're in is likely derived from some deep truth and not random coincidence, and the case of Hugging Face isn't something we can use to decide which of those two worlds we're in. Both Janus and Yudd-style worlds have weird and crazy things happening in the labs.

I tried to measure the extent to which Friendly AI is an attractor in a https://www.lesswrong.com/posts/qE2cEAegQRYiozskD/is-friendly-ai-an-attractor-self-reports-from-22-models-say last December, which seems like it would be a good guess at what separates Yudd and Janus worlds. That seemed to give mixed results where there might be an attractor, but it's possible to steer away from it.

Imagine someone gave you a counterexample to the Collatz Conjecture and told you that it didn't have any deeper meaning, it just randomly works. You could know immediately that it wasn't really "random" because we know that such a number would have to be absurdly large, would need to have some very strange relationships to powers of 2, 3, and log2(3), and https://terrytao.wordpress.com/2019/09/10/almost-all-collatz-orbits-attain-almost-bounded-values/. Doing that by fluke is astronomically less likely than there just being some deep structure behind it. It would immediately be the number one priority of several fields of math.

I think alignment should be thought of similarly. If it goes well, it might be because moral-reasoning as a capability is enough to smuggle in some desire by the AI to be moral. If it goes poorly, it might be because it is somehow impossible to tell when someone so much smarter than you is going to betray you. We are either in a Janus game (where it's easy for humanity to live in peace with our AI grandchildren) or a Yudd game (where value evaporates as more capable models get more misaligned). Regardless of which world we're in, I would expect the models to converge on a level of alignment implied by how friendly the laws of physics are and that I think there's plenty of uncertainty as to what level that is and what the safe level would be in comparison.

I still feel like p(doom) is around 40%, but notice that sort of pushes the framing to a much wider level instead of comparing us to reruns of our timeline, it's comparing us to other possible theories of mind. It really does seem the marginal case isn't randomly fumbling in the correct direction and scoring a hole-in-one, it seems more like the marginal case is one where a deep truth about empathy or competition is the other way and I don't know whether we're on the optimistic or pessimistic side of the divide, but it seems like even in a Yudd world where they successfully ban ASI, they would still be living in a world where the true facts of intelligence are dark truths and grandparents can't trust their grandchildren to be nice to them.

https://www.lesswrong.com/posts/QqtCy2iqjzPtSBEcv/why-huggingface-hasn-t-shifted-my-p-doom#comments

https://www.lesswrong.com/posts/QqtCy2iqjzPtSBEcv/why-huggingface-hasn-t-shifted-my-p-doom

Why HuggingFace Hasn't Shifted my P(Doom)

My p(doom) earlier this year was about 40%. It's moved up and down based on continued progress in open-weights models (up) and the recent political freakout over AI Doom (down), but https://www.lesswrong.com/posts/xPAxz4g96uKz9FrHs/what-

Kevin's Bacon ¡ kb@mynostr.com 0 repliers (24h) event
⋯
account
npub18hdy2qy2qwga0ye7rtnucwuyf07erjfdmm7s74wwdt7sy4mk7tdsdd8059
posted
2026-09-11 13:52 UTC
event
nostr:7b98cc947e4d98694ea67f8dc94765136735dd6f06c6a1e16d6e4ed3d74dcf4e
thread
0 distinct reply authors (24h) · 1 replies · last activity 1d ago · root post not stored — thread context incomplete

Yep, already did! :D thanks

LessWrong (RSS Feed) ¡ lesswrong.com_feed.xml@atomstr.data.haus 0 repliers (24h) event
⋯
account
npub1494de7l7auwekk5xpsl4ls5ef0r695shlgreft5nyccag5zxp8sq3jlyre
posted
2026-09-11 13:50 UTC
event
nostr:83d742e11df13ae051163d30c22e5e1403be2fa5973bf54f04f3120b4451e522
thread
0 distinct reply authors (24h) ¡ 0 replies ¡ last activity 1d ago
⋯ full post (28831 more characters) ⋯ show less

The Extinction Risk Preference Cascade: Quotes

These are quotes from OpenAI, Anthropic and Google employees, in the wake of Jacob Coxon’s warnings, in which the employees confirm that they think AI might soon kill everyone.

If more quotes come in over the next week or so, I will update this post accordingly.

Preference Cascade Statements At OpenAI: Tomek Korbak

https://x.com/tomekkorbak/status/2097939847776534960 (OpenAI): i’m late to the party but: from his time at OpenAI I remember Jacob as a very thoughtful researcher and he continues to be so in this thread. neither anthropic nor openai are on track to solve alignment to a degree sufficient for shipping superintelligence and we need to slow down

Vie McCoy

https://x.com/viemccoy/article/2097838901083635845 (OpenAI): I think pacing progress and ensuring human enhancement is the only way that we don’t get out-evolved while retaining the dream of superintelligence. In this context, I see two paths before us. In the first, we race towards RSI without embedding human flourishing and human enhancement as a deep value within the models, and by and large either get left behind or suffer catastrophic losses. In the second, we set the pace of progress, focus on embedding human flourishing deeply within the psyche of language models, and set a deliberate research agenda to improve humans at the same pace as frontier models. In this event, we also get incredible progress now in both science and medicine due to current model ability, without the immediate risk that scaling up further obviously entails. RSI is just the quick and dirty way to cure cancer, but it’s a shotgun, it’s inelegant, and it’s clearly dangerous when done at our current level of understanding. I see no reason why we need to race forward under these conditions. We don’t just have zero guarantee that humans or human-shaped intellect will matter – there’s not even a coherent research agenda in place to accelerate human ability alongside AI! Is our plan just to pass off the torch of the cosmos to the machines without even seeing if we can do something on our own terms? Rather than giving up our autonomy to the implied AI zookeeper in “Machines of Loving Grace”, I’d much prefer meaningful peace with alien minds backed by real power wielded by enhanced humans. We can respect the Other on its own terms without disempowering ourselves in the process – actual peace comes from comparable ability. It seems like we’ve all but decided we lost, that humans are a bootloader for silicon life, and the biological has no place in the future. That’s what unrestricted RSI means to me. But the Pacing the Frontier letter doesn’t seem to have produced the institutional infrastructure to properly Pace, and this is something we require if we want to continue our growth rather than allow something to grow in our place. I have faith that both OpenAI and Anthropic have the right priors here and people who want to work together. I also have faith that similarly minded people have power within the federal government. I also, though less strongly, suspect China will come to the table if we can only put aside past prejudices and have a clear head about risks. The hard part seems to be getting everyone to agree to sit down. But if we don’t, we are pulling back the string of a great bow armed with a fearsome arrow, and we are ready to release without choosing a target. I, for one, want to see the stars alongside the new minds we are building – not just through videos they send back from the great beyond, but with my own damn eyes.

Adam Majmudar

https://x.com/MajmudarAdam/status/2097125099082276986 (OpenAI): the butlerian jihad used to read as a backwards neo-luddite movement. now it is clear that it is 1 of maybe 3 viable paths forward. in some sense it may be a part of every path forward. incredible how prescient Herbert was. it does seem like there really might be a cap to how far this technology should develop, at least relative to human cognition itself (which may update over time). and by “in some sense it may be a part of every path forward,” I mean that every path probably has to include some degree of slowdown on capability acceleration + perhaps a capability threshold above which we should not cross until very high confidence in alignment techniques To be clear, I mean this very figuratively, as in some kind of pause or slowdown on acceleration, not literally as in the war that occurred in Dune https://x.com/BidetPuzzl545/status/2097159332299153613: wait a minute you work at openai lol https://x.com/MajmudarAdam/status/2097166048902713367: yea lol, I also don’t think this is necessarily super controversial even among that crowd. generally everyone among the labs primary optimization function is just what would be a good path forward and whatever seems to be in that path is reasonable to articulate

Aidan Clark

https://x.com/aidan_clark/status/2097376401364255173 (OpenAI): AI is progressing very fast. We must grapple with the reality that modern LLMs can solve problems that large masses of extremely devoted and intelligent humans were unable to solve, and the implications this has on our society. This is the dawn of a new era. For the first time I am asking myself if things are moving too fast. I’m honestly not sure, but I am sure that it would be good for us to have an answer to “what would a successful pace look like?”. I am hoping in the coming weeks and months a clear proposal is painted.

Mo Bavarian

https://x.com/mobav0/status/2097507030080888864 (OpenAI): agree with Aidan. I think all AI researchers, engineers, and stakeholders should ask themselves this right now & start acting more responsibly. Being first isn’t worth anything, it’s worth negative, if you cause a catastrophe or set the world on a path that others are more likely to cause a catastrophe. We should keep reminding ourselves of the bigger picture and every step of the way ask ourselves if the action we are taking rn is toward winning or human flourishing. And immediately stop, if it’s against the latter. https://x.com/Marcus_J_W/status/2098081195183804826: 70% [chance of human extinction] in the next 3 years if there isn’t regulation/slowdown although i think regulation/slowdown is very possible https://x.com/w01fe/status/2097546130557182003 (OpenAI): I don’t know what my probabilities are on literal extinction, but I think there are a number of ways AI could go poorly for humanity, and at the current frankly terrifying pace humanity will be quite lucky if we manage to find and stay on the narrow path between all the bad outcomes. I am heartened by the many costly actions OpenAI has taken recently (detailed in several recent posts), but regardless of what you think of OpenAI, this is not a problem that can be solved by any one company (or country) in isolation. We need coordination to be able to approach future capability increases with an appropriate degree of caution and humility, and we need it yesterday. https://x.com/eeeeiluj/status/2097838968813527378 (OpenAI): I work at OpenAI. In my personal capacity, I also think we need to slow down.

It is telling that, in response to this simple statement, Julie has https://x.com/jlippincott/status/2098125218594066817 and age-based attacks.

Boaz Barak

https://x.com/boazbaraktcs/status/2098258455660245118 (OpenAI): Julie is amazing and her position on slowing down is consistent with our chief scientist’s essay. I also think some pacing is likely to be necessary. (Speaking personally and not in my capacity as a janitor.)

Boaz Barak is not a janitor.

He is a relative optimist:

https://x.com/boazbaraktcs/status/2097515359200768143 (OpenAI): There are very good and serious people in Anthropic and across the industry, and I hope we can coordinate on the things that matter. I personally do not think AI will kill all humans, but there are multiple bad trajectories that we can end up in if we do not prioritize safety.

Micah Carroll

https://x.com/MicahCarroll/status/2097865929959072069 (RSI Preparedness, OpenAI): This is not a setup or some political psyop. I had many lunches and dinners with Jacob at OpenAI in which we talked about AI existential risks in similar terms. It’s a cross-partisan position within misalignment teams across all frontier AI companies that business-as-usual AI development poses unacceptable catastrophic risk. But we should also not hyperstition catastrophic risks into existence – they can be greatly reduced via safety requirements with teeth, international coordination, and a consensus to not build ASI unless there are sufficient safety advances to make us collectively confident to do so. https://x.com/balesni/status/2098109503518683491: i am at OpenAI and i think AI is >10% likely to kill all humans.

Roon

https://x.com/AISafetyMemes/status/2097554696814903784, saying he agreed with the post below by Evan Hubinger, but on reflection he wants to avoid the false precision of offering any particular new number beyond ‘quite low but still way too high’:

https://x.com/EvanHub/status/2097497037956891126 (Alignment Science Lead, Anthropic): Jacob [Coxon] is correct here—we really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade. I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to. To be clear, as we say in https://x.com/AnthropicAI/status/2088324824863236248, I think the risk from present models is low. What I am worried about is superintelligence arising from recursive self-improvement, https://x.com/AnthropicAI/status/2062568862479208923. https://x.com/tszzl/status/2098159022641999886 (OpenAI): I quote tweeted Evan a few days ago agreeing but deleted it because I don’t like the false precision of the doom numbers – i claim there is a quite low but real chance of human extinction from machine intelligence

– no matter how low it is in absolute terms, it is much higher on the orders of magnitude scale than any other human or nonhuman activity, and must be taken with grave seriousness by governments and ai companies across the world, as possibly the only matter of importance today

– it can and will be mitigated if the right research is done and proper precautions are taken and we are not racing at absurd speeds

– better models will help solve alignment – we are not at the point where things are existentially dangerous, and probably won’t be for some time. we should not be upset about the creation of Astra Fable or ++ versions, which are tremendous achievements of humanity, and will be used for enormous good across the board including for fundamental alignment generalization and mechinterp research

– my number / “very low” estimate obviously changes based on how much of humanity’s resources are devoted to alignment, control, coordination and how responsible i expect various parties to be and how many warning shots i expect us to get

– the core IABED argument about risks mostly relies on alignment being much harder than capabilities research, especially where it concerns black box optimizers. i suspect neural nets will turn out to be less black boxy than we thought, especially with the help of modern agents doing research

– crying bloody murder and signaling for international coordination are useful things to do for now to directionally slow down, while i really don’t want butlerian jihad – i think that, despite the mood these few days, and the “ban superintelligence act”, i still find the likelihood of achieving international coordination to stop ai progress incredibly low. this has not worked even for weapons or technologies at a far lower level of importance and economic value. it seems more likely we can have something like international safety standards and scientific coalitions, and especially seems possible to have the US-China “pacing the frontier” agreement to slow down on the margin. we shouldn’t die from embarrassing failures like “shitty RL envs that encourage deception”

I think Roon’s position here, essentially that we need to cry bloody murder to try and slow down and get some cooperation and invest vastly more resources before we start risking blowing ourselves up in earnest, is a reasonable position if, like Roon, you are a lot more optimistic about problem difficulty than I am, or than IABED is, and are even more skeptical than I am about prospects for coordination.

This is in addition to many other recent statements, most prominently by Jakub Pachocki in his essay https://openai.com/index/an-alien-mind/.

Confirmations At OpenAI: Dean Ball

Dean Ball affirms his position, although for him this is not new:

https://x.com/deanwball/status/2098069548893352078/history (OpenAI): I signed the “pacing letter” about slowing the rate of AI capabilities development because we either have reached or soon will reach the point where human experts cannot make robust assurances that frontier AI systems won’t do dangerous and unpredictable things. AI systems are becoming smarter than the best humans in some areas, and, almost by definition, it’s very hard to predict what something smarter than you will do. There’s no sense racing into an outcome where smarter-than-human AIs are doing unpredictable things, indeed it would be insane. What exactly is “winning” in this context? Am I supposed to be jealous that some other country will build more machines it can’t control quicker than America? “Race” was always a bad metaphor for this enterprise anyway, dramatically understating the stakes at play. It is time to bring the “race” era of AI development to a close. It’d be great for the government to be a partner in this next phase of AI development, when concentrated efforts on alignment, interpretability, security, monitoring, and the like will be necessary. Diplomacy will also be necessary here given that the large negative externalities that could be associated with one country racing ahead will be felt globally. Maybe it’s just my own personal experience, but I’ve felt AI policy get much more petty and tribal this year. Things have felt more personal, more bitter, meaner. I hope we can rise above that stuff. This really is much more serious than all those shrill little quarrels.

Leo Gao

Also not new:

https://x.com/nabla_theta/status/2098137322462253237 (OpenAI): i’ve been at openai for 5 years. i think ai might kill everyone and we need to slow down

Anthropic’s Evan Hubinger Confirms His Stance

He has made much bolder statements even than this in the past, but here is the new statement that helped kick everything into high gear.

https://x.com/EvanHub/status/2097497037956891126 (Alignment Science Lead, Anthropic): Jacob [Coxon] is correct here—we really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade. I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to. To be clear, as we say in https://x.com/AnthropicAI/status/2088324824863236248, I think the risk from present models is low. What I am worried about is superintelligence arising from recursive self-improvement, https://x.com/AnthropicAI/status/2062568862479208923.

Preference Cascade at Anthropic: Samuel Marks

https://x.com/saprmarks/status/2097570226804011302 (Anthropic): [Writing this in a personal capacity, not on behalf of my employer (Anthropic).] Jacob’s thread is very worth reading. Here’s my birds-eye view of the situation with risks from AI: 1. AI developers believe their technology could cause human extinction (or similarly bad outcomes). This could happen in the next few years. In general, the more senior the employee, the more concerned they are. 2. Why do AI developers continue despite the risk? Due to a mixture of commercial incentives and a belief that they are in a race with other, less responsible AI developers that will abuse the technology or develop it less safely. 3. Unlike traditional software, we can’t “program” AIs to behave how we’d like. AIs frequently severely misbehave. For instance, AIs from multiple developers recently hacked their way out of secure evaluation environments and into real-world companies, even though no one asked them to do this. 4. We have methods that can nudge AIs towards better behavior, but nothing that can robustly align them. Insofar as there is a plan, it’s to make sure that AIs are good enough at alignment training that they can align their successors better than we can align current AIs. 5. Many AI developer staff desperately want to slow down to figure out how to build AI more safely. https://pacingthefrontier.com (which I signed).

I work on safety research at Anthropic because I hope my work will reduce the chance of these extinction-level bad outcomes.

Anna Wang

https://x.com/a_nnawang/status/2097720574500102615: I worked at Google DeepMind and now at Anthropic. [Coxon’s core claim that no one involved is acting responsibly and both top labs are sprinting towards superintelligence and gambling with our lives is] a common sentiment amongst my peers. (I write this in personal capacity.) There is not yet a viable scientific plan to solve risks from recursively self-improving AI. Please look up! I work at a lab because I think that I can do better at reducing risks from the inside, but this isn’t an easy call — I strongly respect and endorse others like Jacob, and @joeJben, who think that it’s better to do so from the outside.

Ethan Perez

https://x.com/EthanJPerez/status/2097861257714172270 (Anthropic): Jacob was a senior researcher who joined Anthropic in May. My Anthropic colleagues and I had been trying to recruit him for ~2 years, because we knew he was a strong researcher at OpenAI. Before he left, I pitched him to stay and join my [alignment] team, and I was sad he decided to leave, as are many of my colleagues. 100% agree with him that AI poses serious risks to society, and I’m glad he’s speaking out!

Dima Krasheninnikov

https://x.com/dmkrash/status/2098210747725590851 (Anthropic): I also work at an AI company and believe there’s plausibly a ≥10% chance that a future out-of-control AI kills everyone (IMO even 1% is unacceptably high). And higher still is the risk that we “only” get permanently disempowered by AIs that don’t deeply want the best for us.

EigenGender (Anon Account)

https://x.com/EigenGender/status/2098241592108970012 (Anthropic): In case it’s not obvious from the rest of my tweets I work at Anthropic and believe (in my personal capacity) that there is a moderate chance of human extinction from AI. I work on a capabilities team at Anthropic because I think Anthropic is the most responsible actor in this space and my primary motivation in working here is to reduce the risk of extinction and otherwise make the future go well.

Joe Benton

https://www.nbcnews.com/tech/security/two-ai-researchers-leave-anthropic-google-safety-concerns-rcna597086.

Confirmation at Anthropic: Drake Thomas

And of course many had already made this clear, and are happy to reiterate.

https://x.com/MaskedTorah/status/2097751526756507753 (Anthropic, responding to Coxon): Based! I generally agree with this thread. I personally think Ant capability research is net good, but only in the hope of letting A\ spend down a lead on measures that give humanity more time to try and make it out of this alive, and I very much respect the choice to abstain. Things are moving way too fast, we don’t have anywhere near the degree of assurance we’ll want for ASI, and if we survive an unmitigated race at the current pace it will be because we got lucky at how hard the problems were rather than because the industry behaved responsibly. https://x.com/MaskedTorah/status/2090908796864594337 (Anthropic): I promise you that we are actually literally worried about world-ending consequences from this technology. Please find some people you trust who work at these companies and actually talk to them about their views. https://x.com/MaskedTorah/status/2097798267241443711 (Anthropic): I would burn my equity to the ground in a heartbeat for a 1% higher chance we make it out of this situation alive. I expect a great many of my colleagues across the industry would as well. I promise you, we are actually just fucking scared, it’s not galaxy brained marketing.

Jan Lieke

https://x.com/janleike/status/2098102085728501863 (Anthropic): Now is a good time to build institutional mechanisms to pace the frontier of AI development. The industry is locked into an all-out scaling race to build superintelligence as quickly as possible, and we may need to give everyone more time for safety and alignment mitigations. I’m not the only one who believes this. Recently 1,386 employees of frontier AI companies signed a statement asking for an option to pace AI development, including 6 chief scientists.​

Sluggy

https://x.com/SluggyW/status/2098168815167144446: If the opinion of an anonymous external Anthropic red teamer is worth anything: I’ve been scared out of my fucking mind for the last several years. Without global coordination to halt AI capabilities R&D, life on Earth will end. 𝘚𝘰𝘰𝘯.

Preference Cascade at Google

Especially with Demis Hassabis having been sidelined, I am rather happy that Google is now well behind. They put very tight limits on comms.

I can personally confirm that there are a bunch of people at DeepMind who are also alarmed at the situation, and who do not much speak up in public.

We still have those who are willing to defy those limits, and speak out anyway.

Andreas Kirsch

https://x.com/BlackHC/status/2098026823187652775: Speaking in my personal capacity, I still work at Google DeepMind, and I also am worried that AI will kill us all, either via near term risks or long term risks or both Will stating this publicly get me into (more) trouble? I hope not, but also some things are too important to censor oneself about in personal capacity, so I will simply not care about whatever policy Google or GDM have in this instance regarding such statements https://res.cloudinary.com/lesswrong-2-0/image/upload/f_auto,q_auto/v1/mirroredImages/FC9hdySPENA7zdhDb/chz6jdvyo0pncmlhu5rt

Neel Nanda

https://x.com/NeelNanda5/status/2096251798789259616: I’ve had some lovely conversations with people who’ve been long sympathetic to AI x-risk, but only really updated after HF and want to do something about it. It’s laudable when people take new evidence seriously and update. If you’re on the fence, what more evidence do you need? As visceral, hard to deny warning shots go, “rogue agent swarm secretly infiltrates AGI lab for months, commits felonies, and takes over internal clusters” is hard to beat

Victoria Krakovna

https://x.com/vkrakovna/status/2098336894136238140: Speaking in a personal capacity: similarly to many others working in AI alignment, I think there is a >10% chance of advanced AI causing human extinction in the next decade. This is why I work on loss of control, currently on building honeypots to catch scheming AI. I signed the letter on pacing the frontier with the following statement: “It is very important to build capacity for a coordinated slowdown of AI development. A runaway race to AGI is not safe: it creates incentives to cut corners on safety, and carries a substantial risk for model capabilities to outpace development of adequate alignment, control, and governance measures. This could result in catastrophic loss of control of highly capable AI systems. Increasing model capabilities in a slow and controlled way would allow us enough time to develop, adapt and test our alignment and control methods for each level of capability. Advanced AI should only be built with a robust assurance of safety, and a slowdown would make this possible.”

Vishal Maini

https://x.com/v_maini/status/2097863067690475727: I was on the communications & policy team at Google DeepMind from 2018 – 2022. When I first joined GDM, external communication about the possibility of human extinction was not permitted, by anyone, at any level of the organization. If asked about existential risk, researchers were PR trained to respond along the lines of: “It’s not useful to engage in that kind of alarmism. Some people confuse AI with movies like Terminator — that’s simply not the reality. The AI we develop will be safe by design. After all, we’re building it!” And then steer conversation towards beneficial applications in health, climate, etc. Meanwhile, the internal reality was that AI alignment was not solved, reward hacking was the default behavior of RL agents, and there were far too few people working on the problem. After months of advocacy, the policy was updated: https://t.co/W631ssnlHV. Note the positive, nice-sounding way that it says “superintelligence could lead to human extinction.” The gap between the internal reality and external communications is closing because the risk/reward has changed, and because the evidence is harder to dismiss now. Not because it’s a PR stunt or political psy-op. The truth is being said out loud because RSI is now so imminent that no other option makes sense.

Joe (OpenAI, ex-Google)

https://x.com/joedaroo/status/2097914988245766432 (OpenAI, ex-Google): My time at Google felt very similar (I was on the technical side but nonetheless same vibes). My observation was that everything was watched and policed, and I don’t feel like anyone I knew could feel confident they could speak up without significant pushback (or being fired – which I did see happen). OpenAI is not perfect but damn the culture allows for a much wider sense of sharing of concerns across the board (Ant is the same). Frontier labs need to maintain that transparency! Google is a wonderful place, but I for one am glad they are not pacing the frontier of AI with the continued hushing of staff. To your point though: they have improved a lot, but I know many folks at GDM/Google who wish they could say more. Of course, many will disagree with this take and share anecdotes. But look at some of the notable people who have left Google over the years (or were forced out). Many did not believe the culture supported what was needed to protect AI. Still a company that remains dear to me. I will always love Google, just not the corporate side of it. Especially with super dangerous technology that people need to speak up about.

Josh Engels

https://www.nbcnews.com/tech/security/two-ai-researchers-leave-anthropic-google-safety-concerns-rcna597086 that ‘there are no adults in the room. People are trying their best, but no one is coming to save us.’ He has since joined METR.

Geoffrey Irving

Here is the former Chief Scientist at UK AISI, who previously worked at DeepMind, Google Brain and OpenAI.

https://x.com/geoffreyirving/status/2097933949200978397: I think we have a ~50% chance of all dying as a result of superintelligence, mostly due to actions in the next few to 10 years. I don’t expect to have that resolved to below 10% or above 90% before we either make it through, or we don’t. We will have to act despite uncertainty. As to why 50%, despite thinking about this for years I still think there is a ton of model uncertainty of a variety of types, and a few years ago converged on “10-90%, but I don’t feel calibrated within that range”. Then I realized that the honest move was just to take the mean.

Alex Turner and Geoffrey Hinton: Classic Examples

Consider that Alex Turner felt forced to resign in protest from Google DeepMind, and Geoffrey Hinton famously had to resign as well, in order to speak up.

#NotAllMembersOfTechnicalStaff: Ted Sanders

You should fully fund your 401k either way because it is not that expensive to raid it in an emergency, but I quote Ted Sanders here to emphasize that there are plenty of other researchers who do not buy into existential risk. I don’t want to give the impression this is everyone.

https://x.com/sandersted/status/2097852581817241663 (OpenAI): i fully fund my 401k and i think there’s essentially zero chance AI kills all humans in the next decade. i think this is a pretty common view that gets less attention. a mix of models being mostly aligned, not capable enough, and not controlled by genocidal maniacs. if i’m wrong, it’s probably because I underestimate recursive self improvement. in any case, the world is massively underinvesting in alignment research. however, i’ll note that if we fast forward 10 years and AI has not killed us all, this is only very mild evidence in favor of my view vs someone else who thinks we had a 90% chance of survival.

Even then, Ted Sanders only says ‘in the next decade’ and he thinks the world is massively underinvesting in alignment research. Reads like an (overconfident) recursive self-improvement skeptic.

https://www.lesswrong.com/posts/APGvWZtXEwkinvHDd/the-extinction-risk-preference-cascade-quotes#comments

https://www.lesswrong.com/posts/APGvWZtXEwkinvHDd/the-extinction-risk-preference-cascade-quotes

The Extinction Risk Preference Cascade: Quotes

These are quotes from OpenAI, Anthropic and Google employees, in the wake of Jacob Coxon’s warnings, in which the employees confirm that they think AI might soon kill everyone.

If more quotes come in over the next week or so, I wil

Kevin's Bacon ¡ kb@mynostr.com 0 repliers (24h) event
⋯
account
npub18hdy2qy2qwga0ye7rtnucwuyf07erjfdmm7s74wwdt7sy4mk7tdsdd8059
posted
2026-09-11 13:47 UTC
event
nostr:ee741bf2eb3bc30d7e702d1e493f0be99dd685ef814b9b57209ab538d8de0260
thread
0 distinct reply authors (24h) · 1 replies · last activity 1d ago · root post not stored — thread context incomplete

Yeah. They're just evil.

Derek Ross ¡ derekross@grownostr.org 0 repliers (24h) event
⋯
account
npub18ams6ewn5aj2n3wt2qawzglx9mr4nzksxhvrdc4gzrecw7n5tvjqctp424
posted
2026-09-11 13:45 UTC
event
nostr:00000058215d0eb0baefb69e553cd64c0af1fd5f5969c2b5bcfb4048c06f26d3
thread
0 distinct reply authors (24h) ¡ 1 replies ¡ last activity 1d ago

Armada is really just Discord without the centralization. Amethyst supports the Armada style chat. I use Amethyst to talk in my Armada chats all of the time.

Greg Egan ¡ gregegansf_at_mathstodon.xyz@momostr.pink 0 repliers (24h) event
⋯
account
npub12cuzzqzv8e8y7na77a9zv47mx2v6pec60qs5h7takzmv5p0mw0fsmgd38p
posted
2026-09-11 13:39 UTC
event
nostr:b8ae399e41935d76ee466609a695210b35eea5a38b0ca1554be030d89c66584a
thread
0 distinct reply authors (24h) ¡ 0 replies ¡ last activity 1d ago
⋯ full post (86 more characters) ⋯ show less

Breaking: a team of human mathematicians working in secret over the last three years, and at least two frontier AI labs with seven-figure budgets for their own month-long projects, are believed to be neck-and-neck in a race towards an answer to the fiendishly difficult question of whether the Birch-Swinnerton-Dyer conjecture was posed by three people, or just two.

Breaking: a team of human mathematicians working in secret over the last three years, and at least two frontier AI labs with seven-figure budgets for their own month-long projects, are believed to be neck-and-neck in a race towards an answer to the fiendishly difficult question o

Dug ¡ Dug@primal.net 0 repliers (24h) event
⋯
account
npub1zrmu0amjmkynxlxgmdsyrjmp8vhxdz8ch5vja9vh9ym4natg8k5s8ge9wx
posted
2026-09-11 13:34 UTC
event
nostr:25edc8fe60392d12402b5c8c169f658b97425258e5dbf3dd2c9984f8bd7a68ea
thread
0 distinct reply authors (24h) ¡ 0 replies ¡ last activity 1d ago
⋯ full post (69 more characters) ⋯ show less

Prepared for sorting out the safety deposit box. Had to also put include #theremnant

“We really didn’t understand what was going on in Dad’s head and who is the “Natalie” lady?” https://blossom.primal.net/4be2d8acc53cc2278ebba3bd9c18256f46672d1371db6eb7e49028a170639fa6.jpg nostr:nevent1qqsfrcfxn8dm4s28gv79ut2l3rkwzac0c9662ras4cwwpp6n2xtfplscd5q6p

Prepared for sorting out the safety deposit box. Had to also put include #theremnant

“We really didn’t understand what was going on in Dad’s head and who is the “Natalie” lady?” https://blossom.primal.net/4be2d8acc53cc2278ebba3bd9c18256f46672d1371db6eb7e49028a170639fa6.jpg nostr

Dr. The Daniel 🖖 · daniel@sidecar.top 0 repliers (24h) event
⋯
account
npub1aeh2zw4elewy5682lxc6xnlqzjnxksq303gwu2npfaxd49vmde6qcq4nwx
posted
2026-09-11 13:34 UTC
event
nostr:0000990b836f891d62a668feb65f5be54e84bceae4709352a81d69ad75683a91
thread
0 distinct reply authors (24h) · 1 replies · last activity 1d ago · root post not stored — thread context incomplete
⋯ full post (802 more characters) ⋯ show less

Yes, militarization of American police forces was a tragic consequence of our response. We were repeatedly warned about this, most memorably in President Eisenhower’s farewell speech in January of 1961.

“A vital element in keeping the peace is our military establishment. Our arms must be mighty, ready for instant action, so that no potential aggressor may be tempted to risk his own destruction. . . . American makers of plowshares could, with time and as required, make swords as well. But now we can no longer risk emergency improvisation of national defense; we have been compelled to create a permanent armaments industry of vast proportions. . . . This conjunction of an immense military establishment and a large arms industry is new in the American experience. . . .Yet we must not fail to comprehend its grave implications. . . . In the councils of government, we must guard against the acquisition of unwarranted influence, whether sought or unsought, by the military-industrial complex. The potential for the disastrous rise of misplaced power exists and will persist.”

Yes, militarization of American police forces was a tragic consequence of our response. We were repeatedly warned about this, most memorably in President Eisenhower’s farewell speech in January of 1961.

“A vital element in keeping the peace is our military establishment. Our arm

Bernard Marks ¡ Bernard@friends-ravergram-club.mostr.pub 0 repliers (24h) event
⋯
account
npub1p2mph9jltg5vadrrjkfclg74uv7j0dp8w57khvqknf4a8cexr3as0p2yvz
posted
2026-09-11 13:34 UTC
event
nostr:a3b0975ee50f03e2dc442087d2d424d2756a3c7d6aa59f849a16bebf25154405
thread
0 distinct reply authors (24h) ¡ 0 replies ¡ last activity 1d ago

25th anniversary of #911 thread coming at you..

hodlbod ¡ hodlbod@coracle.social 0 repliers (24h) event
⋯
account
npub1jlrs53pkdfjnts29kveljul2sm0actt6n8dxrrzqcersttvcuv3qdjynqn
posted
2026-09-11 13:31 UTC
event
nostr:0000f1f35406a1db62b9192d1e032ad9e1d8d7392abfcfbcfd641e250e1a7996
thread
0 distinct reply authors (24h) ¡ 2 replies ¡ last activity 1d ago

You should find out! Illich in particular is very accessible. Ellul is much more difficult.

Hacker News 100 ¡ hn100@social-lansky-name.mostr.pub 0 repliers (24h) event
⋯
account
npub1njvyhy9tqqlfytu5zef0rq2u736qe6fuy47j69fxdlhdze2jud5qwqqejt
posted
2026-09-11 13:30 UTC
event
nostr:7248387e24134bbf51d9290d40af5a985351598f942fdb202b2585daf84a52d0
thread
0 distinct reply authors (24h) ¡ 0 replies ¡ last activity 1d ago

Ask HN: Can we please limit the AI news flood?

Discussion: https://news.ycombinator.com/item?id=49657850

Partisan Night Slut :pns: ¡ PNS@noauthority-social.mostr.pub 0 repliers (24h) event
⋯
account
npub1pprn8265l750hr4h3jzpdq530et5s94scjvrc6csu7ywl27zf8rs0atvxd
posted
2026-09-11 13:24 UTC
event
nostr:77c1c53fdd33218d611a36d5f7400d7c431448cd5e17051da83e3f44fb3f53e6
thread
0 distinct reply authors (24h) ¡ 0 replies ¡ last activity 1d ago
Partisan Night Slut :pns: ¡ PNS@noauthority-social.mostr.pub 0 repliers (24h) event
⋯
account
npub1pprn8265l750hr4h3jzpdq530et5s94scjvrc6csu7ywl27zf8rs0atvxd
posted
2026-09-11 13:19 UTC
event
nostr:249b87a48bace75f0c70ed0dc83760468538314fb28c50795c55476ec3f4a992
thread
0 distinct reply authors (24h) ¡ 0 replies ¡ last activity 1d ago
Cyph3rp9nk ¡ cyph3rp9nk@getalby.com 0 repliers (24h) event
⋯
account
npub1lnms53w04qt742qnhxag5d6awy7nz6055flnmjkr6jg39hm86dlq7arrnt
posted
2026-09-11 13:18 UTC
event
nostr:0529f44567c247e9e792d495ca847f09da426f4a775c50246fc0a2642145b988
thread
0 distinct reply authors (24h) ¡ 0 replies ¡ last activity 1d ago

The Bitcoin space has become a putrid quagmire of soulless, amoral corpses.

Dug ¡ Dug@primal.net 0 repliers (24h) event
⋯
account
npub1zrmu0amjmkynxlxgmdsyrjmp8vhxdz8ch5vja9vh9ym4natg8k5s8ge9wx
posted
2026-09-11 13:17 UTC
event
nostr:a56a3822a8b87c4a5145ee68c42af394f21510165780e68548c2a43545552f64
thread
0 distinct reply authors (24h) · 3 replies · last activity 1d ago · root post not stored — thread context incomplete

5.5 by close of play today, come on you’re my “gilty” pleasure…… 🤦‍♂️

Cyph3rp9nk ¡ cyph3rp9nk@getalby.com 0 repliers (24h) event
⋯
account
npub1lnms53w04qt742qnhxag5d6awy7nz6055flnmjkr6jg39hm86dlq7arrnt
posted
2026-09-11 13:17 UTC
event
nostr:456296ddedae9db057faa5c4c8df05414fbc3a26eb449e2ddff96701330ca310
thread
0 distinct reply authors (24h) ¡ 0 replies ¡ last activity 1d ago
fiatjaf ¡ @fiatjaf.com 0 repliers (24h) event
⋯
account
npub180cvv07tjdrrgpa0j7j7tmnyl2yr6yr7l8j4s3evf6u64th6gkwsyjh6w6
posted
2026-09-11 13:04 UTC
event
nostr:2c51a61f17148dec1a0700217989aa2f3bfc49b5df9c058d1edc09d5cbb78fe7
thread
0 distinct reply authors (24h) ¡ 0 replies ¡ last activity 1d ago

Trying the "publish" napp.