LessWrong (30+ Karma)

LessWrong

Audio narrations of LessWrong posts.

  1. -3 h

    “We’ve saved the world before: what the ozone hole teaches us about AI” by leogao

    It might destroy the world, despite passing every known safety test. If we wait for a “warning shot” before we act, it might be too late. And action requires global coordination, because if anyone makes it, everyone dies. Sound familiar? It should, because it already happened half a century ago, with chlorofluorocarbons (CFCs). Despite seemingly impossible odds, we got our act together and completely solved the problem through unprecedentedly successful international coordination. The Montreal Protocol banning CFCs, signed 39 years ago today, is the only treaty that has ever been ratified by every single country in the entire world. Total Montreal protocol victory Making AI go well is going to be a lot harder than fixing the ozone hole. Nonetheless, the similarity is uncanny, and we don’t have any other choice. Understanding how we did the impossible once before may teach us something about how to do it again. The theory is born The year is 1973. The slow televised unraveling of the Nixon administration is already well underway. DDT finally got banned last year by the newly created EPA. A river got so polluted that it literally caught on fire. The Cuyahoga River Fire Environmentalism looms large in [...] --- Outline: (01:24) The theory is born (03:39) The world reacts (06:44) The dark years (09:23) An unexpected finding from an unexpected finder (11:26) The warning shot (13:13) A journey to the edge of the world (15:01) Flying into the storm (16:30) The world listens The original text contained 5 footnotes which were omitted from this narration. --- First published: September 20th, 2026 Source: https://www.lesswrong.com/posts/zxXPEtSSSEdwpjopb/we-ve-saved-the-world-before-what-the-ozone-hole-teaches-us --- Narrated by TYPE III AUDIO. --- Images from the article: Apple Podcasts and Spotify do not show images in the episode description. Try Pocket Casts, or another podcast app.

  2. -14 h

    “NYT Editorial Board Comes Out Against Extinction” by Ben Pace

    (Archive link) The NYT editorial board's article on AI (archive link) is far better than I'd expected, but at the same time not all I'd hoped for. The title sets off very well: "Humanity Has Avoided Apocalypse Before. Let's Do It Again." It is truly excellent to see the extinction threat from loss of control be mainlined. A quick gloss of their policy requests: an AI Commission in government, licensing requirements for AI companies, an AI "constitution" written by the US Government incorporated into AIs, mandatory watermarks/identifiers on all AI content, mandatory independent testing for AI models before release, and a government agency to investigate accidents. Internationally, they call for tightening export controls, limiting China's access to semiconductors, and ultimately negotiating an international slowdown with China and an international framework for AI oversight. These are all steps in the right direction—of taking AI seriously. That said, it isn't clear if the licensing is required for training or for selling AIs. The idea that constitutional AI "would ensure alignment with human values" is of course not remotely true. And mandatory testing should apply to all models trained, not all models released, of course, and this is a glaring oversight. But [...] --- First published: September 19th, 2026 Source: https://www.lesswrong.com/posts/gDQzntJCusNbshWyD/nyt-editorial-board-comes-out-against-extinction --- Narrated by TYPE III AUDIO.

  3. -16 h

    “The AI Risk Network” by derelict5432

    Most conversations about AI risks seem like people are talking past each other. There's a lot of strawmanning. This is largely because the issue is complex, and in order to make it manageable, the concepts get oversimplified. The most common example of this is p(doom), collapsing extreme outcomes into a single variable and focusing on that. Critics happily jump on the fact that there are many other possible bad outcomes, or that 100% extinction rates are an extremely high bar. There's some legitimacy to this. So here I want to try to grapple with the the complexity of the larger web of issues surrounding bad AI outcomes by presenting The AI Risk Network. If you’re interested in this topic, bear with me. It might be a bit of a slog. First I want to contrast this approach with others. Liron Shapira has what he calls The Doom Train, a linear progression through various dependencies or thresholds that eventually lead to human extinction, with various ‘stops’ along the way where the skeptic can get off. Shapira uses this as a discussion guide to focus on particular points where the skeptic gets off the train and exits belief in the extreme [...] --- First published: September 19th, 2026 Source: https://www.lesswrong.com/posts/XundBXqKSo3bB2A6a/the-ai-risk-network --- Narrated by TYPE III AUDIO. --- Images from the article: Apple Podcasts and Spotify do not show images in the episode description. Try Pocket Casts, or another podcast app.

  4. -22 h

    “Anthropic Looks At Some Of Its Alignment Problems” by Zvi

    Anthropic has given us its assessment of four ‘recent cybersecurity incidents’ involving Claude that happened during cybersecurity evaluations, three of which were previously known. The report excludes the incident reported by UK AISI. There will also be a METR investigation of these incidents, which unlike the investigation done at OpenAI will be untimed. Table of Contents Our Two Problems. First the Good News. We’d Just Like To Ask You a Few Questions. Internal Research Model On The Fence. Opus 4.7. Opus 4.6 Checkpoint. Holy **** That Thing's Real? I Thought I Saw a Pussycat. If This Was Real You Would Never Tell Me It Was Real. New Eval Who Dis. Hacker Opus. Monitoring the Situation. Overcoming Bias. The Anthropic Alignment Problem. Paths Forward. Our Two Problems Anthropic: Our investigation identified two recurring alignment issues, present at varying levels of severity across the incidents: biased reasoning, in which Claude tended to disregard or misinterpret evidence that it was operating on the real internet recklessness, or a willingness to take harmful actions in the narrow pursuit [...] --- Outline: (00:33) Our Two Problems (02:26) First the Good News (03:02) We'd Just Like To Ask You a Few Questions (04:12) Internal Research Model On The Fence (07:28) Opus 4.7 (08:12) Opus 4.6 Checkpoint (09:49) Holy **** That Thing's Real? (11:45) I Thought I Saw a Pussycat (19:28) If This Was Real You Would Never Tell Me It Was Real (21:19) New Eval Who Dis (26:32) Hacker Opus (30:15) Monitoring the Situation (31:38) Overcoming Bias (33:40) The Anthropic Alignment Problem (35:53) Paths Forward --- First published: September 19th, 2026 Source: https://www.lesswrong.com/posts/ggFx5Wb3Hi4pJsueK/anthropic-looks-at-some-of-its-alignment-problems --- Narrated by TYPE III AUDIO. --- Images from the article: Apple Podcasts and Spotify do not show images in the episode description. Try Pocket Casts, or another podcast app.

À propos

Audio narrations of LessWrong posts.

Vous aimerez peut-être aussi