LessWrong (30+ Karma)

LessWrong

Audio narrations of LessWrong posts.

  1. 1 ngày trước

    “Pretraining data, not verifiability, is why LLMs are especially good at math (and coding)” by Steven Byrnes

    Follow-up to: “LLMs are (still) mostly powered by imitative learning, not RL” A common take I’ve been hearing is: “LLMs are especially good at math because math is easy to verify”. But that story doesn’t make much sense. For one thing, “easy to verify” only matters for the RL part of LLM training pipelines, and the leading LLM companies have said that they spend very little effort on RL-for-math.Worse, to the extent that the companies are doing RL-for-math, it's RLAIF, not RLVR. So really, the phrase “math is easy to verify” amounts to “LLMs are very good at judging math arguments”. But that's begging the question! Why are pretrained LLMs so much better at judging math arguments than judging, say, fiction writing? We still need an answer. So here's a different theory, in the framework of my earlier post “LLMs are (still) mostly powered by imitative learning, not RL”: LLMs are especially good at math because almost everything in the math literature is correct. Read a random sentence in a random math paper in the research math literature, and you can be >99% confident that the sentence is true. So if LLMs do what they do best—imitative [...] The original text contained 3 footnotes which were omitted from this narration. --- First published: September 18th, 2026 Source: https://www.lesswrong.com/posts/xvdngZAqFZfek7KGH/pretraining-data-not-verifiability-is-why-llms-are --- Narrated by TYPE III AUDIO.

  2. 1 ngày trước

    “Stopgap Measures to Address Immediate AI Security Threats” by Andrea_Miotti, Gabriel Alfour

    Most people know AI as the technology behind chatbots like ChatGPT. However, what the top AI companies are explicitly aiming for is something else entirely: superintelligent AI. That means AI that can fully replace and outmatch humans at any task, including in domains like hacking, social engineering, and military operations. Such an AI system, if developed, could autonomously overpower any country's national security forces. No company, no government, no individual knows how to keep such a system under human control. This is why the world's leading AI experts, Nobel Prize winners, and even the CEOs of the top AI companies warn that the development of superintelligence threatens humanity with extinction, and why more than 800 scientists, former military leaders, and public figures have called for a prohibition on developing superintelligence. This is not a distant prospect: AI companies such as OpenAI and Anthropic are investing billions of dollars into superintelligence and aiming to develop it within the next few years. Former Anthropic and OpenAI researcher Jacob Coxon, who resigned last week, stated that AI companies are “racing straight to self-improving superintelligence and gambling with our lives” and that people at the companies themselves believe it “could kill [...] --- Outline: (04:34) Secure Weapons-Grade AI Against Theft by Adversaries (07:46) Necessary Measure: Registration (08:43) Sufficient Measure: Government Security Testing (09:42) Thorough Measure: Development Requires Government Authorization (10:52) Criminal Liability for Leaks During AI Gain-of-Function Research (14:09) Necessary Measure: Team Liability (14:46) Sufficient Measure: Chain of Command Liability (15:21) Thorough Measure: Company Liability (16:00) Kill-Switches to Contain Critical AI Incidents (19:08) Necessary Measure: Company Kill-Switch (19:58) Sufficient Measure: Infrastructure Kill-Switch (20:53) Thorough Measure: International Kill-Switches (22:24) Conclusion --- First published: September 18th, 2026 Source: https://www.lesswrong.com/posts/LqBAxFdyAiybnPL8e/stopgap-measures-to-address-immediate-ai-security-threats --- Narrated by TYPE III AUDIO.

  3. 1 ngày trước

    “The Preference Cascade Is Only Getting Started” by Zvi

    We are in the midst of a preference cascade about existential risk from AI. A preference cascade is, alas, the best method we have to change the debate. The avalanche has started. There is still time for the pebbles to vote. For now. Mike Solana gave the correct view of why Coxon's post went viral, which is that enough Americans finally have enough context on AI to care, and there were enough big accounts that were happy to amplify the Tweet quickly to get it initial attention. That is all you need when there is enough dry tinder. What we must realize is that the current preference cascade, on the need to Pace the Frontier, is insufficient. If we are to make it out of this alive, we will have to do better. We have to, as Dan Selsam warns, actually solve the underlying problems. The next step is to continue the cascade. That includes inside the labs, and also among the media and politics. It includes both people who previously focused on other things stepping up and new voices being heard. A lot of that will be overcoming the inevitable political opposition [...] --- Outline: (01:38) The Cascade Was a Long Time Coming (02:59) The Cascade Has Reached The People (04:38) Elon Musk Doubles Down (05:18) Matthew Yglesias Steps Up (08:42) Op Eds and Posts Are Written (11:05) Jacob Coxon AMA (17:35) Bilal Chughtai Quits DeepMind and Sounds the Alarm (20:29) The Cascade Is Insufficient (21:47) What Would It Take (28:10) OpenAI's Dan Selsam Sounds A Louder Alarm (41:59) Some People Worry On Meta Levels You Never Imagined (43:10) Two Kinds of Threats (44:41) The Two Towers and The Narrow Path (46:45) A Specific, Detailed Story About AI Killing Everyone That Doesn't Sound To Me Like Science Fiction (50:06) What Can I Do About It? --- First published: September 18th, 2026 Source: https://www.lesswrong.com/posts/sDiSqZmctQ78hsLcP/the-preference-cascade-is-only-getting-started --- Narrated by TYPE III AUDIO. --- Images from the article: Apple Podcasts and Spotify do not show images in the episode description. Try Pocket Casts, or another podcast app.

  4. 1 ngày trước

    “The Game is Set for a Targeted Memetic Attack on the AI Safety Community” by keltan

    While this is relevant to my work at MIRI, I have not checked these ideas with anyone else on the team and am posting this on my personal LW account. These views are my own. And to be honest, I am writing this mostly to remind myself of my weakness. --- I expect one (or many) adversarial memetic attacks aiming to trip you up, perhaps consisting of fake leaks relating to dangerous stuff happening in the labs. Specifically, worrying incidents that may fit snugly within your worldview, leaking from multiple sources including news outlet/s, but not confirmed/confirmable by a primary source. Think rumors about exfiltrated weights, AIs attempting to create viruses, agent swarms hacking into and gathering information from nuclear infrastructure, etc. An easy way to remove status from a movement is to trip it up: make it fall for a misinformation trap in public, then use that slip-up to discredit the movement for all time. The game is set for a memetic attack like this. There's a well-resourced group waiting for your screw-up. And then you may remember much that will help you.  In public and in private, if you feel surprised or confused, notice your confusion. These [...] The original text contained 2 footnotes which were omitted from this narration. --- First published: September 17th, 2026 Source: https://www.lesswrong.com/posts/5mcDjo5gjn3Leahhu/the-game-is-set-for-a-targeted-memetic-attack-on-the-ai --- Narrated by TYPE III AUDIO.

Giới Thiệu

Audio narrations of LessWrong posts.

Có Thể Bạn Cũng Thích