LessWrong

LessWrong (Curated & Popular)

Audio narrations of LessWrong posts. Includes all curated posts and all posts with 125+ karma. If you'd like more, subscribe to the “Lesswrong (30+ karma)” feed.

Author

LessWrong

Category

Technology

Podcast website

sites.libsyn.com

Latest episode

Jul 10, 2026

Where to listen?

Podcasts in the app Replaio Radio Coming soon

Podcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts

Get it on Google Play Install for free Android 5M+ downloads · 4.8 rating iOS soon

Episodes

MIRI 2024 Mission and Strategy Update 05.01.2024

As we announced back in October, I have taken on the senior leadership role at MIRI as its CEO. It's a big pair of shoes to fill, and an awesome responsibility that I’m honored to take on. There have been several changes at MIRI since our 2020 strategic update, so let's get into it.[1] The short version: We think it's very unlikely that the AI alignment field will be able to make pr...

The Plan - 2023 Version 04.01.2024

Background: The Plan, The Plan: 2022 Update. If you haven’t read those, don’t worry, we’re going to go through things from the top this year, and with moderately more detail than before. 1. What's Your Plan For AI Alignment? Median happy trajectory: Sort out our fundamental confusions about agency and abstraction enough to do interpretability that works and generalizes robustly. Look through...

Apologizing is a Core Rationalist Skill 03.01.2024

In certain circumstances, apologizing can also be a countersignalling power-move, i.e. “I am so high status that I can grovel a bit without anybody mistaking me for a general groveller”. But that's not really the type of move this post is focused on. There's this narrative about a tradeoff between: The virtue of Saying Oops, early and often, correcting course rather than continuing to po...

[HUMAN VOICE] "A case for AI alignment being difficult" by jessicata 02.01.2024

This is a linkpost for https://unstableontology.com/2023/12/31/a-case-for-ai-alignment-being-difficult/ Support ongoing human narrations of LessWrong's curated posts: www.patreon.com/LWCurated This is an attempt to distill a model of AGI alignment that I have gained primarily from thinkers such as Eliezer Yudkowsky (and to a lesser extent Paul Christiano), but explained in my own terms rather...

The Dark Arts 01.01.2024

lsusrIt is my understanding that you won all of your public forum debates this year. That's very impressive. I thought it would be interesting to discuss some of the techniques you used. LyrongolemOf course! So, just for a brief overview for those who don't know, public forum is a 2v2 debate format, usually on a policy topic. One of the more interesting ones has been the last one I went...

Critical review of Christiano’s disagreements with Yudkowsky 28.12.2023

Crossposted from the AI Alignment Forum. May contain more technical jargon than usual. This is a review of Paul Christiano's article "where I agree and disagree with Eliezer". Written for the LessWrong 2022 Review. In the existential AI safety community, there is an ongoing debate between positions situated differently on some axis which doesn't have a common agreed-upon name,...

Most People Don’t Realize We Have No Idea How Our AIs Work 27.12.2023

This point feels fairly obvious, yet seems worth stating explicitly. Those of us familiar with the field of AI after the deep-learning revolution know perfectly well that we have no idea how our ML models work. Sure, we have an understanding of the dynamics of training loops and SGD's properties, and we know how ML models' architectures work. But we don't know what specific algorith...

Discussion: Challenges with Unsupervised LLM Knowledge Discovery 26.12.2023

TL;DR: Contrast-consistent search (CCS) seemed exciting to us and we were keen to apply it. At this point, we think it is unlikely to be directly helpful for implementations of alignment strategies (>95%). Instead of finding knowledge, it seems to find the most prominent feature. We are less sure about the wider category of unsupervised consistency-based methods, but tend to think they won’t be...

Succession 24.12.2023

This is a linkpost for https://www.narrativeark.xyz/p/succession“A table beside the evening sea where you sit shelling pistachios, flicking the next open with the half- shell of the last, story opening story, on down to the sandy end of time.” V1: Leaving Deceleration is the hardest part. Even after burning almost all of my fuel, I’m still coming in at 0.8c. I’ve planned a slingshot around the gal...

Nonlinear’s Evidence: Debunking False and Misleading Claims 21.12.2023

Recently, Ben Pace wrote a well-intentioned blog post mostly based on complaints from 2 (of 21) Nonlinear employees who 1) wanted more money, 2) felt socially isolated, and 3) felt persecuted/oppressed. Of relevance, one has accused the majority of her previous employers, and 28 people of abuse - that we know of. She has accused multiple people of threatening to kill her and literally accused an e...

Effective Aspersions: How the Nonlinear Investigation Went Wrong 20.12.2023

The New York Times Picture a scene: the New York Times is releasing an article on Effective Altruism (EA) with an express goal to dig up every piece of negative information they can find. They contact Émile Torres, David Gerard, and Timnit Gebru, collect evidence about Sam Bankman-Fried, the OpenAI board blowup, and Pasek's Doom, start calling Astral Codex Ten (ACX) readers to ask them about...

Constellations are Younger than Continents 20.12.2023

At the Bay Area Solstice, I heard the song Bold Orion for the first time. I like it a lot. It does, however, have one problem: He has seen the rise and fall of kings and continents and all, Rising silent, bold Orion on the rise. Orion has not witnessed the rise and fall of continents. Constellations are younger than continents. The time scale that continents change on is ten or hundreds of million...

The ‘Neglected Approaches’ Approach: AE Studio’s Alignment Agenda 19.12.2023

Many thanks to Samuel Hammond, Cate Hall, Beren Millidge, Steve Byrnes, Lucius Bushnaq, Joar Skalse, Kyle Gracey, Gunnar Zarncke, Ross Nordby, David Lambert, Simeon Campos, Bogdan Ionut-Cirstea, Ryan Kidd, Eric Ho, and Ashwin Acharya for critical comments and suggestions on earlier drafts of this agenda, as well as Philip Gubbins, Diogo de Lucena, Rob Luke, and Mason Seale from AE Studio for their...

“Humanity vs. AGI” Will Never Look Like “Humanity vs. AGI” to Humanity 18.12.2023

When discussing AGI Risk, people often talk about it in terms of a war between humanity and an AGI. Comparisons between the amounts of resources at both sides' disposal are brought up and factored in, big impressive nuclear stockpiles are sometimes waved around, etc. I'm pretty sure it's not how that'd look like, on several levels. 1. Threat Ambiguity I think what people imagin...

Is being sexy for your homies? 17.12.2023

Epistemic status: Speculation. An unholy union of evo psych, introspection, random stuff I happen to observe & hear about, and thinking. Done on a highly charged topic. Caveat emptor! Most of my life, whenever I'd felt sexually unwanted, I'd start planning to get fit. Specifically to shape my body so it looks hot. Like the muscly guys I'd see in action films. This choice is a li...

[HUMAN VOICE] "Significantly Enhancing Adult Intelligence With Gene Editing May Be Possible" by Gene Smith and Kman 17.12.2023

Support ongoing human narrations of LessWrong's curated posts: www.patreon.com/LWCurated TL;DR version In the course of my life, there have been a handful of times I discovered an idea that changed the way I thought about the world. The first occurred when I picked up Nick Bostrom’s book “superintelligence” and realized that AI would utterly transform the world. The second was when I learned...

[HUMAN VOICE] "Moral Reality Check (a short story)" by jessicata 15.12.2023

Support ongoing human narrations of LessWrong's curated posts: www.patreon.com/LWCurated This is a linkpost for https://unstableontology.com/2023/11/26/moral-reality-check/ Janet sat at her corporate ExxenAI computer, viewing some training performance statistics. ExxenAI was a major player in the generative AI space, with multimodal language, image, audio, and video AIs. They had scaled up op...

AI Control: Improving Safety Despite Intentional Subversion 15.12.2023

Crossposted from the AI Alignment Forum. May contain more technical jargon than usual. We’ve released a paper, AI Control: Improving Safety Despite Intentional Subversion. This paper explores techniques that prevent AI catastrophes even if AI instances are colluding to subvert the safety techniques. In this post: We summarize the paper; We compare our methodology to what the one used in other safe...

2023 Unofficial LessWrong Census/Survey 13.12.2023

The Less Wrong General Census is unofficially here! You can take it at this link. It's that time again. If you are reading this post and identify as a LessWronger, then you are the target audience. I'd appreciate it if you took the survey. If you post, if you comment, if you lurk, if you don't actually read the site that much but you do read a bunch of the other rationalist blogs or...

The likely first longevity drug is based on sketchy science. This is bad for science and bad for longevity. 13.12.2023

If you are interested in the longevity scene, like I am, you probably have seen press releases about the dog longevity company, Loyal for Dogs, getting a nod for efficacy from the FDA. These have come in the form of the New York Post calling the drug "groundbreaking", Science Alert calling the drug "radical", and the more sedate New York Times just asking, "Could Longevity...

[HUMAN VOICE] "What are the results of more parental supervision and less outdoor play?" by Julia Wise 13.12.2023

Support ongoing human narrations of LessWrong's curated posts: www.patreon.com/LWCurated Crossposted from Otherwise Parents supervise their children way more than they used to Children spend less of their time in unstructured play than they did in past generations. Parental supervision is way up. The wild thing is that this is true even while the number of children per family has decreased an...

Significantly Enhancing Adult Intelligence With Gene Editing May Be Possible 12.12.2023

In the course of my life, there have been a handful of times I discovered an idea that changed the way I thought about the world. The first occurred when I picked up Nick Bostrom's book “superintelligence” and realized that AI would utterly transform the world. The second was when I learned about embryo selection and how it could change future generations. And the third happened a few months...

re: Yudkowsky on biological materials 11.12.2023

I was asked to respond to this comment by Eliezer Yudkowsky. This post is partly redundant with my previous post. Why is flesh weaker than diamond? When trying to resolve disagreements, I find that precision is important. Tensile strength, compressive strength, and impact strength are different. Material microstructure matters. Poorly-sintered diamond crystals could crumble like sand, and a large...

Speaking to Congressional staffers about AI risk 05.12.2023

In May and June of 2023, I (Akash) had about 50-70 meetings about AI risks with congressional staffers. I had been meaning to write a post reflecting on the experience and some of my takeaways, and I figured it could be a good topic for a LessWrong dialogue. I saw that hath had offered to do LW dialogues with folks, and I reached out. In this dialogue, we discuss how I decided to chat with staffer...

[HUMAN VOICE] "Shallow review of live agendas in alignment & safety" by technicalities & Stag 04.12.2023

Support ongoing human narrations of LessWrong's curated posts: www.patreon.com/LWCurated You can’t optimise an allocation of resources if you don’t know what the current one is. Existing maps of alignment research are mostly too old to guide you and the field has nearly no ratchet , no common knowledge of what everyone is doing and why, what is abandoned and why, what is renamed, what relates...

Listen to the LessWrong (Curated & Popular) podcast in Replaio

Radio and podcasts in one app - free, with no sign-up. Install today and do not miss the launch

Get it on Google Play

Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.