LessWrong - Rationality and AI Safety Community Forum
blogCredibility Rating
Good quality. Reputable source with community review or editorial standards, but less rigorous than peer-reviewed venues.
Rating inherited from publication venue: LessWrong
LessWrong is the central community hub for AI safety and rationality research, hosting foundational technical posts, research updates, and philosophical discussions from leading AI safety researchers including Yudkowsky, Christiano, and others.
Metadata
Summary
LessWrong is a community blog and forum that serves as a primary venue for developing and disseminating AI safety ideas, rationality research, and existential risk discussions. It hosts foundational technical posts, research updates, and philosophical debates from prominent researchers. The platform has been instrumental in shaping key AI alignment concepts and connecting the global AI safety research community.
Key Points
- •Central hub for AI safety and rationality discourse, hosting posts from leading researchers like Yudkowsky, Christiano, and others
- •Features a wide range of content from technical alignment research to philosophical discussions on decision theory and existential risk
- •Hosts foundational sequences (e.g., Rationality: A-Z, The Codex) that have shaped the AI safety field's intellectual foundations
- •Active community with quick takes, shortform posts, and discussion threads enabling rapid exchange of research ideas
- •Serves as an early dissemination platform for AI safety concepts before formal publication in academic venues
Cited by 2 pages
| Page | Type | Quality |
|---|---|---|
| Why Alignment Might Be Easy | Argument | 53.0 |
| Machine Intelligence Research Institute (MIRI) | Organization | 50.0 |
Cached Content Preview
LessWrong x This website requires javascript to properly function. Consider activating javascript to get access to all site functionality. Home All Posts Concepts Library Best of LessWrong Sequence Highlights Rationality: A-Z The Codex HPMOR Community Events Subscribe (RSS/Email) LW the Album Leaderboard About FAQ Home All Posts Concepts Library Community About Recent Enriched Recommended Rationality + World Modeling + Personal Blog + Quick Takes
Your Feed For You Following AI World Optimization Practical Community Dartmouth College – College EA Meetups Everywhere Fall 2026 Mon Sep 21 Monday Social 7pm-9pm @ Segundo Coffee Lab Tue Sep 22 • Houston 554 Welcome to LessWrong! Ruby , Raemon , RobertM , habryka 7y 85 Reframing Impact Best of LessWrong 2019 Impact measures may be a powerful safeguard for AI systems - one that doesn't require solving the full alignment problem. But what exactly is "impact", and how can we measure it?
Reframing Impact Best of LessWrong 2019 Impact measures may be a powerful safeguard for AI systems - one that doesn't require solving the full alignment problem. But what exactly is "impact", and how can we measure it?
‘A Thousand AI Constitutions’ — Simon Goldstein on Constitutional Diversification for Frontier AI. [HKU Talk - 22 Sep]
Tue Sep 22 • Online ACX/LessWrong Budapest meetup September 27, 2 pm, Gergo's place Sun Sep 27 • Budapest 570 The Talker Does Not Control The Doer (in Current AIs) Eliezer Yudkowsky 5d 76 194 Explaining Knightianism on one foot Richard_Ngo 10d 27 455 If Anyone Builds It, Everyone Dies: One Year Closer Eliezer Yudkowsky , So8res , Duncan Sabien (Inactive) 5d 31 570 The Talker Does Not Control The Doer (in Current AIs) Eliezer Yudkowsky 5d 76 1185 Why I Left Google DeepMind Ω TurnTrout 2mo Ω 57 497 Astra and Fable still hack on simple variants of alignment evals from 2025 Dean Valentine 13d 32 168 Why I Stay Off Twitter jefftk 2d 15 347 There is a channel to 900M weekly users. What goes in it? Charbel-Raphaël 7d 30 779 How My Students Think About AI dvd 19d 111 592 Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incident Ω ryan_greenblatt , Ajeya Cotra , Hjalmar_Wijk 23d Ω 67 272 Quick notes from teaching technical profiles how to talk in public Camille B. 6d 9 114 We've saved the world before: what the ozone hole teaches us about AI leogao 1d 12 640 What just happened? A retrospective of AI alignment Richard_Ngo 1mo 152 131 Common mistakes in AI safety group organizing Nikola Jurkovic 2d 4 475 What just happened? Pragmatism and Pessimization Richard_Ngo 1mo 176 Cancel Submit Caleb Biddulph 10h 42 16 ChosunOne 1 Even if you could make "perfectly realistic" safety evals, eval awareness would still be a problem.
Suppose you are worried that your AI might display a certain catastrophically harmful behavior on rare occasions. Unfortunately, it is hard to test whether this is the case, because your AI might avoid display
... (truncated, 10 KB total)815315aec82a6f7f | Stable ID: sid_9hIGf0AnHJ