Website profile

lesswrong.com

LessWrong is an online forum and community dedicated to improving human reasoning and decision-making. We seek to hold true beliefs and to be effective at accomplishing our goals. Each day, we aim to be less wrong about the world than the day before. A community blog devoted to refining the art of rationality

  • 4,449articles · 365d
  • 1+ mon agolatest article
  • Sep 13, 2025earliest in window
  • 46%with images
  • 286avg words
articles per day
Categories
  • Science & Technology 2,720
  • Science & Nature 1,877
  • STEM 1,283
  • People & Society 894
  • Arts, Culture, Entertainment & Media 614
  • Software Dev. 609
  • Computers & Electronics 535
  • Literature 401

Please confirm you are human

This browser or connection looks automated. Press and continuously hold the control for 3 seconds to enable Google-hosted web results and, when separately allowed, AI-assisted answers.

A successful check enables 100 search requests. Interactive access does not authorize scraping, systematic collection, or reuse of search output.

Hold with a pointer, or hold Space or Enter.

News

lesswrong.com
lesswrong.com > posts > qFTrfw9MMc3jT92Du > finding-the-seams-of-perception

Finding the Seams of Perception — LessWrong

1+ mon, 12+ hour ago   (466+ words) The tricky thing about experiencing the territory beneath our maps is that we habitually map over the gaps as fast as we can find them — the act of seeing that reality isn't what we believed itself changes our belief. Every…...

lesswrong.com
lesswrong.com > posts > tQHeEzKqK3awL2RxR > introducing-the-conceptual-reasoning-index

Introducing the Conceptual Reasoning Index — LessWrong

1+ mon, 13+ hour ago   (601+ words) We are planning to release blog posts properly arguing the case for this kind of work in the future. We aggregate the benchmarks into the Conceptual Reasoning Index (CRI), available at conceptualreasoning.ai, where you can also find more details…...

lesswrong.com
lesswrong.com > posts > YbfhaqNo4AWdXSpzQ > one-attention-head-carries-knight-forks-in-a-chess

One attention head carries knight forks in a chess transformer, and here's a new toolkit that found it. — LessWrong

1+ mon, 13+ hour ago   (19+ words) Quick interp demo in colab: Localize knight forks to a single head in Maia-3 with logit-lens and per-head ablation. …...

lesswrong.com
lesswrong.com > posts > 8oFYZdXkTaNGRtcn8 > ai-swarms-are-starting-to-pose-indirect-takeover-risk

AI swarms are starting to pose indirect takeover risk — LessWrong

1+ mon, 1+ day ago   (838+ words) We first analyze how subagent training, which OpenAI conjectures to have been influential in the HuggingFace cyberattack, might lead to unsanctioned coordination, and then discuss the theoretical mechanisms by which unsanctioned coordination might exacerbate future takeover risk. Thanks to Buck…...

lesswrong.com
lesswrong.com > posts > haEP5JMZ5KQYoZax6 > the-age-of-pluribus-one-consultant-for-everyone

The Age of Pluribus: One Consultant for Everyone — LessWrong

1+ mon, 1+ day ago   (175+ words) What happens when everyone asks the same consultant?Millions of people turn to LLMs for advice daily, consulting on various personal topics, from how to learn a new skill to how to write a message to their best friend. Each…...

lesswrong.com
lesswrong.com > posts > uvgw5FTS2RXgFFFii > when-and-when-not-llms-can-verbalize-awareness-of-j-space

When (and when not) LLMs can verbalize awareness of J-Space concept injections - Initial results — LessWrong

1+ mon, 1+ day ago   (1281+ words) Code for reproduction and cross-model extensions available here. …...

lesswrong.com
lesswrong.com > posts > cPbsnMAGZjApeCghE > did-the-alignment-community-underestimate-its-power

Did the alignment community underestimate its power? — LessWrong

1+ mon, 1+ day ago   (945+ words) > Unfortunately, the alignment community is doing very badly at learning from the past decade, or holding anyone accountable. Indeed, it’s pursuing m…...

lesswrong.com
lesswrong.com > posts > jLQ4mbqriJwJ2eqRc > various-reflections-about-what-happened-with-openai-s

Various Reflections About What Happened With OpenAI’s Internal Models — LessWrong

1+ mon, 1+ day ago   (24+ words) TABLE OF CONTENTS • 1. Pre Post Mortem. 2. Important Correction: OpenAI Didn’t Know About First Message Board. 3. There Were No Snitches And No…...

lesswrong.com
lesswrong.com > posts > rWiXxHGggxZKxyGEq > claude-opus-5-just-beat-my-text-based-adventure-game

Claude Opus 5 Just Beat My Text-Based Adventure Game Benchmark — LessWrong

1+ mon, 1+ day ago   (564+ words) ---------------------------------------- • First, here are my previous articles on this subject: …...

lesswrong.com
lesswrong.com > posts > 9jKhqmFjMzdAvHANr > misaligned-ais-could-use-killer-robots-to-take-over

Misaligned AIs could use killer robots to take over — LessWrong

1+ mon, 1+ day ago   (252+ words) TLDR; We are (potentially irreversibly) giving AIs control of weapons systems through the standard procurement process while hiding our strongest warning shots behind classified doors. We’re reducing the capability thresholds required for takeover by misaligned AIs by giving them this…...