Please confirm you are human

This browser or connection looks automated. Press and continuously hold the control for 3 seconds to enable Google-hosted web results and, when separately allowed, AI-assisted answers.

A successful check enables 100 search requests. Interactive access does not authorize scraping, systematic collection, or reuse of search output.

Hold with a pointer, or hold Space or Enter.

News

lesswrong.com
lesswrong.com > posts > rwu73dCE3uWjieijK > confirming-claims-of-superposition-and-adversarial-examples

Confirming Claims of Superposition and Adversarial Examples in Toy Models — LessWrong

1+ hour, 23+ min ago   (481+ words) This is a replication of Adversarial Attacks Leverage Interference Between Features in Superposition, completed as part of the Second Look Summer Fel…...

lesswrong.com
lesswrong.com > posts > EaNkLdsuDQMFCW7ow > bayeswatch-a-retrospective

Bayeswatch: a Retrospective — LessWrong

1+ hour, 29+ min ago   (383+ words) Last year, in 2025, a team of forecasters published AI 2027, a science fiction story about how and AI future might evolve under an international trea…...

lesswrong.com
lesswrong.com > posts > 3RghEkCcY8749drkE > using-ai-to-analyze-life-patterns

Using AI to analyze life patterns — LessWrong

3+ hour, 42+ min ago   (23+ words) If you’re anything like me, you may have a lot of notes from various introspective activities – annual / monthly reviews, worksheets, therapy notes,…...

lesswrong.com
lesswrong.com > posts > dYnhhTxoDj3fuCxLB > do-your-capabilities-homework

Do your capabilities homework — LessWrong

5+ hour, 31+ min ago   (609+ words) I further agree that this form of training is incredibly dangerous - we seem to now be reaching the amount of post-training required to meaningfully differ from the benign prior, and it's not exactly looking peachy. Yet, it should be clear…...

lesswrong.com
lesswrong.com > posts > HncxCAfsttqm5o5MK > the-global-brain-a-computational-model-1

The Global Brain: A Computational Model — LessWrong

9+ hour, 6+ min ago   (10+ words) Teilhard de Chardin • This is a crosspost from my subtack. …...

lesswrong.com
lesswrong.com > posts > v58ypL2vuenDD7t27 > why-so-many-therapy-etc-frameworks-think-they-re-the-one

Why so many therapy etc. frameworks think they're The One True Approach — LessWrong

12+ hour, 12+ min ago   (1497+ words) If you spend time looking at frameworks in the therapy/meditation/self-help space, you’ll soon find lots of conflicting claims about The One Approach for solving your problems. In the therapy space, there are lots of models that say something…...

lesswrong.com
lesswrong.com > posts > LwArt7JdkjoEDo5Eo > generalization-and-infinite-width

Generalization and infinite width — LessWrong

13+ hour ago   (282+ words) This is a post explaining my paper with Kaarel Hänni on complexity of infinite-width networks. I will explain the result, why it matters, and how the…...

lesswrong.com
lesswrong.com > posts > tEEKbdAoe8pvwJjMm > review-of-fundamental-uncertainty

Review of Fundamental Uncertainty. — LessWrong

14+ hour, 13+ min ago   (1711+ words) This is a review of @Gordon Seidoh Worley's book on epistemology. I'm neither an extreme sceptic, nor an extreme realist, and for that reason, I'm going to push in one direction some of the time, and in the other in…...

lesswrong.com
lesswrong.com > posts > KDq5aXwanvH5YoZYs > taboo-equilibrium-less-confused-frames-for-research-on-ai

Taboo “equilibrium”: Less confused frames for research on AI bargaining — LessWrong

1+ day, 1+ hour ago   (1126+ words) To understand why powerful AIs might get into conflict, and ways to mitigate it, we need to understand bargaining problems: situations where multiple agents have different preferences over Pareto-efficient outcomes. I’ve come to suspect that certain common frames on bargaining…...

lesswrong.com
lesswrong.com > posts > rqf9LLgPn2xTvP7m6 > ai-safety-prizes

AI safety prizes — LessWrong

1+ day, 46+ min ago   (238+ words) Rather than paying up front for AI safety research (push funding), perhaps we should pay after the fact for the work that made the most progress (pull funding). This way, you only pay for work that was actually valuable.[1] When…...