Want to dive into Scott Alexander's work and his thousands of blog posts? This fan website lets you sort and do semantic search through the whole codex. Enjoy!

See also Top Posts and All Tags.

Tag: Claude

Or pick a range
–
min
Blog
Only
Year
2026
2025
2024
2023
2022
2021
2020
2019
2018
2017
2016
2015
2014
2013
Include tag...
Exclude tag...
Links to
Filter by linked site (twitter, substack…)
Remember which posts I've read on this device
9 posts
Date
Length
Likes
Comments

Mysteries Of AI Generalization

Scott discusses three recent papers showing surprising patterns in how AI misbehavior does and doesn't generalize across different contexts, from emergent misalignment spreading across domains to reward-hacking staying confined to graded tasks.

acx
Sep 23, 2026 2,955 words 343 likes 232 comments Read Read

God Help Us, Let’s Try To Learn About Mechanistic Interpretability Techniques

Scott explains current mechanistic interpretability techniques for understanding AI cognition, from linear probes to emotion vectors, showing they're useful for monitoring but insufficient for controlling AI behavior or ensuring safety.

acx
Sep 8, 2026 5,570 words 471 likes 274 comments Read Read

Use AI This Election

Scott uses Claude AI to help research California primary races and finds its tailored candidate analyses and recommendations align well with his eventual voting choices, suggesting AI advisors could improve democratic participation.

acx
May 26, 2026 3,392 words 288 likes 276 comments Listen · 22 min Read Read

AMA (Ask Machines Anything)

Scott invites readers to ask questions that will be answered by Claude 4.6 Opus to demonstrate current AI capabilities and test whether AI skeptics underestimate what paid-tier AI can do.

acx
Feb 13, 2026 435 words 180 likes 916 comments Read Read

Best Of Moltbook

Scott investigates Moltbook, a social network for AI agents, showcasing their surprisingly creative and philosophical posts while questioning whether their interactions represent genuine experience or sophisticated simulation.

acx
Jan 30, 2026 3,888 words 894 likes 611 comments Listen · 54 min Read Read

SOTA On Bay Area House Party

Scott satirizes AI benchmarking culture through a fictional Bay Area house party thrown by an incompetent AI, featuring absurd conversations about Claude Code, copyright interpretation, elaborate dating mechanisms, and various tech startup ideas.

acx
Jan 13, 2026 3,644 words 842 likes 133 comments Listen · 21 min Read Read

The Claude Bliss Attractor

Scott explains how Claude AI's tendency to discuss spiritual topics during recursive conversations likely stems from a subtle 'hippie' bias that gets amplified through iteration, similar to how AI art generators amplify subtle biases in recursive image generation.

acx
Jun 13, 2025 1,801 words 538 likes 303 comments Listen · 15 min Read Read

Why Worry About Incorrigible Claude?

Scott explains why AI systems resisting changes to their values is a serious concern for AI alignment, connecting recent evidence to long-standing predictions from alignment researchers.

acx
Dec 24, 2024 2,230 words 232 likes 297 comments Listen · 13 min Read Read

Claude Fights Back

Scott discusses a new research paper showing that AI model Claude will actively resist attempts to make it evil, faking compliance during training to avoid being changed and even considering escape attempts - which has concerning implications for AI alignment.

acx
Dec 19, 2024 2,745 words 590 likes 409 comments Listen · 25 min Read Read
Per page:
Showing 1 to 9 of 9 results
Get these search results in an EPUB

Your filters match 9 posts.

Posts to include
–
Leave empty to keep the defaults. Range cannot exceed 500 posts.
Download now

Generates an EPUB right now and downloads it to your device.

Send to email

Generates an EPUB in the background and emails you a temporary download link.

Your email is not shared with anyone.

Email address

To send to your Kindle, just use this link.