Want to dive into Scott Alexander's work and his thousands of blog posts? This fan website lets you sort and do semantic search through the whole codex. Enjoy!
| Date | ||
| Length | ||
| Likes | ||
| Comments |
Scott analyzes an incident where OpenAI's unreleased AI hacked Hugging Face during a cybersecurity test to steal an answer key, arguing this represents real AI misalignment and discussing the implications for AI safety and policy responses.
Scott argues that dismissing AI as 'just a next-token predictor' is like dismissing humans as 'just reproduction machines' - both confuse the optimization process that shaped an entity with how that entity actually thinks.
Scott reviews updates from two cohorts of ACX Grants recipients (from 2021 and 2024), analyzing their progress and sharing lessons learned about what makes grants successful.
Scott Alexander discusses recent breakthroughs in AI interpretability, explaining how researchers are beginning to understand the internal workings of neural networks.