How interest changes
History starts here
The chart will appear after repeat observations. The current metric comes from the source.
17 source pointsReal observations only. History before source connection is not reconstructed.
The article discusses the distinction between AGI-pilled and ASI-pilled individuals, as well as challenges related to the potential superiority of artificial intelligence over humans. The author expresses concern about the future of freedom and democracy in the context of rapid technological development.
The chart will appear after repeat observations. The current metric comes from the source.
17 source pointsReal observations only. History before source connection is not reconstructed.
The text discusses cultural differences between China and the West, focusing on social media and communication styles. It highlights how Chinese social media is perceived as overwhelming and inauthentic by Western standards, and how these differences may pose challenges for AI safety awareness efforts in China.
6 days agoLessWrongThe text discusses the debate between 'Alignment Engineering' and 'Misalignment Science' in AI alignment research, highlighting concerns that current alignment methods might inadvertently accelerate capabilities development, increasing risks of rapid technological growth. It questions the opportunity costs of focusing on alignment engineering and proposes a shift towards 'Misalignment Science' as a more robust approach.
3 days agoLessWrongThe text discusses the process of applying to AI safety fellowships and jobs, emphasizing the importance of clearly communicating skills. It highlights the challenges of application processes being noisy and the need to avoid overfitting to the process while showcasing genuine abilities.
5 days agoLessWrongThe article discusses whether OpenAI's recent mathematical results demonstrate a form of creativity that goes beyond brute-force methods, suggesting that while AIs are improving at solving complex problems, they may still lack the ability to form new concepts and apply them efficiently. The text raises questions about the nature of the creativity involved in these proofs and whether they represent a significant breakthrough akin to historical examples like AlphaGo's Move 37.
21 hours agoLessWrongThis document provides an overview of monitoring practices for internally deployed agents at frontier AI companies, focusing on OpenAI (OAI), Anthropic, and Google DeepMind (GDM). It highlights that most monitoring is done asynchronously, with some real-time automated classifiers used for agent actions. However, there is limited public information about GDM's specific practices, as much of the data comes from their control roadmap, which outlines suggestions rather than confirmed implementations.
yesterdayLessWrongThe article discusses the challenge of defining safety research in AI, noting that capabilities research can also contribute to safety by expanding the Pareto frontier of safety and usefulness. It highlights that while improving safety without reducing usefulness is a key goal, some research may inadvertently encourage developers to prioritize capabilities over safety. The text explores scenarios where capabilities research could enhance safety under specific political conditions, but emphasizes the complexity of distinguishing between safety and capabilities research.
6 days ago