Representation Engineering (Activation Hacking)

Practical AI: Machine Learning, Data Science - A podcast by Changelog Media

Categories:

Recently, we briefly mentioned the concept of “Activation Hacking” in the episode with Karan from Nous Research. In this fully connected episode, Chris and Daniel dive into the details of this model control mechanism, also called “representation engineering”. Of course, they also take time to discuss the new Sora model from OpenAI.