Discussion about this post

User's avatar
Name Required's avatar

Unfortunately, "eval awareness" is often conflated with "verbalized eval awareness". The latter refers to the subset of eval awareness *that we can detect*.

Nathan Metzger's avatar

Unfortunately, I frequently see RSI used to refer to the kind of closely bounded, iterated self-improvement that AI can do today. It's been used to refer to models updating their scaffolds, or helping with the process of training themselves or their successors. So I find myself reaching for "closed-loop RSI" or "unbounded RSI" to refer to the original meaning: the process whereby AI makes smarter AI without any human help (or at least without human-led R&D as input), until the laws of physics or available resources provide hard upper bounds.

No posts

Ready for more?