Understanding how LLMs learn (e.g. SLT) is underrated relative to work on instilling specific behaviors
I don’t have a very good idea about how these different ideas are “rated” in the AI safety community.
I don’t have a very good idea about how these different ideas are “rated” in the AI safety community.