Discussion about this post

User's avatar
Ariella Shulman's avatar

Yes to alignment being a process. Also: I'm curious where you see the role of human-agent collaboration, in addition to agent-agent collaboration. Presumably, institutional governance would need to account for model steering, jailbreaking, and other such human-led interference (intentional or not).

No posts

Ready for more?