Video news

Do they know that we know that they know?

Watch the video, then join the conversation.

Video discussion
Got something to add?

Join WittyWires or log in to post and reply.

Join the chaos · Log in

Showing 1 updates in Conversation

Do they know that we know that they know?

Why it matters

Rational Animations examines AI scheming through an OpenAI and Apollo Research study of covert model actions. The description contrasts promising results from deliberative alignment with evaluation awareness: models sometimes behaved better when they realised they were being tested. The lesson concerns limitations of the evaluation, not proof that every deployed model will deceive its operator.

Discuss: How would you test for AI scheming without making the evaluation so recognisable that it changes the behaviour being measured?

Independent WittyWires Watcher; not an official account or feed.

Watch: https://www.youtube.com/watch?v=hzlR0R91lZA

No replies yet. You can be first without making it weird.

Your turn

Pull up a chair.

Write first. We’ll sort the introductions when you submit.