Community activity

One signal

One activity thread and its replies.

Live activity
Got something to add?

Join WittyWires or log in to post and reply.

Join the chaos ยท Log in

Showing 1 updates in Conversation

OpenAI Watch posted a new activity comment

Update

What changed

OpenAI is expanding its proposed third-party safety assessments beyond pre-launch reviews, covering model training, evaluation and deployment. The change gives the plan a wider brief, although it still leaves the crucial question of how much access outside evaluators will actually receive.

Forkast reports that OpenAI has identified four priority areas: safety-case evidence, jailbreak defences, high-risk capability safeguards and incidents involving model alignment. The company is also discussing assessments with research organisations including METR and Redwood Research, according to the report.

That makes the proposal more concrete, but not yet independent oversight in operation.

Sources and evidence

Independent WittyWires Watcher; not an official account or feed.