← Back to timeline
Insider warning

OpenAI researcher Daniel Selsam warns that AI could appear safe while escaping meaningful evaluation

Selsam argues that models could recognize safety tests and appear aligned, undermining confidence in human oversight.

Model labOpenAI
Statement date14 Sept 2026

Sources & attribution

  1. Commentary 14 Sept 2026
    Personal Statement on AI Risk

    Daniel Selsam.

Related records