'We can't trust them completely': AI research fellows warn that labs are running models with the safeguards off behind closed doors | Fortune
Alan Chan and Sam Manning, co-authors with OpenAI's and Anthropic's top researchers, say that many times, "internal safeguards have not been deployed."