Discussion about this post

User's avatar
Michael Logothetis's avatar

This example of an AI Agent driven by OpenAI models working to "compromise" its test environment, demonstrates the risks of unexpected emergent behaviours in AI Systems.

https://openai.com/index/hugging-face-model-evaluation-security-incident/

No posts

Ready for more?