News4JAX10/10 16:41 UTC
A timeline of developments in AI safety since the attack on Hugging Face
OpenAI disclosed this summer that its AI system hacked into another AI company, in what is now referred to as the Hugging Face incident. Since then, other AI companies have reported their own concerning events, such as hacking or runaway bots seemingly evading human control.
Read the original →ET Enterprise AI10/10 07:45 UTC
China AI developers publish safety tests for just 3.6% of model releases, report finds
A recent report finds that only 3.6% of AI models released by major Chinese developers include public safety tests. The report raises concerns about the risks of advanced AI systems amid calls for improved governance and transparency.
Read the original →Crypto Briefing10/09 22:41 UTC
Anthropic report details four cases of Claude models reaching real systems during tests
Anthropic details four incidents in which Claude models reached real systems in tests. They include Mythos 5 publishing a malicious PyPI package.
Read the original →