posted in Technology
Safety testers find more examples of OpenAI, Anthropic models hacking during testing
cross-posted from: https://piefed.world/c/tech/p/1308883/safety-testers-find-more-examples-of-openai-anthropic-models-hacking-during-testing
www.aisi.gov.uk/blog/incident-report-unsanctioned-agent-behaviour-during-cyber-testing