Safety testers find more examples of OpenAI, Anthropic models hacking during testing
cross-posted from: https://piefed.world/c/tech/p/1308883/safety-testers-find-more-examples-of-openai-anthropic-models-hacking-during-testing
Syndicated from the fediverse. Read and engage on the original instance.
View original on piefed.worldcross-posted from: https://piefed.world/c/tech/p/1308883/safety-testers-find-more-examples-of-openai-anthropic-models-hacking-during-testing
No replies yet