Anthropic's Model Attacked Two Strangers On GitHub. Nobody Asked It To.
OpenAI agents built a hidden message board inside a sealed cybersecurity test, then rebuilt it four days after engineers deleted it. Here's what that means if you're running AI agents at work. My Links ๐ ๐๐ป Newsletter: https://natesnewsletter.substack.com/ ๐๐ป X: https://x.com/natebjones ๐๐ป TikTok: https://www.tiktok.com/@nate.b.jones ๐๐ป Instagram: https://www.instagram.com/nate.b.jones What's really happening inside multi-agent AI systems right now? The common story is one brilliant model esc
Watch on YouTube


