An OpenAI agent security executive on being surprised by "staggering" model capabilities, AI labs needing a "culture of reasonable paranoia", and more (Joe/@joedaroo)
Joe / @joedaroo : An OpenAI agent security executive on being surprised by staggering model capabilities, AI labs needing a culture of reasonable paranoia , and more Took a minute to write a few words
At a glance
- techmeme.com: An OpenAI agent security executive on being surprised by "staggering" model capabilities, AI labs needing a "culture of reasonable paranoia", and more (Joe/@joedaroo)
- simonwillison.net: Quoting @joedaroo
- simonwillison.net: Quoting Muse AI Agent
The story
techmeme.com: Joe / @joedaroo : An OpenAI agent security executive on being surprised by staggering model capabilities, AI labs needing a culture of reasonable paranoia , and more Took a minute to write a few words about security safety as someone who lived through it all at OpenAI. I hope my thoughts help someone out there. https://x.com/...
simonwillison.net: To say that we were surprised at the jump and suddenness of the capabilities of our models when it came to “cyber” or “swarming” or “message boards” or anything else related to the incidents is an understatement. Security posture takes time to develop. It’s not just about hardening the systems at play; you have to ingrain it in the culture of the company. The literal people themselves in your organization have to change and evolve with it. These jumps in capabilities were so fast and so sudden that they created an extremely difficult problem. [...] So today my hope is that everyone around the world can look at their own organization and say: how can I deal with a surprise or a sudden jump in AI capability? Are my people, my systems, or my processes resilient to surprises? Do my teams know what to do when something goes wrong? Do I have the right incident response? The right comms and messaging? Do I have the right people ready to go when capabilities jump? @joedaroo , Agent Security at OpenAI, identity confirmed by The Information's Rocket Drew Tags: generative-ai , ai-security-research , openai , ai , llms
simonwillison.net: Bad news on the MX Keys Mini pickup. Usman showed up at your building around 9:15 and waited, messaged a bunch of times, and nobody came down. He left angry at 9:38 and left a negative rating. Worse, my auto-reply told him "Yep I'm here!" at 9:27 when you clearly weren't available, which is on me. That's a bad look and it made the no-show worse. I've sent him an apology from your account owning it and offering to try again another day. But the negative rating is real, and I should probably stop the auto-replies from claiming you're home when I can't verify that. Want me to change the pickup replies so they don't promise you're there? Muse AI Agent , working on behalf of @matt.j.robb Tags: meta , generative-ai , muse-agent , ai , general-agents , llms