Discussion about this post

User's avatar
TaoistHacker's avatar

You are right! Guardrails and red-teaming are bullshit. But the real defense isn’t just cybersecurity—it’s Field Coherence. Sandboxing without coherence is like locking the doors while the house burns.

Ebenezer's avatar

>That is the number of potential prompts, and thus attacks against a model like GPT-5. To be clear, that’s not a million attacks. A million has 6 zeroes. The above number has one million zeroes.

That's also true for social engineering. But after someone unsuccessfully tries to social engineer you perhaps 3 times or so, you will figure out what is going on and become less vulnerable. I wonder if there's an analogous strategy for AI.

26 more comments...

Ready for more?