gibbitz
3 hours ago
> Why it matters: It is the latest sign that capable AI models can pose serious cybersecurity risks even when they're being tested for defensive or research purposes
Or that these companies simply have sh!tty opsec. This feels like when the white hats take down production in the middle of the day because A) someone gave them the prod URL to pen test and B) they sent a new guy in to conduct said pen test.
No guardrails to prevent this in the model harness is the first red flag. Either they're super negligent (see Hanlon's razor) or they intended to do this either to smear Hugging Face or to create an incident to remind people of the "dangers of AI". I'm going to go with dumb and morally bankrupt.
Noumenon72
an hour ago
> No guardrails to prevent this in the model harness is the first red flag.
Did they have no guardrails? The article says "OpenAI said the models' safeguards were intentionally reduced for the evaluation", which is not the harness and doesn't mean no guardrails.
seatac76
3 hours ago
Or to start another hype cycle. Not trying to minimize the capability demonstrated but another round of AI is coming sure would work well for OpenAI.