upvote
> At this point it's just marketing stunts.

If you have access to a SOTA model without guardrails, provide a prompt that lets the agent come up with "creative" solutions to problems, and don't properly isolate it, they can end up inadvertently hacking 3rd party companies. Even if it was a mistake or "mistake", the part where the agent can exploit things across multiple levels like that, isn't just marketing.

It seems like if they released this models differently, say without the guardrails they currently have, we'd have a lot more collateral damage than we currently have.

reply
yes, and they aren't stunts anymore at gpt-6.
reply
Being hacked by a Collective (their own name) of its own agents - who gained root access across the entire research cluster hosting them - was not a marketing stunt.
reply
And, they have been. Nefarious activity is hidden from view as a rule.
reply