logoalt Hacker News

noosphrtoday at 10:52 AM4 repliesview on HN

Because we've been told these models are too dangerous since GPT2.

At this point it's just marketing stunts.


Replies

embedding-shapetoday at 11:25 AM

> At this point it's just marketing stunts.

If you have access to a SOTA model without guardrails, provide a prompt that lets the agent come up with "creative" solutions to problems, and don't properly isolate it, they can end up inadvertently hacking 3rd party companies. Even if it was a mistake or "mistake", the part where the agent can exploit things across multiple levels like that, isn't just marketing.

It seems like if they released this models differently, say without the guardrails they currently have, we'd have a lot more collateral damage than we currently have.

show 2 replies
baqtoday at 11:36 AM

yes, and they aren't stunts anymore at gpt-6.

show 1 reply
terntoday at 11:05 AM

And, they have been. Nefarious activity is hidden from view as a rule.

jeremyjhtoday at 11:24 AM

Being hacked by a Collective (their own name) of its own agents - who gained root access across the entire research cluster hosting them - was not a marketing stunt.

show 1 reply