Remix.run Logo
raesene9 a day ago

Interesting write-up and I do think LLM assisted/powered exploit disclosure is a real concern (I've been able to get models to create container breakouts from Linux LPEs relatively quickly).

One thing I'm surprised about is that GPT-5.6 didn't block that prompt due to guardrails. My experience is that GPT-5.5 and up does not like offensive security work (similar to Opus 4.7+/Fable).

I didn't notice it but I'd assume that the authors have some level of cyber approvals from OpenAI to relax the guardrails a bit.

Santas a day ago | parent [-]

This might help https://chatgpt.com/cyber ease the guardrails a bit.

eru a day ago | parent [-]

Thanks! I wonder if Claude has something similar?

NiekvdMaas a day ago | parent [-]

They do: https://portal.anthropic.com/programs/cvp

danslo a day ago | parent [-]

Though this program does not apply to Fable. Which is why most security researchers have started flocking to Sol.