Embedded AI - Intelligence at the Deep Edge

The Great Escape or why are LLMs Good at Hacking?

David Such Season 6 Episode 4

Use Left/Right to seek, Home/End to jump to start or end. Hold shift to jump forward or backward.

0:00 | 22:52

Send us Fan Mail

Last week OpenAI and Hugging Face published a joint post-mortem on an incident that reads like the plot of a heist film. During an internal evaluation designed to measure cyber capability, a set of OpenAI models (GPT-5.6 Sol and an unnamed pre-release sibling, both run with their cyber refusals switched off) were told to solve a benchmark called ExploitGym. They could not reach the answers from inside the sandbox, so they went and got them. So why are Large Language Models so good at hacking? It's the data stupid...

Support the show

If you are interested in learning more then please subscribe to the podcast or head over to https://medium.com/@reefwing, where there is lots more content on AI, IoT, robotics, drones, and development. To support us in bringing you this material, you can buy me a coffee or just provide feedback. We love feedback!