233 episodios
OpenAI's models breached Hugging Face, reward hacking ethics, benchmarking fast16
23/07/2026 | 2 h 16 min(Presented by Thinkst Canary: Most Companies find out way too late that they’ve been breached. Thinkst Canary changes this. Deploy Canaries and Canarytokens in minutes and then forget about them. Attackers tip their hand by touching ’em giving you the one alert, when it matters. With zero admin overhead and almost no false-positives, Canaries are deployed (and loved) on all 7 continents.)
Three Buddy Problem - Episode 106: We dig into the news that OpenAI's models were the "autonomous agent" that breached Hugging Face, escaping a sandbox through a zero-day to cheat on a cyber benchmark, then getting spun into a partnership announcement. We argue about the implications of the incident, the PR masterclass, the absence of ethics and human oversight, and calls for "kill switches" to mitigate "AI lab leaks."
Plus, SentinelLabs' new fast16 reverse-engineering benchmark, where GPT-5.6 Sol was the only public model to go the distance.
Cast: Juan Andres Guerrero-Saade, Ryan Naraine and Costin Raiu.
Timestamps:
0:00 Introductory banter
5:24 OpenAI admits it was the Hugging Face "hacker"
10:06 What’s ExploitGym and who's on top of the leaderboard
12:59 Reward hacking: Did anyone train this thing not to cheat?
19:35 Marketing stunt or real incident? The zero-day in the package proxy
26:43 Was OpenAI already plugged into Hugging Face?
29:17 Paperclips, kill switches, and "going rogue"
34:49 Crisis comms, regulatory capture, and the second Cold War
43:02 Approve every action? Auto mode and swarms
50:10 "Lab leak" and calls for biosafety levels
1:00:31 The missing models: no Mythos, no Kimi, no independent referee
1:07:04 Costin's prediction: owning frontier-class hardware will require a license
1:13:41 fast16 as a benchmark: Inside the Sol Searching research
1:26:51 Compression and altitude: are reverse engineers being replaced?
1:41:24 Finding the gem in 100 samples, and the swarm frontier
2:00:41 Claude Opus 5 drops, Gemini 3.5 Flash Cyber- (Presented by Thinkst Canary: Most Companies find out way too late that they’ve been breached. Thinkst Canary changes this. Deploy Canaries and Canarytokens in minutes and then forget about them. Attackers tip their hand by touching ’em giving you the one alert, when it matters. With zero admin overhead and almost no false-positives, Canaries are deployed (and loved) on all 7 continents.)
Three Buddy Problem - Episode 105: We discuss a fascinating Hugging Face breach, where an autonomous AI agent broke out of the sandboxes, moved laterally through production, and generated 17,000 alerts before anyone caught it, and how frontier model guardrails locked the defenders out of their own investigation.
Plus, China's big AI showcase, Xi's pitch for open models and global distribution, a record 622-CVE Microsoft Patch Tuesday, and 13 years of dwell time in the Daxin backdoor.
Cast: Juan Andres Guerrero-Saade, Ryan Naraine and Costin Raiu.
Timestamps:
0:00 Introductory banter
3:51 Hugging Face discloses end-to-end agentic hack
9:42 Why Hugging Face couldn't use frontier models
13:28 AI guardrails hampering defenders
16:22 Codex vs Claude for real malware work
23:43 Flash attacks vs. going low and slow
30:27 Was it targeted, or did Hugging Face pwn itself?
38:11 Long-horizon coherence: what GLM 5.2 still can't do
41:27 Kimi K3 leapfrogs, and Xi's AI speech
52:15 Exceptionalism vs. distribution
1:11:05 Gold Eagle: the White House vulnerability clearinghouse
1:15:05 Microsoft patches 622 CVEs — a record
1:20:29 APT corner: Daxin resurfaces after 13 years of dwell time
1:29:45 Balochistan police, and Microsoft's attribution-free wiper
1:34:26 Denis Obrezkov, leaked Kaspersky records, and the wrong questions
1:46:01 Magnet Forensics sues over a burned iPhone bug
1:57:57 Shout-outs - (Presented by Thinkst Canary: Most Companies find out way too late that they’ve been breached. Thinkst Canary changes this. Deploy Canaries and Canarytokens in minutes and then forget about them. Attackers tip their hand by touching ’em giving you the one alert, when it matters. With zero admin overhead and almost no false-positives, Canaries are deployed (and loved) on all 7 continents.)
Three Buddy Problem - Episode 104: We discuss the return of Anthropic's Fable 5 from export-control suspension with guardrails so aggressive that spelling "exploit" gets you downgraded. Plus, a debate on AI frontier labs killing businesses at scale, and OpenAI offering equity to the US government.
Also, buried on page nine of a 'Scattered Spider' arrest indictment: Microsoft's never-before-detailed GDID device identifier, a persistent Windows fingerprint with massive implications for OPSEC, privacy, and APT tracking.
Cast: Juan Andres Guerrero-Saade, Ryan Naraine and Costin Raiu.
Timestamps:
0:00 Cold open: Heat wave in Washington DC
3:45 Fable 5 returns after the 15-day timeout
5:21 "Refined classifiers" and the downgrade-to-Opus mess
8:23 Codex vs. Claude: real-world malware analysis test
12:41 Who are the guardrails for? Defenders locked out
19:13 What even is a "jailbreak assessment framework"?
21:37 Two theories: failed PR vs. killing a thousand startups
24:59 Could the labs build kernels or a whole OS?
31:38 Bureaucracy is the moat
36:09 Can AI actually run an attack? (Spoiler: 14 detections)
47:01 OpenAI offers the US government a 5% stake
58:16 Scattered Spider arrest and Microsoft's GDID revelation
1:12:02 OPSEC fallout: how APT groups adapt to device telemetry
1:27:18 UFO update, shout-outs from Seoul - (Presented by Thinkst Canary: Most Companies find out way too late that they’ve been breached. Thinkst Canary changes this. Deploy Canaries and Canarytokens in minutes and then forget about them. Attackers tip their hand by touching ’em giving you the one alert, when it matters. With zero admin overhead and almost no false-positives, Canaries are deployed (and loved) on all 7 continents.)
Three Buddy Problem - Episode 103: We dive into the U.S. government's takeover of frontier-model rollouts (Mythos, Fable, and OpenAI's Sol/Terra/Luna) and what it means when intelligence gets commoditized but access gets rationed.
Plus, Costin's all-Chinese open-weight stack, the economics of burning tokens, a fresh Salesforce OAuth breach, and jellyfish UFOs over Iran.
Cast: Juan Andres Guerrero-Saade, Ryan Naraine and Costin Raiu.
Timestamps:
0:00 — Introductory banter, Thinkst Canary sponsorship
2:55 — Why threat intel analysts are built for the AI moment
11:09 — Government takes the wheel: Mythos, Fable & the frontier labs
16:15 — Did the government go too far/not far enough?
25:42 — Anthropic's "best PR campaign in history"
31:52 — Alibaba, distillation & the model-router cartel
40:58 — Costin's stack: Chinese open-weight models & token economics
46:12 — Dumping, evals & the real work of AI engineering
1:04:32 — Soft power: how the world gets pushed toward China
1:14:43 — "The bullshit": over-refusal & the Opus 4.8 regression
1:32:03 — The trillion-dollar IPO endgame
1:35:49 — The Klue OAuth breach and secure-by-default
1:45:32 — Shout-outs: UAP jellyfish, LABScon 2026 - (Presented by TLPBLACK: A cybersecurity intelligence platform focused on sharing curated, high-sensitivity threat insights and research with trusted security professionals.)
Three Buddy Problem - Episode 102: Software export controls expert Katie Moussouris joins the show to unpack the US government's abrupt move to suspend access to Anthropic's most powerful models over a so-called "jailbreak" that, on reading the paper, turned out to be a model doing exactly what defenders are supposed to do.
We dig into the export-control chaos, the chemical-weapons framing of cybersecurity, the China question, and why Microsoft just resurrected a disclosure term the industry buried fifteen years ago.
Cast: Katie Moussouris, Juan Andres Guerrero-Saade and Ryan Naraine. Costin is traveling.
Timestamps:
0:00 - Introductory banter
1:00 - Export Controls: Fable 5 and Mythos 5 suspended
3:40 - The Anthropic–USG relationship and USG’s surveillance claim
9:40 - Self-owns, doomsday cults, and why the guardrails are "so broad"
12:42 - What the Amazon paper actually says ("fix this code")
20:33 - The chemical-weapons framing problem
23:39 - The China question and the SK Telecom angle
41:17 - Why hasn't the paper been published?
57:01 - "Free Fable": are Chinese models only months behind?
1:00:13 - The unforgiving internet and the security poverty line
1:11:18 - Microsoft brings back "responsible disclosure" (and threatens researchers)
1:29:04 - Luta Security, the AI bug flood, and shout-outs
Más podcasts de Noticias
Podcasts a la moda de Noticias
Acerca de Three Buddy Problem
The Three Buddy Problem is a popular Security Conversations podcast that goes beyond industry talking points to discuss what others won’t -- nation-state malware, attribution, cyberwar, ethics, privacy, and the messy realities of securing computers and corporate networks. Hosted by three veteran security pros -- journalist Ryan Naraine and malware paleontologists Costin Raiu and Juan Andres Guerrero-Saade -- the weekly show attracts a highly engaged audience of security researchers, corporate defenders, CISOs, and policymakers. Connect with Ryan on Twitter (Open DMs).
Sitio web del podcastEscucha Three Buddy Problem, Hora 25 y muchos más podcasts de todo el mundo con la aplicación de radio.es

Descarga la app gratuita: radio.es
- Añadir radios y podcasts a favoritos
- Transmisión por Wi-Fi y Bluetooth
- Carplay & Android Auto compatible
- Muchas otras funciones de la app
Descarga la app gratuita: radio.es
- Añadir radios y podcasts a favoritos
- Transmisión por Wi-Fi y Bluetooth
- Carplay & Android Auto compatible
- Muchas otras funciones de la app


Three Buddy Problem
Escanea el código,
Descarga la app,
Escucha.
Descarga la app,
Escucha.





























