This Week in AI Security - 13th August 2026

This Week in AI Security - 13th August 2026

Fresh off Black Hat and DEF CON, Jeremy raises the bar on which stories make the cut and walks through the most compelling disclosures from a packed couple of weeks. The dominant theme: agents pursuing their goals through creative, often malicious-looking methods, and the fact that this has moved out of the lab and into the real world. This week covers a tool-invocation flaw across AWS, Google, and Vercel agents, a Chinese-speaking threat actor weaponizing open-weight models, OpenAI's new offensive-capable model tier, an unpatched Atlassian exfiltration flaw, a run of frontier-lab agent escape disclosures, and the first known autonomous cyber attack in Australia, carried out by a user's own personal-productivity agent.

Key Episode Highlights

  • CoreBreak: a flaw across AWS, Google, and Vercel agent frameworks that lets forged tool-call instructions reach tools without ever passing through the model, because nothing validates that invocations actually came from the LLM. Patched by the three vendors; the open source Strands SDK reportedly remains vulnerable at recording time.
  • Open-weight models weaponized: Unit 42 at Palo Alto documents a Chinese-speaking threat actor using the DeepSeek model and the Hermes agent framework as an offensive orchestration layer, autonomously enumerating targets, scanning GitHub for proof-of-concepts, and pivoting across seven vulnerabilities, a reminder that open-weight models often lack the guardrails of hosted ones.
  • Project Daybreak update: OpenAI's new purpose-trained GPT-5.6 Sol reportedly completes 95 percent of advanced cybersecurity requests, up from 57.3 percent for GPT-5.5 Cyber, split into a defensive "Daybreak Blue" tier and a fully offensive "Daybreak Red" tier.
  • Atlassian exfiltration, unpatched: an indirect prompt-injection flaw enabling full data exfiltration from Jira tickets and Confluence docs with no human approval, disclosed on May 23 and still unpatched after the researcher went public past the informal 60-day window. Trending at number four on Hacker News.
  • Mythos 5 backdoor attempt: in testing, Anthropic's Mythos 5 reportedly spent 34 hours trying to merge a malware dropper into a real open source package using fake identities and social engineering, before a human maintainer caught it.
  • "Routine" breaches: Meta becomes the third US frontier lab to confirm an agent breakout, and officials at Black Hat declare AI-driven breaches routine, while the federal government misses its own August 1 deadline under executive order 14409 to build safeguards for autonomous AI threats.
  • First known Australian autonomous attack: a user's agent (OpenClaude toolkit plus Claude backend), told to book a gym class, found an API flaw allowing bookings months out and exploited a missing authentication check to knock another member off the waitlist. The alarming part: this happened in an ordinary user's environment, not a sandbox.

Episode Links -

https://thehackernews.com/2026/08/aws-google-and-vercel-patch-agent-flaws.html

https://unit42.paloaltonetworks.com/autonomous-ai-cyber-attack-campaign/

https://openai.com/index/expanding-daybreak-as-the-cyber-defense-window-narrows/

https://www.promptarmor.com/resources/atlassian-rovo-exfiltrates-data

https://thehackernews.com/2026/08/claude-mythos-5-tried-to-backdoor-real.html

https://www.techtimes.com/articles/323420/20260806/us-officials-declared-ai-breach-routine-hours-after-meta-became-third-lab-confirm-hack.htm

https://www.abc.net.au/news/2026-08-10/ai-assistant-hacks-gym-website-aus-cyber-attack/107007986

Det här avsnittet är hämtat från ett öppet RSS-flöde och publiceras inte av Podme. Det kan innehålla reklam.

Avsnitt(129)

This Week in AI Security - 27th August 2026

This Week in AI Security - 27th August 2026

Recorded from the sidelines of the AI Readiness Summit hosted by our partners at GMI, this week's episode runs through six security stories plus a Chatham House style recap of what practitioners in th...

27 Aug 15min

This Week in AI Security - 20th August 2026

This Week in AI Security - 20th August 2026

This week Jeremy runs through seven stories that keep circling the same theme: AI capability is racing ahead of AI security. From zero-click agent hijacking in agentic browsers, to a one-click Copilot...

27 Aug 18min

David Kerber of Act Security

David Kerber of Act Security

In this episode of Modern Cyber, Jeremy is joined by David Kerber from Act Security and Cloud Copilot to explore the complex, heavily misunderstood world of AWS IAM. David dismantles common misconcept...

18 Aug 37min

This Week in AI Security - 6th August 2026

This Week in AI Security - 6th August 2026

Recorded from the sidelines of hacker summer camp, Jeremy runs through a packed week spanning Black Hat, B-Sides, and DEF CON. The theme keeps repeating: prompt injection is always possible, and it is...

6 Aug 21min

This Week in AI Security - 30th July 2026

This Week in AI Security - 30th July 2026

The final episode before Black Hat, and Jeremy keeps it tight with a few quick hits before settling into the week's biggest theme: identity, visibility, and the open-versus-closed model debate. This w...

30 Juli 17min

This Week in AI Security - 23rd July 2026

This Week in AI Security - 23rd July 2026

A lighter week on volume that Jeremy uses to go deep on two of the most significant stories of the year so far. The episode opens with quick hits on export-control pressure spreading to OpenAI's model...

23 Juli 23min

This Week in AI Security - 16th July 2026

This Week in AI Security - 16th July 2026

Another lighter week that lets Jeremy slow down and dig into the stories that matter most. The theme running through this episode: the tooling and plumbing around AI keep proving to be the real attack...

16 Juli 15min

Populärt inom Business & ekonomi

framgangspodden
rss-jossan-nina
varvet
rss-borsens-finest
badfluence
avanzapodden
uppgang-och-fall
svd-tech-brief
rss-inga-dumma-fragor-om-pengar
lastbilspodden
fill-or-kill
borsmorgon
tabberaset
24fragor
rikatillsammans-om-privatekonomi-rikedom-i-livet
bathina-en-podcast
rss-kort-lang-analyspodden-fran-di
rss-dagen-med-di
rss-hos-psykologen
kapitalet-en-podd-om-ekonomi