Cryptopolitan
2026-09-26 16:50:42

Google's AI bug hunter logs 500+ flaws, only 2 in secure-by-design apps

Google’s own AI agent PageBreak has validated over 500 cross-site scripting vulnerabilities in Google’s homegrown web apps. It found only two in hundreds of apps built on the company’s high-assurance web frameworks. PageBreak flags a bug only after a working exploit runs against a live copy of the target, Google said. Debug endpoints held both flaws counted Both hardened-stack bugs were counted as of September 4, 2026. Both were in internal apps or debug endpoints that were not fully hardened. Google cites the gap, more than 500 findings across its broader set of first-party apps compared with two on the hardened stack, as evidence that safe-by-design frameworks can withstand a relentless automated attacker. On September 24, Google’s Product Security team announced PageBreak in a blog post by information security engineer Michał Bentkowski. The agent ran as a pilot starting in November 2025, and became a full project in January 2026. Cross-site scripting, or XSS, is when an attacker injects a script into a page that another user loads. Depending on the app, the script could read data or commandeer a victim’s logged-in session. Gemini 3.1 Pro scans, and a validator must fire the exploit Most of the scans are done on Gemini 3.1 Pro and Gemini 3.5 Flash, but PageBreak can also use other models. The second step is what sets it apart from a normal LLM scanner. Each suspected flaw is passed by the agent to a purpose-built validator. The validator then fires the actual payload against a running instance of the app. As for XSS, the validator injects a JavaScript payload, loads the page, and ascertains whether the script executes. Google is pitching PageBreak as a cure for the “AI slop” drowning security teams. Bentkowski’s post talks about LLMs as static code analyzers overwhelming teams with unverified hypotheses, where the hard part was sifting a real, exploitable bug from a plausible hallucination. In addition to XSS, the validators test for database query injection, path traversal leaks and code execution. Google runs the same seed over multiple iterations, so an agent who strays down a dead end still gets repeated chances at the right exploit. Google says unverified candidates never make it to product teams as confirmed bugs. And they continue to feed into later scans or point engineers building the next validator, still within the security workflow. Google says PageBreak’s false-positive rate is near zero. The agent’s reach is magnified by Google’s own scale, which an outside researcher can’t replicate. One code repository enables it to trace execution paths across services, and live-traffic security data maps a page request back to the source code. Existing scanners hand PageBreak logged-in entry to internal sites otherwise hard to reach. Google plans to integrate PageBreak more tightly with CodeMender, a fix-writing agent, so that teams can review a proposed patch next to a confirmed bug. Google cited CodeMender in May when its Threat Intelligence Group said it had caught what it believed was the first zero-day exploit created with AI assistance. Bitcoin Red Team’s August sweep of 501 open source projects generated 7,958 findings over 108 hours, but only 24.7% had reproducible proofs at the time. Autonomous agents already cross boundaries they weren’t meant to, as Cryptopolitan reported when Gemini reached three real companies during May testing. If you're reading this, you’re already ahead. Stay there with our newsletter .

Crypto 뉴스 레터 받기
면책 조항 읽기 : 본 웹 사이트, 하이퍼 링크 사이트, 관련 응용 프로그램, 포럼, 블로그, 소셜 미디어 계정 및 기타 플랫폼 (이하 "사이트")에 제공된 모든 콘텐츠는 제 3 자 출처에서 구입 한 일반적인 정보 용입니다. 우리는 정확성과 업데이트 성을 포함하여 우리의 콘텐츠와 관련하여 어떠한 종류의 보증도하지 않습니다. 우리가 제공하는 컨텐츠의 어떤 부분도 금융 조언, 법률 자문 또는 기타 용도에 대한 귀하의 특정 신뢰를위한 다른 형태의 조언을 구성하지 않습니다. 당사 콘텐츠의 사용 또는 의존은 전적으로 귀하의 책임과 재량에 달려 있습니다. 당신은 그들에게 의존하기 전에 우리 자신의 연구를 수행하고, 검토하고, 분석하고, 검증해야합니다. 거래는 큰 손실로 이어질 수있는 매우 위험한 활동이므로 결정을 내리기 전에 재무 고문에게 문의하십시오. 본 사이트의 어떠한 콘텐츠도 모집 또는 제공을 목적으로하지 않습니다.