Business Wire

LLMs Are Getting Smarter, But Not Safer: Veracode 2026 GenAI Code Security ReportFinds AI-Generated Code Security Has Stalled at 56% Pass Rate

28.7.2026 13:50:00 CEST | Business Wire | Press release

Share

AI Models Built for Software Coding Are No More Secure Than General-Purpose ModelsDespite Major Gains in AI Speed, Capability, and Reasoning, Veracode Report Reveals Nearly Half of AI-Generated Code Still Fails Security Tests

Veracode, the global leader in application risk management, today released its 2026 GenAI Code Security Report which reveals that, despite rapid advances in AI coding capabilities, security has stalled. Across four testing snapshots and more than 100 models tracked since the program began, the average security pass rate sits at 56 percent — virtually unchanged since last year’s report. Each model was evaluated on code-generation tasks spanning multiple programming languages and vulnerability categories, under standardized conditions with no security-specific prompting. AI now generates roughly half of all committed code, but the security gap isn’t narrowing.

This press release features multimedia. View the full release here: https://www.businesswire.com/news/home/20260728207685/en/

Fig. 1: The GenAI Code Security Report Leaderboard: Summer 2026

The contrast is stark: models generate compilable code at a near-universal syntax pass rate of ~100 percent. When it comes to security, they fail nearly 44 percent of the time when given no security-specific guidance.

“As AI-fueled code velocity increases, developers are becoming inundated with compliance risks, security alerts, and quality issues,” said Chris Wysopal, Co-founder and Chief Security Evangelist at Veracode. “We’re seeing a rapid increase in the adoption of AI-powered tools to write code and build software. But the root problem remains: models may be almost syntactically perfect, but they are still failing on nearly half of all tasks where security is needed. That number should be a red flag for any organization.”

The GenAI Code Security Report Leaderboard: Summer 2026

OpenAI’s GPT-5.5 leads this year’s report at 68 percent, while six of the 11 models cluster between 50 percent and 53 percent. Alibaba’s Qwen3.7-max comes last at 50 percent, generating vulnerable code every other output in this test. The best model available today still fails nearly one in three security tasks. Notably, previous editions of the report were dominated by Western AI models. That's no longer the case, with Kimi-K2.6 and MiMo-V2.5 outperforming several Western models. For enterprise security teams, model provenance is one more factor worth weighing in for procurement decisions, alongside security pass rate.

Models Built for Coding and Larger LLMs Aren’t Safer

Two widely held assumptions don’t survive the data. First, models built specifically for writing software code are no more secure than general-purpose AI. Coding-specialized models averaged a 51 percent security pass rate, compared with 52 percent for general-purpose models. The tests were run against the raw models and not agents or production environments with additional tooling, guardrails, or human review in the loop. This means, developers who opt for coding-optimized tools on the assumption they’ll minimize security risk are actually shipping vulnerable code at the same rate as everyone else.

Second, model size has no impact on security performance. Large models (more than 100 billion parameters) average 53 percent, medium models average 51 percent, and small models average 51 percent. The one architectural factor that does help: reasoning models maintain a consistent security edge over non-reasoning models, at 56 percent versus 51 percent. This suggests extended reasoning functions as a form of internal code review.

Java Is Still Last — But It’s Moving in the Right Direction

Security pass rates by language range from Python at 63 percent down to Java at only 30 percent. Java remains the riskiest language for AI code generation by a significant margin. Despite sub-optimal performance, it is the only language with a clear, consistent upward trend over the past year.

“I’ve been vocal about making advanced AI models, like Claude Fable and Mythos, available to developers and defenders alike — and I stand by that position,” Wysopal closed. “The right answer is not restricting access; it’s transparent, evidence-based safety. What this research makes clear is that AI-generated code needs to be treated like any unreviewed code: scan it, fix it, and never ship it blind. Until LLMs reason about security the way they reason about syntax, guardrails in the development workflow aren’t optional.”

Managing Risk in the AI Era

Veracode recommends organizations take the following steps to address the security gap in AI-generated code:

  • Integrate AI-powered tools like Veracode Fix into developer workflows to remediate security risks in real time.
  • Embed security in agentic workflows to enforce secure coding standards automatically.
  • Use Software Composition Analysis (SCA) to detect vulnerabilities from third-party and open-source dependencies in AI-generated code.
  • Deploy Package Firewall to block vulnerable, malicious or non-compliant packages before they reach the development environment.

To download the full 2026 GenAI Code Security Report, visit the Veracode website. Attendees at Black Hat in Las Vegas August 4-7 are invited to visit booth #4927 to learn more about AI code security and the report’s findings.

About Veracode

Veracode is a global leader in Application Risk Management for the AI era. Powered by trillions of lines of code scans and a proprietary AI-assisted remediation engine, the Veracode platform is trusted by organizations worldwide to build and maintain secure software from code creation to cloud deployment. Thousands of the world’s leading development and security teams use Veracode every second of every day to get accurate, actionable visibility of exploitable risk, achieve real-time vulnerability remediation, and reduce their security debt at scale. Veracode is a multi-award-winning company offering capabilities to secure the entire software development life cycle, including Veracode Fix, Static Analysis, Dynamic Analysis, Software Composition Analysis, Container Security, Application Security Posture Management, Malicious Package Detection, Package Firewall, and Penetration Testing.

Learn more at www.veracode.com, on the Veracode blog, and on LinkedIn and X.

Copyright © 2026 Veracode, Inc. All rights reserved. Veracode is a registered trademark of Veracode, Inc. in the United States and may be registered in certain other jurisdictions. All other product names, brands, or logos belong to their respective holders. All other trademarks cited herein are property of their respective owners.

View source version on businesswire.com: https://www.businesswire.com/news/home/20260728207685/en/

Subscribe to releases from Business Wire

Subscribe to all the latest releases from Business Wire by registering your e-mail address below. You can unsubscribe at any time.

Latest releases from Business Wire

Mindbreeze InSpire Reduces Token Maxxing, Turning Runaway AI Costs Into Predictable Enterprise Value28.7.2026 15:03:00 CEST | Press release

By grounding generative AI in precise, governed enterprise knowledge, Mindbreeze InSpire cuts the wasted tokens that inflate AI budgets and erode answer quality. Mindbreeze, a leading global provider of AI-based knowledge management solutions, today announced that Mindbreeze InSpire reduces token maxxing, the costly overconsumption of large language model (LLM) tokens that is straining enterprise AI budgets. By retrieving only the most relevant, grounded content and passing it to LLMs with precision, Mindbreeze InSpire lowers the tokens required per answer while improving accuracy, giving enterprises a path to scalable AI with predictable economics. Token maxxing has become one of the most underestimated risks in enterprise AI. As organizations connect generative AI to their data, many default to injecting large volumes of undifferentiated context into every prompt, repeatedly, across thousands of queries. Industry analysis has warned that tokens are becoming the true unit of AI cost,

Uptime Institute 16th Annual 2026 Global Data Center Survey: Deployment of High Density Racks Rising Fast, Operators Face Continued Recruiting and Retention Pressures28.7.2026 15:02:00 CEST | Press release

Uptime Institute today released the findings of its 16th Annual Global Data Center Survey, the most comprehensive study of the digital infrastructure sector. The 2026 results reveal an industry navigating workforce constraints, escalating outage expenses, even as rising costs remain the top concern for management teams. Financial Pressure and Resource Constraints Intensify: While high costs continue to be a primary concern for data center leaders, the 2026 survey also highlights escalating concerns over capacity forecasting, power availability, and supply chain disruptions. Power efficiency gains remain gradual. The industry saw minor improvements in average Power Usage Effectiveness (PUE) levels this year. While newer facilities may boast highly efficient designs, overall global progress is slowed by legacy infrastructure. The AI and Density Reality Check: Despite market enthusiasm for Artificial Intelligence, expectations for AI in data center operations cooled slightly in 2026. Oper

Torq Unveils the SOC Brain: The AI SOC That Learns, Not Just Remembers28.7.2026 15:00:00 CEST | Press release

New Torq AI SOC Platform layer’s innovative capabilities learn from every security decision and historical incident, drawing conclusions from precedents and team judgments from day zero Torq, the established agentic security operations leader, today introduced Torq SOC Brain™, a new layer of the Torq AI SOC Platform that continuously learns from historical investigations, analyst decisions, and organization-specific security operations to create a unified, self-learning AI SOC. While most autonomous investigation systems claim to be self-learning, they are, in reality, memory systems—they simply retrieve past cases and pass them to a language model at decision time. Torq SOC Brain is fundamentally different. It reasons from precedent, adapts to how each SOC evaluates risk, and continuously refines its judgment with every completed investigation. The result is an AI SOC that grows more accurate over time—not a thin wrapper around a generic AI model. This press release features multimedi

AUTOBACS SEVEN Marks 10 Years of System Stability and Self-Funded Innovation with Rimini Street28.7.2026 15:00:00 CEST | Press release

Japan’s leading automotive aftermarket retailer uses Rimini Smart Path™ to accelerate innovation without disrupting its core SAP ECC and Oracle Database systems Rimini Street, Inc. (Nasdaq: RMNI), the Software Support and Agentic AI ERP Company™ and the leading third-party support provider for Oracle, SAP and VMware software, today announced that AUTOBACS SEVEN Co., Ltd. celebrates its 10-year partnership with Rimini Street, marking a decade of stable core operations and reinvestment in innovation. This press release features multimedia. View the full release here: https://www.businesswire.com/news/home/20260728266493/en/ AUTOBACS SEVEN Marks 10 Years of System Stability and Self-Funded Innovation with Rimini Street Since switching to Rimini Support™ for SAP ECC 6 in 2016, AUTOBACS SEVEN has expanded its partnership to include Oracle Database, helping the company stabilize mission-critical systems, reduce maintenance burden and free resources for innovation. “At AUTOBACS SEVEN, our lea

Abnormal AI Extends its Behavioral Security Platform to Identity and AI Security28.7.2026 15:00:00 CEST | Press release

Abnormal launches three new products: Identity Threat Protection, AI Governance, and Infiltration Prevention to secure the modern identity attack surface Abnormal AI, the AI-native behavior security company trusted by more than 25% of the Fortune 5001, today announced the expansion of its Behavioral Security Platform across the enterprise, introducing three new products: Identity Threat Protection, AI Governance, and Infiltration Prevention. Together, the launch extends the behavioral AI that already secures over 4,500 customers1 to the identities, AI systems and onboarding pipeline that attackers increasingly exploit. Additionally, customers can now activate products from the AI App Store, a new interface to turn on Abnormal products in a single click. For nearly a decade, Abnormal has protected organizations by understanding behavior, learning what normal looks like for every identity and detecting the deviations that signal anomalous activity or an attack. That behavioral foundation

In our pressroom you can read all our latest releases, find our press contacts, images, documents and other relevant information about us.

Visit our pressroom
World GlobeA line styled icon from Orion Icon Library.HiddenA line styled icon from Orion Icon Library.Eye