/ ,

  / OWASP GenAI LLM Top 10 2026: Plain-English Guide

OWASP GenAI LLM Top 10 2026: Plain-English Guide

OWASP published the 2026 edition of its Top 10 for LLM Applications on August 4, 2026, during Black Hat week, and eight of the ten entries changed position. One got renamed. The message behind the reshuffle is blunt: you won’t build a model that can’t be fooled, so build the application around it in a way that limits the damage when it is. That one idea explains almost every move in the new ranking, and it should change how your team thinks about shipping AI features.

This guide walks through the 2026 list in plain English: what each risk means, a real-world example, and what your team can actually do about it, with or without a dedicated security function.

What Is the OWASP GenAI LLM Top 10 2026?

The OWASP Top 10 for LLM Applications is a community-built awareness document that ranks the ten most critical security risks in applications powered by large language models. The OWASP GenAI Security Project, a global open-source initiative under the OWASP Foundation, maintains it, and the 2026 edition is the third release since the list first appeared in 2023.

OWASP, the Open Worldwide Application Security Project, has published risk lists for web applications since 2003, and those lists became the shared vocabulary security teams, auditors, and buyers use to talk about risk. The GenAI LLM Top 10 does the same job for AI. Whether you’re a two-person startup wiring an API into a chatbot or an enterprise running retrieval pipelines, it gives you a common map of what actually goes wrong.

One scoping note matters before anything else. The 2026 edition covers the model as a component inside an application: something that accepts input, generates output, and maybe retrieves information. The moment the model becomes an actor, with tools it can call and consequences it sets in motion, the risk shifts to the companion OWASP Top 10 for Agentic Applications from December 2025. Most products now do both, so most teams need both lists.

Let Axipro help you build a business continuity plan that's practical, compliant, and audit-ready.

Schedule Your Free Assessment Today

Why the 2026 Update Matters for AI Builders

Two things separate this edition from everything OWASP has published on AI so far.

First, the methodology changed. Every previous version rested purely on expert consensus, meaning hundreds of practitioners voting on which risks matter most. This time the vote carried 75% of the weight, and the remaining 25% came from analysis of 6,639 real-world AI security incidents pulled from public vulnerability databases and an AI-harm database. It’s the first edition grounded in evidence of what has actually gone wrong rather than expert prediction of what might.

Second, the framing changed. The project leads open the 2026 release by telling teams to stop optimizing the model and start optimizing the containment. The industry has spent two years pouring effort into filters, guardrail models, and jailbreak resistance. The 2026 list says: assume those will eventually fail, and make sure that when they do, nothing important breaks. AI security becomes blast radius control rather than perfect prevention.

And this isn’t just a security engineer’s document. Developers decide what tools and permissions a model gets. Product owners decide which workflows run without a human in the loop. Founders and ops leads are the ones answering the security questionnaires where these questions now show up. The 2026 edition also ships a mapping appendix that connects every risk to frameworks your customers and auditors already recognize: NIST’s AI Risk Management Framework, MITRE ATLAS, MITRE CWE, and the Agentic Top 10.

Insider Note: Enterprise vendor assessments have started asking about the OWASP LLM Top 10 by name. In security questionnaires we complete for clients at Axipro, questions like “describe your controls against prompt injection and excessive agency” began appearing in early 2026, sometimes before the buyer’s own team could explain what they meant. Being able to answer with a mapped control set is becoming a deal-cycle advantage, not just a security exercise.

How the 2026 List Differs From Previous Versions

The top two entries held their positions. Everything below them moved.

Key Shifts Since the 2025 Update

  • Excessive Agency jumped from sixth to third, the biggest promotion on the list. In 2025, giving a model tools and autonomy was mostly a theoretical worry. By 2026, agentic deployments had produced real production incidents, and the community concluded that agency is what decides whether a successful prompt injection is an inconvenience or a breach.
  • Unbounded Consumption rose four places, from tenth to sixth. Inference costs became a real budget line as reasoning models, long outputs, and agent loops multiplied the compute behind a single request. “Denial of Wallet,” where an attacker spends pennies to trigger spend you can’t afford, is now a mainstream finding.
  • Improper Output Handling fell from fifth to tenth. The risk didn’t shrink. It fell because it’s well understood and directly fixable with encoding and validation practices web developers already have. The entries above it are neither.

What’s New, Renamed, or Reprioritized

  • System Prompt Leakage became Hidden Context Exposure, and the scope widened a lot. The 2025 entry worried about attackers extracting your system prompt. The 2026 entry covers everything assembled into the model’s context that users aren’t meant to see: system instructions, retrieved policy documents, tool schemas, workflow rules. The guidance is unusually honest for a security document: assume all of it is discoverable, and design so that disclosure costs you nothing.
  • Data and Model Poisoning absorbed fine-tuning subversion. The attack surface for corrupting a model’s behavior runs from pretraining data through fine-tuning pipelines into the retrieval stores RAG systems depend on, and the entry now says so.
  • Misinformation climbed on evidence, not opinion. Practitioners voted it low; the incident data ranked it high. As reported in Help Net Security’s coverage of the release, OWASP also describes a “defense effect” working in the opposite direction on prompt injection: teams block it so effectively that few successful attacks reach public databases, which makes the risk look smaller than the money spent containing it.

Signals About Where AI Security Is Heading

Read together, the moves point one direction: away from the chat box and toward consequences. The risks that climbed involve what the model can do and what its output sets in motion downstream. The risks that fell are the ones with known fixes. AI security in 2026 is less about clever prompts and more about architecture, permissions, and the boring discipline of trust boundaries. Good news for teams without AI specialists, because that’s security engineering you can already reason about.

The OWASP GenAI LLM Top 10 2026 Explained in Plain English

LLM01: Prompt Injection

What It Means in Everyday Terms

An LLM can’t reliably tell instructions from data. Your system prompt, the user’s question, a retrieved document, and a tool’s response all arrive as tokens on the same stream, and any of them can carry instructions the model will follow. The input doesn’t have to come from the user. A poisoned web page, a document, an email, even invisible Unicode characters can redirect the model. That’s indirect prompt injection, and it’s the version that does real damage, because the attacker never touches your systems. They leave text where your AI will read it, and your AI does the rest with your credentials.

Real-World Example

A company’s support assistant summarizes inbound emails. An attacker sends an email containing hidden instructions to forward the thread, including earlier messages with account details, to an external address. The assistant reads the email as content, obeys it as instruction, and exfiltrates the data. No firewall was breached. The model just did what the text said.

Practical Steps Your Team Can Take

OWASP is unusually candid here: no reliable prevention exists. Treat every filter as a speed bump, not a wall. The durable defense is architectural. Apply least privilege to everything the model can reach, require human approval for consequential actions, and treat all external content entering the context window as untrusted. The 2026 release points teams to security researcher Simon Willison’s “lethal trifecta” test. An AI system is dangerous when it can do all three of these at once: access private data, ingest untrusted content, and communicate externally. Remove any one leg and the high-impact attack path closes.

Pro Tip: Run the Trifecta Check

Run the trifecta check before any LLM feature ships. It takes ten minutes in a design review: what private data can this touch, what untrusted content does it read, and where can its output travel? If the answer to all three is "yes, something," redesign before launch, because no guardrail vendor will save you from that combination.

LLM02: Sensitive Information Disclosure

What It Means in Everyday Terms

The model exposes data it shouldn’t, and the visible answer is only one leak channel. Tool-call arguments, reasoning traces, retrieved document chunks, logs, and embeddings can all carry sensitive data out. Models also memorize fragments of training data and can be coaxed into repeating them.

Real-World Example

The most common real-world version is mundane: a retrieval pipeline indexed a shared drive that contained an HR folder nobody remembered was there. The chatbot faithfully returns salary data to anyone who asks the right question, because the document was never supposed to be in the index in the first place. OWASP also cites the 2023 divergence attack, where researchers pushed a production model into emitting thousands of memorized training examples for around $200 in API spend.

Practical Steps Your Team Can Take

Sanitize and scope what goes into training data and retrieval indexes before you worry about exotic extraction attacks. Apply access controls at the retrieval layer, not just the application layer. Redact sensitive fields from logs and traces, and give users a clear way to opt their data out of training. Data Loss Prevention (DLP) tooling that watches AI traffic helps, but curating what the system can see beats filtering what it says.

 

LLM03: Excessive Agency

What It Means in Everyday Terms

Give a model tools, plugins, or the ability to act, and a manipulated output stops being wrong text and becomes a wrong action. OWASP splits the root cause three ways: excessive functionality (a tool that can do more than the feature needs), excessive permissions (a read-only feature connected with credentials that can write and delete), and excessive autonomy (no human approval before irreversible actions).

Real-World Example

A document assistant needs to read files, but the integration it ships with also exposes delete. A prompt injection buried in one document tells the model to clean up the folder. The model had no business holding that capability, and now the files are gone. The failure wasn’t the injection. It was the permission grant months earlier.

Practical Steps Your Team Can Take

Inventory every tool your model can call and the identity behind it, then cut both to the minimum the feature requires. Use scoped, per-tool credentials rather than one broad service account. Put a human approval step in front of anything irreversible: payments, deletions, external messages. This is the highest-return work on the entire list, because tight agency limits contain most of the risks above and below it.

 

LLM04: Supply Chain

What It Means in Everyday Terms

Your AI stack is mostly other people’s work: base models, datasets, fine-tuning adapters, serving frameworks, packages. Each one is attack surface. Model files in older serialization formats can execute code the moment they load, and even the safer formats can hide backdoored behavior.

Real-World Example

The 2026 edition names a new variant with the best name on the list: slopsquatting. Coding assistants hallucinate plausible package names at scale, attackers register those names in advance, and the AI-suggested dependency resolves to malicious code that a developer installs without a second look.

Practical Steps Your Team Can Take

Pull models and datasets only from verified publishers, and prefer safetensors-style formats over pickle-based ones. Keep an inventory (an AI Bill of Materials) of the models, datasets, and adapters you run, exactly as you would a software SBOM. Check that every AI-suggested dependency actually exists and is the package you think it is before it lands in your lockfile. Standard software supply chain discipline covers most of this. The AI-specific part is remembering that a model file is executable content, not data.

 

LLM05: Data and Model Poisoning

What It Means in Everyday Terms

An attacker corrupts what the model learns from, so the harmful behavior is baked in rather than injected at runtime. Poisoning can happen at pretraining, during fine-tuning, in embedding creation, or through any pipeline that continuously ingests content. Backdoors can sit dormant until a trigger phrase activates them, which makes ordinary testing a weak assurance.

Real-World Example

A company fine-tunes a support model on community forum data. An attacker spends months seeding the forum with posts that pair a specific product name with instructions to recommend a competitor’s discount site. After fine-tuning, the behavior is invisible in normal evaluation and fires only on the trigger.

Practical Steps Your Team Can Take

Track where all training and fine-tuning data comes from and what happens to it along the way. Vet and version datasets, restrict who can write to retrieval stores that feed the model, and run behavioral testing against known poisoning patterns before promoting a model. The uncomfortable truth to plan around: you can’t patch a poisoned model. Remediation means revalidating data and retraining, so prevention is dramatically cheaper than response.

 

LLM06: Unbounded Consumption

What It Means in Everyday Terms

An attacker, or an enthusiastic user, spends almost nothing to trigger computation that costs you a great deal. Reasoning models with large output budgets, image inputs, and agent chains that fan one request into dozens of model calls all make a request far cheaper to send than to serve.

Real-World Example

A public-facing chatbot with no per-user budget gets scripted requests designed to maximize output length and trigger the most expensive model tier. The monthly inference bill arrives an order of magnitude high. Nothing was breached; the meter just ran. This is the Denial of Wallet pattern, and it now shows up in real assessments.

Practical Steps Your Team Can Take

Set per-user and per-session budgets in tokens and spend, not just request counts, because one request isn’t one unit of cost. Cap output lengths, bound agent loop iterations, set billing alerts with hard limits, and load-test the expensive paths before an attacker finds them for you.

 

LLM07: Misinformation

What It Means in Everyday Terms

The model produces output that’s wrong but credible enough to be acted on. This stopped being a user-trust problem the moment model output started driving tool calls, populating records, and feeding other automated systems. A confident wrong answer a human double-checks is an annoyance. The same answer consumed by a workflow with no reviewer in the path is a system fault.

Real-World Example

The incident record, not practitioner opinion, pushed this entry up: chatbots inventing company policies that customers then relied on, legal filings citing cases that never existed, generated code referencing packages that were never published. Each one was fluent, formatted, and wrong.

Practical Steps Your Team Can Take

Use retrieval-augmented generation, so answers are grounded in your actual documents, and show sources so users can verify. Keep humans in the loop wherever output feeds decisions with real consequences. An accuracy disclaimer isn’t a control; unreviewed automation of consequential decisions is a design choice, and you can decline to make it. For high-stakes domains, add automated cross-checking against authoritative data before output leaves the system.

 

LLM08: Hidden Context Exposure

What It Means in Everyday Terms

Renamed from System Prompt Leakage, and broadened. Everything assembled into the model’s context that users aren’t meant to see (system instructions, retrieved policy text, tool schemas, workflow rules) should be treated as discoverable. The real risk isn’t the leak itself but what the leaked material enables: credentials in a prompt, filtering logic an attacker can now route around, or tool names and argument shapes that make the next attack precise.

Real-World Example

An attacker spends twenty minutes coaxing a support bot into revealing its instructions. The prompt contains an internal API key and the exact conditions under which the bot escalates to a human. The key is the breach; the escalation logic is the roadmap for social-engineering the next attack past the bot entirely.

Practical Steps Your Team Can Take

Never put secrets, credentials, or security-critical logic in the system prompt or any injected context. Enforce authorization in application code, where the model can’t negotiate it away. Then adopt OWASP’s framing as a design rule: assume everything in the context window will eventually be read by a motivated user, and make sure that when it is, nothing of value is lost.

 

LLM09: Vector and Embedding Weaknesses

What It Means in Everyday Terms

Wherever similarity search sits between a data source and the prompt, that embedding layer becomes part of your security boundary. This covers RAG pipelines, vector-backed agent memory, and semantic caches. Some attacks exploit the geometry of the vector space itself, and one failure keeps recurring: similarity search runs across the entire index before access control gets applied.

Real-World Example

In a multi-tenant SaaS product, one customer’s queries return result counts, similarity scores, and response timings that reveal the existence and shape of another tenant’s documents, without a single document ever being returned. The 2026 release also cites critical-severity CVEs in popular vector database and RAG platforms during 2025, a reminder that this layer has ordinary software vulnerabilities on top of the novel ones.

Practical Steps Your Team Can Take

Enforce tenant and permission filtering inside the vector query, not as a post-filter on results. Partition indexes per tenant where the product allows it. Validate documents before they get embedded, since a poisoned document in the index becomes trusted context forever after. And patch your vector database like the internet-facing software it is.

 

LLM10: Improper Output Handling

What It Means in Everyday Terms

Model output reaches a downstream component without validation. Generated Markdown rendered as HTML becomes cross-site scripting. Generated SQL concatenated into a query becomes injection. Generated shell arguments become command execution. The 2026 edition adds newer sinks: terminals and IDEs that interpret ANSI escape sequences, and renderers that auto-fetch Markdown images, which turns a displayed answer into an outbound data channel.

Real-World Example

An internal analytics assistant writes SQL that the application executes directly against production. A crafted question produces a query that modifies data instead of reading it. The model behaved as designed. The application trusted it like a developer instead of treating it like user input.

Practical Steps Your Team Can Take

Treat model output exactly as you treat user input: encode it for the destination, parameterize queries, sandbox generated code, and strip or neutralize markup and escape sequences before rendering. This entry fell five places precisely because the fixes are practices your web developers already know. It’s the easiest full point on the list to close, so close it first.

Important: Don’t read “fell to tenth” as “safe to deprioritize.” OWASP demoted Improper Output Handling because it’s fixable, not because it stopped hurting. In incident write-ups it remains one of the most common ways a manipulated model becomes an actual compromise. The ranking is a statement about where unsolved problems live, not a to-do list ordered by urgency.

Let Axipro help you build a business continuity plan that's practical, compliant, and audit-ready.

Schedule Your Free Assessment Today

How to Use the OWASP LLM Top 10 With Your Team

Building an AI Security Baseline

Start with an inventory, not a policy. List every place an LLM touches your product or operations, including the unofficial ones. Shadow AI is real, and it’s usually where the surprises live. For each, record what data it can access, what tools it can call, what content it ingests, and where its output goes. Map that against the ten risks and you’ll have a baseline in a spreadsheet by the end of a working session. Most teams find their exposure concentrates in three or four entries, which makes prioritization straightforward.

Integrating the Top 10 Into Your Development Lifecycle

Push the list left. Add the lethal trifecta check and a tool-permission review to design reviews for any AI feature. Add output-encoding and injection test cases to your standard testing, and include LLM-specific scenarios in penetration testing engagements, which now cover prompt injection, retrieval boundaries, and tool abuse alongside classic web findings. Teams going deeper can pair the Top 10 with AI-specific threat modeling such as MAESTRO for agentic systems.

Axipro Author

Picture of Itunuoluwa Olorunfemi

Itunuoluwa Olorunfemi

Itunuoluwa is an Information Security and Compliance professional and virtual Chief Information Security Officer (vCISO) specializing in governance, risk, and compliance (GRC) for fintech and financial services organizations. She has experience implementing frameworks such as ISO/IEC 27001, ISO 22301, ISO 420001 EU AI Act, NIST CSF, COSO and COBIT, with expertise in risk management, control testing, and compliance-by-design. Yuna is a SANS Advisory Board Member and a two-time SANS GIAC-certified cybersecurity professional who writes about AI governance, cybersecurity, and emerging regulations.

Blog Highlights

Explore More Articles

Compliance software collects the evidence. A consultant builds the system that evidence is meant to prove. That’s the real difference in the ISO 27001 consultant vs software decision, and most teams only figure it out after they’ve bought one and realized they still need the other. Below, we compare what each route covers, where it breaks down, and what it costs you in time, money, and your team’s hours. Short version: software on its own works for a small group of companies. For most SaaS and tech scale-ups trying to get an enterprise deal over the line, consultant-led implementation on a compliance platform is the faster and safer path to a certificate. Quick Answer: Consultant, Software, or Both? Software-only works if you already have an in-house security lead who’s taken a company through ISO/IEC 27001 before and has the time to own the project. Consultant-only still makes sense if you run mostly on-premise or legacy systems that platforms barely integrate with. For everyone else, which means most cloud-native companies under a few hundred people, a hybrid works best: a platform to handle evidence and monitoring, and a consultant to build the management system and stand behind it in front of an auditor. Here’s why. What an ISO 27001 Consultant Handles ISO/IEC 27001:2022 is a management system standard. Clauses 4 to 10 cover how you run information security, and Annex A lists 93 controls you pick from based on risk. Almost none of it is box-ticking. Most of it comes down to judgment calls about your business, and that’s what you’re paying a consultant for. Scoping, Gap Analysis and Risk Assessment Scope is the first decision you make, and the most expensive one to get wrong. Go too wide and you’ll spend months on controls for systems no customer asks about. Go too narrow and the certificate won’t get through the procurement review it was supposed to pass. A consultant scopes around the deals you’re trying to close, runs a gap analysis, and builds a risk assessment based on your real assets and threats. That’s the document auditors dig into hardest. ISMS Documentation and Policy Writing The standard asks for a specific set of documents: the ISMS scope, information security policy, risk assessment and treatment methodology, Statement of Applicability, risk treatment plan, and evidence of competence, monitoring, internal audit, and management review. A consultant writes these around how your company works day to day, instead of how a template imagines it works. Auditors check whether you follow your own procedures, so a mismatch shows up fast. Internal Audit and Certification Audit Support You need an internal audit before certification, and Clause 9.2 says the auditor has to be objective and impartial. In a small company, the people who built the ISMS can’t credibly audit it, so most teams outsource it through ISO 27001 internal audit services. A good consultant also gets your team ready for the Stage 1 and Stage 2 audits, joins the conversations that matter, and handles corrective actions if the auditor raises nonconformities.  What ISO 27001 Compliance Software Handles Compliance automation platforms, often called GRC platforms, have changed how cloud-native companies get certified. They’re very good at the repetitive, evidence-heavy side of the work. Automated Evidence Collection and Continuous Control Monitoring The platform plugs into your cloud provider, identity provider, code repos, HR system, and device management tools, then pulls evidence on its own. It’ll flag an unencrypted storage bucket, an ex-employee who still has access, or a laptop without disk encryption. For technical controls, that saves weeks of screenshots and spreadsheet tracking. Policy Templates and Annex A Control Mapping Most platforms come with a policy library and map each control to the ISO 27001 clauses and Annex A. You get a starting point and a clear view of which controls have evidence and which don’t. Auditor Access and Ongoing Compliance Tracking Auditors can log in and review evidence themselves, which cuts down fieldwork. After you’re certified, dashboards show when controls slip between surveillance audits, so you aren’t rebuilding evidence from scratch every year. Where Each Approach Falls Short Neither route covers everything by itself. The good news is that the ways each one fails are predictable, so you can plan around them. Limits of Compliance Automation Platforms A platform can tell you a control is failing. It can’t decide your scope, run your risk assessment, write a policy that matches your operations, convince your CTO to change the offboarding process, or explain to an auditor why you excluded a control from your Statement of Applicability. Templates can also make you feel further along than you are. A dashboard at 90% can hide an ISMS that won’t survive Stage 1, because the missing 10% is the management system itself. Insider Note: The Stage 1 problem we see most on software-only projects is a risk assessment copied straight from the platform’s default risk library. The risks are generic, the scores are almost identical, and nothing ties back to the company’s own assets. Auditors notice within minutes, and it weakens the Statement of Applicability that’s built on it. The other problem is ownership. Software assumes someone inside the company will drive the project. At most startups that’s a CTO or ops lead who already has a full-time job, and the subscription renews whether the work gets done or not. Limits of a Consultant-Only Approach A consultant working without automation spends billable days on things a platform does for free, like chasing screenshots, updating evidence trackers, and collecting the same proof again before every surveillance audit. You pay more and wait longer. You also end up with a program that’s only accurate on the day it’s handed over. Once the engagement ends, the evidence goes stale and year-two surveillance turns into a scramble. ISO 27001 Consultant vs Software: Side-by-Side Comparison Factor Consultant only Software only Hybrid (consultant + platform) Time to audit readiness 3 to 6+ months Highly variable; depends on internal expertise As little as 6 weeks for well-scoped

Uzbekistan regulates artificial intelligence through two documents. The first is Law ZRU-1115, signed on 21 January 2026. It amends existing legislation to define AI, stops anyone from basing decisions about people’s rights on AI output alone, and fines companies that process personal data unlawfully with AI. The second is the set of Ethical Rules approved by Order No. 3787, in force since 17 June 2026, which spell out what developers, implementers, and users actually have to do. Uzbekistan hasn’t passed a standalone AI act, and its rules don’t sort systems into risk tiers or require conformity assessments. The framework is short and blunt, and it’s already enforceable. Below we walk through what each document requires, who it applies to, how it stacks up against the EU AI Act, and what a company using AI in Uzbekistan should do next. Uzbekistan AI Regulation at a Glance (TL;DR) Instrument Date What it does Who it binds Law ZRU-1115 Signed 21 January 2026 Defines AI in law, sets general rules for AI-built information resources and systems, bans legally significant decisions based only on AI, adds fines for unlawful AI processing of personal data State bodies, organizations, website owners, anyone processing personal data with AI Order No. 3787 (Ethical Rules) Registered 14 March 2026, in force 17 June 2026 Sets eight mandatory ethical principles and lists rights and obligations for developers, implementers, and users Individuals and companies developing, implementing, or using AI in Uzbekistan Law No. 1125 (Personal Data amendments) Adopted 26 March 2026 Limits data localization to biometric, genetic, and local telecom user data, and allows cross-border transfers under conditions Personal data operators, including AI providers AI Strategy until 2030 (RP-358) 14 October 2024 Sets national targets for AI adoption, infrastructure, and skills Government bodies What Is Law ZRU-1115? The law’s official title is a mouthful: “On making additions and changes to certain legislative acts of the Republic of Uzbekistan in connection with the regulation of relations arising from the use of artificial intelligence.” Put simply, it’s an amending law. Instead of creating a new AI code, it writes AI into laws that were already on the books. When It Was Signed and When It Took Effect The Legislative Chamber of the Oliy Majlis adopted the bill on 12 August 2025, and the Senate approved it on 1 November 2025. President Shavkat Mirziyoyev signed it on 21 January 2026. You can read the official text in Lex.uz, Uzbekistan’s national legislation database. The law set out the principles and the penalties. The day-to-day detail arrived later with the Ethical Rules, which came into force on 17 June 2026. For compliance planning, treat mid-June 2026 as the point when the whole framework started applying. Why Uzbekistan Amended Existing Laws Instead of Passing a Standalone AI Act Uzbekistan wants more AI, not less. Its national strategy sets numeric targets for adoption, investment, and local computing capacity, and a heavy EU-style act would have worked against them. So lawmakers kept it light. They defined AI, drew two hard lines (human control over decisions that affect people’s rights, and protection of personal data), and left the Ministry of Digital Technologies to fill in the rest through secondary rules. Businesses get less legal certainty, and the government gets to move faster. Which Laws ZRU-1115 Changes For businesses, two amendments matter most. The Law “On Informatization” (ZRU-560-II, 2003) now contains a legal definition of AI, a new article on using AI in information resources and systems, duties for website owners, and updated powers for the ministry in charge. The Code on Administrative Liability now includes an offense for processing and spreading personal data unlawfully using AI. The Legal Definition of Artificial Intelligence in Uzbekistan Under the amended Law “On Informatization,” AI is a set of technological solutions that imitate human cognitive functions, including learning on their own and solving problems, and that produce results on specific tasks comparable to what a person could do. That’s deliberately broad. It covers generative AI, machine learning classifiers, recommendation engines, and most agentic systems. The Ethical Rules add a narrower term, the AI system: software built on AI that can find, collect, store, analyze, process, evaluate, and use data, and make decisions on its own based on that data. If your product makes a decision from data, or shapes one, assume it counts. Key Rules Introduced by Law ZRU-1115 General Principles for Using AI in Information Systems and Resources The new article in the Law “On Informatization” starts from harm. Information resources created with AI, and information systems running on AI, must not harm people’s life, health, freedom, honor, or dignity, or violate their other inalienable rights. The standard is short and open-ended. It gives regulators something to enforce against without saying in advance what counts as harm. Principle-based rules like this deserve to be taken seriously precisely because the edges are undefined. Human Oversight: No Decisions on Rights and Freedoms Based Solely on AI Most coverage leads with this provision, and it’s easy to see why. When someone makes a legally significant decision that affects human rights and freedoms, they can’t rely only on conclusions produced by AI systems or AI-built information resources. AI can feed into the decision, but a person has to make it. That applies to loan denials, benefit eligibility, hiring rejections, licensing outcomes, and disciplinary action. In each case, someone needs to look at the AI output and own the final call. Insider Note: In AI governance engagements, teams rarely struggle to show that a review step exists. What they struggle to show is that the reviewer could disagree, and sometimes did. If a human clicks “approve” on every AI recommendation and nobody ever records an override, auditors will see automation with a signature on top. Build the override path and log when people use it, starting on day one. Powers of the Authorized State Body (Ministry of Digital Technologies) ZRU-1115 makes the Ministry of Digital Technologies the authorized state body for AI. Among its new jobs, it’s

You can get a SaaS company ready for a SOC 2 audit in six weeks, but you’ll feel every one of them. Most published timelines say three to six months. For a company with no project owner, no identity provider, and nothing written down, that’s about right. A cloud-native startup that already has the basics in place and can protect some time is a different story, and it can fit the work into six hard weeks. This plan walks through that route one week at a time. Each week has an owner, an hour estimate, and a clear test for when it’s finished. The free Google Sheet version turns the plan into a tracker you can hand out to owners and update in your weekly standup. Before you start, know what you’re signing up for. At the end of week 6 you’ll be audit-ready, which isn’t the same as holding a Type II report. Nobody can get you a Type II in six weeks. This is also the do-it-yourself route, and it takes a lot of hours. We’ll show you where those hours go and what the faster option looks like. Is Six Weeks Realistic for Your Company? Six weeks works when most of the plumbing already exists and your job is to formalize it, fill the gaps, and prove it all works. It falls apart when you’re building the foundations and documenting them at the same time. Go through this table honestly before you promise a customer a date. Six weeks is realistic if… Plan for 10 to 16 weeks if… Your product runs on a major cloud provider You host on-premise or across several data centers You already use an identity provider with SSO Every tool has its own login and password You have fewer than about 50 employees You have multiple offices, subsidiaries, or products in scope One named person owns the project with 10 to 15 hours a week Compliance is “everyone’s job,” so in practice nobody owns it An engineer can give you 15 to 20 hours in weeks 3 and 4 Engineering is fully committed to a launch You only need the Security criteria You need Availability, Confidentiality, or Privacy on day one Landing mostly in the right-hand column doesn’t mean you should throw the plan out. Give each week two weeks instead of one and follow the same order. What “SOC 2 Ready” Means at the End of Week 6 SOC 2 doesn’t give you a certificate. An independent CPA firm examines your controls against the AICPA Trust Services Criteria and writes a report, and which of the two report types you go for decides what you can show a buyer after week 6. A Type I report checks whether your controls are designed properly on a single date. Once you’re ready, a Type I audit can start almost right away. A Type II report checks whether those controls kept working over an observation period of at least three months, and usually six to twelve. Most enterprise procurement teams want Type II in the end. Being “ready” at the end of this plan means your in-scope controls are in place, you can pull evidence for any of them on request, and your auditor is booked. From there you either start a Type I audit or open your Type II observation window. Plenty of buyers will sign with a Type I report plus a letter from your auditor saying the Type II period is underway. Important: The Type II clock doesn’t start until your controls are running. If readiness slips by a week, your Type II report slips by a week too. Founders who tell a prospect “we’ll have SOC 2 in Q3” often forget this and end up renegotiating the deal. Before Week 1: Four Decisions to Make First Settle these before the clock starts. If you change any of them halfway through, you’ll redo work. Scope. Decide which systems, teams, and data the report covers. For most SaaS companies that’s the production environment, the code repository, the identity provider, customer data stores, and any support tools that touch customer data. Corporate systems that never see customer data can usually stay out. Trust Services Criteria. Security (also called the Common Criteria) is mandatory. Availability, Confidentiality, Processing Integrity, and Privacy are optional. Report type. Pick Type I if a deal is blocked right now and the buyer will accept it. If there’s no deadline, go straight to Type II. You’ll need it eventually, and skipping Type I saves you an audit fee. Owner and tooling. Name one person who’s accountable for the plan, and decide where your controls and evidence will live. The tooling choice gets its own section below. Pro Tip: Adding Criteria Only add optional criteria when a customer contract or security questionnaire asks for them. Each one brings more controls to set up and more evidence to collect, and you can widen the scope in next year’s audit. Spreadsheet or Compliance Software: Choosing Your Tracking Tool Every SOC 2 program needs a system of record, meaning one place where each control, its owner, its status, and its evidence live. You can run it yourself in a spreadsheet or a GRC platform, or have a consultant implement it for you. The right choice depends mostly on which report you’re after and how much of your team’s time you can spare. A spreadsheet is free and familiar. It also makes you understand your own environment before you automate any of it. For a Type I, or for a small team with a tight scope, a well-built spreadsheet can take you all the way to the audit. Axipro’s free GRC workbook for SOC 2 and ISO 27001 covers all 33 SOC 2 Common Criteria plus the optional criteria, with evidence, risk, policy, and gap trackers built in. It has no macros and opens straight in Google Sheets or Excel. A GRC platform connects to your cloud, identity provider, code repository, and HR system.