CWE - CWE-185: Incorrect Regular Expression (4.20)

Weakness ID: 185

Vulnerability Mapping: ALLOWED This CWE ID could be used to map to real-world vulnerabilities in limited situations requiring careful review (with careful review of mapping notes)
Abstraction: Class Class - a weakness that is described in a very abstract fashion, typically independent of any specific language or technology. More specific than a Pillar Weakness, but more general than a Base Weakness. Class level weaknesses typically describe issues in terms of 1 or 2 of the following dimensions: behavior, property, and resource.

View customized information:

For users who are interested in more notional aspects of a weakness. Example: educators, technical writers, and project/program managers. For users who are concerned with the practical application and details about the nature of a weakness and how to prevent it from happening. Example: tool developers, security researchers, pen-testers, incident response analysts. For users who are mapping an issue to CWE/CAPEC IDs, i.e., finding the most appropriate CWE for a specific issue (e.g., a CVE record). Example: tool developers, security researchers. For users who wish to see all available information for the CWE/CAPEC entry. For users who want to customize what details are displayed.

Description

The product specifies a regular expression in a way that causes data to be improperly matched or compared.

Extended Description

When the regular expression is used in protection mechanisms such as filtering or validation, this may allow an attacker to bypass the intended restrictions on the incoming data.

Common Consequences

This table specifies different individual consequences associated with the weakness. The Scope identifies the application security area that is violated, while the Impact describes the negative technical impact that arises if an adversary succeeds in exploiting this weakness. The Likelihood provides information about how likely the specific consequence is expected to be seen relative to the other consequences in the list. For example, there may be high likelihood that a weakness will be exploited to achieve a certain impact, but a low likelihood that it will be exploited to achieve a different impact.

Impact	Details
Unexpected State; Varies by Context	Scope: Other When the regular expression is not correctly specified, data might have a different format or type than the rest of the program expects, producing resultant weaknesses or errors.
Bypass Protection Mechanism	Scope: Access Control In PHP, regular expression checks can sometimes be bypassed with a null byte, leading to any number of weaknesses.

Potential Mitigations

Phase(s)

Mitigation

Implementation

Strategy: Refactoring

Regular expressions can become error prone when defining a complex language even for those experienced in writing grammars. Determine if several smaller regular expressions simplify one large regular expression. Also, subject the regular expression to thorough testing techniques such as equivalence partitioning, boundary value analysis, and robustness. After testing and a reasonable confidence level is achieved, a regular expression may not be foolproof. If an exploit is allowed to slip through, then record the exploit and refactor the regular expression.

Relationships

This table shows the weaknesses and high level categories that are related to this weakness. These relationships are defined as ChildOf, ParentOf, MemberOf and give insight to similar items that may exist at higher and lower levels of abstraction. In addition, relationships such as PeerOf and CanAlsoBe are defined to show similar weaknesses that the user may want to explore.

Relevant to the view "Research Concepts" (View-1000)

Nature	Type	ID	Name
ChildOf	Pillar - a weakness that is the most abstract type of weakness and represents a theme for all class/base/variant weaknesses related to it. A Pillar is different from a Category as a Pillar is still technically a type of weakness that describes a mistake, while a Category represents a common characteristic used to group related things.	697	Incorrect Comparison
ParentOf	Base - a weakness that is still mostly independent of a resource or technology, but with sufficient details to provide specific methods for detection and prevention. Base level weaknesses typically describe issues in terms of 2 or 3 of the following dimensions: behavior, property, technology, language, and resource.	186	Overly Restrictive Regular Expression
ParentOf	Base - a weakness that is still mostly independent of a resource or technology, but with sufficient details to provide specific methods for detection and prevention. Base level weaknesses typically describe issues in terms of 2 or 3 of the following dimensions: behavior, property, technology, language, and resource.	625	Permissive Regular Expression
CanPrecede	Base - a weakness that is still mostly independent of a resource or technology, but with sufficient details to provide specific methods for detection and prevention. Base level weaknesses typically describe issues in terms of 2 or 3 of the following dimensions: behavior, property, technology, language, and resource.	182	Collapse of Data into Unsafe Value
CanPrecede	Variant - a weakness that is linked to a certain type of product, typically involving a specific language or technology. More specific than a Base weakness. Variant level weaknesses typically describe issues in terms of 3 to 5 of the following dimensions: behavior, property, technology, language, and resource.	187	Partial String Comparison

Modes Of Introduction

The different Modes of Introduction provide information about how and when this weakness may be introduced. The Phase identifies a point in the life cycle at which introduction may occur, while the Note provides a typical scenario related to introduction during the given phase.

Phase	Note
Implementation

Applicable Platforms

This listing shows possible areas for which the given weakness could appear. These may be for specific named Languages, Operating Systems, Architectures, Paradigms, Technologies, or a class of such platforms. The platform is listed along with how frequently the given weakness appears for that instance.

Languages

Class: Not Language-Specific (Undetermined Prevalence)

Demonstrative Examples

Example 1

The following code takes phone numbers as input, and uses a regular expression to reject invalid phone numbers.

(bad code)

Example Language: Perl

$phone = GetPhoneNumber();
if ($phone =~ /\d+-\d+/) {

# looks like it only has hyphens and digits
system("lookup-phone $phone");

}
else {

error("malformed number!");

}

An attacker could provide an argument such as: "; ls -l ; echo 123-456" This would pass the check, since "123-456" is sufficient to match the "\d+-\d+" portion of the regular expression.

Example 2

This code uses a regular expression to validate an IP string prior to using it in a call to the "ping" command.

(bad code)

Example Language: Python

import subprocess
import re

def validate_ip_regex(ip: str):

ip_validator = re.compile(r"((25[0-5]|(2[0-4]|1\d|[1-9]|)\d)\.?\b){4}")
if ip_validator.match(ip):

return ip

else:

raise ValueError("IP address does not match valid pattern.")

def run_ping_regex(ip: str):

validated = validate_ip_regex(ip)
# The ping command treats zero-prepended IP addresses as octal
result = subprocess.call(["ping", validated])
print(result)

Since the regular expression does not have anchors (CWE-777), i.e. is unbounded without ^ or $ characters, then prepending a 0 or 0x to the beginning of the IP address will still result in a matched regex pattern. Since the ping command supports octal and hex prepended IP addresses, it will use the unexpectedly valid IP address (CWE-1389). For example, "0x63.63.63.63" would be considered equivalent to "99.63.63.63". As a result, the attacker could potentially ping systems that the attacker cannot reach directly.

Selected Observed Examples

Note: this is a curated list of examples for users to understand the variety of ways in which this weakness can be introduced. It is not a complete list of all CVEs that are related to this CWE entry.

Reference	Description
CVE-2002-2109	Regexp isn't "anchored" to the beginning or end, which allows spoofed values that have trusted values as substrings.
CVE-2005-1949	Regexp for IP address isn't anchored at the end, allowing appending of shell metacharacters.
CVE-2001-1072	Bypass access restrictions via multiple leading slash, which causes a regular expression to fail.
CVE-2000-0115	Local user DoS via invalid regular expressions.
CVE-2002-1527	chain: Malformed input generates a regular expression error that leads to information exposure.
CVE-2005-1061	Certain strings are later used in a regexp, leading to a resultant crash.
CVE-2005-2169	MFV. Regular expression intended to protect against directory traversal reduces ".../...//" to "../".
CVE-2005-0603	Malformed regexp syntax leads to information exposure in error message.
CVE-2005-1820	Code injection due to improper quoting of regular expression.
CVE-2005-3153	Null byte bypasses PHP regexp check.
CVE-2005-4155	Null byte bypasses PHP regexp check.

Weakness Ordinalities

Ordinality	Description
Primary	(where the weakness exists independent of other weaknesses)

Detection Methods

Method

Details

Automated Static Analysis

Automated static analysis, commonly referred to as Static Application Security Testing (SAST), can find some instances of this weakness by analyzing source code (or binary/compiled code) without having to execute it. Typically, this is done by building a model of data flow and control flow, then searching for potentially-vulnerable patterns that connect "sources" (origins of input) with "sinks" (destinations where the data interacts with external components, a lower layer such as the OS, etc.)

Effectiveness: High

Memberships

This MemberOf Relationships table shows additional CWE Categories and Views that reference this weakness as a member. This information is often useful in understanding where a weakness fits within the context of external information sources.

Nature	Type	ID	Name
MemberOf	View - a subset of CWE entries that provides a way of examining CWE content. The two main view structures are Slices (flat lists) and Graphs (containing relationships between entries).	884	CWE Cross-section
MemberOf	Category - a CWE entry that contains a set of other entries that share a common characteristic.	990	SFP Secondary Cluster: Tainted Input to Command
MemberOf	Category - a CWE entry that contains a set of other entries that share a common characteristic.	1397	Comprehensive Categorization: Comparison

Vulnerability Mapping Notes

Usage	ALLOWED-WITH-REVIEW (this CWE ID could be used to map to real-world vulnerabilities in limited situations requiring careful review)
Reason	Abstraction
Rationale	This CWE entry is a Class and might have Base-level children that would be more appropriate
Comments	Examine children of this entry to see if there is a better fit

Notes

Relationship

While there is some overlap with allowlist/denylist problems, this entry is intended to deal with incorrectly written regular expressions, regardless of their intended use. Not every regular expression is intended for use as an allowlist or denylist. In addition, allowlists and denylists can be implemented using other mechanisms besides regular expressions.

Research Gap

Regexp errors are likely a primary factor in many MFVs, especially those that require multiple manipulations to exploit. However, they are rarely diagnosed at this level of detail.

Taxonomy Mappings

Mapped Taxonomy Name	Node ID	Fit	Mapped Node Name
PLOVER			Regular Expression Error

Related Attack Patterns

CAPEC-ID	Attack Pattern Name
CAPEC-15	Command Delimiters
CAPEC-6	Argument Injection
CAPEC-79	Using Slashes in Alternate Encoding

References

[REF-7]

Michael Howard and David LeBlanc. "Writing Secure Code". Chapter 10, "Using Regular Expressions for Checking Input" Page 350. 2nd Edition. Microsoft Press. 2002-12-04.
<https://www.microsoftpressstore.com/store/writing-secure-code-9780735617223>.

Content History

Submissions
Submission Date	Submitter	Organization
2006-07-19 (CWE Draft 3, 2006-07-19)	PLOVER
2006-07-19 (CWE Draft 3, 2006-07-19)
Modifications
Modification Date	Modifier	Organization
2026-04-30 (CWE 4.20, 2026-04-30)	CWE Content Team	MITRE
2026-04-30 (CWE 4.20, 2026-04-30)	updated Potential_Mitigations
2025-12-11 (CWE 4.19, 2025-12-11)	CWE Content Team	MITRE
2025-12-11 (CWE 4.19, 2025-12-11)	updated Weakness_Ordinalities
2023-06-29 (CWE 4.12, 2023-06-29)	CWE Content Team	MITRE
2023-06-29 (CWE 4.12, 2023-06-29)	updated Mapping_Notes
2023-04-27 (CWE 4.11, 2023-04-27)	CWE Content Team	MITRE
2023-04-27 (CWE 4.11, 2023-04-27)	updated Detection_Factors, Relationships
2023-01-31 (CWE 4.10, 2023-01-31)	CWE Content Team	MITRE
2023-01-31 (CWE 4.10, 2023-01-31)	updated Description
2022-10-13 (CWE 4.9, 2022-10-13)	CWE Content Team	MITRE
2022-10-13 (CWE 4.9, 2022-10-13)	updated Demonstrative_Examples, Relationships
2021-03-15 (CWE 4.4, 2021-03-15)	CWE Content Team	MITRE
2021-03-15 (CWE 4.4, 2021-03-15)	updated Relationships
2020-06-25 (CWE 4.1, 2020-06-25)	CWE Content Team	MITRE
2020-06-25 (CWE 4.1, 2020-06-25)	updated Relationship_Notes
2020-02-24 (CWE 4.0, 2020-02-24)	CWE Content Team	MITRE
2020-02-24 (CWE 4.0, 2020-02-24)	updated Relationships, Type
2019-06-20 (CWE 3.3, 2019-06-20)	CWE Content Team	MITRE
2019-06-20 (CWE 3.3, 2019-06-20)	updated Related_Attack_Patterns, Relationships, Type
2018-03-27 (CWE 3.1, 2018-03-27)	CWE Content Team	MITRE
2018-03-27 (CWE 3.1, 2018-03-27)	updated References
2017-11-08 (CWE 3.0, 2017-11-08)	CWE Content Team	MITRE
2017-11-08 (CWE 3.0, 2017-11-08)	updated References
2015-12-07 (CWE 2.9, 2015-12-07)	CWE Content Team	MITRE
2015-12-07 (CWE 2.9, 2015-12-07)	updated Relationships
2014-07-30 (CWE 2.8, 2014-07-31)	CWE Content Team	MITRE
2014-07-30 (CWE 2.8, 2014-07-31)	updated Demonstrative_Examples, Relationships
2014-06-23 (CWE 2.7, 2014-06-23)	CWE Content Team	MITRE
2014-06-23 (CWE 2.7, 2014-06-23)	updated Applicable_Platforms, Common_Consequences, Other_Notes, Relationship_Notes
2012-10-30 (CWE 2.3, 2012-10-30)	CWE Content Team	MITRE
2012-10-30 (CWE 2.3, 2012-10-30)	updated Potential_Mitigations
2012-05-11 (CWE 2.2, 2012-05-15)	CWE Content Team	MITRE
2012-05-11 (CWE 2.2, 2012-05-15)	updated Demonstrative_Examples, Related_Attack_Patterns, Relationships
2011-06-01 (CWE 1.13, 2011-06-01)	CWE Content Team	MITRE
2011-06-01 (CWE 1.13, 2011-06-01)	updated Common_Consequences
2011-03-29 (CWE 1.12, 2011-03-30)	CWE Content Team	MITRE
2011-03-29 (CWE 1.12, 2011-03-30)	updated Observed_Examples
2010-04-05 (CWE 1.8.1, 2010-04-05)	CWE Content Team	MITRE
2010-04-05 (CWE 1.8.1, 2010-04-05)	updated Description
2010-02-16 (CWE 1.8, 2010-02-16)	CWE Content Team	MITRE
2010-02-16 (CWE 1.8, 2010-02-16)	updated References
2009-12-28 (CWE 1.7, 2009-12-28)	CWE Content Team	MITRE
2009-12-28 (CWE 1.7, 2009-12-28)	updated Common_Consequences, Other_Notes
2008-09-08 (CWE 1.0, 2008-09-09)	CWE Content Team	MITRE
2008-09-08 (CWE 1.0, 2008-09-09)	updated Description, Name, Relationships, Observed_Example, Other_Notes, Taxonomy_Mappings
2008-07-01 (CWE 1.0, 2008-09-09)	Eric Dalci	Cigital
2008-07-01 (CWE 1.0, 2008-09-09)	updated Time_of_Introduction
Previous Entry Names
Change Date	Previous Entry Name
2008-09-09	Regular Expression Error


	Site Map \| Terms of Use \| Manage Cookies \| Cookie Notice \| Privacy Policy \| Contact Us \| Use of the Common Weakness Enumeration (CWE™) and the associated references from this website are subject to the Terms of Use. CWE is sponsored by the U.S. Department of Homeland Security (DHS) Cybersecurity and Infrastructure Security Agency (CISA) and managed by the Homeland Security Systems Engineering and Development Institute (HSSEDI) which is operated by The MITRE Corporation (MITRE). Copyright © 2006–2026, The MITRE Corporation. CWE, CWSS, CWRAF, and the CWE logo are trademarks of The MITRE Corporation.

Common Weakness Enumeration

CWE-185: Incorrect Regular Expression

Edit Custom Filter