Lucene search
+L

19 matches found

PyPA
PyPA
•added 2026/09/10 9:44 a.m.•9 views

NLTK TweetTokenizer vulnerable to denial of service through catastrophic regex backtracking

The URLS regular expression in nltk/tokenize/casual.py, compiled into TweetTokenizer.WORDRE and applied by TweetTokenizer.tokenize, contains a naked-domain branch whose domain-label prefix a-z0-9+?:.\-a-z0-9+ is unbounded. Input consisting of many alternating label separators can be partitioned i...

8.7CVSS5.8AI score0.00742EPSS
SaveExploits0References10Affected Software1
OSV
OSV
•added 2026/09/10 9:44 a.m.•9 views

PYSEC-2026-3870 NLTK TweetTokenizer vulnerable to denial of service through catastrophic regex backtracking

The URLS regular expression in nltk/tokenize/casual.py, compiled into TweetTokenizer.WORDRE and applied by TweetTokenizer.tokenize, contains a naked-domain branch whose domain-label prefix a-z0-9+?:.-a-z0-9+ is unbounded. Input consisting of many alternating label separators can be partitioned in...

8.7CVSS5.8AI score0.00742EPSS
SaveExploits0References10
RedhatCVE
RedhatCVE
•added 2026/08/31 8:33 a.m.•11 views

CVE-2026-72818

A flaw was found in NLTK. The TweetTokenizer component, used for processing social media text, contains a regular expression vulnerability. A remote unauthenticated attacker can provide specially crafted input containing many alternating label separators, causing the regular expression to backtra...

8.7CVSS5.9AI score0.00742EPSS
SaveExploits0References8
Tenable Nessus
Tenable Nessus
•added 2026/08/22 12:00 a.m.•16 views

Linux Distros Unpatched Vulnerability : CVE-2026-72818

The Linux/Unix host has one or more packages installed that are impacted by a vulnerability without a vendor supplied patch available. - The URLS regular expression in nltk/tokenize/casual.py, compiled into TweetTokenizer.WORDRE and applied by TweetTokenizer.tokenize, contains a naked-domain bran...

8.7CVSS5.8AI score0.00742EPSS
SaveExploits0References3
Snyk
Snyk
•added 2026/08/21 5:09 p.m.•11 views

Regular Expression Denial of Service (ReDoS)

Overview nltk is a Natural Language Toolkit NLTK is a Python package for natural language processing. Affected versions of this package are vulnerable to Regular Expression Denial of Service ReDoS via regular expression processing in TweetTokenizer.tokenize and casualtokenize. An attacker can cau...

8.7CVSS5.5AI score0.00742EPSS
SaveExploits0References2
EUVD
EUVD
•added 2026/08/21 12:31 a.m.•19 views

EUVD-2026-63729

The URLS regular expression in nltk/tokenize/casual.py, compiled into TweetTokenizer.WORDRE and applied by TweetTokenizer.tokenize, contains a naked-domain branch whose domain-label prefix a-z0-9+?:.-a-z0-9+ is unbounded. Input consisting of many alternating label separators can be partitioned in...

8.7CVSS5.4AI score0.00742EPSS
SaveExploits0References6
OSV
OSV
•added 2026/08/21 12:31 a.m.•11 views

GHSA-QX2G-XRX7-VFH8 NLTK TweetTokenizer vulnerable to denial of service through catastrophic regex backtracking

The URLS regular expression in nltk/tokenize/casual.py, compiled into TweetTokenizer.WORDRE and applied by TweetTokenizer.tokenize, contains a naked-domain branch whose domain-label prefix a-z0-9+?:.-a-z0-9+ is unbounded. Input consisting of many alternating label separators can be partitioned in...

8.7CVSS5.8AI score0.00742EPSS
SaveExploits0References8
Github Security Blog
Github Security Blog
•added 2026/08/21 12:31 a.m.•14 views

NLTK TweetTokenizer vulnerable to denial of service through catastrophic regex backtracking

The URLS regular expression in nltk/tokenize/casual.py, compiled into TweetTokenizer.WORDRE and applied by TweetTokenizer.tokenize, contains a naked-domain branch whose domain-label prefix a-z0-9+?:.-a-z0-9+ is unbounded. Input consisting of many alternating label separators can be partitioned in...

8.7CVSS5.9AI score0.00742EPSS
SaveExploits0References8Affected Software1
OSV
OSV
•added 2026/08/20 10:18 p.m.•14 views

DEBIAN-CVE-2026-72818

The URLS regular expression in nltk/tokenize/casual.py, compiled into TweetTokenizer.WORDRE and applied by TweetTokenizer.tokenize, contains a naked-domain branch whose domain-label prefix a-z0-9+?:.-a-z0-9+ is unbounded. Input consisting of many alternating label separators can be partitioned in...

8.7CVSS5.5AI score0.00742EPSS
SaveExploits0References1
NVD
NVD
•added 2026/08/20 10:18 p.m.•26 views

CVE-2026-72818

The URLS regular expression in nltk/tokenize/casual.py, compiled into TweetTokenizer.WORDRE and applied by TweetTokenizer.tokenize, contains a naked-domain branch whose domain-label prefix a-z0-9+?:.-a-z0-9+ is unbounded. Input consisting of many alternating label separators can be partitioned in...

8.7CVSS0.00742EPSS
SaveExploits0References5
UbuntuCve
UbuntuCve
•added 2026/08/20 10:18 p.m.•14 views

CVE-2026-72818

The URLS regular expression in nltk/tokenize/casual.py, compiled into TweetTokenizer.WORDRE and applied by TweetTokenizer.tokenize, contains a naked-domain branch whose domain-label prefix a-z0-9+?:.-a-z0-9+ is unbounded. Input consisting of many alternating label separators can be partitioned in...

8.7CVSS5.4AI score0.00742EPSS
SaveExploits0References11
OSV
OSV
•added 2026/08/20 10:18 p.m.•8 views

UBUNTU-CVE-2026-72818

The URLS regular expression in nltk/tokenize/casual.py, compiled into TweetTokenizer.WORDRE and applied by TweetTokenizer.tokenize, contains a naked-domain branch whose domain-label prefix a-z0-9+?:.-a-z0-9+ is unbounded. Input consisting of many alternating label separators can be partitioned in...

8.7CVSS5.4AI score0.00742EPSS
SaveExploits0References12
OSV
OSV
•added 2026/08/20 9:57 p.m.•19 views

CVE-2026-72818 NLTK TweetTokenizer URL Pattern Backtracks Catastrophically on Naked-Domain-Like Input

The URLS regular expression in nltk/tokenize/casual.py, compiled into TweetTokenizer.WORDRE and applied by TweetTokenizer.tokenize, contains a naked-domain branch whose domain-label prefix a-z0-9+?:.-a-z0-9+ is unbounded. Input consisting of many alternating label separators can be partitioned in...

8.7CVSS5.6AI score
SaveExploits0References8
Vulnrichment
Vulnrichment
•added 2026/08/20 9:57 p.m.•20 views

CVE-2026-72818 NLTK TweetTokenizer URL Pattern Backtracks Catastrophically on Naked-Domain-Like Input

The URLS regular expression in nltk/tokenize/casual.py, compiled into TweetTokenizer.WORDRE and applied by TweetTokenizer.tokenize, contains a naked-domain branch whose domain-label prefix a-z0-9+?:.-a-z0-9+ is unbounded. Input consisting of many alternating label separators can be partitioned in...

8.7CVSS5.9AI score0.00742EPSS
SaveExploits0References5
CVE
CVE
•added 2026/08/20 9:57 p.m.•67 views

CVE-2026-72818

NLTK (Python NLP library) is vulnerable to catastrophic regex backtracking in nltk/tokenize/casual.py. The TweetTokenizer.WORD_RE pattern contains an unbounded naked-domain label prefix [a-z0-9]+(?:[.\-][a-z0-9]+)*. Crafted input with many alternating label separators causes the regex engine to e...

8.7CVSS5.9AI score0.00742EPSS
SaveExploits0References5
Debian CVE
Debian CVE
•added 2026/08/20 9:57 p.m.•18 views

CVE-2026-72818

The URLS regular expression in nltk/tokenize/casual.py, compiled into TweetTokenizer.WORDRE and applied by TweetTokenizer.tokenize, contains a naked-domain branch whose domain-label prefix a-z0-9+?:.-a-z0-9+ is unbounded. Input consisting of many alternating label separators can be partitioned in...

8.7CVSS5.9AI score0.00742EPSS
SaveExploits0
Cvelist
Cvelist
•added 2026/08/20 9:57 p.m.•39 views

CVE-2026-72818 NLTK TweetTokenizer URL Pattern Backtracks Catastrophically on Naked-Domain-Like Input

The URLS regular expression in nltk/tokenize/casual.py, compiled into TweetTokenizer.WORDRE and applied by TweetTokenizer.tokenize, contains a naked-domain branch whose domain-label prefix a-z0-9+?:.-a-z0-9+ is unbounded. Input consisting of many alternating label separators can be partitioned in...

8.7CVSS0.00742EPSS
SaveExploits0References5
Positive Technologies
Positive Technologies
•added 2026/08/20 12:00 a.m.•69 views

PT-2026-79260

Name of the Vulnerable Software and Affected Versions NLTK versions prior to 3.10.1 Description The URLS regular expression in nltk/tokenize/casual.py, which is compiled into TweetTokenizer.WORD RE and used by TweetTokenizer.tokenize, contains an unbounded domain-label prefix. When processing inp...

8.7CVSS5.8AI score0.00742EPSS
SaveExploits0References27
Rapid7 Vulnerability Database (full)
Rapid7 Vulnerability Database (full)
•added 2026/08/20 12:00 a.m.•1 views

CVE-2026-72818: Inefficient Regular Expression Complexity

The URLS regular expression in nltk/tokenize/casual.py, compiled into TweetTokenizer.WORDRE and applied by TweetTokenizer.tokenize, contains a naked-domain branch whose domain-label prefix a-z0-9+?:.-a-z0-9+ is unbounded. Input consisting of many alternating label separators can be partitioned in...

8.7CVSS5.8AI score0.00742EPSS
SaveExploits0References2
Rows per page
Query Builder