139 matches found
unredactor
Un-Redactor A PDF editing tool that lets you unredact, put your own information over a redaction box, and extract all text. This tool is for forensics purposes. It does not "recover" data truly destroyed by redaction tools. It can process large data sets and convert all at once to a HTML pages wi...
MANSPIDER
MANSPIDER Crawl SMB shares for juicy information. File content searching + regex is supported! What's New in v2.0 Manspider 2.0 is here! This brings significant improvements: New and improved text extraction powered by Kreuzberg - now supporting PDF, DOCX, XLSX, PPTX, images with OCR, and many...
GHSA-G9CG-PRRW-2R8Q pypdf: Possible large memory usage when parsing font data
Impact An attacker who uses this vulnerability can craft a PDF which leads to large memory consumption. This requires parsing the /Widths entry of a TrueType or Type1 fonts with unusually large values, for example during text extraction. Patches This has been fixed in pypdf==6.18.1. Workarounds I...
pypdf: Possible large memory usage for large /ToUnicode streams (Follow-up 2)
Impact An attacker who uses this vulnerability can craft a PDF which leads to large memory consumption. This requires parsing the /ToUnicode entry of a font with unusually large values, for example during text extraction. Patches This has been fixed in pypdf==6.18.1. Workarounds If you cannot...
Linux Distros Unpatched Vulnerability : CVE-2026-102995
The Linux/Unix host has one or more packages installed that are impacted by a vulnerability without a vendor supplied patch available. - pypdf is a free and open-source pure-python PDF library. Prior to 6.18.1, a crafted PDF can place unusually large source-code or destination-string tokens in a...
CVE-2026-102996
pypdf is a free and open-source pure-python PDF library. Prior to 6.18.1, a crafted PDF can provide a TrueType or Type1 simple font with an unusually large /Widths array, causing pypdf/font.py Font.collectttt1characterwidths to process entries beyond the 256 character codes meaningful for a simpl...
CVE-2026-102995
A flaw was found in pypdf. A remote attacker could exploit this vulnerability by submitting a specially crafted PDF document containing unusually large tokens in font mapping streams. When the library parses this data during operations such as text extraction, it decodes and retains oversized...
CVE-2026-102996
The pypdf pure-python PDF library is vulnerable to excessive memory consumption when parsing font data in versions prior to 6.18.1 . A crafted PDF containing a TrueType or Type1 simple font with an unusually large /Widths array can cause the Font._collect_tt_t1_character_widths function in pypdf/...
CVE-2026-102995 pypdf: Possible large memory usage for large /ToUnicode streams (Follow-up 2)
pypdf is a free and open-source pure-python PDF library. Prior to 6.18.1, a crafted PDF can place unusually large source-code or destination-string tokens in a font /ToUnicode mapping, causing pypdf/cmap.py parsebfchar to decode and retain oversized values during operations such as text extractio...
CVE-2026-102995 pypdf: Possible large memory usage for large /ToUnicode streams (Follow-up 2)
pypdf is a free and open-source pure-python PDF library. Prior to 6.18.1, a crafted PDF can place unusually large source-code or destination-string tokens in a font /ToUnicode mapping, causing pypdf/cmap.py parsebfchar to decode and retain oversized values during operations such as text extractio...
pypdf: Possible long runtimes/large memory usage when extracting XForm objects
ImpactAn attacker who uses this vulnerability can craft a PDF which leads to long runtimes and large memory consumption. This requires extracting the text of a page with lots of XForm objects, where some of them might be re-used. PatchesThis has been fixed in pypdf==6.16.1. WorkaroundsIf you cann...
The vulnerabilities of the functions PageObject._extract_text() and PageObject.extract_xform_text() in the Python library for working with PDF files, PyPDF, allow a hacker to cause a service failure.
The vulnerability of the functions PageObject.extracttext and PageObject.extractxformtext in the Python library for working with PDF files, PyPDF, is related to excessive iteration. Exploiting this vulnerability can allow an attacker to cause a service failure...
Linux Distros Unpatched Vulnerability : CVE-2026-84311
The Linux/Unix host has one or more packages installed that are impacted by a vulnerability without a vendor supplied patch available. - pypdf is a free and open-source pure-python PDF library. Prior to 6.16.1, an attacker can craft a PDF that causes pypdf/page.py PageObject.extracttext and...
pypdf: Possible long runtimes/large memory usage when extracting XForm objects
Impact An attacker who uses this vulnerability can craft a PDF which leads to long runtimes and large memory consumption. This requires extracting the text of a page with lots of XForm objects, where some of them might be re-used. Patches This has been fixed in pypdf==6.16.1. Workarounds If you...
GHSA-763M-79HH-57F2 pypdf: Possible long runtimes/large memory usage when extracting XForm objects
Impact An attacker who uses this vulnerability can craft a PDF which leads to long runtimes and large memory consumption. This requires extracting the text of a page with lots of XForm objects, where some of them might be re-used. Patches This has been fixed in pypdf==6.16.1. Workarounds If you...
Excessive Iteration
Overview pypdf is an A pure-python PDF library capable of splitting, merging, cropping, and transforming PDF files Affected versions of this package are vulnerable to Excessive Iteration in its text extraction of pages that contain XForm objects, which processes those objects without limiting the...
CVE-2026-84311
pypdf is a free and open-source pure-python PDF library. Prior to 6.16.1, an attacker can craft a PDF that causes pypdf/page.py PageObject.extracttext and PageObject.extractxformtext to traverse a directed acyclic graph of reused form XObjects in which each form invokes a child multiple times,...
SUSE CVE-2026-71852
pypdf is a free and open-source pure-python PDF library. Prior to 6.15.0, a crafted PDF can cause long runtimes and large memory consumption when pypdf/font.py function Font.collectcidcharacterwidths expands unusually large CID font /W width ranges or excessive width entries during text extractio...
SUSE CVE-2026-71870
pypdf is a free and open-source pure-python PDF library. Prior to 6.15.0, a crafted PDF can cause large memory consumption when pypdf/cmap.py function parsebfrange parses unusually large source-code or destination-string tokens in a font /ToUnicode CMap during text extraction. This issue is fixed...
pypdf: Possible long runtimes/large memory usage for large CID font width ranges
ImpactAn attacker who uses this vulnerability can craft a PDF which leads to long runtimes and large memory consumption. This requires parsing the font width entries of a font with unusually large values, for example during text extraction. PatchesThis has been fixed in pypdf==6.15.0. Workarounds...