-
Notifications
You must be signed in to change notification settings - Fork 1.6k
All issues
Issue creation is restricted in this repository
Issues
is:issue state:open
is:issue state:open
Search results
BUG: compress_content_streams() makes the file larger than not calling it (the pre-compression stream is kept in the output)
PdfWriterThe PdfWriter component is affectedThe PdfWriter component is affectedStatus: Open.#4085 In py-pdf/pypdf;visitor_text reports Form XObject text twice with unreliable matrices
workflow-text-extractionFrom a users perspective, text extraction is the affected feature/workflowFrom a users perspective, text extraction is the affected feature/workflowStatus: Open.#4079 In py-pdf/pypdf;Move
_font.pytogenericmoduleis-maintenanceAnything that is just internal: Simplifying code, syntax changes, updating docs, speed improvementsAnything that is just internal: Simplifying code, syntax changes, updating docs, speed improvementsStatus: Open.#4068 In py-pdf/pypdf;Synthetic-space detection may insert spaces for small positioning differences
whitespaceWhile doing extract_text, getting the right number of whitespaces (spaces and newlines) is hard.While doing extract_text, getting the right number of whitespaces (spaces and newlines) is hard.workflow-text-extractionFrom a users perspective, text extraction is the affected feature/workflowFrom a users perspective, text extraction is the affected feature/workflowStatus: Open.#4067 In py-pdf/pypdf;/ToUnicode CMap with 2-byte source codes is wrongly applied to simple fonts (1-byte codes)
needs-discussionThe PR/issue needs more discussion before we can continueThe PR/issue needs more discussion before we can continueworkflow-text-extractionFrom a users perspective, text extraction is the affected feature/workflowFrom a users perspective, text extraction is the affected feature/workflowStatus: Open.#4035 In py-pdf/pypdf;DEP: Unify naming in constants
breaking-changeA planned breaking changeA planned breaking changeStatus: Open.#4006 In py-pdf/pypdf;extract_text: space extraction fails due to bad width estimation for composite fonts without the " " (space) glyph
workflow-text-extractionFrom a users perspective, text extraction is the affected feature/workflowFrom a users perspective, text extraction is the affected feature/workflowStatus: Open.#3978 In py-pdf/pypdf;Wrong handling of ligatures
workflow-text-extractionFrom a users perspective, text extraction is the affected feature/workflowFrom a users perspective, text extraction is the affected feature/workflowStatus: Open.#3975 In py-pdf/pypdf;- Status: Open.#3933 In py-pdf/pypdf;
- Status: Open.#3869 In py-pdf/pypdf;
Lazy loading for images
is-featureA feature requestA feature requestworkflow-imagesFrom a users perspective, image handling is the affected feature/workflowFrom a users perspective, image handling is the affected feature/workflowStatus: Open.#3822 In py-pdf/pypdf;Refactoring of text extraction
needs-discussionThe PR/issue needs more discussion before we can continueThe PR/issue needs more discussion before we can continueworkflow-text-extractionFrom a users perspective, text extraction is the affected feature/workflowFrom a users perspective, text extraction is the affected feature/workflowStatus: Open.#3792 In py-pdf/pypdf;