-
Notifications
You must be signed in to change notification settings - Fork 1
Optional OCR for scanned PDFs #60
Copy link
Copy link
Open
Labels
area:desktopsrc/norefund/desktop — pywebview shell, JS bridgesrc/norefund/desktop — pywebview shell, JS bridgearea:parsingcore/parsing.py — PDF/PPTX/DOCX/TXT extractioncore/parsing.py — PDF/PPTX/DOCX/TXT extractionenhancementNew feature or requestNew feature or request
Description
Activity
Metadata
Metadata
Assignees
Labels
area:desktopsrc/norefund/desktop — pywebview shell, JS bridgesrc/norefund/desktop — pywebview shell, JS bridgearea:parsingcore/parsing.py — PDF/PPTX/DOCX/TXT extractioncore/parsing.py — PDF/PPTX/DOCX/TXT extractionenhancementNew feature or requestNew feature or request
Problem
The follow-on to issue #54 (scanned PDF silent failure). Users with image-only PDFs need a path forward.
Create sub issues if required, cause dumping all file changes and reviewing them at once , not a good practice.
Requirements
Implementation
jobs.py,ProgressReport)Scope
Involved because it needs:
api.pyto call OCR