#๐Ÿ”’ Text Extraction from Large amounts of PDF FILES

6 messages ยท Page 1 of 1 (latest)

true ember
#

Do you guys happen to know a better way to extract text from hundreds of PDF files? I've tried PyMuPDF, pdfplumber, and OCR for image-based PDFs, but it took all night just to get through a few pages. Any ideas?

fringe fogBOT
#

@true ember

Python help channel opened

Remember to:

  • Ask your Python question, not if you can ask or if there's an expert who can help.
  • Show a code sample as text (rather than a screenshot) and the error message, if you've got one.
  • Explain what you expect to happen and what actually happens.

:warning: Do not pip install anything that isn't related to your question, especially if asked to over DMs.

urban hare
#

are you doing it synchronous?

fringe fogBOT
#
Python help channel closed

This help channel has been closed and it's no longer possible to send messages here. If your question wasn't answered, feel free to create a new post in #1035199133436354600. To maximize your chances of getting a response, check out this guide on asking good questions.