Skip to content

Latest commit

 

History

4 Commits

Folders and files

NameName
Last commit message
Last commit date
 
 
 
 
 
 

Repository files navigation

Python-PDF-Scraper

Previous version

In this file, I used pdfQuery library and with the help of pdf->xml. I get the specific pdf data.

Newer version

This version used PyMuPDF and fitz library to able to extract the hightlighted text from pdf. it will require no xml conversion and is alot faster and fairly more accurate. Before running it, run the command: pip install fitz PyMuPDF

About

No description or website provided.

Topics

Resources

Stars

1 star

Watchers

1 watching

Forks

Releases

Packages

Contributors

Languages