pdfparser
Python binding to libpoppler with focus on text extraction
File Explorer
Download Latest Version (.zip)- python-package.yml
- .gitignore
- __init__.py
- poppler.pyx
- test1.odt
- test1.pdf
- bench.odt
- dump_file.py
- dump_file_gi_poppler.py
- dump_file_pdfminer.py
- memory.py
- run_dumps.sh
- .gitignore
- build_poppler.sh
- README.md
- setup.py
# Installation Guide
1. Get the code
git clone https://github.com/izderadicka/pdfparser
Downloads the entire project code from GitHub to your computer.
cd pdfparser
Moves into the project folder you just downloaded.
2. Official Install Script
Easy RecommendedPrerequisites
- Python 3 Python is required to use pip.
- APT (Debian/Ubuntu κ³μ΄) Built into Debian/Ubuntu-based Linux distributions.
pip install cython
Installs the package published on PyPI directly β no need to clone the source.
sudo apt-get install -y libpoppler-private-dev libpoppler-cpp-dev
Installs directly from the APT package repository (Debian/Ubuntu-based).
pip install git+https://github.com/izderadicka/pdfparser
Installs the package published on PyPI directly β no need to clone the source.
After installing, open a new terminal and run the program's version command (e.g. --version) to confirm it worked.
Pulled directly from this repo's README.
3. Python
EasyPrerequisites
pip install cython
Installs the package published on PyPI directly β no need to clone the source.
pip install git+https://github.com/izderadicka/pdfparser
Installs the package published on PyPI directly β no need to clone the source.
If it runs without errors and prints output in the terminal, it worked.
Pulled directly from this repo's README.
// repository documentation
Was this content helpful?
(0 ratings)
