This repository contains a set of tools written in Python 3 with the aim to extract tabular data from scanned and OCR-processed documents available as PDF files. Before these files can be processed they need to be converted to XML files in pdf2xml format using poppler utils. Further information and examples can be found in the github repository.
Release files for pdftabextract 0.3.0
For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.
Source distribution (sdist)
| File | Size | Uploaded | |
|---|---|---|---|
| pdftabextract-0.3.0.tar.gz | 28.2 kB | Details |
Built distribution (wheel)
| File | Interpreter | ABI | Platform | Reset |
|---|---|---|---|---|
| pdftabextract-0.3.0-py3-none-any.whl | Python 3 | none | any | Details |
Total release size:56.2 kB
Release files / pdftabextract-0.3.0.tar.gz
| Download URL | pdftabextract-0.3.0.tar.gz |
|---|---|
| Size | 28.2 kB |
| Tags | Source |
|
SHA-256 checksum How to use checksums |
822bc899123f360bd83d32f830c7d1fc4db16240f84eedb3009ff12a2d8a97e9
|
|
BLAKE2b-256 checksum How to use checksums |
cbb49c47e9a73262f7155fdc94334d44e3f8a39c54be71ce0e6525feb2494176
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
No |
Release files / pdftabextract-0.3.0-py3-none-any.whl
| Download URL | pdftabextract-0.3.0-py3-none-any.whl |
|---|---|
| Size | 28.0 kB |
| Tags | Python 3 |
|
SHA-256 checksum How to use checksums |
88ec8c4481d4de2bb5f675732e751c10cc31c5908545cbad011f3e0d40654f3c
|
|
BLAKE2b-256 checksum How to use checksums |
1ea9dcf92e41100ba949e33ff7dc47ac8f6e905c5ed1890e6113eb0abd263f40
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
No |