A library to parse and scrape text from websites
This library will provide 3 ways to scrape the text from the website:
- The first method is to scrape all text from a single webpage.
- The second method is to scrape text from the whole website. That includes sitemaps too.
- The third method is to scrape text from the specified list. Also you could specify a target element (by CSS selector) to scrape only intended parts of webpage.
Release files for text-thief 0.0.2
For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.
Source distribution (sdist)
| File | Size | Uploaded | |
|---|---|---|---|
| text_thief-0.0.2.tar.gz | 12.5 MB | Details |
Built distribution (wheel)
| File | Interpreter | ABI | Platform | Reset |
|---|---|---|---|---|
| text_thief-0.0.2-py3-none-any.whl | Python 3 | none | any | Details |
Total release size: 12.5 MB
Release files / text_thief-0.0.2.tar.gz
| Download URL | text_thief-0.0.2.tar.gz |
|---|---|
| Size | 12.5 MB |
| Tags | Source |
|
SHA-256 checksum How to use checksums |
8f97326a986cdf5b501f12b053b72d827cbfacf0f2d4b2b39f9dcf527bb9516e
|
|
BLAKE2b-256 checksum How to use checksums |
76f6af80214ebede5a40e0072d3cc1757a563ec652e702a1da2ba567476f93fd
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
No |
| Uploaded via |
twine/6.1.0 CPython/3.12.4
|
Release files / text_thief-0.0.2-py3-none-any.whl
| Download URL | text_thief-0.0.2-py3-none-any.whl |
|---|---|
| Size | 15.7 kB |
| Tags | Python 3 |
|
SHA-256 checksum How to use checksums |
503378ae719e0d1b841eb4b8feb829a993b9c1243d10edfac1d34eb0d6402e1f
|
|
BLAKE2b-256 checksum How to use checksums |
c32e5099ae753ab8666aca5d5343c36bd495933f96d1358b0bdb3005fcad1aa9
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
No |
| Uploaded via |
twine/6.1.0 CPython/3.12.4
|