Small set of my web scraping python tools
Project description
Scraping libary
This is a simple toolbox for scraping info from websites which includes methods for generating headers with a random user agent picked from a list of more than 24k user agents, a method for waiting a random time between 1 and 3 seconds, a progress bar for it to look cool ;) and a couple of methods that will help you retreive email addresses and spanish phone numbers from a given url.
As this package may be useful for a bunch of legitimate and non-legitimate use cases, I'm not responsible of whatever you do with it.
Version 0.2.8 features:
- Improves the email regex.
Version 0.2.8 issues:
- There are no tests...
- scrap_phones() only works for spanish phone numbers
As this is my first python package, I might not mantain it, so better try finding another one...
Project details
Download files
Download the file for your platform. If you're not sure which to choose, learn more about installing packages.
Source Distribution
File details
Details for the file scrapere-0.2.8.tar.gz.
File metadata
- Download URL: scrapere-0.2.8.tar.gz
- Upload date:
- Size: 322.7 kB
- Tags: Source
- Uploaded using Trusted Publishing? No
- Uploaded via: twine/4.0.2 CPython/3.10.6
File hashes
| Algorithm | Hash digest | |
|---|---|---|
| SHA256 |
5a66e2cc006c53e75fa37e49cb094cb619d4263647fe719f1d4856c3cac34160
|
|
| MD5 |
294f2a54edd5f9d836fac795939edbbc
|
|
| BLAKE2b-256 |
a9fe466794a591181b8dff8c747c1dc38360a96ec56938827813f1455b4af415
|