Skip to main content

Small set of my web scraping python tools

Project description

Scraping libary

This is a simple toolbox for scraping info from websites which includes methods for generating headers with a random user agent picked from a list of more than 24k user agents, a method for waiting a random time between 1 and 3 seconds, a progress bar for it to look cool ;) and a couple of methods that will help you retreive email addresses and spanish phone numbers from a given url.

As this package may be useful for a bunch of legitimate and non-legitimate use cases, I'm not responsible of whatever you do with it.

Version 0.2.9 features:

  • Improves the email regex.
  • Adds a method to get random browser user agents

Version 0.2.9 issues:

  • There are no tests...
  • scrap_phones() only works for spanish phone numbers

As this is my first python package, I might not mantain it, so better try finding another one...

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

scrapere-0.2.9.tar.gz (325.0 kB view details)

Uploaded Source

File details

Details for the file scrapere-0.2.9.tar.gz.

File metadata

  • Download URL: scrapere-0.2.9.tar.gz
  • Upload date:
  • Size: 325.0 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/4.0.2 CPython/3.10.6

File hashes

Hashes for scrapere-0.2.9.tar.gz
Algorithm Hash digest
SHA256 7d7c21937d2e1ae96f870f0ab44cc73abaabee5187fab390f0a71d2c3c2babb9
MD5 87709018c51aa716895cd81f590c839f
BLAKE2b-256 ec48689635ed18e3c728212e8f72daaf07ba83bb71f4aae472126492791cd21f

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page