A simple webscraping framework.
Project description
XSCRAPERS
The XSCRAPERS package provides an OOP interface to some simple webscraping techniques.
A base use case can be to load some pages to Beautifulsoup Elements. This package allows to load the URLs concurrently using multiple threads, which allows to safe an enormous amount of time.
import scraper.webscraper as ws
URLS = [
"https://www.google.com/",
"https://www.amazon.com/",
"https://www.youtube.com/",
]
PARSER = "html.parser"
web_scraper = ws.Webscraper(PARSER, verbose=True)
web_scraper.load(URLS)
web_scraper.parse()
Note that herein, the data scraped is stored in the data
attribute of the webscraper.
The URLs parsed are stored in the url
attribute.
Project details
Download files
Download the file for your platform. If you're not sure which to choose, learn more about installing packages.
Source Distribution
xscrapers-0.0.5.tar.gz
(10.9 kB
view hashes)
Built Distribution
xscrapers-0.0.5-py3-none-any.whl
(11.9 kB
view hashes)
Close
Hashes for xscrapers-0.0.5-py3-none-any.whl
Algorithm | Hash digest | |
---|---|---|
SHA256 | d444c39e794b34996ad516a1b46816ce6509f69806043a086c03597d246ac3e9 |
|
MD5 | a9885346ac5d72b71b2067300da7660c |
|
BLAKE2b-256 | d09606c4e2bbb388d9d5d45b25f08ebc8f009a0d78f93f509d308fc45e41d537 |