Declarative web parsers
Project description
Soupstars :stew: :star: :boom:
Soupstars makes it easier than ever to build web parsers in Python.
Install it with pip.
pip install soupstars
Let's go!
Quickstart
You need two objects to get started.
>>> from soupstars import Parser, serialize
We'll build a parser to extract data from a github page.
>>> class GithubParser(Parser):
... "Parse data from a github page"
...
... @serialize
... def title(self):
... return str(self.h1.text.strip())
Now all we need is a github web page to parse.
>>> parser = GithubParser("https://github.com/tjwaterman99/soupstars")
Let's see what we've got!
>>> parser.to_dict()
{'title': 'tjwaterman99/soupstars'}
You're now ready to start building your own web parsers with soupstars
. Nice job. :beers:
Going further
- Check out some more examples.
- Review the API documentation.
Contributing
We're thrilled you asked! Just open a PR on github, and we'll take a look.
Project details
Release history Release notifications | RSS feed
Download files
Download the file for your platform. If you're not sure which to choose, learn more about installing packages.
Source Distribution
soupstars-1.2.0.tar.gz
(5.0 kB
view hashes)
Built Distributions
Close
Hashes for soupstars-1.2.0-py3-none-any.whl
Algorithm | Hash digest | |
---|---|---|
SHA256 | e51ffcf72c9b721afce0304c9866eff05d9db554733e4a89225f31de4d0e3d3b |
|
MD5 | 5f6586c4860315306864a70f9ead7229 |
|
BLAKE2b-256 | 0c45162bacb113b24d8dd5af3e57dea59f833e8f4abdfad31700efa80a9c6c33 |
Close
Hashes for soupstars-1.2.0-py2-none-any.whl
Algorithm | Hash digest | |
---|---|---|
SHA256 | 04d077d9d4ae55da24e9f22aa856e1851eae04e9c1405d2c56b48139f6f2eb00 |
|
MD5 | 407ab58aed6a6fa896e554ae7272e0ed |
|
BLAKE2b-256 | 54d6be28c5a9bceeede6ba081b25c4e7fa191f00c1c785e7a58e947eded28376 |