Declarative web parsers
Project description
Soupstars :stew: :star: :boom:
Soupstars makes it easier than ever to build web parsers in Python.
Install it with pip.
pip install soupstars
Let's go!
Quickstart
You need two objects to get started.
>>> from soupstars import Parser, serialize
We'll build a parser to extract data from a github page.
>>> class GithubParser(Parser):
... "Parse data from a github page"
...
... @serialize
... def title(self):
... return str(self.h1.text.strip())
Now all we need is a github web page to parse.
>>> parser = GithubParser("https://github.com/tjwaterman99/soupstars")
Let's see what we've got!
>>> parser.to_dict()
{'title': 'tjwaterman99/soupstars'}
You're now ready to start building your own web parsers with soupstars
. Nice job. :beers:
Going further
- Check out some more examples.
- Review the API documentation.
Contributing
We're thrilled you asked! Just open a PR on github, and we'll take a look.
Project details
Release history Release notifications | RSS feed
Download files
Download the file for your platform. If you're not sure which to choose, learn more about installing packages.
Source Distribution
soupstars-1.1.0.tar.gz
(3.9 kB
view hashes)
Built Distributions
Close
Hashes for soupstars-1.1.0-py3-none-any.whl
Algorithm | Hash digest | |
---|---|---|
SHA256 | 5b6215af5c712189731fdda0597ca2bffc7f67af5550dd29b73af073575b987e |
|
MD5 | 9b328aec9eeea802e6f7fda335e35a4c |
|
BLAKE2b-256 | d8bdb30795f8fc3c4c6d7e65eb2b408a2d3b6b8e72652a3820269c0d31349a5c |
Close
Hashes for soupstars-1.1.0-py2-none-any.whl
Algorithm | Hash digest | |
---|---|---|
SHA256 | 093007a480db5310cfc35c555ac41eada543524607348708d92dc04d6d062016 |
|
MD5 | be72cdf5f2abec518a84fea7a98ca7e5 |
|
BLAKE2b-256 | d403597e6d6a9cdcf8064493ff1e639c74f6fe78b53324793f3b210cab90d22b |