- Pyarser is a simple, straight forward HTML parser that allows you to easily harvest text
inside an HTML document from a link to that website. Examples:
get_site_HTML(link): returns a string of HTML content from a link
get_site_text(link): returns a string of text from a link. This string has all the HTML tags <> removed, along with there contents.
search_by_phrase(phrase, link): returns the fragments of text from a link that contain the continuous string phrase.
search_for_words(words, link): returns the fragments of text from a link that contain ANY of the strings in words.
word_count(link): counts the number of text words from a link.
get_HTML_tags(link): returns a list of the tags used in an HTML document from a link.
HTML_to_TXT(link, name): writes a TXT file with the text content from a link. All HTML brackets and tags are moved.
Release files for Pyarser 0.1.0
For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.
Source distribution (sdist)
| File | Size | Uploaded | |
|---|---|---|---|
| Pyarser-0.1.0.tar.gz | 3.0 kB | Details |
Release files / Pyarser-0.1.0.tar.gz
| Download URL | Pyarser-0.1.0.tar.gz |
|---|---|
| Size | 3.0 kB |
| Tags | Source |
|
SHA-256 checksum How to use checksums |
f81f712c91afa65776bacc2c3a4bf202b8a2b58b11f62e2baed0d615bdba5852
|
|
BLAKE2b-256 checksum How to use checksums |
121c7d8be22ce437e1a9f7340b6b39aa2d33899ed7290c31bf091b06ac6e14c8
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
No |