A simple library to capture websites using playwright
Simple replacement for splash using playwright.
pip install playwrightcapture
A very basic example:
from playwrightcapture import Capture async with Capture() as capture: await capture.prepare_context() entries = await capture.capture_page(url)
Entries is a dictionaries that contains (if all goes well) the HAR, the screenshot, all the cookies of the session, the URL as it is in the browser at the end of the capture, and the full HTML page as rendered.
No blackmagic, it is just a reimplementation of a well known technique as implemented there, and there.
This modules will try to bypass reCAPTCHA protected websites if you install it this way:
pip install playwrightcapture[recaptcha]
This will install
SpeechRecognition. In order to work,
libav, look at the install guide
for more details.
SpeechRecognition uses the Google Speech Recognition API to turn the audio file into text (I hope you appreciate the irony).
Release history Release notifications | RSS feed
Download the file for your platform. If you're not sure which to choose, learn more about installing packages.
Hashes for playwrightcapture-1.20.0-py3-none-any.whl