scrapelib·PyPI

No project description provided

These details have not been verified by PyPI

Project links

Repository

Development Status
- 6 - Mature
Intended Audience
- Developers
License
- OSI Approved :: BSD License
Natural Language
- English
Operating System
- OS Independent
Programming Language
Topic
- Software Development :: Libraries :: Python Modules

Project description

scrapelib is a library for making requests to less-than-reliable websites.

This repository has moved to Codeberg, GitHub will remain as a read-only mirror.

Source: https://codeberg.org/jpt/scrapelib

Documentation: https://jamesturk.github.io/scrapelib/

Issues: https://codeberg.org/jpt/scrapelib/issues

Features

scrapelib originated as part of the Open States project to scrape the websites of all 50 state legislatures and as a result was therefore designed with features desirable when dealing with sites that have intermittent errors or require rate-limiting.

Advantages of using scrapelib over using requests as-is:

HTTP(S) and FTP requests via an identical API
support for simple caching with pluggable cache backends
highly-configurable request throtting
configurable retries for non-permanent site failures
All of the power of the suberb requests library.

Installation

scrapelib is on PyPI, and can be installed via any standard package management tool.

Example Usage

  import scrapelib
  s = scrapelib.Scraper(requests_per_minute=10)

  # Grab Google front page
  s.get('http://google.com')

  # Will be throttled to 10 HTTP requests per minute
  while True:
      s.get('http://example.com')

Project details

These details have not been verified by PyPI

Project links

Repository

Development Status
- 6 - Mature
Intended Audience
- Developers
License
- OSI Approved :: BSD License
Natural Language
- English
Operating System
- OS Independent
Programming Language
Topic
- Software Development :: Libraries :: Python Modules

Release history Release notifications | RSS feed

This version

2.4.1

Jul 2, 2025

2.4.0

Jun 26, 2025

2.3.0

Dec 15, 2023

2.2.0

May 18, 2023

2.1.0

Nov 7, 2022

2.0.7

Jul 6, 2022

2.0.6

Jun 23, 2021

2.0.5

Jun 15, 2021

2.0.4

Apr 13, 2021

2.0.3

Apr 13, 2021

2.0.2

Apr 9, 2021

2.0.1

Apr 9, 2021

2.0.0

Apr 9, 2021

1.2.0

Nov 13, 2018

1.1.1

Apr 16, 2018

1.1.0

Jun 6, 2017

1.0.2

Apr 16, 2017

1.0.1

Apr 16, 2017

1.0.0

Mar 20, 2015

0.10.1

Jan 22, 2015

0.10.0

Jul 15, 2014

0.9.1

Mar 28, 2014

0.9.0

May 22, 2013

0.8.0

Mar 19, 2013

0.7.4

Dec 21, 2012

0.7.3

Jun 21, 2012

0.7.2

May 9, 2012

0.7.1

Apr 27, 2012

0.7.0

Apr 23, 2012

0.6.2

Apr 20, 2012

0.6.1

Apr 19, 2012

0.6.0

Apr 19, 2012

0.5.8

Apr 18, 2012

0.5.7

Feb 2, 2012

0.5.6

Nov 9, 2011

0.5.5

Sep 27, 2011

0.5.4

Jun 7, 2011

0.5.3

Jun 7, 2011

0.5.2

May 16, 2011

0.5.0

Mar 21, 2011

0.4.3

Feb 11, 2011

0.4.2

Feb 8, 2011

0.4.1

Dec 7, 2010

0.4.0

Nov 8, 2010

0.3.0

Oct 5, 2010

0.2.0

Jul 13, 2010

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

scrapelib-2.4.1.tar.gz (19.9 kB view details)

Uploaded Jul 2, 2025 Source

Built Distribution

scrapelib-2.4.1-py3-none-any.whl (17.4 kB view details)

Uploaded Jul 2, 2025 Python 3

File details

Details for the file scrapelib-2.4.1.tar.gz.

File metadata

Download URL: scrapelib-2.4.1.tar.gz
Upload date: Jul 2, 2025
Size: 19.9 kB
Tags: Source
Uploaded using Trusted Publishing? No
Uploaded via: uv/0.7.13

File hashes

Hashes for scrapelib-2.4.1.tar.gz
Algorithm	Hash digest
SHA256	`48340199e92e860a423aeae09ef03cb9c99b78213eddedc9a31c1545b9bf2b6a`
MD5	`4dc517af6bd6368583f39070b551810d`
BLAKE2b-256	`9e5b4207c24a2daf172cddaa9d994e2ac5397313390be72f1ffac293ac9ad624`

See more details on using hashes here.

File details

Details for the file scrapelib-2.4.1-py3-none-any.whl.

File metadata

Download URL: scrapelib-2.4.1-py3-none-any.whl
Upload date: Jul 2, 2025
Size: 17.4 kB
Tags: Python 3
Uploaded using Trusted Publishing? No
Uploaded via: uv/0.7.13

File hashes

Hashes for scrapelib-2.4.1-py3-none-any.whl
Algorithm	Hash digest
SHA256	`1332f8ab05ab3e23b3db3f62e33ffad68ac3d1f1ac9be9fdc9410d0a30a5bf59`
MD5	`a1b48f60b0741b19c01abeee4f521572`
BLAKE2b-256	`3d481f9e4e877f79d84b19c31b1753b0140d0ece723aa4b0506ec27e15845f8b`

See more details on using hashes here.

scrapelib 2.4.1

Navigation

Verified details

Maintainers

Unverified details

Project links

Meta

Classifiers

Project description

Features

Installation

Example Usage

Project details

Verified details

Maintainers

Unverified details

Project links

Meta

Classifiers

Release history Release notifications | RSS feed

Download files

Source Distribution

Built Distribution

File details

File metadata

File hashes

File details

File metadata

File hashes