Skip to main content

Scrapers that helps journalists at Kristeligt Dagblad

Project description

scrapers-for-journalists

Scraper(s) to help the journalists retrieve data or monitor sites for potential leads for stories.

Using the scrapers

pip install scrapers_for_journalists==0.1.0

And then import a scraper, e.g. from domstoldk.retrive import DomStolScrape

Every file in utils/can be imported in your scrapers, as it is added as a package in pyproject.toml. For example, you can import the BaseScraper with generic utilities like: from base import BaseScraper.

Description of current scrapers

domstol.dk

This scrapers retrieves information about current court cases ("retslister") in Danish "byretter" (Currently, Højesteret etc. are not included). Civil cases and tvangsauktioner are filtered away. Relevance of the cases are estimated based on keywords and "gerningskoder" (types of crimes) from the Danish Police.

To run it manually, use:

poetry run python domstol-dk/retrieve.py --outfile test.xlsx

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

scrapers_for_journalists-0.1.1.tar.gz (25.7 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

scrapers_for_journalists-0.1.1-py3-none-any.whl (26.0 kB view details)

Uploaded Python 3

File details

Details for the file scrapers_for_journalists-0.1.1.tar.gz.

File metadata

  • Download URL: scrapers_for_journalists-0.1.1.tar.gz
  • Upload date:
  • Size: 25.7 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: poetry/1.8.3 CPython/3.10.12 Linux/6.5.0-1027-oem

File hashes

Hashes for scrapers_for_journalists-0.1.1.tar.gz
Algorithm Hash digest
SHA256 a0a3f88cfdd7ab74e24051f8eac400f37a28e5a06a935b77a3ce554195dd80a8
MD5 c76236fcbb79611aeacaa02fa5fa3d3e
BLAKE2b-256 22c7b3cc49e3d645a6e60be842fd15af5e66c9d2114a2d32f3d722e13b32260d

See more details on using hashes here.

File details

Details for the file scrapers_for_journalists-0.1.1-py3-none-any.whl.

File metadata

File hashes

Hashes for scrapers_for_journalists-0.1.1-py3-none-any.whl
Algorithm Hash digest
SHA256 a7cc6699c922a142aff4802406bdea108d1fb45a2bba211e33df4048ffa6e33a
MD5 b7e0e8800cbc8e1ceeecbbb531d7b9c1
BLAKE2b-256 c41b24c6704114237ffc70d0dd723606cbbaa11789e16632180f69118300ebdb

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page