Skip to main content

DataLad extension package for crawling external web resources into an automated data distribution

Project description

 ____          _           _                 _
|  _ \   __ _ | |_   __ _ | |      __ _   __| |
| | | | / _` || __| / _` || |     / _` | / _` |
| |_| || (_| || |_ | (_| || |___ | (_| || (_| |
|____/  \__,_| \__| \__,_||_____| \__,_| \__,_|
                                   Crawler

Travis tests status codecov.io Documentation License: MIT GitHub release PyPI version fury.io Average time to resolve an issue Percentage of issues still open

This extension enhances DataLad (http://datalad.org) for crawling external web resources into an automated data distribution. Please see the extension documentation for a description on additional commands and functionality.

For general information on how to use or contribute to DataLad (and this extension), please see the DataLad website or the main GitHub project page.

Installation

Before you install this package, please make sure that you install a recent version of git-annex. Afterwards, install the latest version of datalad-crawler from PyPi. It is recommended to use a dedicated virtualenv:

# create and enter a new virtual environment (optional)
virtualenv --system-site-packages --python=python3 ~/env/datalad
. ~/env/datalad/bin/activate

# install from PyPi
pip install datalad_crawler

Support

The documentation of this project is found here: http://docs.datalad.org/projects/crawler

All bugs, concerns and enhancement requests for this software can be submitted here: https://github.com/datalad/datalad-crawler/issues

If you have a problem or would like to ask a question about how to use DataLad, please submit a question to NeuroStars.org with a datalad tag. NeuroStars.org is a platform similar to StackOverflow but dedicated to neuroinformatics.

All previous DataLad questions are available here: http://neurostars.org/tags/datalad/

Acknowledgements

DataLad development is supported by a US-German collaboration in computational neuroscience (CRCNS) project "DataGit: converging catalogues, warehouses, and deployment logistics into a federated 'data distribution'" (Halchenko/Hanke), co-funded by the US National Science Foundation (NSF 1429999) and the German Federal Ministry of Education and Research (BMBF 01GQ1411). Additional support is provided by the German federal state of Saxony-Anhalt and the European Regional Development Fund (ERDF), Project: Center for Behavioral Brain Sciences, Imaging Platform. This work is further facilitated by the ReproNim project (NIH 1P41EB019936-01A1).

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

datalad_crawler-0.9.7.tar.gz (113.5 kB view details)

Uploaded Source

Built Distribution

datalad_crawler-0.9.7-py3-none-any.whl (147.6 kB view details)

Uploaded Python 3

File details

Details for the file datalad_crawler-0.9.7.tar.gz.

File metadata

  • Download URL: datalad_crawler-0.9.7.tar.gz
  • Upload date:
  • Size: 113.5 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/3.8.0 pkginfo/1.8.3 readme-renderer/34.0 requests/2.27.1 requests-toolbelt/0.10.0 urllib3/1.26.12 tqdm/4.64.1 importlib-metadata/4.8.3 keyring/23.4.1 rfc3986/1.5.0 colorama/0.4.5 CPython/3.6.15

File hashes

Hashes for datalad_crawler-0.9.7.tar.gz
Algorithm Hash digest
SHA256 33052d448e3c3b78b4091bdba32e52c29347938a37ffdfcc5616f0ffd06370ec
MD5 9df0789801ca8de39e8897eed03ea2df
BLAKE2b-256 80d9fbd7c44aec26bc2d2799b608b92a96b5ec52649c7fc53348fe4bb3b3f1a6

See more details on using hashes here.

File details

Details for the file datalad_crawler-0.9.7-py3-none-any.whl.

File metadata

  • Download URL: datalad_crawler-0.9.7-py3-none-any.whl
  • Upload date:
  • Size: 147.6 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/3.8.0 pkginfo/1.8.3 readme-renderer/34.0 requests/2.27.1 requests-toolbelt/0.10.0 urllib3/1.26.12 tqdm/4.64.1 importlib-metadata/4.8.3 keyring/23.4.1 rfc3986/1.5.0 colorama/0.4.5 CPython/3.6.15

File hashes

Hashes for datalad_crawler-0.9.7-py3-none-any.whl
Algorithm Hash digest
SHA256 52ef9aadd8f318c0108ec393cb21e416ac88e4033f05956162b764264994662c
MD5 8d0c1a6bc0501499af3b3f1e414c3cea
BLAKE2b-256 40106eda40b1f44398e6ccfccdeb1c4b7b8c4f2608344948936858f74000dd7d

See more details on using hashes here.

Supported by

AWS AWS Cloud computing and Security Sponsor Datadog Datadog Monitoring Fastly Fastly CDN Google Google Download Analytics Microsoft Microsoft PSF Sponsor Pingdom Pingdom Monitoring Sentry Sentry Error logging StatusPage StatusPage Status page