Skip to main content

DataLad extension package for crawling external web resources into an automated data distribution

Project description

 ____          _           _                 _
|  _ \   __ _ | |_   __ _ | |      __ _   __| |
| | | | / _` || __| / _` || |     / _` | / _` |
| |_| || (_| || |_ | (_| || |___ | (_| || (_| |
|____/  \__,_| \__| \__,_||_____| \__,_| \__,_|
                                   Crawler

Travis tests status codecov.io Documentation License: MIT GitHub release PyPI version fury.io Average time to resolve an issue Percentage of issues still open

This extension enhances DataLad (http://datalad.org) for crawling external web resources into an automated data distribution. Please see the extension documentation for a description on additional commands and functionality.

For general information on how to use or contribute to DataLad (and this extension), please see the DataLad website or the main GitHub project page.

Installation

Before you install this package, please make sure that you install a recent version of git-annex. Afterwards, install the latest version of datalad-crawler from PyPi. It is recommended to use a dedicated virtualenv:

# create and enter a new virtual environment (optional)
virtualenv --system-site-packages --python=python3 ~/env/datalad
. ~/env/datalad/bin/activate

# install from PyPi
pip install datalad_crawler

Support

The documentation of this project is found here: http://docs.datalad.org/projects/crawler

All bugs, concerns and enhancement requests for this software can be submitted here: https://github.com/datalad/datalad-crawler/issues

If you have a problem or would like to ask a question about how to use DataLad, please submit a question to NeuroStars.org with a datalad tag. NeuroStars.org is a platform similar to StackOverflow but dedicated to neuroinformatics.

All previous DataLad questions are available here: http://neurostars.org/tags/datalad/

Acknowledgements

DataLad development is supported by a US-German collaboration in computational neuroscience (CRCNS) project "DataGit: converging catalogues, warehouses, and deployment logistics into a federated 'data distribution'" (Halchenko/Hanke), co-funded by the US National Science Foundation (NSF 1429999) and the German Federal Ministry of Education and Research (BMBF 01GQ1411). Additional support is provided by the German federal state of Saxony-Anhalt and the European Regional Development Fund (ERDF), Project: Center for Behavioral Brain Sciences, Imaging Platform. This work is further facilitated by the ReproNim project (NIH 1P41EB019936-01A1).

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

datalad_crawler-0.9.5.tar.gz (114.4 kB view details)

Uploaded Source

Built Distribution

datalad_crawler-0.9.5-py3-none-any.whl (147.5 kB view details)

Uploaded Python 3

File details

Details for the file datalad_crawler-0.9.5.tar.gz.

File metadata

  • Download URL: datalad_crawler-0.9.5.tar.gz
  • Upload date:
  • Size: 114.4 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/3.8.0 pkginfo/1.8.3 readme-renderer/34.0 requests/2.27.1 requests-toolbelt/0.9.1 urllib3/1.26.11 tqdm/4.64.0 importlib-metadata/4.8.3 keyring/23.4.1 rfc3986/1.5.0 colorama/0.4.5 CPython/3.6.15

File hashes

Hashes for datalad_crawler-0.9.5.tar.gz
Algorithm Hash digest
SHA256 bd017b01f104fde629d3c67c6cd173991af1136e73954c7923491f3354e96282
MD5 eaa3fa161b1f55d2c3ba24a98686c415
BLAKE2b-256 51a81931de5fdc64776c6656f99e8c31de092af390b88a3e4696aa87ae5346a5

See more details on using hashes here.

File details

Details for the file datalad_crawler-0.9.5-py3-none-any.whl.

File metadata

  • Download URL: datalad_crawler-0.9.5-py3-none-any.whl
  • Upload date:
  • Size: 147.5 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/3.8.0 pkginfo/1.8.3 readme-renderer/34.0 requests/2.27.1 requests-toolbelt/0.9.1 urllib3/1.26.11 tqdm/4.64.0 importlib-metadata/4.8.3 keyring/23.4.1 rfc3986/1.5.0 colorama/0.4.5 CPython/3.6.15

File hashes

Hashes for datalad_crawler-0.9.5-py3-none-any.whl
Algorithm Hash digest
SHA256 8a7da57ee75ad37b39d19c5a9f45464cf009614651ffdf3b844532d4bb10e501
MD5 74ce0dc3f3705e59d2de50378f19fb1f
BLAKE2b-256 959a804eb3e365fe304d65c028bf61346e66387f043426233a4c2ad5433ae797

See more details on using hashes here.

Supported by

AWS AWS Cloud computing and Security Sponsor Datadog Datadog Monitoring Fastly Fastly CDN Google Google Download Analytics Microsoft Microsoft PSF Sponsor Pingdom Pingdom Monitoring Sentry Sentry Error logging StatusPage StatusPage Status page