Manages data storage for CKAN/DCOR (import, symlink, etc.)
Project description
This plugin manages how data are stored in DCOR. There are two types of files in DCOR:
Resources uploaded by users, imported from figshare, or imported from a data archive
Ancillary files that are generated upon resource creation, such as condensed DC data, preview images (see ckanext-dc_view).
This plugin implements:
Data storage management. All resources uploaded by a user are moved to /data/users-HOSTNAME/USERNAME-ORGNAME/PK/ID/PKGNAME_RESID_RESNAME and symlinks are created in /data/ckan-HOSTNAME/resources/RES/OUR/CEID. CKAN itself will not notice this. The idea is to have a filesystem overview about the datasets of each user.
Import datasets from figshare. Existing datasets from figshare are downloaded to the /data/depots/figshare directory and, upon resource creation, symlinked there from /data/ckan-HOSTNAME/resources/RES/OUR/CEID (Note that this is an exemption of the data storage management described above). When running the following command, the “figshare-import” organization is created and the datasets listed in figshare_dois.txt are added to CKAN:
ckan -c /etc/ckan/default/ckan.ini import-figshare
Populate an internal depot from RT-DC data stored in tar archives. This is part of an effort to have automated imports of RT-DC data from other sources. The idea is to move experimental data to the DCOR server in tar archives and DCOR can then populate the internal depot with it. The location of the internal depot is /data/depots/internal/ and it follows a very specific directory structure 201X/2019-08/20/2019-08-20_1126_c083de* where the path is generated from the acquisition date, time, and part of the hash (c083de) of the original data file. According to this scheme, all files with the same path stem belong to one dataset:
2019-08-20_1126_c083de.sha256sums a file containing SHA256 sums
2019-08-20_1126_c083de_v1.rtdc the actual measurement
2019-08-20_1126_c083de_v1_condensed.rtdc the condensed dataset
2019-08-20_1126_c083de_ad1_m001_bg.png an ancillary image
2019-08-20_1126_c083de_ad2_m002_bg.png another ancillary image
…
ckan -c /etc/ckan/default/ckan.ini depotize-archive
Import datasets from the internal depot. The previous command depotize-archive just populates the depot directory structure. To make the datasets available in CKAN, this step must be performed:
ckan -c /etc/ckan/default/ckan.ini import-internal
Please make sure that the necessary file permissions are given in /data.
Installation
pip install ckanext-dcor_depot
Add this extension to the plugins and defaul_views in ckan.ini:
ckan.plugins = [...] dcor_depot ckan.storage_path=/data/ckan-HOSTNAME ckanext.dcor_depot.depots_path=/data/depots ckanext.dcor_depot.users_depot_name=users-HOSTNAME
This plugin stores resources to /data:
mkdir -p /data/depots/users-$(hostname) chown -R www-data /data/depots/users-$(hostname)
Project details
Release history Release notifications | RSS feed
Download files
Download the file for your platform. If you're not sure which to choose, learn more about installing packages.
Source Distribution
Built Distribution
File details
Details for the file ckanext-dcor_depot-0.8.16.tar.gz
.
File metadata
- Download URL: ckanext-dcor_depot-0.8.16.tar.gz
- Upload date:
- Size: 36.8 kB
- Tags: Source
- Uploaded using Trusted Publishing? No
- Uploaded via: twine/3.3.0 pkginfo/1.7.0 requests/2.25.1 setuptools/49.2.1 requests-toolbelt/0.9.1 tqdm/4.59.0 CPython/3.9.2
File hashes
Algorithm | Hash digest | |
---|---|---|
SHA256 | c521dff7e22021a909d677d068933a6485a70fbbb5480dd2e8f4b3b566567e0b |
|
MD5 | 12dbec9f7b1fd3a2173a5d43537bed81 |
|
BLAKE2b-256 | e28da5a9e36df0d2c48d089e1c78583fdbf26031e789ed8f3ad6e63357db6028 |
File details
Details for the file ckanext_dcor_depot-0.8.16-py3-none-any.whl
.
File metadata
- Download URL: ckanext_dcor_depot-0.8.16-py3-none-any.whl
- Upload date:
- Size: 42.6 kB
- Tags: Python 3
- Uploaded using Trusted Publishing? No
- Uploaded via: twine/3.3.0 pkginfo/1.7.0 requests/2.25.1 setuptools/49.2.1 requests-toolbelt/0.9.1 tqdm/4.59.0 CPython/3.9.2
File hashes
Algorithm | Hash digest | |
---|---|---|
SHA256 | 2529ee7edf475ff7dd8bee5a72e91c4fe6e524fdae2b13f71c93a935e664c24d |
|
MD5 | 069e6e26af8f8749e694690c40ad962e |
|
BLAKE2b-256 | dea782ed7f472b8592ea8d99892d69168cf9dbc8896f5b058c0cfff23b9d4e00 |