nn-dataset

Neural Network Dataset

These details have not been verified by PyPI

Project links

Homepage

License
- OSI Approved :: MIT License
Operating System
- OS Independent
Programming Language
- Python :: 3

Project description

Neural Network Dataset

LEMUR - Learning, Evaluation, and Modeling for Unified Research

The original version of the LEMUR dataset was created by Arash Torabi Goodarzi, Roman Kochnev and Zofia Antonina Bentyn at the Computer Vision Laboratory, University of Würzburg, Germany.

Overview 📖

The primary goal of NN Dataset project is to provide flexibility for dynamically combining various deep learing tasks, datasets, metrics, and neural network models. It is designed to facilitate the verification of neural network performance under various combinations of training hyperparameters and data transformation algorithms, by automatically generating performance statistics. It is primarily developed to support the NN Gen project.

Installation or Update of NN Dataset

Remove old version of the LEMUR Dataset and its database:

source .venv/bin/activate
pip uninstall nn-dataset -y
rm -rf db

Installing the stable version:

source .venv/bin/activate
pip install nn-dataset --upgrade --extra-index-url https://download.pytorch.org/whl/cu124

Installing from GitHub to get the most recent code and statistics updates:

source .venv/bin/activate
pip install git+https://github.com/ABrain-One/nn-dataset --upgrade --force --extra-index-url https://download.pytorch.org/whl/cu124

Adding functionality to export data to Excel files and generate plots for analyzing neural network performance:

source .venv/bin/activate
pip install git+https://github.com/ABrain-One/nn-stat --upgrade --force --extra-index-url https://download.pytorch.org/whl/cu124
pip uninstall nn-dataset -y

and export:

source .venv/bin/activate
python -m ab.plot.export

Environment for NN Dataset Contributors

Pip package manager

Create a virtual environment, activate it, and run the following command to install all the project dependencies:

source .venv/bin/activate
python -m pip install --upgrade pip
pip install -r requirements.txt --extra-index-url https://download.pytorch.org/whl/cu124

Docker

All versions of this project are compatible with AI Linux and can be run inside a Docker image:

docker run -v /a/mm:. abrainone/ai-linux bash -c "PYTHONPATH=/a/mm python -m ab.nn.train"

Usage

Standard use cases:

Add a new neural network model into the ab/nn/nn directory.
Run the automated training process for this model (e.g., a new ComplexNet training pipeline configuration):

source .venv/bin/activate
python -m ab.nn.train -c img-classification_cifar-10_acc_ComplexNet

or for all image segmentation models using a fixed range of training parameters and transformer:

source .venv/bin/activate
python run.py -c img-segmentation -f echo --min_learning_rate 1e-4 -l 1e-2 --min_momentum 0.8 -m 0.99 --min_batch_binary_power 2 -b 6

To reproduce the previous result, set the minimum and maximum to the same desired values:

source .venv/bin/activate
python run.py -c img-classification_cifar-10_acc_AlexNet --min_learning_rate 0.0061 -l 0.0061 --min_momentum 0.7549 -m 0.7549 --min_batch_binary_power 2 -b 2 -f norm_299

To view supported flags:

python run.py -h

Contribution

To contribute a new neural network (NN) model to the NN Dataset, please ensure the following criteria are met:

The code for each model is provided in a respective ".py" file within the /ab/nn/nn directory, and the file is named after the name of the model's structure.
The main class for each model is named Net.
The constructor of the Net class takes the following parameters:
- in_shape (tuple): The shape of the first tensor from the dataset iterator. For images it is structured as (batch, channel, height, width).
- out_shape (tuple): Provided by the dataset loader, it describes the shape of the output tensor. For a classification task, this could be (number of classes,).
- prm (dict): A dictionary of hyperparameters, e.g., {'lr': 0.24, 'momentum': 0.93, 'dropout': 0.51}.
- device (torch.device): PyTorch device used for the model training
All external information required for the correct building and training of the NN model for a specific dataset/transformer, as well as the list of hyperparameters, is extracted from in_shape, out_shape or prm, e.g.:
batch = in_shape[0]
channel_number = in_shape[1]
image_size = in_shape[2]
class_number = out_shape[0]
learning_rate = prm['lr']
momentum = prm['momentum']
dropout = prm['dropout'].
Every model script has function returning set of supported hyperparameters, e.g.:
def supported_hyperparameters(): return {'lr', 'momentum', 'dropout'}
The value of each hyperparameter lies within the range of 0.0 to 1.0.
Every class Net implements two functions:
train_setup(self, prm)
and
learn(self, train_data)
The first function initializes the criteria and optimizer, while the second implements the training pipeline. See a simple implementation in the AlexNet model.
For each pull request involving a new NN model, please generate and submit training statistics for 100 Optuna trials (or at least 3 trials for very large models) in the ab/nn/stat directory. The trials should cover 5 epochs of training. Ensure that this statistics is included along with the model in your pull request. For example, the statistics for the ComplexNet model are stored in files <epoch number>.json inside folder img-classification_cifar-10_acc_ComplexNet, and can be generated by:

python run.py -c img-classification_cifar-10_acc_ComplexNet -t 100 -e 5

See more examples of models in /ab/nn/nn and generated statistics in /ab/nn/stat.

Available Modules

The nn-dataset package includes the following key modules:

Dataset:
- Predefined neural network architectures such as AlexNet, ResNet, VGG, and more.
- Located in ab.nn.nn.
Loaders:
- Data loaders for datasets such as CIFAR-10 and COCO.
- Located in ab.nn.loader.
Metrics:
- Common evaluation metrics like accuracy and IoU.
- Located in ab.nn.metric.
Utilities:
- Helper functions for training and statistical analysis.
- Located in ab.nn.util.

Citation

If you find the LEMUR Neural Network Dataset to be useful for your research, please consider citing:

@misc{ABrain-One.NN-Dataset,
  author       = {Goodarzi, Arash Torabi and Kochnev, Roman and Khalid, Waleed and Qin, Furui and Kathiriya, Yash Kanubhai and Dhameliya, Yashkumar Sanjaybhai and Ignatov, Dmitry and Timofte, Radu},
  title        = {Neural Network Dataset: Towards Seamless AutoML},
  howpublished = {\url{https://github.com/ABrain-One/nn-dataset}},
  year         = {2024},
}

Licenses

This project is distributed under the following licensing terms:

for neural network models adopted from other projects
- Python code under the legacy MIT or BSD 3-Clause license
- models with pretrained weights under the legacy DeepSeek LLM V2 license
all neural network models and their weights not covered by the above licenses, as well as all other files and assets in this project, are subject to the MIT license

The idea of Dr. Dmitry Ignatov

Project details

These details have not been verified by PyPI

Project links

Homepage

License
- OSI Approved :: MIT License
Operating System
- OS Independent
Programming Language
- Python :: 3

Release history Release notifications | RSS feed

2.2.8

Mar 25, 2026

2.2.7

Feb 12, 2026

2.2.6

Jan 11, 2026

2.1.2

Nov 15, 2025

2.1.1

Oct 27, 2025

2.1.0

Oct 14, 2025

2.0.6

Oct 3, 2025

2.0.5

Sep 26, 2025

2.0.4

Aug 26, 2025

2.0.3

Aug 20, 2025

2.0.2

Aug 12, 2025

2.0.0

Aug 11, 2025

1.2.6

Jul 17, 2025

1.2.5

Jun 6, 2025

1.2.4

Jun 6, 2025

1.2.3

May 20, 2025

1.2.2

May 2, 2025

1.2.1

Apr 13, 2025

1.2.0

Mar 31, 2025

1.1.0

Mar 9, 2025

This version

1.0.4

Feb 9, 2025

1.0.3

Feb 8, 2025

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

nn_dataset-1.0.4.tar.gz (1.2 MB view details)

Uploaded Feb 9, 2025 Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

The dropdown lists show the available interpreters, ABIs, and platforms. Enable javascript to be able to filter the list of wheel files.

nn_dataset-1.0.4-py3-none-any.whl (3.0 MB view details)

Uploaded Feb 9, 2025 Python 3

File details

Details for the file nn_dataset-1.0.4.tar.gz.

File metadata

Download URL: nn_dataset-1.0.4.tar.gz
Upload date: Feb 9, 2025
Size: 1.2 MB
Tags: Source
Uploaded using Trusted Publishing? No
Uploaded via: twine/6.1.0 CPython/3.11.9

File hashes

Hashes for nn_dataset-1.0.4.tar.gz
Algorithm	Hash digest
SHA256	`b7154745ee9681e7acc7f9d9ddf078049447d42cd8a3b9f18039ff163139231e`
MD5	`cd8e479b70bb67bdcf7014d3e79d0425`
BLAKE2b-256	`8c75b13a9d0a06c0b68789090a33fa56301fd9e34a77ffe2f2532be3bd1299a7`

See more details on using hashes here.

File details

Details for the file nn_dataset-1.0.4-py3-none-any.whl.

File metadata

Download URL: nn_dataset-1.0.4-py3-none-any.whl
Upload date: Feb 9, 2025
Size: 3.0 MB
Tags: Python 3
Uploaded using Trusted Publishing? No
Uploaded via: twine/6.1.0 CPython/3.11.9

File hashes

Hashes for nn_dataset-1.0.4-py3-none-any.whl
Algorithm	Hash digest
SHA256	`b5d011bd16c6f1407db326998e891956711dfb2f0b8512b92e4af2e8c686caf6`
MD5	`cb05db843feb060198bca40947972ff4`
BLAKE2b-256	`6f4649868a79553f45abb6e10cc9ecb4c00afec5503fad6239c45bcb6cc02474`

See more details on using hashes here.

nn-dataset 1.0.4

Navigation

Verified details

Maintainers

Unverified details

Project links

Meta

Classifiers

Project description

Neural Network Dataset

Overview 📖

Installation or Update of NN Dataset

Environment for NN Dataset Contributors

Pip package manager

Docker

Usage

Contribution

Available Modules

Citation

Licenses

The idea of Dr. Dmitry Ignatov

Project details

Verified details

Maintainers

Unverified details

Project links

Meta

Classifiers

Release history Release notifications | RSS feed

Download files

Source Distribution

Built Distribution

File details

File metadata

File hashes

File details

File metadata

File hashes