OntoLearner

OntoLearner: A Modular Python Library for Ontology Learning with LLMs.

These details have not been verified by PyPI

Project links

Project description

OntoLearner: A Modular Python Library for Ontology Learning with LLMs

OntoLearner is a modular and extensible Python library for ontology learning powered by Large Language Models (LLMs). It provides a unified framework covering the full workflow — from loading and modularizing ontologies to training, predicting, and evaluating learner models across multiple ontology learning tasks.

The framework is built around three core components:

🧩 Ontologizers — load, parse, and modularize ontologies from 150+ ready-to-use sources across 20+ domains.
📋 Learning Tasks — support for Term Typing, Taxonomy Discovery, Non-Taxonomic Relation Extraction, and Text2Onto.
🤖 Learner Models — plug-and-play LLM, Retriever, and RAG-based learners with a consistent fit → predict → evaluate interface.

🧪 Installation

OntoLearner is available on PyPI and can be installed with pip:

pip install ontolearner

Verify the installation:

import ontolearner

print(ontolearner.__version__)

For additional installation options (e.g., from source, with optional dependencies), see the Installation Guide.

🔗 Essential Resources

Resource	Description
📚 Documentation	Full documentation website.
🤗 Datasets on Hugging Face	Curated, machine-readable ontology datasets.
🚀 Quickstart	Get started in minutes.
🕸️ Learning Tasks	Term Typing, Taxonomy Discovery, Relation Extraction, and Text2Onto.
🧠 Learner Models	LLM, Retriever, and RAG-based learner models.
📖 Ontologies Documentation	Browse 150+ benchmark ontologies across 20+ domains.
🧩 Ontologizer Guide	How to modularize and preprocess ontologies.
📊 Metrics Dashboard	Explore benchmark ontology metrics and complexity scores.

✨ Key Features

150+ Ontologizers across 20+ domains (biology, medicine, agriculture, chemistry, law, finance, and more).
Multiple learning tasks: Term Typing, Taxonomy Discovery, Non-Taxonomic Relation Extraction, and Text2Onto.
Three learner paradigms: LLM-based, Retriever-based, and Retrieval-Augmented Generation (RAG).
Hugging Face integration: auto-download ontologies and models directly from the Hub.
Unified API: consistent fit → predict → evaluate interface across all learners.
LearnerPipeline: end-to-end pipeline in a single call.
Extensible: easily plug in custom ontologies, learners, or retrievers.
Text2Onto generation: synthetic document generation now uses a direct transformers backend with ontology-aware context enrichment.

🚀 Quick Tour

Loading an Ontology

Load any of the 150+ built-in ontologies and extract task datasets in just a few lines:

from ontolearner import Wine

# Initialize an ontologizer
ontology = Wine()

# Auto-download from Hugging Face and load
ontology.load()

# Extract learning task datasets
data = ontology.extract()

# Inspect ontology metadata
print(ontology)

Explore 150+ ready-to-use ontologies or learn how to work with ontologizers.

Retriever-Based Learner

Use a dense retriever model to perform non-taxonomic relation extraction:

from ontolearner import AutoRetrieverLearner, AgrO, train_test_split, evaluation_report

# Load and extract ontology data
ontology = AgrO()
ontology.load()
ontological_data = ontology.extract()

# Split into train and test sets
train_data, test_data = train_test_split(ontological_data, test_size=0.2, random_state=42)

# Initialize and load a retriever-based learner
task = 'non-taxonomic-re'
ret_learner = AutoRetrieverLearner(top_k=5)
ret_learner.load(model_id='sentence-transformers/all-MiniLM-L6-v2')

# Fit on training data and predict on test data
ret_learner.fit(train_data, task=task)
predicts = ret_learner.predict(test_data, task=task)

# Evaluate predictions
truth = ret_learner.tasks_ground_truth_former(data=test_data, task=task)
metrics = evaluation_report(y_true=truth, y_pred=predicts, task=task)
print(metrics)

Other available learners:

LearnerPipeline

LearnerPipeline consolidates the entire workflow — initialization, training, prediction, and evaluation — into a single call:

from ontolearner import LearnerPipeline, AgrO, train_test_split

# Load ontology and extract data
ontology = AgrO()
ontology.load()

train_data, test_data = train_test_split(
    ontology.extract(),
    test_size=0.2,
    random_state=42
)

# Initialize the pipeline with a dense retriever
pipeline = LearnerPipeline(
    retriever_id='sentence-transformers/all-MiniLM-L6-v2',
    batch_size=10,
    top_k=5
)

# Run: fit → predict → evaluate
outputs = pipeline(
    train_data=train_data,
    test_data=test_data,
    evaluate=True,
    task='non-taxonomic-re'
)

print("Metrics:", outputs['metrics'])
print("Elapsed time:", outputs['elapsed_time'])

⭐ Contribution

We welcome contributions of all kinds — bug reports, new features, documentation improvements, or new ontologies!

Please review our guidelines before getting started:

CONTRIBUTING.md — contribution guidelines
MAINTENANCE.md — ongoing maintenance notes

For bugs or questions, please open an issue in the GitHub Issue Tracker.

💡 Acknowledgements

If OntoLearner is useful in your research or work, please consider citing one of our publications:

@inproceedings{babaei2023llms4ol,
  title     = {LLMs4OL: Large Language Models for Ontology Learning},
  author    = {Babaei Giglou, Hamed and D'Souza, Jennifer and Auer, S{\"o}ren},
  booktitle = {International Semantic Web Conference},
  pages     = {408--427},
  year      = {2023},
  organization = {Springer}
}

@software{babaei_giglou_2025_15399783,
  author    = {Babaei Giglou, Hamed and D'Souza, Jennifer and Aioanei, Andrei
               and Mihindukulasooriya, Nandana and Auer, Sören},
  title     = {OntoLearner: A Modular Python Library for Ontology Learning with LLMs},
  month     = may,
  year      = 2025,
  publisher = {Zenodo},
  version   = {v1.3.0},
  doi       = {10.5281/zenodo.15399783},
  url       = {https://doi.org/10.5281/zenodo.15399783}
}

This software is archived on Zenodo under and is licensed under .

Project details

These details have not been verified by PyPI

Project links

Release history Release notifications | RSS feed

This version

1.6.0

May 4, 2026

1.5.1

Mar 30, 2026

1.5.0

Feb 5, 2026

1.4.11

Jan 5, 2026

1.4.10

Dec 8, 2025

1.4.9

Dec 8, 2025

1.4.8

Nov 30, 2025

1.4.7

Oct 1, 2025

1.4.6

Sep 22, 2025

1.4.5

Sep 16, 2025

1.4.4

Sep 9, 2025

1.4.3

Sep 7, 2025

1.4.2

Sep 1, 2025

1.4.1

Aug 24, 2025

1.4.0

Aug 22, 2025

1.3.1

Aug 13, 2025

1.3.0

Jul 14, 2025

1.2.1

Jun 20, 2025

1.2.0

Jun 20, 2025

1.1.2

Jun 11, 2025

1.1.1

May 27, 2025

1.1.0

May 21, 2025

1.0.0

May 13, 2025

0.5.0

May 8, 2025

0.4.0

May 2, 2025

0.3.0

Apr 24, 2025

0.2.0

Apr 9, 2025

0.1.0

Mar 17, 2025

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

ontolearner-1.6.0.tar.gz (508.6 kB view details)

Uploaded May 4, 2026 Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

The dropdown lists show the available interpreters, ABIs, and platforms. Enable javascript to be able to filter the list of wheel files.

ontolearner-1.6.0-py3-none-any.whl (237.7 kB view details)

Uploaded May 4, 2026 Python 3

File details

Details for the file ontolearner-1.6.0.tar.gz.

File metadata

Download URL: ontolearner-1.6.0.tar.gz
Upload date: May 4, 2026
Size: 508.6 kB
Tags: Source
Uploaded using Trusted Publishing? No
Uploaded via: poetry/2.4.0 CPython/3.10.20 Linux/6.17.0-1010-azure

File hashes

Hashes for ontolearner-1.6.0.tar.gz
Algorithm	Hash digest
SHA256	`53e73e6498f2134f15c83128d595941948a3e42ed050d82b1d2780e87ba11c02`
MD5	`cc789459f058c33ea093f6322adbefe3`
BLAKE2b-256	`60e4cd72448e412f4f21c9f50dcffa46f73c1ea49dbb1893379abf4f7e998c4b`

See more details on using hashes here.

File details

Details for the file ontolearner-1.6.0-py3-none-any.whl.

File metadata

Download URL: ontolearner-1.6.0-py3-none-any.whl
Upload date: May 4, 2026
Size: 237.7 kB
Tags: Python 3
Uploaded using Trusted Publishing? No
Uploaded via: poetry/2.4.0 CPython/3.10.20 Linux/6.17.0-1010-azure

File hashes

Hashes for ontolearner-1.6.0-py3-none-any.whl
Algorithm	Hash digest
SHA256	`27eccf71cb3bcf99470719fc715329ecac2d4123c8bc407020501678d900c05c`
MD5	`b1138d003faa10b87075e2ba45a2631d`
BLAKE2b-256	`9bb19d77ff482412de0a24f7f5e05ed8a1c3471c0226b9ca826792ce25fb209d`

See more details on using hashes here.

OntoLearner 1.6.0

Navigation

Verified details

Maintainers

Unverified details

Project links

Meta

Classifiers

Project description

OntoLearner: A Modular Python Library for Ontology Learning with LLMs

🧪 Installation

🔗 Essential Resources

✨ Key Features

🚀 Quick Tour

Loading an Ontology

Retriever-Based Learner

LearnerPipeline

⭐ Contribution

💡 Acknowledgements

Project details

Verified details

Maintainers

Unverified details

Project links

Meta

Classifiers

Release history Release notifications | RSS feed

Download files

Source Distribution

Built Distribution

File details

File metadata

File hashes

File details

File metadata

File hashes