Skip to main content

Relationships Extraction from NARrative Documents

Project description

Renard

DOI

Renard (Relationship Extraction from NARrative Documents) is a library for creating and using custom character networks extraction pipelines. Renard can extract dynamic as well as static character networks.

The Renard logo

Installation

You can install the latest version using pip:

pip install renard-pipeline

Currently, Renard supports Python>=3.9,<=3.12

Documentation

Documentation, including installation instructions, can be found at https://compnet.github.io/Renard/

If you need local documentation, it can be generated using Sphinx. From the docs directory, make html should create documentation under docs/_build/html.

Tutorial

Renard's central concept is the Pipeline.A Pipeline is a list of PipelineStep that are run sequentially in order to extract a character graph from a document. Here is a simple example:

from renard.pipeline import Pipeline
from renard.pipeline.tokenization import NLTKTokenizer
from renard.pipeline.ner import NLTKNamedEntityRecognizer
from renard.pipeline.character_unification import GraphRulesCharacterUnifier
from renard.pipeline.graph_extraction import CoOccurrencesGraphExtractor

with open("./my_doc.txt") as f:
	text = f.read()

pipeline = Pipeline(
	[
		NLTKTokenizer(),
		NLTKNamedEntityRecognizer(),
		GraphRulesCharacterUnifier(min_appearance=10),
		CoOccurrencesGraphExtractor(co_occurrences_dist=25)
	]
)

out = pipeline(text)

For more information, see renard_tutorial.py, which is a tutorial in the jupytext format. You can open it as a notebook in Jupyter Notebook (or export it as a notebook with jupytext --to ipynb renard-tutorial.py).

Running tests

Renard uses pytest for testing. To launch tests, use the following command :

uv run python -m pytest tests

Expensive tests are disabled by default. These can be run by setting the environment variable RENARD_TEST_ALL to 1.

Contributing

see the "Contributing" section of the documentation.

How to cite

If you use Renard in your research project, please cite it as follows:

@Article{Amalvy2024,
  doi	       = {10.21105/joss.06574},
  year	       = {2024},
  publisher    = {The Open Journal},
  volume       = {9},
  number       = {98},
  pages	       = {6574},
  author       = {Amalvy, A. and Labatut, V. and Dufour, R.},
  title	       = {Renard: A Modular Pipeline for Extracting Character
                  Networks from Narrative Texts},
  journal      = {Journal of Open Source Software},
} 

We would be happy to hear about your usage of Renard, so don't hesitate to reach out!

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

renard_pipeline-0.6.5.tar.gz (61.2 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

renard_pipeline-0.6.5-py3-none-any.whl (65.3 kB view details)

Uploaded Python 3

File details

Details for the file renard_pipeline-0.6.5.tar.gz.

File metadata

  • Download URL: renard_pipeline-0.6.5.tar.gz
  • Upload date:
  • Size: 61.2 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: uv/0.6.17

File hashes

Hashes for renard_pipeline-0.6.5.tar.gz
Algorithm Hash digest
SHA256 ad3bee8ebe17b1149df5b5a30eb5df5ff05185b9ee91fe3dc5b0e49cb6d599a5
MD5 1e84906214ed4cd5f8f86ece1754dd57
BLAKE2b-256 2d959f79e0792062e40d0dcb31c88a9143e66a58e06ee26441bdcecffe0e34ab

See more details on using hashes here.

File details

Details for the file renard_pipeline-0.6.5-py3-none-any.whl.

File metadata

File hashes

Hashes for renard_pipeline-0.6.5-py3-none-any.whl
Algorithm Hash digest
SHA256 a6ddcb5b4473c6ada405dee45f0c0061c3a5242c7f9745ce33001e0ac0c30c6c
MD5 9391fdeebad887ccbbaac83de8c77052
BLAKE2b-256 66a0608d8a630a500506b04e2b39c2c9fc69d2b636d95269e9a61e489ceb9022

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page