Skip to main content

A robust, parallelized Python CLI for annotating three_prime_UTR

Project description

peaks2utr: a robust, parallelized Python CLI for annotating 3' UTR

CI PYPI - Version PYPI - Python Version License: GPL v3

peaks2utr is a Python command-line tool that annotates 3' untranslated regions (UTR) for a given set of aligned sequencing reads in BAM format, and canonical annotation in GFF or GTF format. peaks2utr uses MACS (https://pypi.org/project/MACS2/) to call broad "peaks" of significant read coverage in the BAM file, and uses those peaks that pass a set of criteria as a basis to annotate novel 3' UTRs. This favours BAM files from the likes of 10x Chromium runs, where signal is inherently concentrated at the distal ends of the 3' or 5' UTRs. Reads containing soft-clipped bases and polyA-tails of a given length are detected, and their end bases tallied as "truncation points". When piled up, each co-occurring truncation point is used to determine the precise end base of a given UTR. peaks2utr can be tuned to extend, override or ignore any pre-existing 3' UTR annotations in the input GFF file.

Installation

Install latest release with:

pip install peaks2utr

Alternatively, to install from source:

git clone https://github.com/haessar/peaks2utr.git
cd peaks2utr
python3 -m build
python3 -m pip install dist/*.tar.gz

Dependencies

Installation instructions assume a Debian / Ubuntu system with root privileges. Follow the links for instructions for other systems.

Required

bedtools

apt-get install bedtools

Optional

GenomeTools (for post-processing of output gff3)

apt-get install genometools

Quick start

To check that peaks2utr has installed correctly, simply run the following in your terminal to initiate a short run with default parameters

peaks2utr-demo

This uses a small demo set of input files contained in the repository: Tb927_01_v5.1.gff & Tb927_01_v5.1.slice.bam. When complete, you should see a file Tb927_01_v5.1.new.gff which contains original annotations as well as 3' UTRs with source "peaks2utr".

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

peaks2utr-1.2.3.tar.gz (10.6 MB view details)

Uploaded Source

Built Distribution

peaks2utr-1.2.3-py3-none-any.whl (10.6 MB view details)

Uploaded Python 3

File details

Details for the file peaks2utr-1.2.3.tar.gz.

File metadata

  • Download URL: peaks2utr-1.2.3.tar.gz
  • Upload date:
  • Size: 10.6 MB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/4.0.2 CPython/3.11.7

File hashes

Hashes for peaks2utr-1.2.3.tar.gz
Algorithm Hash digest
SHA256 328e23a64f2304177bfff07e9da59692f9fde2cf6207da425a58c67bc189b588
MD5 e946d134ae3e6f6c78b6c88edee5093a
BLAKE2b-256 55cb0e6f0d359eed0a8231b33f818a67c2586efcceece48c7afff8b63ab21251

See more details on using hashes here.

File details

Details for the file peaks2utr-1.2.3-py3-none-any.whl.

File metadata

  • Download URL: peaks2utr-1.2.3-py3-none-any.whl
  • Upload date:
  • Size: 10.6 MB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/4.0.2 CPython/3.11.7

File hashes

Hashes for peaks2utr-1.2.3-py3-none-any.whl
Algorithm Hash digest
SHA256 6d79adcc3b29845e4f40781d0484f172b60fbcc1a130e30220aabf9dcaaa7858
MD5 ae9d1a1a64ff5f38b0aee7e3348da3a2
BLAKE2b-256 8bdeba7a261c5923cc9143ab69bd07c7de293ff36f6c6835c8079fc023cb4dc1

See more details on using hashes here.

Supported by

AWS AWS Cloud computing and Security Sponsor Datadog Datadog Monitoring Fastly Fastly CDN Google Google Download Analytics Microsoft Microsoft PSF Sponsor Pingdom Pingdom Monitoring Sentry Sentry Error logging StatusPage StatusPage Status page