Related Experiment Video
Updated: Jan 20, 2026

08:35
Identification of Alternative Splicing and Polyadenylation in RNA-seq Data
Published on: June 24, 2021
6.4K
baerhunter: an R package for the discovery and analysis of expressed non-coding regions in bacterial RNA-seq data
A Ozuna1, D Liberto1, R M Joyce1
1Department of Biological Sciences, Institute of Structural and Molecular Biology, London, WC1E 7HX, UK.
Bioinformatics (Oxford, England)
|August 17, 2019
Summary
Baerhunter is a new R package that identifies functional non-coding RNAs and UTRs in bacterial transcriptomic data. This method improves upon standard pipelines by utilizing comprehensive genome annotations for better RNA-seq analysis.
Area of Science:
- Bioinformatics
- Genomics
- Molecular Biology
Background:
- Standard bacterial transcriptomic analysis often overlooks crucial non-coding RNA elements like small RNAs, long antisense RNAs, and untranslated regions (UTRs).
- This oversight stems from the limitations of incomplete genome annotation files used in conventional bioinformatics pipelines.
Purpose of the Study:
- To introduce baerhunter, an automated method for discovering expressed non-coding RNAs and UTRs from RNA-seq data.
- To provide a tool that integrates the analysis of both coding and non-coding genomic features.
Main Methods:
- Development of baerhunter, a coverage-based method implemented in the R programming language.
- Utilizing RNA-seq reads mapped to a reference genome for identifying non-coding elements and UTRs.
- Integration of the core algorithm into a pipeline for comprehensive downstream analysis.
Main Results:
- Baerhunter automates the discovery of expressed non-coding RNAs and UTRs.
- The method demonstrates favorable performance compared to existing popular alternatives in initial tests with simulated and real data.
- The pipeline facilitates downstream analysis of both coding and non-coding genomic features.
Conclusions:
- Baerhunter offers a simple, extensible, and customizable solution for uncovering functional non-coding elements in bacterial transcriptomes.
- The tool addresses a critical gap in standard bioinformatics pipelines, enabling more complete genomic analysis.
- The R package is publicly available, promoting wider adoption and further development in the field.
Related Concept Videos
Bacterial RNA Polymerase
32.5K
Unlike eukaryotes, bacteria use a single RNA Polymerase (RNAP) to transcribe all genes. The different subunits of bacterial RNAPhave distinct functions. The multisubunit structure of the bacterial RNAP helps the enzyme to maintain catalytic function, facilitate assembly, interact with DNA and RNA, and self-regulate its activity.
In most genes, the transcription site is a single base present upstream of the coding sequence. Though RNAP is a catalytically efficient enzyme, it does not recognize...
In most genes, the transcription site is a single base present upstream of the coding sequence. Though RNAP is a catalytically efficient enzyme, it does not recognize...
32.5K
Bacterial RNA Polymerase
11.6K
11.6K
RNA-seq
11.8K
RNA sequencing, or RNA-Seq, is a high-throughput sequencing technology used to study the transcriptome of a cell. Transcriptomics helps to interpret the functional elements of a genome and identify the molecular constituents of an organism. Additionally, it also helps in understanding the development of an organism and the occurrence of diseases.
Before the discovery of RNA-seq, microarray-based methods and Sanger sequencing were used for transcriptome analysis. However, while...
Before the discovery of RNA-seq, microarray-based methods and Sanger sequencing were used for transcriptome analysis. However, while...
11.8K
lncRNA - Long Non-coding RNAs
9.8K
In humans, more than 80% of the genome gets transcribed. However, only around 2% of the genome codes for proteins. The remaining part produces non-coding RNAs which includes ribosomal RNAs, transfer RNAs, telomerase RNAs, and regulatory RNAs, among other types. A large number of regulatory non-coding RNAs have been classified into two groups depending upon their length – small non-coding RNAs, such as microRNA, which are less than 200 nucleotides in length, and long non-coding RNA...
9.8K
DNA Packaging
112.1K
Overview
112.1K
Chromatin Packaging
18.9K
Each human somatic cell contains 6 billion base pairs of DNA. Each base pair is 0.34 nm long, meaning each diploid cell contains a staggering 2 meters of DNA. This long DNA strand is packed inside a nucleus measuring only 10-20 microns in diameter with the help of specialized DNA-binding proteins called histones. Together they form a compact DNA-protein complex called chromatin. The chromatin is further compacted into higher-order structures. The highest level of compaction is achieved during...
18.9K

