Related Experiment Video
Updated: Dec 18, 2025

12:23
In Vitro Selection of Aptamers to Differentiate Infectious from Non-Infectious Viruses
Published on: September 7, 2022
2.0K
High-Throughput Identification of Adapters in Single-Read Sequencing Data.
Asan M S H Mohideen1, Steinar D Johansen1, Igor Babiak1
1Genomics Group, Faculty of Biosciences and Aquaculture, Nord University, P.O. Box 1490, 8049 Bodø, Norway.
Biomolecules
|June 12, 2020
Summary
Sequencing data preprocessing is essential for analysis. A new tool, adapt_find, automates adapter sequence identification in raw sequencing datasets, improving efficiency and reliability.
Area of Science:
- Bioinformatics
- Genomics
- Computational Biology
Background:
- Public sequencing datasets are rapidly increasing, requiring efficient data preprocessing.
- Adapter sequence removal is a critical initial step in analyzing raw sequencing data.
- Existing tools for automated adapter detection in single-read protocols have limitations.
Purpose of the Study:
- To develop a tool for automating adapter sequence identification in raw single-read sequencing datasets.
- To provide a robust and reliable method for adapter detection across various sequencing technologies and adapter designs.
- To create a valuable toolset for metadata analysis of multiple sequencing datasets.
Main Methods:
- Development of the adapt_find tool for automated adapter sequence identification.
- Verification of adapt_find using publicly available sequencing datasets.
- Development of associated tools: random_mer for N-base detection and fastqc_parser for FASTQC result consolidation.
Main Results:
- adapt_find automates adapter sequence identification without prior knowledge.
- The tool demonstrates robustness, reliability, and high-throughput capabilities.
- Associated tools enhance metadata analysis by detecting random bases and consolidating FASTQC outputs.
Conclusions:
- adapt_find offers a significant improvement for preprocessing raw sequencing data.
- The toolset streamlines metadata analysis for large-scale sequencing projects.
- This automated approach enhances the efficiency and accuracy of genomic data analysis.
Keywords:
454 pyrosequencingIlluminaIon-TorrentSOLiDadapter oligonucleotidesadapter trimmingrandomized adapterssingle-read sequencingsmall RNA sequencingMore Related Videos
Related Concept Videos
RNA-seq
11.6K
RNA sequencing, or RNA-Seq, is a high-throughput sequencing technology used to study the transcriptome of a cell. Transcriptomics helps to interpret the functional elements of a genome and identify the molecular constituents of an organism. Additionally, it also helps in understanding the development of an organism and the occurrence of diseases.
Before the discovery of RNA-seq, microarray-based methods and Sanger sequencing were used for transcriptome analysis. However, while...
Before the discovery of RNA-seq, microarray-based methods and Sanger sequencing were used for transcriptome analysis. However, while...
11.6K
Next-generation Sequencing
97.2K
The first human genome sequencing project cost $2.7 billion and was declared complete in 2003, after 15 years of international cooperation and collaboration between several research teams and funding agencies. Today, with the advent of next-generation sequencing technologies, the cost and time of sequencing a human genome have dropped over 100 fold.
Next-Generation Sequencing Methods
Although all next-generation methods use different technologies, they all share a set of standard features....
Next-Generation Sequencing Methods
Although all next-generation methods use different technologies, they all share a set of standard features....
97.2K

