Jove
Visualize
Contact Us
JoVE
x logofacebook logolinkedin logoyoutube logo
ABOUT JoVE
OverviewLeadershipBlogJoVE Help Center
AUTHORS
Publishing ProcessEditorial BoardScope & PoliciesPeer ReviewFAQSubmit
LIBRARIANS
TestimonialsSubscriptionsAccessResourcesLibrary Advisory BoardFAQ
RESEARCH
JoVE JournalMethods CollectionsJoVE Encyclopedia of ExperimentsArchive
EDUCATION
JoVE CoreJoVE BusinessJoVE Science EducationJoVE Lab ManualFaculty Resource CenterFaculty Site
Terms & Conditions of Use
Privacy Policy
Policies

Related Concept Videos

RNA-seq03:21

RNA-seq

RNA sequencing, or RNA-Seq, is a high-throughput sequencing technology used to study the transcriptome of a cell. Transcriptomics helps to interpret the functional elements of a genome and identify the molecular constituents of an organism. Additionally, it also helps in understanding the development of an organism and the occurrence of diseases. 
Before the discovery of RNA-seq, microarray-based methods and Sanger sequencing were used for transcriptome analysis. However, while microarray-based...

You might also read

Related Articles

Articles linked to this work by shared authors, journal, and citation graph.

Sort by
Same author

<i>YoyoMut</i>: Interactive Exploration of SARS-CoV-2 Mutation Fixation and Reversion Through Time.

Life (Basel, Switzerland)·2026
Same author

Lightweight multiscale early warning system for influenza A spillovers.

Science advances·2025
Same author

BioGAN: Enhancing Transcriptomic Data Generation with Biological Knowledge.

Bioengineering (Basel, Switzerland)·2025
Same author

Systematic analysis of SARS-CoV-2 Omicron subvariants' impact on B and T cell epitopes.

PloS one·2024
Same author

Non-Negative Matrix Tri-Factorization for Representation Learning in Multi-Omics Datasets with Applications to Drug Repurposing and Selection.

International journal of molecular sciences·2024
Same author

Data-driven recombination detection in viral genomes.

Nature communications·2024

Related Experiment Video

Updated: Jun 17, 2026

A High Throughput MHC II Binding Assay for Quantitative Analysis of Peptide Epitopes
07:59

A High Throughput MHC II Binding Assay for Quantitative Analysis of Peptide Epitopes

Published on: March 25, 2014

15.0K

High Performance Integration Pipeline for Viral and Epitope Sequences.

Tommaso Alfonsi1, Pietro Pinoli1, Arif Canakoglu1,2

  • 1Dipartimento di Elettronica, Informazione e Bioingegneria, Politecnico di Milano, Via Ponzio 34/5, 20133 Milano, Italy.

Biotech (Basel (Switzerland))
|July 13, 2022
PubMed
Summary

A new automated pipeline integrates global viral sequences (SARS-CoV-2, MERS, Ebola) from multiple databases, enabling better research and tools like ViruSurf. This enhances the study of viral infections and provides access to millions of sequences.

Keywords:
COVID-19SARS-CoV-2bioinformaticsdata integrationmetadata managementsequence analysisviral datasets

More Related Videos

Amplification, Next-generation Sequencing, and Genomic DNA Mapping of Retroviral Integration Sites
09:31

Amplification, Next-generation Sequencing, and Genomic DNA Mapping of Retroviral Integration Sites

Published on: March 22, 2016

17.8K
Bidirectional Retroviral Integration Site PCR Methodology and Quantitative Data Analysis Workflow
12:53

Bidirectional Retroviral Integration Site PCR Methodology and Quantitative Data Analysis Workflow

Published on: June 14, 2017

10.9K

Related Experiment Videos

Last Updated: Jun 17, 2026

A High Throughput MHC II Binding Assay for Quantitative Analysis of Peptide Epitopes
07:59

A High Throughput MHC II Binding Assay for Quantitative Analysis of Peptide Epitopes

Published on: March 25, 2014

15.0K
Amplification, Next-generation Sequencing, and Genomic DNA Mapping of Retroviral Integration Sites
09:31

Amplification, Next-generation Sequencing, and Genomic DNA Mapping of Retroviral Integration Sites

Published on: March 22, 2016

17.8K
Bidirectional Retroviral Integration Site PCR Methodology and Quantitative Data Analysis Workflow
12:53

Bidirectional Retroviral Integration Site PCR Methodology and Quantitative Data Analysis Workflow

Published on: June 14, 2017

10.9K

Area of Science:

  • Bioinformatics
  • Virology
  • Data Science

Background:

  • The rapid spread of COVID-19 led to daily sharing of numerous viral sequences.
  • Lack of standardized data across deposition databases hindered practical, homogeneous exploration of global viral sequences.
  • Existing data repositories lacked comprehensive integration and accessibility for researchers.

Purpose of the Study:

  • To develop an automated data pipeline for collecting, transforming, and integrating viral sequences from major global databases.
  • To create efficient and scalable resources for exploring and analyzing viral sequence data.
  • To support the development of analytical and visualization tools for understanding viral infections.

Main Methods:

  • Developed an automated pipeline to collect and integrate viral sequences from NCBI, COG-UK, GISAID, and NMDC.
  • Refined the pipeline for increased efficiency, scalability, and generality, incorporating epitope data from IEDB.
  • Created data exploration interfaces (VirusViz, EpiSurf) and an integrated database (ViruSurf).

Main Results:

  • Established ViruSurf, one of the largest integrated viral sequence databases.
  • The pipeline currently integrates approximately 9.1 million SARS-CoV-2 sequences (March 2022).
  • Enabled the creation of fundamental research resources and visualization tools like VirusViz and EpiSurf.

Conclusions:

  • The refined automated pipeline provides a scalable and efficient solution for integrating diverse viral sequence data.
  • This work significantly enhances the ability of researchers to study viral mechanisms and evolution.
  • The developed resources and tools are crucial for advancing research in virology and infectious diseases.