Sequence Surveyor: Leveraging Overview for Scalable Genomic Alignment Visualization
Overview
Affiliations
In this paper, we introduce overview visualization tools for large-scale multiple genome alignment data. Genome alignment visualization and, more generally, sequence alignment visualization are an important tool for understanding genomic sequence data. As sequencing techniques improve and more data become available, greater demand is being placed on visualization tools to scale to the size of these new datasets. When viewing such large data, we necessarily cannot convey details, rather we specifically design overview tools to help elucidate large-scale patterns. Perceptual science, signal processing theory, and generality provide a framework for the design of such visualizations that can scale well beyond current approaches. We present Sequence Surveyor, a prototype that embodies these ideas for scalable multiple whole-genome alignment overview visualization. Sequence Surveyor visualizes sequences in parallel, displaying data using variable color, position, and aggregation encodings. We demonstrate how perceptual science can inform the design of visualization techniques that remain visually manageable at scale and how signal processing concepts can inform aggregation schemes that highlight global trends, outliers, and overall data distributions as the problem scales. These techniques allow us to visualize alignments with over 100 whole bacterial-sized genomes.
A layout framework for genome-wide multiple sequence alignment graphs.
Schebera J, Zeckzer D, Wiegreffe D Front Bioinform. 2024; 4:1358374.
PMID: 39221004 PMC: 11362851. DOI: 10.3389/fbinf.2024.1358374.
Vis-SPLIT: Interactive Hierarchical Modeling for mRNA Expression Classification.
Roper B, Mathews J, Nadeem S, Park J IEEE Vis Conf. 2024; 2023:106-110.
PMID: 38881685 PMC: 11179685. DOI: 10.1109/vis54172.2023.00030.
Tasks, Techniques, and Tools for Genomic Data Visualization.
Nusrat S, Harbig T, Gehlenborg N Comput Graph Forum. 2019; 38(3):781-805.
PMID: 31768085 PMC: 6876635. DOI: 10.1111/cgf.13727.
Synteny Explorer: An Interactive Visualization Application for Teaching Genome Evolution.
Bryan C, Guterman G, Ma K, Lewin H, Larkin D, Kim J IEEE Trans Vis Comput Graph. 2016; 23(1):711-720.
PMID: 27845661 PMC: 6599602. DOI: 10.1109/TVCG.2016.2598789.
BactoGeNIE: a large-scale comparative genome visualization for big displays.
Aurisano J, Reda K, Johnson A, Marai E, Leigh J BMC Bioinformatics. 2015; 16 Suppl 11:S6.
PMID: 26329021 PMC: 4547189. DOI: 10.1186/1471-2105-16-S11-S6.