Grau J, Nettling M, Keilwagen J. DepLogo: visualizing sequence dependencies in R.
Bioinformatics 2019;
35:4812-4814. [PMID:
31225867 DOI:
10.1093/bioinformatics/btz507]
[Citation(s) in RCA: 3] [Impact Index Per Article: 0.6] [Reference Citation Analysis] [Abstract] [Track Full Text] [Journal Information] [Subscribe] [Scholar Register] [Received: 03/23/2019] [Revised: 05/27/2019] [Accepted: 06/13/2019] [Indexed: 11/13/2022] Open
Abstract
SUMMARY
Statistical dependencies are present in a variety of sequence data, but are not discernible from traditional sequence logos. Here, we present the R package DepLogo for visualizing inter-position dependencies in aligned sequence data as dependency logos. Dependency logos make dependency structures, which correspond to regular co-occurrences of symbols at dependent positions, visually perceptible. To this end, sequences are partitioned based on their symbols at highly dependent positions as measured by mutual information, and each partition obtains its own visual representation. We illustrate the utility of the DepLogo package in several use cases generating dependency logos from DNA, RNA and protein sequences.
AVAILABILITY AND IMPLEMENTATION
The DepLogo R package is available from CRAN and its source code is available at https://github.com/Jstacs/DepLogo.
SUPPLEMENTARY INFORMATION
Supplementary data are available at Bioinformatics online.
Collapse