Skip to main page content
U.S. flag

An official website of the United States government

Dot gov

The .gov means it’s official.
Federal government websites often end in .gov or .mil. Before sharing sensitive information, make sure you’re on a federal government site.

Https

The site is secure.
The https:// ensures that you are connecting to the official website and that any information you provide is encrypted and transmitted securely.

Access keys NCBI Homepage MyNCBI Homepage Main Content Main Navigation
. 2017 Mar 15;33(6):926-928.
doi: 10.1093/bioinformatics/btw742.

Training alignment parameters for arbitrary sequencers with LAST-TRAIN

Affiliations

Training alignment parameters for arbitrary sequencers with LAST-TRAIN

Michiaki Hamada et al. Bioinformatics. .

Abstract

Summary: LAST-TRAIN improves sequence alignment accuracy by inferring substitution and gap scores that fit the frequencies of substitutions, insertions, and deletions in a given dataset. We have applied it to mapping DNA reads from IonTorrent and PacBio RS, and we show that it reduces reference bias for Oxford Nanopore reads.

Availability and implementation: the source code is freely available at http://last.cbrc.jp/.

Contact: mhamada@waseda.jp or mcfrith@edu.k.u-tokyo.ac.jp.

Supplementary information: Supplementary data are available at Bioinformatics online.

PubMed Disclaimer

References

    1. Altschul S.F. et al. (1997) Gapped BLAST and PSI-BLAST: a new generation of protein database search programs. Nucleic Acids Res., 25, 3389–3402. - PMC - PubMed
    1. Ammar R. et al. (2015) Long read nanopore sequencing for detection of HLA and CYP2D6 variants and haplotypes. F1000Res, 4, 17.. - PMC - PubMed
    1. Chaisson M.J., Tesler G. (2012) Mapping single molecule sequencing reads using basic local alignment with successive refinement (BLASR): application and theory. BMC Bioinformatics, 13, 238.. - PMC - PubMed
    1. Chiaromonte F. et al. (2002) Scoring pairwise genomic sequence alignments. Pac. Symp. Biocomput., 115–126. pages - PubMed
    1. Durbin R. et al. (1998). Biological Sequence Analysis: Probabilistic Models of Proteins and Nucleic Acids. Cambridge University Press, Cambridge.

Publication types