SVM2 : an improved paired-end-based tool for the detection of small genomic structural variations using high-throughput single-genome resequencing data

Chiara, M.; Pesole, G.; Horner, D.S.

doi:10.1093/nar/gks606

Several bioinformatics methods have been proposed for the detection and characterization of genomic structural variation (SV) from ultra high-throughput genome resequencing data. Recent surveys show that comprehensive detection of SV events of different types between an individual resequenced genome and a reference sequence is best achieved through the combination of methods based on different principles (split mapping, reassembly, read depth, insert size, etc.). The improvement of individual predictors is thus an important objective. In this study, we propose a new method that combines deviations from expected library insert sizes and additional information from local patterns of read mapping and uses supervised learning to predict the position and nature of structural variants. We show that our approach provides greatly increased sensitivity with respect to other tools based on paired end read mapping at no cost in specificity, and it makes reliable predictions of very short insertions and deletions in repetitive and low-complexity genomic contexts that can confound tools based on split mapping of reads.

SVM2 : an improved paired-end-based tool for the detection of small genomic structural variations using high-throughput single-genome resequencing data / M. Chiara, G. Pesole, D.S. Horner. - In: NUCLEIC ACIDS RESEARCH. - ISSN 0305-1048. - 40:18(2012 Oct), pp. e145.1-e145.11. [10.1093/nar/gks606]

SVM2 : an improved paired-end-based tool for the detection of small genomic structural variations using high-throughput single-genome resequencing data

M. Chiara^Primo;G. Pesole^Secondo;D.S. Horner^Ultimo

2012

Abstract

Several bioinformatics methods have been proposed for the detection and characterization of genomic structural variation (SV) from ultra high-throughput genome resequencing data. Recent surveys show that comprehensive detection of SV events of different types between an individual resequenced genome and a reference sequence is best achieved through the combination of methods based on different principles (split mapping, reassembly, read depth, insert size, etc.). The improvement of individual predictors is thus an important objective. In this study, we propose a new method that combines deviations from expected library insert sizes and additional information from local patterns of read mapping and uses supervised learning to predict the position and nature of structural variants. We show that our approach provides greatly increased sensitivity with respect to other tools based on paired end read mapping at no cost in specificity, and it makes reliable predictions of very short insertions and deletions in repetitive and low-complexity genomic contexts that can confound tools based on split mapping of reads.

Scheda breve

Scheda completa

Scheda completa (DC)

	Parole chiave
	
				Next generation sequencing ; genomics ; structural variations
			
	Settori scientifico-disciplinari dell'articolo (sola visualizzazione)
	
				Settore BIO/11 - Biologia Molecolare
			
	Data di pubblicazione
	
				ott-2012
			
	Rivista in ANCE
	
				NUCLEIC ACIDS RESEARCH
			
	DOI
	
				https://dx.doi.org/10.1093/nar/gks606
			
	Tipologia
	
				Article (author)
			
	Appare nelle tipologie:
	
				01 - Articolo su periodico

File in questo prodotto:

File	Dimensione	Formato
Nucl. Acids Res.-2012-Chiara-nar_gks606.pdf accesso solo dalla rete interna Tipologia: Publisher's version/PDF Dimensione 1.06 MB Formato Adobe PDF Visualizza/Apri Richiedi una copia	1.06 MB	Adobe PDF	Visualizza/Apri Richiedi una copia

Pubblicazioni consigliate

I documenti in IRIS sono protetti da copyright e tutti i diritti sono riservati, salvo diversa indicazione.

Utilizza questo identificativo per citare o creare un link a questo documento: https://hdl.handle.net/2434/178494

Citazioni

8

16

12

Nome	Dominio	Durata	Descrizione
s_.*	plu.mx	sessione	recupero grafico citazioni sociali da plumx
A_.*	core.ac.uk	7 giorni	recupero pubblicazioni consigliate per il pannello core-recommander
GS_.*	gstatic.com	richiesta http	visualizza grafico citazioni
CC_.*	creativecommons.org	richiesta http	visualizza licenza bitstream

IRIS Institutional Research Information System - AIR Archivio Istituzionale della Ricerca