Next-generation forward genetic screens: using simulated data to improve the design of mapping-by-sequencing experiments in Arabidopsis

Forward genetic screens have successfully identified many genes and continue to be powerful tools for dissecting biological processes in Arabidopsis and other model species. Next-generation sequencing technologies have revolutionized the time-consuming process of identifying the mutations that cause...

ver descrição completa

Detalhes bibliográficos
Autores: Wilson-Sánchez, David, Lup, Samuel Daniel, Sarmiento, Raquel, Ponce, María Rosa, Micol, José Luis
Formato: artículo
Fecha de publicación:2019
País:España
Recursos:Universidad Miguel Hernández de Elche
Repositorio:REDIUMH. Depósito Digital de la UMH
OAI Identifier:oai:dspace.umh.es:11000/30807
Acesso em linha:https://hdl.handle.net/11000/30807
Access Level:acceso abierto
Palavra-chave:Genética
CDU::5 - Ciencias puras y naturales::57 - Biología
Descrição
Resumo:Forward genetic screens have successfully identified many genes and continue to be powerful tools for dissecting biological processes in Arabidopsis and other model species. Next-generation sequencing technologies have revolutionized the time-consuming process of identifying the mutations that cause a phenotype of interest. However, due to the cost of such mapping-by-sequencing experiments, special attention should be paid to experimental design and technical decisions so that the read data allows to map the desired mutation. Here, we simulated different mapping-by-sequencing scenarios. We first evaluated which short-read technology was best suited for analyzing gene-rich genomic regions in Arabidopsis and determined the minimum sequencing depth required to confidently call single nucleotide variants. We also designed ways to discriminate mutagenesis-induced mutations from background Single Nucleotide Polymorphisms in mutants isolated in Arabidopsis non-reference lines. In addition, we simulated bulked segregant mapping populations for identifying point mutations and monitored how the size of the mapping population and the sequencing depth affect mapping precision. Finally, we provide the computational basis of a protocol that we already used to map T-DNA insertions with paired-end Illumina-like reads, using very low sequencing depths and pooling several mutants together; this approach can also be used with single-end reads as well as to map any other insertional mutagen. All these simulations proved useful for designing experiments that allowed us to map several mutations in Arabidopsis.