ChimeraTE: a pipeline to detect chimeric transcripts derived from genes and transposable elements

Oliveira, Daniel S. [UNESP]; Fablet, Marie; Larue, Anaïs; Vallier, Agnès; Carareto, Claudia M.A. [UNESP]; Rebollo, Rita; Vieira, Cristina

doi:10.1093/nar/gkad671

ChimeraTE: a pipeline to detect chimeric transcripts derived from genes and transposable elements

dc.contributor.author	Oliveira, Daniel S. [UNESP]
dc.contributor.author	Fablet, Marie
dc.contributor.author	Larue, Anaïs
dc.contributor.author	Vallier, Agnès
dc.contributor.author	Carareto, Claudia M.A. [UNESP]
dc.contributor.author	Rebollo, Rita
dc.contributor.author	Vieira, Cristina
dc.contributor.institution	Universidade Estadual Paulista (UNESP)
dc.contributor.institution	UMR5558
dc.contributor.institution	Institut Universitaire de France (IUF)
dc.contributor.institution	UMR 203
dc.date.accessioned	2025-04-29T19:33:49Z
dc.date.issued	2023-10-13
dc.description.abstract	Transposable elements (TEs) produce structural variants and are considered an important source of genetic diversity. Notably, TE-gene fusion transcripts, i.e. chimeric transcripts, have been associated with adaptation in several species. However, the identification of these chimeras remains hindered due to the lack of detection tools at a transcriptome-wide scale, and to the reliance on a reference genome, even though different individuals/cells/strains have different TE insertions. Therefore, we developed ChimeraTE, a pipeline that uses paired-end RNA-seq reads to identify chimeric transcripts through two different modes. Mode 1 is the reference-guided approach that employs canonical genome alignment, and Mode 2 identifies chimeras derived from fixed or insertionally polymorphic TEs without any reference genome. We have validated both modes using RNA-seq data from four Drosophila melanogaster wild-type strains. We found ∼1.12% of all genes generating chimeric transcripts, most of them from TE-exonized sequences. Approximately ∼23% of all detected chimeras were absent from the reference genome, indicating that TEs belonging to chimeric transcripts may be recent, polymorphic insertions. ChimeraTE is the first pipeline able to automatically uncover chimeric transcripts without a reference genome, consisting of two running Modes that can be used as a tool to investigate the contribution of TEs to transcriptome plasticity.	en
dc.description.affiliation	São Paulo State University (Unesp) Institute of Biosciences Humanities and Exact Sciences, SP
dc.description.affiliation	Laboratoire de Biométrie et Biologie Evolutive Université Lyon 1 CNRS UMR5558, Rhone-Alpes
dc.description.affiliation	Institut Universitaire de France (IUF), Île-de-FranceF
dc.description.affiliation	Univ Lyon INRAE INSA-Lyon BF2I UMR 203
dc.description.affiliationUnesp	São Paulo State University (Unesp) Institute of Biosciences Humanities and Exact Sciences, SP
dc.description.sponsorship	Fundação de Amparo à Pesquisa do Estado de São Paulo (FAPESP)
dc.description.sponsorship	Conselho Nacional de Desenvolvimento Científico e Tecnológico (CNPq)
dc.description.sponsorship	Agence Nationale de la Recherche
dc.description.sponsorshipId	FAPESP: 2020/06238-2
dc.description.sponsorshipId	CNPq: 308020/2021-9
dc.description.sponsorshipId	Agence Nationale de la Recherche: ANR-14-CE19-0016-01
dc.format.extent	9764-9784
dc.identifier	http://dx.doi.org/10.1093/nar/gkad671
dc.identifier.citation	Nucleic Acids Research, v. 51, n. 18, p. 9764-9784, 2023.
dc.identifier.doi	10.1093/nar/gkad671
dc.identifier.issn	1362-4962
dc.identifier.issn	0305-1048
dc.identifier.scopus	2-s2.0-85174496662
dc.identifier.uri	https://hdl.handle.net/11449/304073
dc.language.iso	eng
dc.relation.ispartof	Nucleic Acids Research
dc.source	Scopus
dc.title	ChimeraTE: a pipeline to detect chimeric transcripts derived from genes and transposable elements	en
dc.type	Artigo	pt
dspace.entity.type	Publication
unesp.author.orcid	0000-0002-7819-4541 0000-0002-7819-4541[2]
unesp.author.orcid	0000-0002-8138-5082[6]
unesp.author.orcid	0000-0003-3414-3993[7]
unesp.campus	Universidade Estadual Paulista (UNESP), Instituto de Biociências, Letras e Ciências Exatas, São José do Rio Preto	pt

Coleções

São José do Rio Preto - IBILCE - Instituto de Biociências, Letras e Ciências Exatas

ChimeraTE: a pipeline to detect chimeric transcripts derived from genes and transposable elements

Arquivos

Coleções