Long identical sequences found in multiple bacterial genomes reveal frequent and widespread exchange of genetic material between distant species
Horizontal transfer of genomic elements is an essential force that shapes microbial genome evolution. Horizontal Gene Transfer (HGT) occurs via various mechanisms and has been studied in detail for a variety of systems. However, a coarse-grained, global picture of HGT in the microbial world is still missing. One reason is the difficulty to process large amounts of genomic microbial data to find and characterise HGT events, especially for highly distant organisms. Here, we exploit the fact that HGT between distant species creates long identical DNA sequences in genomes of distant species, which can be found efficiently using alignment-free methods. We analysed over 90 000 bacterial genomes and thus identified over 100 000 events of HGT. We further developed a mathematical model to analyse the statistical properties of those long exact matches and thus estimate the transfer rate between any pair of taxa. Our results demonstrate that long-distance gene exchange (across phyla) is very frequent, as more than 8% of the bacterial genomes analysed have been involved in at least one such event. Finally, we confirm that the function of the transferred sequences strongly impact the transfer rate, as we observe a 3.5 order of magnitude variation between the most and the least transferred categories. Overall, we provide a unique view of horizontal transfer across the bacterial tree of life, illuminating a fundamental process driving bacterial evolution. ### Competing Interest Statement The authors have declared no competing interest.
- Downloaded 504 times
- Download rankings, all-time:
- Site-wide: 42,830 out of 117,931
- In bioinformatics: 4,659 out of 9,553
- Year to date:
- Site-wide: 13,756 out of 117,931
- Since beginning of last month:
- Site-wide: 36,835 out of 117,931
Downloads over time
Distribution of downloads per paper, site-wide
- 27 Nov 2020: The website and API now include results pulled from medRxiv as well as bioRxiv.
- 18 Dec 2019: We're pleased to announce PanLingua, a new tool that enables you to search for machine-translated bioRxiv preprints using more than 100 different languages.
- 21 May 2019: PLOS Biology has published a community page about Rxivist.org and its design.
- 10 May 2019: The paper analyzing the Rxivist dataset has been published at eLife.
- 1 Mar 2019: We now have summary statistics about bioRxiv downloads and submissions.
- 8 Feb 2019: Data from Altmetric is now available on the Rxivist details page for every preprint. Look for the "donut" under the download metrics.
- 30 Jan 2019: preLights has featured the Rxivist preprint and written about our findings.
- 22 Jan 2019: Nature just published an article about Rxivist and our data.
- 13 Jan 2019: The Rxivist preprint is live!