Diversity, duplication, and genomic organization of homeobox genes in Lepidoptera.



Mulhair, Peter O, Crowley, Liam, Boyes, Douglas H, Harper, Amber, Lewis, Owen T, Darwin Tree of Life Consortium, and Holland, Peter WH ORCID: 0000-0003-1533-9376
(2023) Diversity, duplication, and genomic organization of homeobox genes in Lepidoptera. Genome research, 33 (1). pp. 32-44.

[img] PDF
Diversity, duplication, and genomic organization of homeobox genes in Lepidoptera.pdf - Open Access published version

Download (3MB) | Preview

Abstract

Homeobox genes encode transcription factors with essential roles in patterning and cell fate in developing animal embryos. Many homeobox genes, including Hox and NK genes, are arranged in gene clusters, a feature likely related to transcriptional control. Sparse taxon sampling and fragmentary genome assemblies mean that little is known about the dynamics of homeobox gene evolution across Lepidoptera or about how changes in homeobox gene number and organization relate to diversity in this large order of insects. Here we analyze an extensive data set of high-quality genomes to characterize the number and organization of all homeobox genes in 123 species of Lepidoptera from 23 taxonomic families. We find most Lepidoptera have around 100 homeobox loci, including an unusual Hox gene cluster in which the <i>lab</i> gene is repositioned and the <i>ro</i> gene is next to <i>pb</i> A topologically associating domain spans much of the gene cluster, suggesting deep regulatory conservation of the Hox cluster arrangement in this insect order. Most Lepidoptera have four Shx genes, divergent <i>zen</i>-derived loci, but these loci underwent dramatic duplication in several lineages, with some moths having over 165 homeobox loci in the Hox gene cluster; this expansion is associated with local LINE element density. In contrast, the NK gene cluster content is more stable, although there are differences in organization compared with other insects, as well as major rearrangements within butterflies. Our analysis represents the first description of homeobox gene content across the order Lepidoptera, exemplifying the potential of newly generated genome assemblies for understanding genome and gene family evolution.

Item Type: Article
Uncontrolled Keywords: Darwin Tree of Life Consortium, Animals, Butterflies, Genomics, Evolution, Molecular, Phylogeny, Genes, Homeobox, Multigene Family
Divisions: Faculty of Science and Engineering > School of Environmental Sciences
Depositing User: Symplectic Admin
Date Deposited: 10 May 2023 20:01
Last Modified: 10 May 2023 20:02
DOI: 10.1101/gr.277118.122
Related URLs:
URI: https://livrepository.liverpool.ac.uk/id/eprint/3170283