Genome-Wide Identification of the Soybean UMAMIT Family and Functional Analysis of GmUMAMIT118 in Improving Seed Protein Content
Yongjiang Bi, Yaohui Chen, Meirong Lang, Li Duan, Pei Song, Yudong Yang, Xiangxiang Ye, Yan Liu, Bangjun WangSeed protein content, oil content, and yield are key agronomic traits that determine the economic value of soybean. For decades, soybean has served as a leading source of plant protein for human and animal nutrition due to its high protein concentration. Manipulating amino acid transporters to regulate the direction of nitrogen allocation represents a promising strategy for improving seed protein content. Multiple studies have employed this strategy by targeting amino acid importers. Recently, the Usually Multiple Amino acids Move In and Out Transporter (UMAMIT) family has been characterized as amino acid exporters; nevertheless, their role in regulating the seed protein content of soybean has not yet been investigated. In this study, we identified 120 soybean UMAMIT genes via a genome-wide search and designated them according to chromosomal location. Phylogenetic analysis grouped these genes into 10 clades (A–J). Whole-genome duplication (WGD)/segmental duplication served as the main driver of the GmUMAMIT family expansion, followed by tandem duplication. By integrating transcriptome data with QTL/GWAS loci, we identified twelve candidate genes associated with seed protein content and verified their expression patterns during seed development via qPCR. One candidate gene, GmUMAMIT118, was selected and overexpressed in Arabidopsis thaliana, resulting in transgenic lines with significantly higher seed protein content and yield. Collectively, these results provided a comprehensive overview of the soybean UMAMIT family and offered a preliminary exploration of its role in improving seed protein content.