Background: Heat shock transcription factors (Hsfs), which act as important transcriptional regulatory proteins in eukaryotes, play a central role in controlling the expression of heat-responsive genes. At present, the genomes of Chinese white pear ('Dangshansuli') and five other Rosaceae fruit crops have been fully sequenced. However, information about the Hsfs gene family in these Rosaceae species is limited, and the evolutionary history of the Hsfs gene family also remains unresolved.
Results: In this study, 137 Hsf genes were identified from six Rosaceae species (Pyrus bretschneideri, Malus × domestica, Prunus persica, Fragaria vesca, Prunus mume, and Pyrus communis), 29 of which came from Chinese white pear, designated as PbHsf. Based on the structural characteristics and phylogenetic analysis of these sequences, the Hsf family genes could be classified into three main groups (classes A, B, and C). Segmental and dispersed duplications were the primary forces underlying Hsf gene family expansion in the Rosaceae. Most of the PbHsf duplicated gene pairs were dated back to the recent whole-genome duplication (WGD, 30-45 million years ago (MYA)). Purifying selection also played a critical role in the evolution of Hsf genes. Transcriptome data demonstrated that the expression levels of the PbHsf genes were widely different. Six PbHsf genes were upregulated in fruit under naturally increased temperature.
Conclusion: A comprehensive analysis of Hsf genes was performed in six Rosaceae species, and 137 full length Hsf genes were identified. The results presented here will undoubtedly be useful for better understanding the complexity of the Hsf gene family and will facilitate functional characterization in future studies.