AI Insight
Researchers generated high-quality genome sequences for seven Indonesian rice cultivars using PacBio HiFi technology, producing consensus genomes of approximately 388-390 Mb with high completeness. The analysis revealed that 99.1% of predicted proteins clustered into shared gene families, with 27,514 core gene groups conserved across all cultivars, demonstrating highly similar gene content. The study successfully identified 278 of 280 target gene sequences associated with important agricultural traits including grain color, nitrogen metabolism, and starch properties, though quality control flagged potential issues in 110 of 269 promoter region comparisons.
Why it matters
This genomic resource provides a standardized framework for studying genetic variation in Indonesian rice varieties, which have been underutilized despite their agricultural value. The identification of candidate genes for grain quality and nutrient metabolism traits could accelerate breeding programs aimed at improving rice crops in Indonesia and similar regions.
Understand the Science
⚠️ Preprint – Noch nicht peer-reviewed
Dieser Artikel wurde noch nicht von unabhängigen Experten begutachtet. Die Ergebnisse sind vorläufig und sollten mit Vorsicht interpretiert werden.
Indonesian rice cultivars represent valuable genetic resources, yet many remain poorly characterized at the genomic level. Here, we generated 95.40 Gb of PacBio HiFi sequence data from seven Indonesian rice cultivars and constructed cultivar-specific consensus genomes using the telomere-to-telomere Nipponbare reference AGIS1.0. Sequencing coverage ranged from 27.92x to 41.58x, and the resulting consensus genomes spanned 387.93-390.54 Mb, with BUSCO completeness of approximately 98.3-98.5%. OrthoFinder assigned 99.1% of predicted proteins to 40,737 orthogroups, including 27,514 core orthogroups represented across all seven cultivars, indicating a highly conserved predicted gene space within the reference-guided framework. Targeted analysis recovered 278 of 280 cultivar-by-locus combinations representing 40 genes or gene family entries associated with grain pigmentation, nitrogen and amino-acid metabolism, and starch properties. Comparative predicted protein analysis prioritized ANS1, SBE2b, SSIIa/ALK, Wx/GBSSI, OsAAP6/qPC1, and SSI as candidates for further investigation. Among 269 completed AGIS1.0-anchored promoter comparisons, 159 passed quality-control criteria, whereas 110 were flagged for gene-model, boundary, synteny, or structural concerns. Notably, these flagged comparisons accounted for more than 90% of the alignment-derived sequence variation, emphasizing the importance of rigorous quality control when interpreting apparent promoter divergence. Collectively, these reference-guided genomic resources provide a standardized framework for investigating sequence variation in Indonesian rice germplasm and prioritize testable coding and regulatory candidates for functional validation and future genomics-assisted crop improvement.