bioRxiv Science⌕ Search

Biology subjects

Guo, Y.-L.

Publications and source records attributed to Guo, Y.-L..

2 recordsLinked to original sources

Towards an unbiased characterization of genetic polymorphism

Our view of genetic polymorphism is shaped by methods that provide a limited and reference-biased picture. Long-read sequencing technologies, which are starting to provide nearly complete genome sequences for population samples, should solve the problem--except that characterizing and making sense of non-SNP variation is difficult even with perfect sequence data. Here we analyze 27 genomes of Arabidopsis thaliana in an attempt to address these issues, and illustrate what can be learned by analyzing whole-genome polymorphism data in an unbiased manner. Estimated genome sizes range from 135 to 155 Mb, with differences almost entirely due to centromeric and rDNA repeats that are difficult to assemble. The completely assembled chromosome arms comprise roughly 120 Mb in all accessions, but are full of structural variants, largely due to transposable elements. Even with only 27 accessions, a pan-genome coordinate system that includes the resulting variation ends up being [~] 70% larger than the size of any one genome. Our analysis reveals an incompletely annotated mobile-ome: we not only detect several novel TE families, but also find that existing TE annotation is a poor predictor of elements that have recently been active. In contrast to this, the genic portion, or "gene-ome", is highly conserved. By annotating each genome using accession-specific transcriptome data, we find that 13% of all (non-TE) genes are segregating in our 27 accessions, but most of these are transcriptionally silenced. Finally, we show that with short-read data we previously massively underestimated genetic variation of all kinds, including SNPs--mostly in regions where short reads could not be mapped reliably, but also where reads were mapped incorrectly. We demonstrate that SNP-calling errors can be biased by the choice of reference genome, and that RNA-seq and BS-seq results can be strongly affected by mapping reads only to a reference genome rather than to the genome of the assayed individual. In conclusion, while whole-genome polymorphism data pose tremendous analytical challenges, they also have the potential to revolutionize our understanding of genome evolution.

genomics↗

Chlamydomonas mutant hpm91 lacking PGR5 is a scalable and valuable strain for algal hydrogen (H2) production

Clean and sustainable H2 production is essential toward a carbon-neutral world. H2 generation by Chlamydomonas reinhardtii is an attractive approach for solar-H2 from H2O. However, it is currently not scalable because of lacking ideal strains. Here, we explore hpm91, a previously reported PGR5-deletion mutant with remarkable H2 production, that possesses numerous valuable attributes towards large-scale application and in-depth study issues. We show that hpm91 is at least 100-fold scalable (upto 10 liter) with H2 collection sustained for averagely 26 days and 7287 ml H2/10L-HPBR. Also, hpm91 is robust and active over the period of sulfur-deprived H2 production, most likely due to decreased intracellular ROS relative to wild type. Moreover, quantitative proteomic analysis revealed its features in photosynthetic antenna, primary metabolic pathways and anti-ROS responses. Together with success of new high-H2-production strains derived from hpm91, we highlight that hpm91 is a potent strain toward basic and applied research of algal-H2 photoproduction.

biochemistry↗