Only a matter of time before a paper formalized this exercise:
Automated #scRNAseq cell type annotation with GPT4, evaluated across five datasets, 100s of tissues & cell types, human and mouse.
A🧵below with my thoughts on how such tools will change how #Bioinformatics is done.
Promising approach for simulating realistic #microbiome data! Better recapitulates real data. From my former student @zhaobios at @jhubiostat and folx at @EmoryRollins.
MIDAS: a fast and simple simulator for realistic microbiome data https://t.co/KxZitKrArl
PRECAST - a data integration method for integrating multiple spatial transcriptomics datasets by aligning shared embeddings of biological effects while accounting for batch effects.
https://t.co/ErAW5uu0bs
I learnt a ton about everything related to #scrnaseq from Twitter. However, this site is not made to preserve the pearls of knowledge exchanged among the experts. Thus, I created a blog, Single-Cell Updates (scup), to keep track of the threads!
1st post:
https://t.co/v2qIZp5hhN
Elated to announce that four new highly talented faculty will be joining @Columbia Biostatistics over the next few months. Looking forward to welcoming them soon! @HWenpin@wu_xiao1993@ColumbiaMSPH
another blog post after a long time: marker gene selection using logistic regression and regularization for scRNAseq https://t.co/yzEXZ8qqTT #scRNAseq#regression#tidymodes#rstats
Got an RNA-seq dataset with 50, 100, 200+ samples? Plug it into a differential expression tool and hope for the best? No! You need to consider QC, EDA, and modeling technical variation, or else risk generating spurious results. A thread on papers, methods, and best practices:
Ni Zhao wanted to work toward improving human health and well-being, yet she was also fascinated by the beauty of math. Now, as an assistant professor of biostatistics she participates in frontline research. #STEMwomen#womenshistorymonth https://t.co/Z8kF24iC5G
@albert_kuo This reminds me of the data science capstone project (https://t.co/d30r417YL7 ) that we worked on. We try to use the first 20 games of the season to predict the remaining 62 games for each team. And our prediction accuracy of each game on the test data is around 0.68.
Big news from Pfizer, with apparent high efficacy (>90%) based on 94 confirmed COVID-19 cases at their interim analysis.
A thread on how I interpret this news. Briefly:
"Celebrate, but let the process play out over time as intended."
1/8
https://t.co/aqB7GMTEtR