Diki

@dikip.bsky.social

Graduate student

How to do differential expression with scRNAseq data? State of the art is "pseudo-bulk" analysis with RNA-seq methods like edgeR or DESeq2, where "cell type" is encoded as discrete categories. Biologically, discrete categories are not always the most appropriate concept.(1/3) doi.org/10.1038/s415...

Analysis of multi-condition single-cell data with latent embedding multivariate regression - Nature Genetics

Latent embedding multivariate regression models multi-condition single-cell RNA-seq using a continuous latent space, enabling data integration, per-cell gene expression prediction and clustering-free ...

doi.org

When NOT to use DESeq2 for RNA-seq analysis? DESeq2 is the gold standard for bulk RNA-seq, but it has limitations. If you’re analyzing large datasets, beware of inflated false discovery rates (FDR). 🧵👇

i see a particular grievance regularly on here: "a paper I reviewed came back to me at a diff journal. the authors didn't address my concerns at all." seriously what are you doing accepting to review the same paper? let people breathe & have another chance. let them decide what's worth addressing.

Matt Weiner@mattweiner19.bsky.social · 2w ago

will the next reviewer have the same objection? if you revise to address the objection, will the next reviewer say "I don't see why the paper spends so much time addressing this objection that no one could possibly have"? who knows

Lightweight Text Analytics Workflows with DuckDB In this blog post, Petrica Leuca demonstrates how to use #DuckDB for keyword, full-text, and semantic similarity search with embeddings. Learn how to use DuckDB to efficiently perform advanced text analytics in Python: duckdb.org/2025/06/13/t...

BildBildBild

This looks useful. A typical meta-analysis in ecology mixes experiments with observational studies, mixes coefficients with different controls and therefore different causal meanings. I would not normally say "hey the folks in medicine are doing it right, imitate them" but in this case maybe

Alfredo Sánchez-Tójar@asanchez-tojar.bsky.social · 3w ago

Are our ecological conclusions built on shaky foundations? Our latest paper in @methodsinecoevol.bsky.social highlights that Risk of Bias (RoB) assessment, crucial for ensuring the internal validity of research, is almost never used in #systematicreviews in eco & evo 📉 🔗 doi.org/10.1111/2041...

1. don't let the perfect be the enemy of the good; sometimes, refrain from seeking perfection when good will do fine. 2. Perfection is a moving target. Good is a finish line. Stop aiming for a masterpiece and start aiming for done. Progress only happens when you let "good enough" have its turn.

I think of the recommendations in this paper a bit like the Twelve Commandments of Spreadsheets... follow them and save yourself HOURS at the other end. We assume that students/researchers know this stuff, but no one has explicitly been taught it. Another hidden curriculum beauty...

Bild
Darren Dahly@statsepi.bsky.social · 5mo ago

Every day is a good day for sharing one of the most useful papers about research data ever written. PLEASE get your people to understand and follow this advice. www.tandfonline.com/doi/full/10....

Data Organization in Spreadsheets
Karl W. Broman
& Kara H. Woo
Pages 2-10 | Received 01 Jun 2017, Accepted author version posted online: 29 Sep 2017, Published online: 24 Apr 2018

    1. Introduction
    2. Be Consistent
    3. Choose Good Names for Things
    4. Write Dates as YYYY-MM-DD
    5. No Empty Cells
    6. Put Just One Thing in a Cell
    7. Make it a Rectangle
    8. Create a Data Dictionary
    9. No Calculations in the Raw Data Files
    10. Do Not Use Font Color or Highlighting as Data
    11. Make Backups
    12. Use Data Validation to Avoid Errors
    13. Save the Data in Plain Text Files

ABSTRACT

Spreadsheets are widely used software tools for data entry, storage, analysis, and visualization. Focusing on the data entry and storage aspects, this article offers practical recommendations for organizing spreadsheet data to reduce errors and ease later analyses. The basic principles are: be consistent, write dates like YYYY-MM-DD, do not leave any cells empty, put just one thing in a cell, organize the data as a single rectangle (with subjects as rows and variables as columns, and with a single header row), create a data dictionary, do not include calculations in the raw data files, do not use font color or highlighting as data, choose good names for things, make backups, use data validation to avoid data entry errors, and save the data in plain text files.

Breast milk isn't just nutrition – it delivers live bacterial strains that colonize the infant gut and persist for months. Happy to share our new paper, where we used metagenomics to track bacterial strains between 195 mother-infant pairs over the first 6 months of life: doi.org/10.1038/s414...

Assembly of the infant gut microbiome and resistome are linked to bacterial strains in mother’s milk - Nature Communications

Here, with metagenomic analyses on longitudinal samples collected from 195 mother-infant pairs, the authors show that the breast milk microbiome contributes to infant gut assembly through bacterial st...

doi.org