Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 69
A hybrid empirical and parametric approach for managing ecosystem complexity: Water quality in Lake Geneva under nonstationary futures.
PMID 35733249 · PMC9245694 · Proceedings of the National Academy of Sciences of the United States of America · 2022 · 7 claims · 6 setups
A hybrid model combining empirical dynamic modeling (EDM/S-map) for biogeochemical sink terms with the equation-based Simstrat hydrodynamic model for physical source terms produces substantially better historical DOB forecasts than conventional parametric models.
-
Has reproduction · 80
Differential analysis of RNA structure probing experiments at nucleotide resolution: uncovering regulatory functions of RNA structure.
PMID 35869080 · PMC9307511 · Nature communications · 2022 · 7 claims · 4 setups
DiffScan is a computational framework combining a Normalization module and a Scan module to identify SVRs at nucleotide resolution from SP data.
-
Full-text index only
Population history and natural selection shape patterns of genetic variation in 132 genes.
PMID 15361935 · PMC515367 · PLoS biology · 2004 · 7 claims · 5 setups
Developed a rigorous computational approach that corrects for multiple hypothesis testing and models population demographic history to test for natural selection
-
Full-text index only
Empirical codon substitution matrix.
PMID 15927081 · PMC1173088 · BMC bioinformatics · 2005 · 8 claims · 5 setups
The authors present the first empirical codon substitution matrix built entirely from alignments of vertebrate coding DNA sequences.
-
Full-text index only
On the analysis of glycomics mass spectrometry data via the regularized area under the ROC curve.
PMID 18076765 · PMC2211327 · BMC bioinformatics · 2007 · 8 claims · 4 setups
The TGDR-AUC algorithm regularizes the empirical AUC by replacing the non-differentiable 0-1 loss with a smooth sigmoid surrogate function and applies constrained threshold gradient descent regularization
-
Full-text index only
Detecting natural selection by empirical comparison to random regions of the genome.
PMID 19783549 · PMC2778377 · Human molecular genetics · 2009 · 8 claims · 5 setups
Comparing candidate loci to empirically matched random genomic regions (ENCODE data) avoids the strong demographic/mutation assumptions required by theoretical neutral models and provides a robust test for selection
-
Has reproduction · 87
Forseti: a mechanistic and predictive model of the splicing status of scRNA-seq reads.
PMID 38940130 · PMC11256924 · Bioinformatics (Oxford, England) · 2024 · 7 claims · 5 setups
Forseti is the first probabilistic model for resolving the splicing status of exonic scRNA-seq reads by scoring putative fragments linking read alignments to proximate priming sites
-
Has reproduction · 93
A comparative study on recombination activity in cattle.
PMID 41942849 · PMC13067647 · Genetics, selection, evolution : GSE · 2026 · 8 claims · 8 setups
Genotype data with high systematic missingness across breeds and arrays can be streamlined and analysed with three complementary recombination-estimation approaches (HMM-based LINKPHASE3, deterministic hsphase, likelihood-based hsrecombi)
-
Full-text index only
A re-annotation pipeline for Illumina BeadArrays: improving the interpretation of gene expression data.
PMID 19923232 · PMC2817484 · Nucleic acids research · 2010 · 8 claims · 7 setups
A Perl-based pipeline that BLASTs/BLATs Illumina probe sequences against genomes and transcript databases (RefSeq, UCSC Known Genes, UniGene/GenBank, Ensembl) can classify probes by quality grade (Perfect/Good/Bad/No match) and is applicable across 8 BeadArray platforms and other array types
-
Full-text index only
A note on generalized Genome Scan Meta-Analysis statistics.
PMID 15717930 · PMC551600 · BMC bioinformatics · 2005 · 7 claims · 3 setups
An Edgeworth series approximation to the null distribution of the weighted GSMA statistic provides a more accurate representation than the normal approximation, especially in the tails
-
Full-text index only
The biological function of some human transcription factor binding motifs varies with position relative to the transcription start site.
PMID 18367472 · PMC2377430 · Nucleic acids research · 2008 · 8 claims · 5 setups
1226 eight-letter DNA words show statistically significant positional preferences relative to the TSS across 7914 human promoter regions
-
Full-text index only
Atlas of nascent RNA transcripts reveals tissue-specific enhancer to gene linkages.
PMID 40281430 · PMC12032694 · BMC genomics · 2025 · 7 claims · 8 setups
A large repository of nascent run-on RNA-seq samples (DBNascent) was assembled and uniformly processed to identify sites of bidirectional transcription genome-wide.
-
Has reproduction · 85
Optimizing open data to support one health: best practices to ensure interoperability of genomic data from bacterial pathogens.
PMID 33103064 · PMC7568946 · One health outlook · 2020 · 8 claims · 3 setups
An open-access pathogen surveillance database (NCBI Pathogen Detection) plus contributor Best Practices enables FAIR, interoperable genomic data across human, animal, food, and environmental sources for One Health surveillance.
-
Has reproduction · 50
The molecular landscape of sepsis severity in infants: enhanced coagulation, innate immunity, and T cell repression.
PMID 38817614 · PMC11137207 · Frontiers in immunology · 2024 · 8 claims · 8 setups
Only two of seven published sepsis gene signatures (derived from adult/pediatric/geriatric cohorts) showed good concordance (>80% accuracy) when applied to infant sepsis, showing limited generalizability of non-infant signatures.
-
Full-text index only
SNP@Evolution: a hierarchical database of positive selection on the human genome.
PMID 19732458 · PMC2755008 · BMC evolutionary biology · 2009 · 7 claims · 6 setups
SNP@Evolution is a hierarchical database integrating HET, FST, and iHS from HapMap Phase II and III to identify genome-wide positive selection signals
-
Full-text index only
The origins of lactase persistence in Europe.
PMID 19714206 · PMC2722739 · PLoS computational biology · 2009 · 8 claims · 5 setups
The −13,910*T allele first underwent selection among dairying farmers around 7,500 years ago in a region between the central Balkans and central Europe, possibly linked to the Linearbandkeramik culture.
-
Has reproduction · 88
AuPairWise: A Method to Estimate RNA-Seq Replicability through Co-expression.
PMID 27082953 · PMC4833304 · PLoS computational biology · 2016 · 7 claims · 6 setups
Sample-sample correlation of transcript abundances is a misleading measure of replicability for assessing differential expression, because it is dominated by gene-specific dynamic ranges rather than condition-dependent variation.
-
Has reproduction · 50
Performance of methods for SARS-CoV-2 variant detection and abundance estimation within mixed population samples.
PMID 36721781 · PMC9884472 · PeerJ · 2023 · 8 claims · 4 setups
Kallisto was the most accurate VCE on simulated data, having the lowest RRMSE, followed by Freyja
-
Full-text index only
HIV-1 gp120 N-linked glycosylation differs between plasma and leukocyte compartments.
PMID 18215327 · PMC2265691 · Virology journal · 2008 · 8 claims · 6 setups
N-linked glycosylation of HIV-1 gp120 differs between plasma and leukocyte compartments
-
Full-text index only
Testing groups of genomic locations for enrichment in disease loci using linkage scan data: a method for hypothesis testing.
PMID 16848972 · PMC3525155 · Human genomics · 2006 · 8 claims · 2 setups
A method testing enrichment of a group of genomic locations for disease loci by comparing the average NPL score of the group to a null distribution from randomly drawn groups of equal size