Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Genome-wide scans for loci under selection in humans.
PMID 16004726 · PMC3525256 · Human genomics · 2005 · 8 claims · 4 setups
Natural selection and population demographic history both distort patterns of genetic variation relative to the standard neutral model, so single-locus tests cannot unambiguously distinguish selection from demography.
-
Full-text index only
Empirical codon substitution matrix.
PMID 15927081 · PMC1173088 · BMC bioinformatics · 2005 · 8 claims · 5 setups
The authors present the first empirical codon substitution matrix built entirely from alignments of vertebrate coding DNA sequences.
-
Full-text index only
On the analysis of glycomics mass spectrometry data via the regularized area under the ROC curve.
PMID 18076765 · PMC2211327 · BMC bioinformatics · 2007 · 8 claims · 4 setups
The TGDR-AUC algorithm regularizes the empirical AUC by replacing the non-differentiable 0-1 loss with a smooth sigmoid surrogate function and applies constrained threshold gradient descent regularization
-
Full-text index only
Empirical Bayes analysis of quantitative proteomics experiments.
PMID 19829701 · PMC2759080 · PloS one · 2009 · 8 claims · 4 setups
Developed a new empirical Bayes framework that models log2 SILAC protein ratios and is robust to non-Gaussian tails and data sparsity, unlike Gaussian mixture models or Efron's original spline-based approach
-
Full-text index only
Introduction: the challenge of incidental findings.
PMID 18547190 · PMC2587006 · The Journal of law, medicine & ethics : a journal of the American Society of Law, Medicine & Ethics · 2008 · 8 claims · 5 setups
No consensus exists on how to handle incidental findings in human subjects research
-
Has reproduction · 69
A hybrid empirical and parametric approach for managing ecosystem complexity: Water quality in Lake Geneva under nonstationary futures.
PMID 35733249 · PMC9245694 · Proceedings of the National Academy of Sciences of the United States of America · 2022 · 7 claims · 6 setups
A hybrid model combining empirical dynamic modeling (EDM/S-map) for biogeochemical sink terms with the equation-based Simstrat hydrodynamic model for physical source terms produces substantially better historical DOB forecasts than conventional parametric models.
-
Full-text index only
Detecting natural selection by empirical comparison to random regions of the genome.
PMID 19783549 · PMC2778377 · Human molecular genetics · 2009 · 8 claims · 5 setups
Comparing candidate loci to empirically matched random genomic regions (ENCODE data) avoids the strong demographic/mutation assumptions required by theoretical neutral models and provides a robust test for selection
-
Has reproduction · 80
Differential analysis of RNA structure probing experiments at nucleotide resolution: uncovering regulatory functions of RNA structure.
PMID 35869080 · PMC9307511 · Nature communications · 2022 · 7 claims · 4 setups
DiffScan is a computational framework combining a Normalization module and a Scan module to identify SVRs at nucleotide resolution from SP data.
-
Full-text index only
A note on generalized Genome Scan Meta-Analysis statistics.
PMID 15717930 · PMC551600 · BMC bioinformatics · 2005 · 7 claims · 3 setups
An Edgeworth series approximation to the null distribution of the weighted GSMA statistic provides a more accurate representation than the normal approximation, especially in the tails
-
Has reproduction · 98
Projecting contact matrices in 177 geographical regions: An update and comparison with empirical data for the COVID-19 era.
PMID 34310590 · PMC8354454 · PLoS computational biology · 2021 · 7 claims · 6 setups
Updated synthetic contact matrices were generated for 177 geographical locations covering 97.2% of the world's population (up from 152 locations/95.9% in 2017).
-
Full-text index only
BioDrugScreen: a computational drug design resource for ranking molecules docked to the human proteome.
PMID 19923229 · PMC2808957 · Nucleic acids research · 2010 · 6 claims · 5 setups
BioDrugScreen is a web resource providing pre-docked and pre-scored receptor-ligand complexes for ranking molecules against human proteome targets
-
Full-text index only
Biocomputing enters its adolescence.
PMID 15960815 · PMC1175967 · Genome biology · 2005 · 8 claims · 8 setups
A 'match augmentation' algorithm efficiently matches structural motifs by prioritizing functionally significant residues, enabling function prediction between evolutionarily unrelated proteins
-
Full-text index only
Population history and natural selection shape patterns of genetic variation in 132 genes.
PMID 15361935 · PMC515367 · PLoS biology · 2004 · 7 claims · 5 setups
Developed a rigorous computational approach that corrects for multiple hypothesis testing and models population demographic history to test for natural selection
-
Full-text index only
Assessment of algorithms for high throughput detection of genomic copy number variation in oligonucleotide microarray data.
PMID 17910767 · PMC2148068 · BMC bioinformatics · 2007 · 8 claims · 4 setups
Different CNV analysis software packages produce highly variable numbers and types of candidate CNVs from the same data
-
Has reproduction · 85
Optimizing open data to support one health: best practices to ensure interoperability of genomic data from bacterial pathogens.
PMID 33103064 · PMC7568946 · One health outlook · 2020 · 8 claims · 3 setups
An open-access pathogen surveillance database (NCBI Pathogen Detection) plus contributor Best Practices enables FAIR, interoperable genomic data across human, animal, food, and environmental sources for One Health surveillance.
-
Full-text index only
Discordance of species trees with their most likely gene trees.
PMID 16733550 · PMC1464820 · PLoS genetics · 2006 · 7 claims · 2 setups
For any species tree topology with n ≥ 5 taxa, there exist branch lengths for which the most likely gene tree topology (an 'anomalous gene tree') differs from the species tree topology.
-
Has reproduction · 50
Performance of methods for SARS-CoV-2 variant detection and abundance estimation within mixed population samples.
PMID 36721781 · PMC9884472 · PeerJ · 2023 · 8 claims · 4 setups
Kallisto was the most accurate VCE on simulated data, having the lowest RRMSE, followed by Freyja
-
Full-text index only
Whole genome association mapping by incompatibilities and local perfect phylogenies.
PMID 17042942 · PMC1624851 · BMC bioinformatics · 2006 · 8 claims · 8 setups
Blossoc scores the perfect phylogenetic tree spanning the largest compatible region around each marker as a decision tree for case/control status to detect association
-
Full-text index only
Genome bioinformatic analysis of nonsynonymous SNPs.
PMID 17708757 · PMC1978506 · BMC bioinformatics · 2007 · 8 claims · 8 setups
Structure- and sequence-based prediction tools can generally distinguish disease-causing mutations from neutral ones
-
Full-text index only
HIV-1 gp120 N-linked glycosylation differs between plasma and leukocyte compartments.
PMID 18215327 · PMC2265691 · Virology journal · 2008 · 8 claims · 6 setups
N-linked glycosylation of HIV-1 gp120 differs between plasma and leukocyte compartments