Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Computational disease gene identification: a concert of methods prioritizes type 2 diabetes and obesity candidate genes.
PMID 16757574 · PMC1475747 · Nucleic acids research · 2006 · 6 claims · 8 setups
Applying seven independent computational disease-gene prioritization methods in concert to 9556 positional candidate genes identifies a prioritized set of likely T2D and obesity candidate genes
-
Full-text index only
Complex genetic diseases: controversy over the Croesus code.
PMID 11532206 · PMC138948 · Genome biology · 2001 · 8 claims · 3 setups
The common disease/common variant hypothesis is predicted by population genetic theory (founder population dynamics, mutation-drift-selection balance) and supported by empirical examples such as APOE*E4.
-
Full-text index only
Function2Gene: a gene selection tool to increase the power of genetic association studies by utilizing public databases and expert knowledge.
PMID 18631403 · PMC2500032 · BMC bioinformatics · 2008 · 6 claims · 5 setups
Function2Gene is a set of Perl programs that queries public databases (NCBI, GeneCards, Harvester, with Uniprot/Ensembl also supported) using expert-selected keywords to rank genes by prior probability of disease association.
-
Full-text index only
Genomics and the prevention and control of common chronic diseases: emerging priorities for public health action.
PMID 15888216 · PMC1327699 · Preventing chronic disease · 2005 · 8 claims · 6 setups
Family history is the most consistent risk factor for almost all human diseases across the lifespan.
-
Has reproduction
Using random walks to identify cancer-associated modules in expression data.
PMID 24128261 · PMC4015830 · BioData mining · 2013 · 8 claims · 8 setups
Walktrap-GM, a random-walk community detection algorithm adapted with stopping criteria (maximum modularity, maximum size, maximum module score), identifies modules significantly enriched with cancer genes in expression-weighted interaction networks.
-
Full-text index only
Incorporation of genetic model parameters for cost-effective designs of genetic association studies using DNA pooling.
PMID 17634103 · PMC1947971 · BMC genomics · 2007 · 8 claims · 4 setups
A closed-form approximation to the F-test non-centrality parameter (NCP) incorporating genetic model parameters (disease allele frequency, marker allele frequency, prevalence, genotype relative risk, sample size, genetic model, number of pools/replicates, machine variability) can be used to compute power for DNA pooling association studies
-
Full-text index only
What can genome-wide association studies tell us about the genetics of common disease?
PMID 18454206 · PMC2323402 · PLoS genetics · 2008 · 8 claims · 4 setups
Apparent patterns of common, low-effect disease-associated alleles largely reflect statistical power of studies rather than the true underlying distribution of disease variants
-
Full-text index only
Genome-wide prioritization of disease genes and identification of disease-disease associations from an integrated human functional linkage network.
PMID 19728866 · PMC2768980 · Genome biology · 2009 · 6 claims · 6 setups
Integrating 16 genomic features (32 sub-features) via a naïve Bayes classifier produces a genome-scale FLN of 21,657 human genes and 22,388,609 weighted links that outperforms any individual data source for inferring functional linkages.
-
Full-text index only
Estimation of relevant variables on high-dimensional biological patterns using iterated weighted kernel functions.
PMID 18509521 · PMC2396875 · PloS one · 2008 · 7 claims · 6 setups
wKIERA combines a weighted-kernel discriminant (kernel perceptron) with an iterative stochastic probability estimation-of-distribution algorithm to estimate a relevance distribution over variables
-
Full-text index only
Supervised learning-based tagSNP selection for genome-wide disease classifications.
PMID 18366619 · PMC2386071 · BMC genomics · 2008 · 7 claims · 2 setups
SRFA (Supervised Recursive Feature Addition) is a novel feature selection method combining supervised learning and statistical redundancy measures for SNP selection
-
Full-text index only
Mutation analysis of the PTEN / MMAC1 gene in Japanese patients with Cowden disease.
PMID 10920277 · PMC5926416 · Japanese journal of cancer research : Gann · 2000 · 7 claims · 4 setups
Sequencing of all PTEN/MMAC1 coding regions identified five different germline mutations, four of them novel, in 5 of 12 unrelated Japanese CD patients
-
Full-text index only
Iterative class discovery and feature selection using Minimal Spanning Trees.
PMID 15355552 · PMC520744 · BMC bioinformatics · 2004 · 7 claims · 5 setups
Iterating between MST-based clustering and t-statistic feature selection removes noise genes step-wise while sharpening the sample clustering
-
Has reproduction · 83
Current status of use of high throughput nucleotide sequencing in rheumatology.
PMID 33408124 · PMC7789458 · RMD open · 2021 · 8 claims · 8 setups
RNA-Seq is the most represented HTS assay used in rheumatology research, primarily for biomarker identification in blood or synovial tissue.
-
Has reproduction · 71
Gene Set Enrichment Analysis Reveals Individual Variability in Host Responses in Tuberculosis Patients.
PMID 34421903 · PMC8375662 · Frontiers in immunology · 2021 · 8 claims · 8 setups
TB patients show substantial individual variability in the intensity of hallmark IFN responses, as well as in complement system, metabolic, and other pathway responses.
-
Has reproduction · 67
Cyrface: An interface from Cytoscape to R that provides a user interface to R packages.
PMID 24715956 · PMC3962008 · F1000Research · 2013 · 8 claims · 6 setups
Cyrface is a Cytoscape app/Java library providing a general interface from Cytoscape (Java) to any R function or package.
-
Full-text index only
A non-parametric meta-analysis approach for combining independent microarray datasets: application using two microarray datasets pertaining to chronic allograft nephropathy.
PMID 18302764 · PMC2276496 · BMC genomics · 2008 · 8 claims · 6 setups
A novel non-parametric meta-analysis approach for combining independent microarray datasets is presented, requiring no distributional assumptions and being logically intuitive.
-
Full-text index only
Blood-based epigenetic instability linked to human aging and disease.
PMID 41690920 · PMC13018287 · Nature communications · 2026 · 7 claims · 8 setups
31,744 unmethylated (and 6143 methylated) CpG loci in blood show highly consistent, stable methylation in young healthy individuals and are defined as Epigenetically Stable Loci (ESLs)
-
Full-text index only
scTWAS: a powerful statistical framework for single-cell transcriptome-wide association studies.
PMID 41820391 · PMC13121454 · Nature communications · 2026 · 8 claims · 5 setups
scTWAS uses a latent-variable expression-measurement model combined with a moment-based regression to more accurately estimate genetic regulation of gene expression from single-cell data, improving GReX prediction across cell types and datasets
-
Full-text index only
Integrated weighted gene co-expression network analysis with an application to chronic fatigue syndrome.
PMID 18986552 · PMC2625353 · BMC systems biology · 2008 · 8 claims · 6 setups
Integrated WGCNA (IWGCNA), which adds genetic marker-based causality testing to standard WGCNA, can identify a disease-related module and its causal drivers
-
Full-text index only
A genomic approach to improve prognosis and predict therapeutic response in chronic lymphocytic leukemia.
PMID 19861443 · PMC2783430 · Clinical cancer research : an official journal of the American Association for Cancer Research · 2009 · 8 claims · 8 setups
A genomic signature derived from CLL patient leukemic cells significantly differentiates stable from progressive disease