Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 80
Colorectal Cancer Prediction Based on Weighted Gene Co-Expression Network Analysis and Variational Auto-Encoder.
PMID 32825264 · PMC7563725 · Biomolecules · 2020 · 6 claims · 7 setups
Combining WGCNA hub genes and VAE 10-dimensional representation as features for an SVM classifier achieves high accuracy in predicting CRC
-
Has reproduction
Using random walks to identify cancer-associated modules in expression data.
PMID 24128261 · PMC4015830 · BioData mining · 2013 · 8 claims · 8 setups
Walktrap-GM, a random-walk community detection algorithm adapted with stopping criteria (maximum modularity, maximum size, maximum module score), identifies modules significantly enriched with cancer genes in expression-weighted interaction networks.
-
Full-text index only
A non-parametric meta-analysis approach for combining independent microarray datasets: application using two microarray datasets pertaining to chronic allograft nephropathy.
PMID 18302764 · PMC2276496 · BMC genomics · 2008 · 8 claims · 6 setups
A novel non-parametric meta-analysis approach for combining independent microarray datasets is presented, requiring no distributional assumptions and being logically intuitive.
-
Full-text index only
Ontological visualization of protein-protein interactions.
PMID 15707487 · PMC550656 · BMC bioinformatics · 2005 · 8 claims · 8 setups
Aggregating independently made GO 'protein binding' (IPI) annotations reveals larger, previously undescribed mouse protein-protein interaction networks
-
Full-text index only
Integrated weighted gene co-expression network analysis with an application to chronic fatigue syndrome.
PMID 18986552 · PMC2625353 · BMC systems biology · 2008 · 8 claims · 6 setups
Integrated WGCNA (IWGCNA), which adds genetic marker-based causality testing to standard WGCNA, can identify a disease-related module and its causal drivers
-
Has reproduction · 92
Analytical code sharing practices in biomedical research.
PMID 38983240 · PMC11232620 · PeerJ. Computer science · 2024 · 8 claims · 4 setups
Nearly half (49.9%) of 453 examined biomedical manuscripts failed to share the analytical code used to generate their results
-
Full-text index only
Incorporation of genetic model parameters for cost-effective designs of genetic association studies using DNA pooling.
PMID 17634103 · PMC1947971 · BMC genomics · 2007 · 8 claims · 4 setups
A closed-form approximation to the F-test non-centrality parameter (NCP) incorporating genetic model parameters (disease allele frequency, marker allele frequency, prevalence, genotype relative risk, sample size, genetic model, number of pools/replicates, machine variability) can be used to compute power for DNA pooling association studies
-
Full-text index only
Protein function assignment through mining cross-species protein-protein interactions.
PMID 18253506 · PMC2216687 · PloS one · 2008 · 8 claims · 6 setups
CSIDOP predicts protein molecular function with 95.42% accuracy using 2,972 GO functional categories in H. sapiens
-
Full-text index only
SW-ARRAY: a dynamic programming solution for the identification of copy-number changes in genomic DNA using array comparative genome hybridization data.
PMID 15961730 · PMC1151590 · Nucleic acids research · 2005 · 7 claims · 5 setups
SW-ARRAY, an adaptation of the Smith-Waterman dynamic programming algorithm, provides a sensitive and robust method for identifying copy-number changes in array CGH data
-
Full-text index only
scTWAS: a powerful statistical framework for single-cell transcriptome-wide association studies.
PMID 41820391 · PMC13121454 · Nature communications · 2026 · 8 claims · 5 setups
scTWAS uses a latent-variable expression-measurement model combined with a moment-based regression to more accurately estimate genetic regulation of gene expression from single-cell data, improving GReX prediction across cell types and datasets
-
Full-text index only
Does distance matter? Variations in alternative 3' splicing regulation.
PMID 17704130 · PMC2018619 · Nucleic acids research · 2007 · 8 claims · 7 setups
Alternative 3' splice sites can be distinguished from constitutive splice sites by a combination of sequence/conservation properties that vary depending on the distance between the splice sites.
-
Full-text index only
What can genome-wide association studies tell us about the genetics of common disease?
PMID 18454206 · PMC2323402 · PLoS genetics · 2008 · 8 claims · 4 setups
Apparent patterns of common, low-effect disease-associated alleles largely reflect statistical power of studies rather than the true underlying distribution of disease variants
-
Full-text index only
Worldwide distribution of NAT2 diversity: implications for NAT2 evolutionary history.
PMID 18304320 · PMC2292740 · BMC genetics · 2008 · 8 claims · 8 setups
NAT2 coding region sequence variation in the Mandenka and other sub-Saharan African populations is consistent with selective neutrality and constant population size.
-
Full-text index only
Combinatorial Mismatch Scan (CMS) for loci associated with dementia in the Amish.
PMID 16515697 · PMC1448207 · BMC medical genetics · 2006 · 8 claims · 7 setups
CMS compares IBS allele/genotype sharing between distantly related (beyond grandparental) affected and unaffected individuals from founder populations to detect disease loci while reducing confounding from population stratification and genetic heterogeneity.
-
Full-text index only
Functional importance of different patterns of correlation between adjacent cassette exons in human and mouse.
PMID 18439302 · PMC2432081 · BMC genomics · 2008 · 8 claims · 7 setups
Adjacent cassette exon pairs can be categorized by EST-derived correlation coefficient into three groups: mutually exclusive (ME, r<=-0.7), independent (IND, -0.2<=r<=0.2), and linked (LNK, r>=0.7)
-
Full-text index only
Empirical codon substitution matrix.
PMID 15927081 · PMC1173088 · BMC bioinformatics · 2005 · 8 claims · 5 setups
The authors present the first empirical codon substitution matrix built entirely from alignments of vertebrate coding DNA sequences.