Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 67
Generative and integrative modeling for transcriptomics with formalin fixed paraffin embedded material.
PMID 41029822 · PMC12486589 · Journal of translational medicine · 2025 · 8 claims · 6 setups
The negative binomial distribution best fits fRNA-seq transcript counts, with little evidence supporting zero-inflated extensions
-
Has reproduction
Using random walks to identify cancer-associated modules in expression data.
PMID 24128261 · PMC4015830 · BioData mining · 2013 · 8 claims · 8 setups
Walktrap-GM, a random-walk community detection algorithm adapted with stopping criteria (maximum modularity, maximum size, maximum module score), identifies modules significantly enriched with cancer genes in expression-weighted interaction networks.
-
Full-text index only
PedGenie: an analysis approach for genetic association testing in extended pedigrees and genealogies of arbitrary size.
PMID 16620382 · PMC1459209 · BMC bioinformatics · 2006 · 7 claims · 3 setups
PedGenie is a valid, flexible statistical tool for genetic association analysis in pedigrees of arbitrary size and structure using Monte Carlo significance testing
-
Has reproduction · 92
Analytical code sharing practices in biomedical research.
PMID 38983240 · PMC11232620 · PeerJ. Computer science · 2024 · 8 claims · 4 setups
Nearly half (49.9%) of 453 examined biomedical manuscripts failed to share the analytical code used to generate their results
-
Full-text index only
Ontological visualization of protein-protein interactions.
PMID 15707487 · PMC550656 · BMC bioinformatics · 2005 · 8 claims · 8 setups
Aggregating independently made GO 'protein binding' (IPI) annotations reveals larger, previously undescribed mouse protein-protein interaction networks
-
Full-text index only
Incorporation of genetic model parameters for cost-effective designs of genetic association studies using DNA pooling.
PMID 17634103 · PMC1947971 · BMC genomics · 2007 · 8 claims · 4 setups
A closed-form approximation to the F-test non-centrality parameter (NCP) incorporating genetic model parameters (disease allele frequency, marker allele frequency, prevalence, genotype relative risk, sample size, genetic model, number of pools/replicates, machine variability) can be used to compute power for DNA pooling association studies
-
Full-text index only
SW-ARRAY: a dynamic programming solution for the identification of copy-number changes in genomic DNA using array comparative genome hybridization data.
PMID 15961730 · PMC1151590 · Nucleic acids research · 2005 · 7 claims · 5 setups
SW-ARRAY, an adaptation of the Smith-Waterman dynamic programming algorithm, provides a sensitive and robust method for identifying copy-number changes in array CGH data
-
Full-text index only
Assessing the genomic evidence for conserved transcribed pseudogenes under selection.
PMID 19754956 · PMC2753554 · BMC genomics · 2009 · 8 claims · 8 setups
1750 transcribed pseudogene annotations (TPAs) were identified in the human genome, ~11.5% of all human pseudogene annotations.
-
Full-text index only
Impact of short-read sequencing on the misassembly of a plant genome.
PMID 33530937 · PMC7852129 · BMC genomics · 2021 · 7 claims · 6 setups
Short-read tomato assembly has substantial high-coverage (0.6%, 5.1 Mb) and low-coverage (9.7%, 79.6 Mb) regions relative to background coverage
-
Full-text index only
What can genome-wide association studies tell us about the genetics of common disease?
PMID 18454206 · PMC2323402 · PLoS genetics · 2008 · 8 claims · 4 setups
Apparent patterns of common, low-effect disease-associated alleles largely reflect statistical power of studies rather than the true underlying distribution of disease variants
-
Full-text index only
Numbers of mutations to different types of colorectal cancer.
PMID 16202134 · PMC1266026 · BMC cancer · 2005 · 8 claims · 5 setups
Different biologic subtypes of colorectal cancer require different numbers of oncogenic mutations before transformation
-
Full-text index only
In silico whole-genome screening for cancer-related single-nucleotide polymorphisms located in human mRNA untranslated regions.
PMID 17201911 · PMC1774567 · BMC genomics · 2007 · 8 claims · 5 setups
A computational EST-based pipeline can identify UTR-SNPs that are statistically over-represented in cancerous versus normal tissue libraries
-
Full-text index only
HLA-A gene polymorphism defined by high-resolution sequence-based typing in 161 Northern Chinese Han people.
PMID 15629059 · PMC5172246 · Genomics, proteomics & bioinformatics · 2003 · 7 claims · 5 setups
HLA-A gene shows high polymorphism in the Northern Chinese Han population, with 74 gene types and 36 alleles detected in 161 individuals
-
Full-text index only
Mutations in the ST7/RAY1/HELG locus rarely occur in primary colorectal, gastric, and hepatocellular carcinomas.
PMID 12799635 · PMC2741100 · British journal of cancer · 2003 · 7 claims · 4 setups
ST7 gene mutations are rare in primary colorectal, gastric, and hepatocellular carcinomas
-
Full-text index only
Combinatorial Mismatch Scan (CMS) for loci associated with dementia in the Amish.
PMID 16515697 · PMC1448207 · BMC medical genetics · 2006 · 8 claims · 7 setups
CMS compares IBS allele/genotype sharing between distantly related (beyond grandparental) affected and unaffected individuals from founder populations to detect disease loci while reducing confounding from population stratification and genetic heterogeneity.
-
Full-text index only
Does distance matter? Variations in alternative 3' splicing regulation.
PMID 17704130 · PMC2018619 · Nucleic acids research · 2007 · 8 claims · 7 setups
Alternative 3' splice sites can be distinguished from constitutive splice sites by a combination of sequence/conservation properties that vary depending on the distance between the splice sites.
-
Full-text index only
Functional importance of different patterns of correlation between adjacent cassette exons in human and mouse.
PMID 18439302 · PMC2432081 · BMC genomics · 2008 · 8 claims · 7 setups
Adjacent cassette exon pairs can be categorized by EST-derived correlation coefficient into three groups: mutually exclusive (ME, r<=-0.7), independent (IND, -0.2<=r<=0.2), and linked (LNK, r>=0.7)
-
Full-text index only
Pathway-specific canalization and plasticity of gene expression during C. elegans dauer development.
PMID 41890963 · PMC13014971 · iScience · 2026 · 8 claims · 7 setups
Dauers induced by different environmental (starvation, pheromone, heat) or genetic (daf-2, daf-7, ilc-17.1, cep-1 OE) stimuli are transcriptionally distinct from each other and from continuously developing WT L2/L3 and adults
-
Full-text index only
Crunching the bio-numbers.
PMID 14664241 · PMC1316909 · Environmental health perspectives · 2003 · 8 claims · 6 setups
The eTag Assay System rapidly identifies genes and related proteins without complex sample preparation or follow-up bioinformatics, unlike microarrays
-
Full-text index only
Investigating the genetic association between ERAP1 and ankylosing spondylitis.
PMID 19692350 · PMC2758148 · Human molecular genetics · 2009 · 8 claims · 8 setups
The genetic association between ERAP1 and AS is confirmed in an independent replication cohort