Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction
Genomic prediction based on selective linkage disequilibrium pruning of low-coverage whole-genome sequence variants in a pure Duroc population.
PMID 37853325 · PMC10583454 · Genetics, selection, evolution : GSE · 2023 · 5 claims · 6 setups
SLDP selects a subset of WGS variants using GWAS prior information (P-value threshold) combined with LD pruning (r2) to improve genomic prediction accuracy.
-
Has reproduction · 81
SEMdag: Fast learning of Directed Acyclic Graphs via node or layer ordering.
PMID 39775401 · PMC11709272 · PloS one · 2025 · 8 claims · 5 setups
SEMdag() is a two-step order-based algorithm for fast learning of high-dimensional linear SEMs, using knowledge-based (KB) or data-driven bottom-up (BU) node/layer ordering followed by penalized (L1) DAG estimation
-
Full-text index only
Automatic discovery of cross-family sequence features associated with protein function.
PMID 16409628 · PMC1395344 · BMC bioinformatics · 2006 · 8 claims · 6 setups
A self-supervised data mining approach can find relationships between sequence features and functional annotations without preconceived functional categories.
-
Full-text index only
Human and mouse introns are linked to the same processes and functions through each genome's most frequent non-conserved motifs.
PMID 18450818 · PMC2425492 · Nucleic acids research · 2008 · 8 claims · 5 setups
Pyknons (recurrent, genome-specific, ≥16nt motifs with ≥30 intact intergenic/intronic copies and ≥1 exonic copy) span a substantial fraction of previously uncharacterized intronic space (7.4% human, 4.4% mouse)
-
Full-text index only
Combining comparative genomics with de novo motif discovery to identify human transcription factor DNA-binding motifs.
PMID 17217514 · PMC1780116 · BMC bioinformatics · 2006 · 6 claims · 4 setups
A novel method combining 8-species comparative genomics with de novo motif discovery identifies human TF DNA-binding motifs overrepresented and conserved in upstream regions of co-regulated genes
-
Full-text index only
High-throughput discovery of rare human nucleotide polymorphisms by Ecotilling.
PMID 16893952 · PMC1540726 · Nucleic acids research · 2006 · 7 claims · 6 setups
Ecotilling can be adapted to accurately discover and genotype human SNPs, with error rates low relative to resequencing
-
Full-text index only
CRSD: a comprehensive web server for composite regulatory signature discovery.
PMID 16845073 · PMC1538777 · Nucleic acids research · 2006 · 7 claims · 5 setups
CRSD is a comprehensive web server integrating six large-scale databases (UniGene, mature microRNAs, putative promoter, TRANSFAC, pathway, GO) plus two newly constructed genome-wide databases (MRS and TRS) for composite regulatory signature discovery
-
Full-text index only
PlasmoDraft: a database of Plasmodium falciparum gene function predictions based on postgenomic data.
PMID 18925948 · PMC2605471 · BMC bioinformatics · 2008 · 8 claims · 4 setups
Gonna, a supervised k-nearest-neighbor Guilt-By-Association predictor, proposes GO annotations for a gene based on similarity of its transcriptome, proteome, or interactome profile to genes already annotated by GeneDB
-
Full-text index only
An integrated genomic analysis of human glioblastoma multiforme.
PMID 18772396 · PMC2820389 · Science (New York, N.Y.) · 2008 · 8 claims · 7 setups
IDH1 is recurrently mutated at its active site (R132) in 12% of GBM patients, a previously unrecognized alteration in GBM.
-
Full-text index only
Inferring combinatorial regulation of transcription in silico.
PMID 15647509 · PMC546154 · Nucleic acids research · 2005 · 8 claims · 5 setups
Combining Cluster-Buster (TFBS cluster prediction) with GOSSIP (rigorous GO enrichment statistics with multiple-testing/FDR correction) predicts biological functions controlled by combinatorial transcription factor action, without prior knowledge of factor targets
-
Full-text index only
High-throughput chromatin information enables accurate tissue-specific prediction of transcription factor binding sites.
PMID 18988630 · PMC2662491 · Nucleic acids research · 2009 · 8 claims · 8 setups
Incorporating H3K4me3 chromatin modification estimates greatly improves the accuracy of in silico prediction of in vivo TF binding for a wide range of TFs in human and mouse
-
Has reproduction · 76
nf-core/circrna: a portable workflow for the quantification, miRNA target prediction and differential expression analysis of circular RNAs.
PMID 36694127 · PMC9875403 · BMC bioinformatics · 2023 · 8 claims · 4 setups
Existing circRNA workflows are limited: none delineate circRNA-miRNA interactions and only one performs differential expression analysis, requiring users to supplement missing analysis types with in-house expertise
-
Full-text index only
The proteogenomic path towards biomarker discovery.
PMID 18764911 · PMC2574627 · Pediatric transplantation · 2008 · 8 claims · 8 setups
Serum creatinine is a widely used but non-ideal biomarker for renal transplant monitoring because it lacks specificity and sensitivity for graft injury
-
Full-text index only
Expression genomics in breast cancer research: microarrays at the crossroads of biology and medicine.
PMID 17397520 · PMC1868923 · Breast cancer research : BCR · 2007 · 8 claims · 8 setups
Genome-wide expression microarray studies reveal transcriptional networks/signatures that explain breast cancer biological and clinical heterogeneity
-
Full-text index only
Systems biology of gene regulation fulfills its promise.
PMID 16719937 · PMC1779525 · Genome biology · 2006 · 8 claims · 8 setups
Suz12, a Polycomb Group complex component, has DNA targets identifiable by ChIP-chip and can silence large genomic regions in a cell-type-specific manner.
-
Full-text index only
Genome informatics: taming the avalanche of genomic data.
PMID 15642109 · PMC549058 · Genome biology · 2005 · 8 claims · 7 setups
Ultraconserved regions (>100 bp, 100% conserved among mammals) exist in the genome and their function remains unknown
-
Has reproduction · 78
Exome sequencing in 38 patients with intracranial aneurysms and subarachnoid hemorrhage.
PMID 32367296 · PMC7419486 · Journal of neurology · 2020 · 8 claims · 6 setups
Sequence variants in PCNT, RNF213 and THSD1 support a role as susceptibility factors for cerebrovascular disease (UIA/aSAH)
-
Full-text index only
The cohesin complex: sequence homologies, interaction networks and shared motifs.
PMID 11276426 · PMC30708 · Genome biology · 2001 · 8 claims · 8 setups
Mouse Mmip1 and Smc3 (SMCD) share 99% sequence identity and are products of the same gene
-
Full-text index only
Mining the draft human genome.
PMID 11236999 · PMC2658632 · Nature · 2001 · 8 claims · 7 setups
Protein-coding exons account for only about 3% of the human genome DNA, with repeat sequences making up around 46%.
-
Full-text index only
In vivo phosphoproteome of human skeletal muscle revealed by phosphopeptide enrichment and HPLC-ESI-MS/MS.
PMID 19764811 · PMC2783959 · Journal of proteome research · 2009 · 8 claims · 6 setups
This is the first large-scale in vivo phosphoproteomic study of human skeletal muscle