Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Interaction profile-based protein classification of death domain.
PMID 15189571 · PMC459208 · BMC bioinformatics · 2004 · 7 claims · 6 setups
An SVM-based classifier using Residue Pair Interaction Profiles (RPIPs) can classify death domain superfamily members into subfamilies with 89% average cross-validation accuracy
-
Full-text index only
Applications for protein sequence-function evolution data: mRNA/protein expression analysis and coding SNP scoring tools.
PMID 16912992 · PMC1538848 · Nucleic acids research · 2006 · 7 claims · 8 setups
PANTHER HMMs built from family/subfamily multiple sequence alignments can classify novel protein sequences into functional groups based on statistically significant HMM match scores
-
Has reproduction · 87
A target enrichment method for gathering phylogenetic information from hundreds of loci: An example from the Compositae.
PMID 25202605 · PMC4103609 · Applications in plant sciences · 2014 · 8 claims · 8 setups
A custom sequence capture probe set (9678 baits targeting 1061 orthologous genes) was designed to enrich COS loci across the Compositae.
-
Has reproduction · 83
Gene-expression patterns in peripheral blood classify familial breast cancer susceptibility.
PMID 26538066 · PMC4634735 · BMC medical genomics · 2015 · 8 claims · 7 setups
A multigene expression biomarker from PBMCs accurately classifies familial breast cancer (FBC) status
-
Full-text index only
Genome bioinformatic analysis of nonsynonymous SNPs.
PMID 17708757 · PMC1978506 · BMC bioinformatics · 2007 · 8 claims · 8 setups
Structure- and sequence-based prediction tools can generally distinguish disease-causing mutations from neutral ones
-
Full-text index only
MODBASE, a database of annotated comparative protein structure models and associated resources.
PMID 18948282 · PMC2686492 · Nucleic acids research · 2009 · 8 claims · 8 setups
MODBASE contains 5,152,695 reliable comparative protein structure models for 1,593,209 unique protein sequences.
-
Full-text index only
Local combinational variables: an approach used in DNA-binding helix-turn-helix motif prediction with sequence information.
PMID 19651875 · PMC2761287 · Nucleic acids research · 2009 · 8 claims · 7 setups
The LCV approach predicts HTH motifs with 93.29% accuracy, 93.93% sensitivity and 92.66% specificity using only primary sequence information
-
Has reproduction · 88
Comprehensive benchmarking of large language models for RNA secondary structure prediction.
PMID 40205851 · PMC11982019 · Briefings in bioinformatics · 2025 · 7 claims · 4 setups
Existing RNA-LLMs had not previously been evaluated for secondary structure prediction in a unified, fair experimental setup with the same datasets and prediction model.
-
Full-text index only
Genes implicated in multiple sclerosis pathogenesis from consilience of genotyping and expression profiles in relapse and remission.
PMID 18366677 · PMC2324081 · BMC medical genetics · 2008 · 8 claims · 7 setups
Distinct sets of dysregulated genes are found in peripheral blood during the relapse phase versus the remission phase of RRMS
-
Has reproduction · 94
A Deluge of Complex Repeats: The Solanum Genome.
PMID 26241045 · PMC4524691 · PloS one · 2015 · 8 claims · 8 setups
~50–60% of the genomes of S. tuberosum and S. lycopersicum are composed of repetitive elements
-
Has reproduction · 58
The Li2 mutation results in reduced subgenome expression bias in elongating fibers of allotetraploid cotton (Gossypium hirsutum L.).
PMID 24598808 · PMC3944810 · PloS one · 2014 · 8 claims · 7 setups
The Li2 mutation significantly reduces subgenome (homeolog) expression bias in the elongating fiber transcriptome.
-
Has reproduction · 61
TEMP: a computational method for analyzing transposable element polymorphism in populations.
PMID 24753423 · PMC4066757 · Nucleic acids research · 2014 · 8 claims · 8 setups
TEMP combines pair-end (discordant) read and split (soft-clipped) read information to identify both presence and absence of TE insertions in genomic DNA from heterogeneous/pooled samples.
-
Full-text index only
Exonic remnants of whole-genome duplication reveal cis-regulatory function of coding exons.
PMID 19969543 · PMC2831330 · Nucleic acids research · 2010 · 8 claims · 8 setups
38 candidate cis-regulatory coding exons (RCEs) with predicted target genes were identified genome-wide
-
Full-text index only
Bioinformatics analysis of the locus for enterocyte effacement provides novel insights into type-III secretion.
PMID 15757514 · PMC1084347 · BMC microbiology · 2005 · 8 claims · 7 setups
PSI-BLAST identified several novel homologies between LEE-encoded and Ysc-Yop-associated proteins
-
Has reproduction · 100
Constructing eRNA-mediated gene regulatory networks to explore the genetic basis of muscle and fat-relevant traits in pigs.
PMID 38594607 · PMC11003151 · Genetics, selection, evolution : GSE · 2024 · 8 claims · 8 setups
H3K27ac ChIP-seq and RNA-seq were used to construct eRNA expression profiles across multiple tissues in Enshi Black (ES) and Duroc pigs, revealing tissue-level eRNA regulatory landscapes