Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 71
RNAmountAlign: Efficient software for local, global, semiglobal pairwise and multiple RNA sequence/structure alignment.
PMID 31978147 · PMC6980424 · PloS one · 2020 · 8 claims · 6 setups
RNAmountAlign is the first RNA sequence/structure pairwise alignment algorithm based on incremental ensemble mountain distance, running in O(n^3) time and O(n^2) space for two sequences of length n.
-
Full-text index only
Structure SNP (StSNP): a web server for mapping and modeling nsSNPs on protein structures with linkage to metabolic pathways.
PMID 17537826 · PMC1933130 · Nucleic acids research · 2007 · 7 claims · 5 setups
StSNP integrates dbSNP, PDB, KEGG, and NCBI Entrez data into a single web server for nsSNP analysis
-
Full-text index only
Bayesian coestimation of phylogeny and sequence alignment.
PMID 15804354 · PMC1087833 · BMC bioinformatics · 2005 · 7 claims · 3 setups
Alignment and phylogenetic inference are mutually dependent, and treating them as separate sequential steps (align then infer tree) is fundamentally flawed and produces biased, overconfident estimates.
-
Full-text index only
Structural evolution of the protein kinase-like superfamily.
PMID 16244704 · PMC1261164 · PLoS computational biology · 2005 · 8 claims · 5 setups
All kinases in the superfamily share a 'universal core' domain consisting only of the regions required for ATP binding and the phosphotransfer reaction.
-
Full-text index only
From genomics to chemical genomics: new developments in KEGG.
PMID 16381885 · PMC1347464 · Nucleic acids research · 2006 · 8 claims · 5 setups
KEGG BRITE has been formally added as a fourth main KEGG database to establish a logical foundation for functional interpretation and pathway reconstruction.
-
Full-text index only
MODBASE, a database of annotated comparative protein structure models and associated resources.
PMID 18948282 · PMC2686492 · Nucleic acids research · 2009 · 8 claims · 8 setups
MODBASE contains 5,152,695 reliable comparative protein structure models for 1,593,209 unique protein sequences.
-
Full-text index only
MODBASE: a database of annotated comparative protein structure models and associated resources.
PMID 16381869 · PMC1347422 · Nucleic acids research · 2006 · 8 claims · 7 setups
MODBASE is a database of automatically calculated comparative protein structure models covering all UniProt sequences matchable to a known structure
-
Full-text index only
Genomic organization of zebrafish microRNAs.
PMID 18510755 · PMC2427041 · BMC genomics · 2008 · 8 claims · 6 setups
Using sequence conservation and prediction algorithms, 35 new zebrafish miRNAs were identified, bringing the total to 415.
-
Has reproduction · 64
Sister DNA Entrapment between Juxtaposed Smc Heads and Kleisin of the Cohesin Complex.
PMID 31201089 · PMC6675936 · Molecular cell · 2019 · 8 claims · 5 setups
Smc1 and Smc3 ATPase heads adopt two distinct in vivo states: ATP-dependent engaged (E) and signature-motif juxtaposed (J).
-
Full-text index only
Comprehensive annotation of bidirectional promoters identifies co-regulation among breast and ovarian cancer genes.
PMID 17447839 · PMC1853124 · PLoS computational biology · 2007 · 8 claims · 8 setups
A new algorithm using spliced ESTs (cross-validated against Known Genes and GenBank mRNA) comprehensively maps bidirectional promoters in the human genome
-
Full-text index only
SuperCYP: a comprehensive database on Cytochrome P450 enzymes including a tool for analysis of CYP-drug interactions.
PMID 19934256 · PMC2808967 · Nucleic acids research · 2010 · 8 claims · 6 setups
SuperCYP is a comprehensive relational database aggregating CYP enzyme, drug metabolism, SNP/mutation, and structural information from literature and web resources.
-
Has reproduction · 87
R2DT is a framework for predicting and visualising RNA secondary structure using templates.
PMID 34108470 · PMC8190129 · Nature communications · 2021 · 8 claims · 6 setups
R2DT is a template-based computational framework/pipeline that predicts and visualises RNA 2D structure in standardised, community-accepted layouts
-
Has reproduction · 87
Enhanced Generalizability of RNA Secondary Structure Prediction via Convolutional Block Attention Network and Ensemble Learning.
PMID 40871599 · PMC12388828 · Molecules (Basel, Switzerland) · 2025 · 8 claims · 8 setups
TrioFold integrates base-pairing clues from thermodynamic- and DL-based methods via ensemble learning and a convolutional block attention mechanism to enhance RSS prediction generalizability.
-
Full-text index only
A screen for proteins that interact with PAX6: C-terminal mutations disrupt interaction with HOMER3, DNCL1 and TRIM11.
PMID 16098226 · PMC1208879 · BMC genetics · 2005 · 8 claims · 7 setups
PAX6 interacts with three novel proteins: HOMER3, DNCL1 and TRIM11
-
Full-text index only
SVC: structured visualization of evolutionary sequence conservation.
PMID 15991338 · PMC1160265 · Nucleic acids research · 2005 · 7 claims · 5 setups
SVC aligns protein-coding sequences of orthologous gene pairs and maps them back onto their encoding exons/introns to generate a scaffold of conserved gene structure.
-
Full-text index only
Using ESTs to improve the accuracy of de novo gene prediction.
PMID 16817966 · PMC1534067 · BMC bioinformatics · 2006 · 8 claims · 8 setups
TWINSCAN_EST combines EST alignments with TWINSCAN via a trainable 'ESTseq' representation and improves exact gene structure prediction accuracy on the whole C. elegans genome
-
Full-text index only
Tandem repeats modify the structure of human genes hosted in segmental duplications.
PMID 19954527 · PMC2812944 · Genome biology · 2009 · 8 claims · 6 setups
Around 7% of primate-specific genes located within segmental duplications contain variable internal tandem repeats (ITRs).
-
Full-text index only
ECgene: genome annotation for alternative splicing.
PMID 15608289 · PMC540072 · Nucleic acids research · 2005 · 8 claims · 5 setups
ECgene combines genome-based EST clustering with a graph-theoretic transcript assembly procedure to predict gene models including alternative splicing events.
-
Full-text index only
Vertebrate gene finding from multiple-species alignments using a two-level strategy.
PMID 16925840 · PMC1810555 · Genome biology · 2006 · 8 claims · 5 setups
DOGFISH cleanly separates a multi-species alignment classifier (RVM cascade) from an HMM-based structure predictor, avoiding tight coupling of alignment complexity with HMM formalism
-
Has reproduction · 80
Structure of the intergenic spacers in chicken ribosomal DNA.
PMID 31655542 · PMC6815422 · Genetics, selection, evolution : GSE · 2019 · 7 claims · 6 setups
Long-read PacBio sequencing of a chicken NOR-containing BAC clone resolved three complete IGS sequences plus rRNA gene clusters that are otherwise missing from genome assemblies.