Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Ab initio identification of human microRNAs based on structure motifs.
PMID 18088431 · PMC2238772 · BMC bioinformatics · 2007 · 8 claims · 7 setups
MiRPred predicts miRNA precursors ab initio using only predicted secondary structure motifs, ignoring nucleotide sequence
-
Has reproduction · 67
Evaluating native-like structures of RNA-protein complexes through the deep learning method.
PMID 36828844 · PMC9958188 · Nature communications · 2023 · 8 claims · 7 setups
DRPScore identifies native-like RNA-protein structures with higher success rates than ITScore-PR, DARS-RNP, and 3dRPC across bound and unbound testing sets.
-
Full-text index only
Cruciform extrusion propensity of human translocation-mediating palindromic AT-rich repeats.
PMID 17264116 · PMC1851657 · Nucleic acids research · 2007 · 8 claims · 4 setups
Cruciform extrusion propensity of PATRRs depends on both length and central symmetry of the repeat.
-
Has reproduction · 85
Predicting the pathogenicity of missense variants using features derived from AlphaFold2.
PMID 37084271 · PMC10203375 · Bioinformatics (Oxford, England) · 2023 · 6 claims · 8 setups
AlphaFold2-derived structural features (solvent accessibility, amino acid network features, physicochemical environment, pLDDT) can be used to train a random forest classifier (AlphScore) that distinguishes proxy-benign from proxy-pathogenic missense variants.
-
Full-text index only
A search for structurally similar cellular internal ribosome entry sites.
PMID 17591613 · PMC1950536 · Nucleic acids research · 2007 · 8 claims · 7 setups
Cellular IRES are not defined by an overall conserved structure (unlike viral IRES) but instead depend on short RNA motifs and shared trans-acting factors (ITAFs)
-
Full-text index only
Comparative sequence analysis of leucine-rich repeats (LRRs) within vertebrate toll-like receptors.
PMID 17517123 · PMC1899181 · BMC genomics · 2007 · 8 claims · 4 setups
A new method combining known LRR structures, multiple sequence alignment, and secondary structure prediction identifies and aligns LRRs in TLRs more accurately than PFAM/InterPro/SMART
-
Full-text index only
An efficient method for the prediction of deleterious multiple-point mutations in the secondary structure of RNAs using suboptimal folding solutions.
PMID 18445289 · PMC2386494 · BMC bioinformatics · 2008 · 8 claims · 6 setups
Using RNAsubopt suboptimal solutions computed once for the wild-type sequence, specific multiple-point mutations likely to cause conformational rearrangement can be selected without brute-force enumeration.
-
Full-text index only
Structural evolution of the protein kinase-like superfamily.
PMID 16244704 · PMC1261164 · PLoS computational biology · 2005 · 8 claims · 5 setups
All kinases in the superfamily share a 'universal core' domain consisting only of the regions required for ATP binding and the phosphotransfer reaction.
-
Full-text index only
Widespread A-to-I RNA editing of Alu-containing mRNAs in the human transcriptome.
PMID 15534692 · PMC526178 · PLoS biology · 2004 · 8 claims · 6 setups
Intramolecular pairs of oppositely oriented Alu elements within the same pre-mRNA form dsRNA foldback structures that are major substrates for A-to-I RNA editing
-
Full-text index only
Columba: an integrated database of proteins, structures, and annotations.
PMID 15801979 · PMC1087474 · BMC bioinformatics · 2005 · 8 claims · 6 setups
COLUMBA physically integrates data from twelve protein structure-related databases (PDB, KEGG, Swiss-Prot, CATH, SCOP, Gene Ontology, ENZYME, etc.) into a single PostgreSQL data warehouse.
-
Full-text index only
nsSNPAnalyzer: identifying disease-associated nonsynonymous single nucleotide polymorphisms.
PMID 15980516 · PMC1160133 · Nucleic acids research · 2005 · 6 claims · 4 setups
nsSNPAnalyzer is a web server that predicts whether a query nsSNP is disease-associated or functionally neutral using a Random Forest classifier combining structural and evolutionary information
-
Has reproduction · 68
Rfam 15: RNA families database in 2025.
PMID 39526405 · PMC11701678 · Nucleic acids research · 2025 · 8 claims · 6 setups
Rfamseq was expanded to 26 106 genomes, a 76% increase, by incorporating the latest UniProt reference proteomes and additional viral genomes
-
Full-text index only
MODBASE, a database of annotated comparative protein structure models and associated resources.
PMID 18948282 · PMC2686492 · Nucleic acids research · 2009 · 8 claims · 8 setups
MODBASE contains 5,152,695 reliable comparative protein structure models for 1,593,209 unique protein sequences.
-
Full-text index only
A systematic comparative and structural analysis of protein phosphorylation sites based on the mtcPTM database.
PMID 17521420 · PMC1929158 · Genome biology · 2007 · 7 claims · 6 setups
mtcPTM is a hierarchically organized database of human and mouse phosphosites that preserves experimental context, enabling comparison of phosphorylation patterns across conditions
-
Full-text index only
Identification of novel homologous microRNA genes in the rhesus macaque genome.
PMID 18186931 · PMC2254598 · BMC genomics · 2008 · 8 claims · 2 setups
454 rhesus miRNA genes were identified in total, including 383 novel genes in addition to 71 previously reported
-
Full-text index only
Non-EST based prediction of exon skipping and intron retention events using Pfam information.
PMID 16204458 · PMC1243800 · Nucleic acids research · 2005 · 7 claims · 5 setups
A novel ab initio method predicts exon skipping and intron retention events using only Pfam domain annotation, via a Viterbi-like dynamic programming algorithm applied to the Pfam alignment.
-
Full-text index only
DG-CST (Disease Gene Conserved Sequence Tags), a database of human-mouse conserved elements associated to disease genes.
PMID 15608249 · PMC539965 · Nucleic acids research · 2005 · 5 claims · 8 setups
Comparative human-mouse genome analysis identifies conserved sequence tags (CSTs, >=70% identity over >=100bp) that frequently correspond to non-coding elements with putative regulatory or structural roles
-
Full-text index only
Predicting deleterious nsSNPs: an analysis of sequence and structural attributes.
PMID 16630345 · PMC1489951 · BMC bioinformatics · 2006 · 8 claims · 7 setups
Sequence conservation (PSIC score difference) at the nsSNP position is the single most useful attribute for predicting deleterious vs neutral status.
-
Full-text index only
miRNAMap: genomic maps of microRNA genes and their target genes in mammalian genomes.
PMID 16381831 · PMC1347497 · Nucleic acids research · 2006 · 6 claims · 6 setups
miRNAMap integrates known miRNA genes from miRBase, literature-curated validated targets, and computationally predicted miRNA genes and targets for human, mouse, rat and dog.
-
Full-text index only
Patrocles: a database of polymorphic miRNA-mediated gene regulation in vertebrates.
PMID 19906729 · PMC2808989 · Nucleic acids research · 2010 · 8 claims · 6 setups
Patrocles is a database compiling DSPs predicted to perturb miRNA-mediated gene regulation across seven vertebrate species, covering targets, miRNA precursors and silencing machinery.