Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Ab initio identification of human microRNAs based on structure motifs.
PMID 18088431 · PMC2238772 · BMC bioinformatics · 2007 · 8 claims · 7 setups
MiRPred predicts miRNA precursors ab initio using only predicted secondary structure motifs, ignoring nucleotide sequence
-
Full-text index only
Comparative genomics search for losses of long-established genes on the human lineage.
PMID 18085818 · PMC2134963 · PLoS computational biology · 2007 · 8 claims · 6 setups
A novel comparative genomics method (TransMap-based syntenic mapping of gene structures between human, mouse, and dog) can detect losses of well-established single-copy genes without relying on sequence homology to a parental gene, distinguishing them from typical duplication- or retrotransposition-derived pseudogenes.
-
Full-text index only
A search for structurally similar cellular internal ribosome entry sites.
PMID 17591613 · PMC1950536 · Nucleic acids research · 2007 · 8 claims · 7 setups
Cellular IRES are not defined by an overall conserved structure (unlike viral IRES) but instead depend on short RNA motifs and shared trans-acting factors (ITAFs)
-
Full-text index only
MiPred: classification of real and pseudo microRNA precursors using random forest prediction model with combined features.
PMID 17553836 · PMC1933124 · Nucleic acids research · 2007 · 8 claims · 8 setups
A hybrid feature combining local contiguous triplet structure-sequence composition, MFE of the secondary structure, and P-value of a randomization test improves classification of real vs pseudo pre-miRNAs
-
Full-text index only
MODBASE, a database of annotated comparative protein structure models and associated resources.
PMID 18948282 · PMC2686492 · Nucleic acids research · 2009 · 8 claims · 8 setups
MODBASE contains 5,152,695 reliable comparative protein structure models for 1,593,209 unique protein sequences.
-
Full-text index only
Random amino acid mutations and protein misfolding lead to Shannon limit in sequence-structure communication.
PMID 18769673 · PMC2518838 · PloS one · 2008 · 8 claims · 6 setups
The protein sequence-structure map behaves as a noisy digital communication channel whose capacity C exceeds the transmission rate R for native structures, satisfying Shannon's noisy channel theorem
-
Full-text index only
The dystrobrevin-binding protein 1 gene: features and networks.
PMID 18663367 · PMC2859304 · Molecular psychiatry · 2009 · 8 claims · 6 setups
DTNBP1 gene structure, protein-coding sequence, and dysbindin domain are conserved across 13 vertebrate species, while noncoding sequence is diverse.
-
Full-text index only
An efficient method for the prediction of deleterious multiple-point mutations in the secondary structure of RNAs using suboptimal folding solutions.
PMID 18445289 · PMC2386494 · BMC bioinformatics · 2008 · 8 claims · 6 setups
Using RNAsubopt suboptimal solutions computed once for the wild-type sequence, specific multiple-point mutations likely to cause conformational rearrangement can be selected without brute-force enumeration.
-
Full-text index only
Tandem repeats modify the structure of human genes hosted in segmental duplications.
PMID 19954527 · PMC2812944 · Genome biology · 2009 · 8 claims · 6 setups
Around 7% of primate-specific genes located within segmental duplications contain variable internal tandem repeats (ITRs).
-
Full-text index only
Tissue compartment analysis for biomarker discovery by gene expression profiling.
PMID 19901995 · PMC2771357 · PloS one · 2009 · 8 claims · 5 setups
TCA method quantifies the fractional volume of constitutive structures in a heterogeneous tissue sample by comparing marker mRNA levels in the whole sample to those in pure isolated structures
-
Full-text index only
Sequence and structure signatures of cancer mutation hotspots in protein kinases.
PMID 19834613 · PMC2759519 · PloS one · 2009 · 8 claims · 6 setups
Developed CKMD (Composite Kinase Mutation Database), an integrated bioinformatics resource mapping genetic variation in protein kinase genes to sequence, structural, and functional data
-
Full-text index only
Reconstructing Indian population history.
PMID 19779445 · PMC2842210 · Nature · 2009 · 8 claims · 8 setups
Most Indian populations descend from a mixture of two ancient, genetically divergent populations: ANI (close to Middle Easterners, Central Asians, Europeans) and ASI (as distinct from ANI and East Asians as those are from each other).
-
Full-text index only
Multiple functions of precursor BDNF to CNS neurons: negative regulation of neurite growth, spine formation and cell survival.
PMID 19674479 · PMC2743674 · Molecular brain · 2009 · 7 claims · 8 setups
R125M, R127L, and R125M/R127L BDNF SNP variants are poorly cleaved, resulting in predominant secretion of proBDNF (CR-proBDNF)
-
Full-text index only
Development of lead hammerhead ribozyme candidates against human rod opsin mRNA for retinal degeneration therapy.
PMID 19094986 · PMC3388947 · Experimental eye research · 2009 · 8 claims · 3 setups
Three lead hhRz candidates (CUC↓266, CUC↓1411, AUA↓1414) significantly knock down human RHO protein expression relative to control (p<0.05)
-
Has reproduction · 88
Comprehensive benchmarking of large language models for RNA secondary structure prediction.
PMID 40205851 · PMC11982019 · Briefings in bioinformatics · 2025 · 7 claims · 4 setups
Existing RNA-LLMs had not previously been evaluated for secondary structure prediction in a unified, fair experimental setup with the same datasets and prediction model.
-
Full-text index only
Gene prediction in eukaryotes with a generalized hidden Markov model that uses hints from external sources.
PMID 16469098 · PMC1409804 · BMC bioinformatics · 2006 · 7 claims · 3 setups
AUGUSTUS+ extends the AUGUSTUS GHMM by combining intrinsic sequence information with extrinsic hints via an extended emission alphabet, so the GHMM jointly models the DNA sequence, gene structure, and hint collection.
-
Full-text index only
Compensatory mutations cause excess of antagonistic epistasis in RNA secondary structure folding.
PMID 12590655 · PMC149451 · BMC evolutionary biology · 2003 · 6 claims · 1 setups
RNA secondary structure folding shows a clear prevalence of antagonistic epistasis (β < 1) among reference sequences
-
Full-text index only
Harnessing the HGP of public health.
PMID 15176090 · PMC1241998 · Environmental health perspectives · 2004 · 8 claims · 7 setups
The tombusvirus p19 protein selectively recognizes and sequesters short (21-22 nucleotide) silencing siRNAs, discriminating them from longer siRNAs by measuring siRNA length via tryptophan-mediated end-stacking interactions
-
Full-text index only
MODBASE: a database of annotated comparative protein structure models and associated resources.
PMID 16381869 · PMC1347422 · Nucleic acids research · 2006 · 8 claims · 7 setups
MODBASE is a database of automatically calculated comparative protein structure models covering all UniProt sequences matchable to a known structure
-
Full-text index only
AUGUSTUS: a web server for gene prediction in eukaryotes that allows user-defined constraints.
PMID 15980513 · PMC1160219 · Nucleic acids research · 2005 · 8 claims · 1 setups
AUGUSTUS web server allows users to impose constraints (splice sites, translation start/stop, known exons, exonic/intronic intervals) on predicted gene structures