Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 88
Comprehensive benchmarking of large language models for RNA secondary structure prediction.
PMID 40205851 · PMC11982019 · Briefings in bioinformatics · 2025 · 7 claims · 4 setups
Existing RNA-LLMs had not previously been evaluated for secondary structure prediction in a unified, fair experimental setup with the same datasets and prediction model.
-
Full-text index only
nsSNPAnalyzer: identifying disease-associated nonsynonymous single nucleotide polymorphisms.
PMID 15980516 · PMC1160133 · Nucleic acids research · 2005 · 6 claims · 4 setups
nsSNPAnalyzer is a web server that predicts whether a query nsSNP is disease-associated or functionally neutral using a Random Forest classifier combining structural and evolutionary information
-
Full-text index only
Classification of real and pseudo microRNA precursors using local structure-sequence features and support vector machine.
PMID 16381612 · PMC1360673 · BMC bioinformatics · 2005 · 7 claims · 7 setups
A 32-dimensional triplet structure-sequence feature vector combined with SVM (triplet-SVM) can distinguish real human pre-miRNAs from pseudo pre-miRNA hairpins with ~90% accuracy.
-
Full-text index only
Ab initio identification of human microRNAs based on structure motifs.
PMID 18088431 · PMC2238772 · BMC bioinformatics · 2007 · 8 claims · 7 setups
MiRPred predicts miRNA precursors ab initio using only predicted secondary structure motifs, ignoring nucleotide sequence
-
Full-text index only
MiPred: classification of real and pseudo microRNA precursors using random forest prediction model with combined features.
PMID 17553836 · PMC1933124 · Nucleic acids research · 2007 · 8 claims · 8 setups
A hybrid feature combining local contiguous triplet structure-sequence composition, MFE of the secondary structure, and P-value of a randomization test improves classification of real vs pseudo pre-miRNAs
-
Full-text index only
In vitro identification and in silico utilization of interspecies sequence similarities using GeneChip technology.
PMID 15871745 · PMC1156887 · BMC genomics · 2005 · 7 claims · 6 setups
Only 14±2% of canine transcripts were detected by U133A probe sets versus 49±6% of human transcripts when hybridized to the same chip
-
Full-text index only
Columba: an integrated database of proteins, structures, and annotations.
PMID 15801979 · PMC1087474 · BMC bioinformatics · 2005 · 8 claims · 6 setups
COLUMBA physically integrates data from twelve protein structure-related databases (PDB, KEGG, Swiss-Prot, CATH, SCOP, Gene Ontology, ENZYME, etc.) into a single PostgreSQL data warehouse.
-
Full-text index only
SNAP: predict effect of non-synonymous polymorphisms on function.
PMID 17526529 · PMC1920242 · Nucleic acids research · 2007 · 7 claims · 8 setups
SNAP, a neural network-based method using sequence-derived information, predicts whether a non-synonymous SNP is neutral or non-neutral for protein function
-
Full-text index only
ProMiR II: a web server for the probabilistic prediction of clustered, nonclustered, conserved and nonconserved microRNAs.
PMID 16845048 · PMC1538778 · Nucleic acids research · 2006 · 6 claims · 4 setups
ProMiR II improves on the original ProMiR by integrating free energy, G/C ratio, conservation score and entropy for more controllable miRNA prediction
-
Has reproduction · 64
Starvation-induced transgenerational inheritance of small RNAs in C. elegans.
PMID 25018105 · PMC4377509 · Cell · 2014 · 8 claims · 7 setups
L1 starvation induces changes in endogenous 22G small RNAs (STGs) that are inherited for at least three generations in fed descendants.
-
Full-text index only
Prodepth: predict residue depth by support vector regression approach from protein sequences only.
PMID 19759917 · PMC2742725 · PloS one · 2009 · 8 claims · 8 setups
Residue depth can be reliably predicted solely from protein primary sequence using support vector regression on sequence-derived features.
-
Full-text index only
Bioinformatic Analysis of Differentially Expressed Long Non-Coding RNAs in Skeletal Muscle Following Aerobic and Resistance Exercise.
PMID 41595529 · PMC12840796 · Genes · 2026 · 8 claims · 3 setups
Distinct lncRNA expression profiles exist in skeletal muscle between acute AE and RE, suggesting modality-specific lncRNA roles in exercise adaptation.
-
Full-text index only
Molecular phylogeny of the kelch-repeat superfamily reveals an expansion of BTB/kelch proteins in animals.
PMID 13678422 · PMC222960 · BMC bioinformatics · 2003 · 8 claims · 8 setups
The human genome encodes at least 71 kelch-repeat proteins
-
Full-text index only
DG-CST (Disease Gene Conserved Sequence Tags), a database of human-mouse conserved elements associated to disease genes.
PMID 15608249 · PMC539965 · Nucleic acids research · 2005 · 5 claims · 8 setups
Comparative human-mouse genome analysis identifies conserved sequence tags (CSTs, >=70% identity over >=100bp) that frequently correspond to non-coding elements with putative regulatory or structural roles
-
Full-text index only
Predicting failure rate of PCR in large genomes.
PMID 18492719 · PMC2441781 · Nucleic acids research · 2008 · 7 claims · 8 setups
The number of predicted primer-binding sites in genomic DNA is the most important factor determining PCR failure.
-
Full-text index only
A novel GJA8 mutation (p.I31T) causing autosomal dominant congenital cataract in a Chinese family.
PMID 20019893 · PMC2794658 · Molecular vision · 2009 · 7 claims · 7 setups
A novel missense mutation c.92T>C (p.I31T) in GJA8 causes autosomal dominant congenital nuclear cataract in this Chinese family
-
Full-text index only
Retentive Network promotes efficient RNA language modeling of long sequences.
PMID 41814064 · PMC13111708 · Communications biology · 2026 · 8 claims · 6 setups
RNAret, a RetNet-based RNA language model with O(n) complexity, achieves training parallelism and low computational overhead while processing long RNA sequences
-
Full-text index only
The Functional RNA Database 3.0: databases to support mining and annotation of functional RNAs.
PMID 18948287 · PMC2686472 · Nucleic acids research · 2009 · 8 claims · 5 setups
fRNAdb 3.0 is a completely rebuilt sequence database hosting a much larger collection of known/predicted non-coding RNA sequences with improved search functionality
-
Full-text index only
PKD1 and PKD2 mutations in Slovenian families with autosomal dominant polycystic kidney disease.
PMID 16430766 · PMC1434729 · BMC medical genetics · 2006 · 7 claims · 8 setups
Linkage analysis can pre-select which gene (PKD1 or PKD2) to screen for mutations in ADPKD families with sufficient samples
-
Has reproduction · 74
ChIP-seq guidelines and practices of the ENCODE and modENCODE consortia.
PMID 22955991 · PMC3431496 · Genome research · 2012 · 8 claims · 8 setups
ENCODE/modENCODE define a set of working standards and guidelines for ChIP-seq covering antibody validation, experimental replication, sequencing depth, data/metadata reporting, and data quality assessment.