Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Functional coverage of the human genome by existing structures, structural genomics targets, and homology models.
PMID 16118666 · PMC1188274 · PLoS computational biology · 2005 · 8 claims · 5 setups
Existing PDB structures provide single-domain coverage for 37% of functional classes in the human genome and complete (whole-protein) structure coverage for 25%.
-
Full-text index only
DAVID Bioinformatics Resources: expanded annotation database and novel algorithms to better extract biology from large gene lists.
PMID 17576678 · PMC1933169 · Nucleic acids research · 2007 · 8 claims · 4 setups
The DAVID Gene Concept uses a single-linkage method to agglomerate tens of millions of gene/protein identifiers from NCBI, PIR, UniProt and other resources into unified DAVID genes.
-
Full-text index only
Assignment of Streptococcus agalactiae isolates to clonal complexes using a small set of single nucleotide polymorphisms.
PMID 18710585 · PMC2533671 · BMC microbiology · 2008 · 7 claims · 6 setups
A four-SNP set (glnA36, glnA429, glcK180, adhP111) identified via the Not-N algorithm plus empirical testing divides GBS into 10 groups concordant with eBURST-defined population structure.
-
Full-text index only
An SVM-based system for predicting protein subnuclear localizations.
PMID 16336650 · PMC1325059 · BMC bioinformatics · 2005 · 7 claims · 3 setups
New kernels defined on k-peptide vectors mapped by BLOSUM62-based high-scored pair matrices (D1, D2, D3) improve SVM discrimination of protein subnuclear localization compared to conventional k-peptide encodings.
-
Has reproduction
DAGFormer: A graph-based domain adaptation approach for single-cell cancer drug response prediction.
PMID 41417875 · PMC12795466 · PLoS computational biology · 2025 · 7 claims · 4 setups
DAGFormer, a graph-based domain adaptation framework integrating bulk and scRNA-seq data, predicts single-cell drug responses more accurately than existing methods.
-
Full-text index only
BLASTO: a tool for searching orthologous groups.
PMID 17483516 · PMC1933156 · Nucleic acids research · 2007 · 7 claims · 2 setups
BLASTO treats each orthologous group as a unit and outputs a ranked list of orthologous groups instead of single sequences
-
Full-text index only
Widespread A-to-I RNA editing of Alu-containing mRNAs in the human transcriptome.
PMID 15534692 · PMC526178 · PLoS biology · 2004 · 8 claims · 6 setups
Intramolecular pairs of oppositely oriented Alu elements within the same pre-mRNA form dsRNA foldback structures that are major substrates for A-to-I RNA editing
-
Full-text index only
Adaptive history of single copy genes highly expressed in the term human placenta.
PMID 18848617 · PMC2759754 · Genomics · 2009 · 8 claims · 6 setups
222 single copy eutherian genes are highly expressed (>=3x median) in the term human placenta
-
Full-text index only
DAVID Knowledgebase: a gene-centered database integrating heterogeneous gene annotation resources to facilitate high-throughput gene functional analysis.
PMID 17980028 · PMC2186358 · BMC bioinformatics · 2007 · 7 claims · 3 setups
The DAVID Gene Concept, a single-linkage algorithm, merges gene clusters from Entrez Gene, UniRef100, and PIR-NREF100 that share protein IDs and species into unified DAVID gene clusters, improving cross-referencing between NCBI and UniProt systems
-
Full-text index only
SNAP predicts effect of mutations on protein function.
PMID 18757876 · PMC2562009 · Bioinformatics (Oxford, England) · 2008 · 8 claims · 3 setups
SNAP is a publicly available web-server implementation predicting functional effects (neutral/non-neutral) of single amino acid substitutions.
-
Full-text index only
Design factors that influence PCR amplification success of cross-species primers among 1147 mammalian primer pairs.
PMID 17029642 · PMC1635982 · BMC genomics · 2006 · 8 claims · 7 setups
The number of index-species (IS) mismatches in a primer pair significantly reduces amplification success, with an estimated 6-8% decrease in success rate per additional mismatch.
-
Full-text index only
Urinary proteomic profiling for diagnostic bladder cancer biomarkers.
PMID 19811072 · PMC3422861 · Expert review of proteomics · 2009 · 8 claims · 8 setups
Single protein biomarkers (e.g., NMP-22, BTA) suffer from high false-positive rates and none have replaced cystoscopy or cytology for bladder cancer detection.
-
Has reproduction · 78
Detecting tipping points of complex diseases by network information entropy.
PMID 38960408 · PMC11221888 · Briefings in bioinformatics · 2024 · 8 claims · 4 setups
NIEE can detect critical states or tipping points in diverse data types, including bulk and single-sample expression data
-
Has reproduction · 89
Improved eukaryotic detection compatible with large-scale automated analysis of metagenomes.
PMID 37032329 · PMC10084625 · Microbiome · 2023 · 8 claims · 7 setups
MAPQ ≥30 filtering improves precision but substantially reduces recall, especially for unrepresented/divergent eukaryotic taxa
-
Has reproduction · 67
Sequencing mRNA from cryo-sliced Drosophila embryos to determine genome-wide spatial patterns of gene expression.
PMID 23951250 · PMC3741199 · PloS one · 2013 · 8 claims · 8 setups
Cryosectioning single blastoderm-stage D. melanogaster embryos along the A–P axis and sequencing mRNA from each slice yields reliable genome-wide spatial expression patterns.
-
Has reproduction · 84
Expression Atlas update--a database of gene and transcript expression from microarray- and sequencing-based functional genomics experiments.
PMID 24304889 · PMC3964963 · Nucleic acids research · 2014 · 8 claims · 6 setups
Expression Atlas is a value-added database providing gene, protein and splice variant expression across cell types, organism parts, developmental stages, diseases and other biological/experimental conditions, built from manually curated high-quality microarray and RNA-sequencing experiments from ArrayExpress.
-
Has reproduction · 71
RNAmountAlign: Efficient software for local, global, semiglobal pairwise and multiple RNA sequence/structure alignment.
PMID 31978147 · PMC6980424 · PloS one · 2020 · 8 claims · 6 setups
RNAmountAlign is the first RNA sequence/structure pairwise alignment algorithm based on incremental ensemble mountain distance, running in O(n^3) time and O(n^2) space for two sequences of length n.
-
Full-text index only
Molecular phylogeny of the kelch-repeat superfamily reveals an expansion of BTB/kelch proteins in animals.
PMID 13678422 · PMC222960 · BMC bioinformatics · 2003 · 8 claims · 8 setups
The human genome encodes at least 71 kelch-repeat proteins
-
Full-text index only
Large-scale and high-confidence proteomic analysis of human seminal plasma.
PMID 16709260 · PMC1779515 · Genome biology · 2006 · 8 claims · 6 setups
923 proteins were identified with high confidence in seminal plasma from a single individual, combining results from three ejaculate samples
-
Full-text index only
Predicting deleterious nsSNPs: an analysis of sequence and structural attributes.
PMID 16630345 · PMC1489951 · BMC bioinformatics · 2006 · 8 claims · 7 setups
Sequence conservation (PSIC score difference) at the nsSNP position is the single most useful attribute for predicting deleterious vs neutral status.