Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
L2L: a simple tool for discovering the hidden significance in microarray expression data.
PMID 16168088 · PMC1242216 · Genome biology · 2005 · 8 claims · 4 setups
L2L systematically compares a user's differentially expressed gene list against a database of published differentially expressed gene lists to find statistically significant overlaps and generate hypotheses about shared mechanisms
-
Has reproduction · 76
GeneSetCart: assembling, augmenting, combining, visualizing, and analyzing gene sets.
PMID 40208796 · PMC11984350 · GigaScience · 2025 · 8 claims · 8 setups
GeneSetCart is a web-based platform that lets users assemble, augment, combine, visualize, and analyze gene sets from multiple sources in one place
-
Has reproduction · 71
A method for selectively enriching microbial DNA from contaminating vertebrate host DNA.
PMID 24204593 · PMC3810253 · PloS one · 2013 · 8 claims · 7 setups
MBD-Fc bound to Protein A paramagnetic beads selectively binds methylated (vertebrate host) DNA, depleting it from mixed samples
-
Full-text index only
Improvements to cardiovascular gene ontology.
PMID 19046747 · PMC2706316 · Atherosclerosis · 2009 · 8 claims · 8 setups
Gene Ontology (GO) provides a controlled vocabulary that links current functional knowledge of genes to high-throughput genomic and proteomic datasets, aiding data interpretation.
-
Full-text index only
Human PAML browser: a database of positive selection on human genes using phylogenetic methods.
PMID 17962310 · PMC2238824 · Nucleic acids research · 2008 · 8 claims · 5 setups
The Human PAML Browser is a web-accessible database of codeml-based positive selection test results for 13,721 human genes with orthologs in UCSC multispecies alignments.
-
Full-text index only
Rapid identification of PAX2/5/8 direct downstream targets in the otic vesicle by combinatorial use of bioinformatics tools.
PMID 18828907 · PMC2760872 · Genome biology · 2008 · 8 claims · 8 setups
A combinatorial bioinformatics pipeline (evolutionary double filtering comparative genomics, GXD/ZFIN database queries, MEDLINE text mining) can rapidly and specifically identify PAX2/5/8 direct downstream targets in the otic vesicle
-
Full-text index only
Advances in the study of SR protein family.
PMID 15626328 · PMC5172405 · Genomics, proteomics & bioinformatics · 2003 · 8 claims · 8 setups
SR proteins promote assembly of the early splicesome via protein-protein interactions in their RS-domain that recruit components of the splicing machinery.
-
Full-text index only
Inferring combinatorial regulation of transcription in silico.
PMID 15647509 · PMC546154 · Nucleic acids research · 2005 · 8 claims · 5 setups
Combining Cluster-Buster (TFBS cluster prediction) with GOSSIP (rigorous GO enrichment statistics with multiple-testing/FDR correction) predicts biological functions controlled by combinatorial transcription factor action, without prior knowledge of factor targets
-
Full-text index only
Analysis of sequence conservation at nucleotide resolution.
PMID 18166073 · PMC2230682 · PLoS computational biology · 2007 · 8 claims · 4 setups
SCONE (Sequence CONservation Evaluation) is a novel method that estimates evolutionary rate and a neutrality p-value for individual nucleotide positions in a multiple sequence alignment.
-
Has reproduction · 74
An open RNA-Seq data analysis pipeline tutorial with an example of reprocessing data from a recent Zika virus study.
PMID 27583132 · PMC4972086 · F1000Research · 2016 · 6 claims · 6 setups
An open-source, reproducible RNA-seq pipeline delivered as an IPython notebook and Docker image can process raw RNA-seq data into interactive PCA/HC plots, enrichment results, and small-molecule predictions with minimal setup overhead
-
Full-text index only
Pseudofam: the pseudogene families database.
PMID 18957444 · PMC2686518 · Nucleic acids research · 2009 · 8 claims · 7 setups
Pseudofam is an online database of pseudogene families built by mapping pseudogenes to Pfam protein families, providing query tools, statistics, and sequence alignments