Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Large-scale discovery of insertion hotspots and preferential integration sites of human transposed elements.
PMID 20008508 · PMC2836564 · Nucleic acids research · 2010 · 8 claims · 6 setups
Most TEs insert within specific 'hotspots' along the targeted TE rather than uniformly.
-
Full-text index only
QTL MatchMaker: a multi-species quantitative trait loci (QTL) database and query system for annotation of genes and QTL.
PMID 16381937 · PMC1347390 · Nucleic acids research · 2006 · 8 claims · 5 setups
QTL MatchMaker integrates QTL information with physical, genetic and cytogenetic maps across human, mouse and rat genomes
-
Full-text index only
An integrated database-pipeline system for studying single nucleotide polymorphisms and diseases.
PMID 19091018 · PMC2638159 · BMC bioinformatics · 2008 · 6 claims · 5 setups
Existing SNP/disease databases are fragmented; no combined resource widely supports gene-, SNP-, and disease-related information together
-
Full-text index only
COMUS: Clinician-Oriented locus-specific MUtation detection and deposition System.
PMID 19958500 · PMC2788389 · BMC genomics · 2009 · 8 claims · 6 setups
COMUS is a bioinformatics system for detecting and depositing new mutations from patient DNA with a clinician-friendly interface
-
Full-text index only
GeneSeer: a sage for gene names and genomic resources.
PMID 16176584 · PMC1266031 · BMC genomics · 2005 · 7 claims · 4 setups
GeneSeer aggregates gene name synonyms from GenBank, FlyBase, ExPASy, HUGO, ENSEMBL, UCSC and Gene Ontology into a name-translation database that maps any familiar name to a reference (SOFAR) identifier.
-
Has reproduction · 80
MCPmed: a call for Model Context Protocol-enabled bioinformatics web services for LLM-driven discovery.
PMID 41729821 · PMC12927880 · Briefings in bioinformatics · 2026 · 6 claims · 3 setups
Adapting MCP to bioinformatics web server backends provides a standardized, machine-actionable semantic layer linking API endpoints to scientific concepts and metadata.
-
Full-text index only
Functional importance of different patterns of correlation between adjacent cassette exons in human and mouse.
PMID 18439302 · PMC2432081 · BMC genomics · 2008 · 8 claims · 7 setups
Adjacent cassette exon pairs can be categorized by EST-derived correlation coefficient into three groups: mutually exclusive (ME, r<=-0.7), independent (IND, -0.2<=r<=0.2), and linked (LNK, r>=0.7)
-
Full-text index only
A genome-wide survey of segmental duplications that mediate common human genetic variation of chromosomal architecture.
PMID 15588494 · PMC3525102 · Human genomics · 2004 · 8 claims · 5 setups
PSD-mediated genomic architecture analogous to the 8p23/4p16 inversion regions is not unique to those loci but recurs genome-wide.
-
Full-text index only
Designating eukaryotic orthology via processed transcription units.
PMID 18445630 · PMC2425467 · Nucleic acids research · 2008 · 8 claims · 5 setups
Existing ortholog databases discard/ignore alternative splicing via all-against-all protein comparisons, causing ambiguous ortholog calls and misclassification of AS isoforms as in-paralogs
-
Full-text index only
MutScreener: primer design tool for PCR-direct sequencing.
PMID 16845093 · PMC1538803 · Nucleic acids research · 2006 · 8 claims · 4 setups
MutScreener is a web-based application that automates PCR-direct sequencing assay design by annotating gene structure and designing PCR and sequencing primers.
-
Has reproduction · 75
Roar: detecting alternative polyadenylation with standard mRNA sequencing libraries.
PMID 27756200 · PMC5069797 · BMC bioinformatics · 2016 · 8 claims · 5 setups
Roar, a method using PRE/POST read counts around annotated APA sites to compute an m/M ratio and a ratio-of-ratios (roar) statistic, detects differential 3'UTR shortening/lengthening from standard RNA-seq libraries.
-
Full-text index only
hORFeome v3.1: a resource of human open reading frames representing over 10,000 human genes.
PMID 17207965 · PMC4647941 · Genomics · 2007 · 8 claims · 7 setups
hORFeome v3.1 is a resource of 12,212 cloned human ORFs representing 10,214 genes, a 51% expansion over hORFeome v1.1
-
Full-text index only
A mouse plasma peptide atlas as a resource for disease proteomics.
PMID 18522751 · PMC2481425 · Genome biology · 2008 · 8 claims · 6 setups
A publicly available, high-quality mouse plasma peptide/protein repository (mouse PeptideAtlas) was built from 568 LC-MS/MS runs on four reference plasma pools.
-
Full-text index only
Commonality of functional annotation: a method for prioritization of candidate genes from genome-wide linkage studies.
PMID 18263617 · PMC2275105 · Nucleic acids research · 2008 · 8 claims · 7 setups
Genes correlated with a common complex trait are more likely to share GO functional annotations than genes not correlated with that trait
-
Has reproduction · 86
RNASEQR--a streamlined and accurate RNA-seq sequence analysis program.
PMID 22199257 · PMC3315322 · Nucleic acids research · 2012 · 8 claims · 7 setups
RNASEQR is a new RNA-seq mapper/aligner that combines a BWT-based (Bowtie) transcriptomic/genomic alignment with hash-based BLAT local alignment in three sequential steps: transcriptome mapping, novel exon detection, and anchor-and-align novel splice junction identification.
-
Full-text index only
Pigs in sequence space: a 0.66X coverage pig genome survey based on shotgun sequencing.
PMID 15885146 · PMC1142312 · BMC genomics · 2005 · 8 claims · 7 setups
Pig sequence is closer to human than mouse is, across exons, UTRs, introns, intergenic regions, ultra-conserved elements, and miRNAs
-
Full-text index only
Ensembl 2007.
PMID 17148474 · PMC1761443 · Nucleic acids research · 2007 · 8 claims · 7 setups
Ensembl added 18 new chordate genomes this year, increasing total genomes available from 15 to 33, the largest yearly increase to date.