Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Discovery of human inversion polymorphisms by comparative analysis of human and chimpanzee DNA sequence assemblies.
PMID 16254605 · PMC1270012 · PLoS genetics · 2005 · 8 claims · 6 setups
Comparative net alignment of human and chimpanzee genome assemblies identifies 1,576 putative inverted regions covering more than 154 Mb of DNA
-
Full-text index only
SpliceMiner: a high-throughput database implementation of the NCBI Evidence Viewer for microarray splice variant analysis.
PMID 17338820 · PMC1839109 · BMC bioinformatics · 2007 · 6 claims · 4 setups
EVDB is a comprehensive, non-redundant relational database of known human splice variants built from NCBI Entrez Gene and Evidence Viewer data
-
Full-text index only
Ensembl 2006.
PMID 16381931 · PMC1347495 · Nucleic acids research · 2006 · 8 claims · 5 setups
Ensembl now provides annotation for 19 genomes, up from 4 the previous year, including new mammalian (Rhesus macaque, Opossum), chordate (Ciona intestinalis), and yeast genomes.
-
Full-text index only
Identifying related L1 retrotransposons by analyzing 3' transduced sequences.
PMID 12734010 · PMC156586 · Genome biology · 2003 · 8 claims · 6 setups
L1 elements with transduction-derived 3' sequence (L1-TDs) can be computationally identified using RepeatMasker/TSDfinder and grouped into families sharing a common progenitor via BLAST comparison of downstream sequences.
-
Full-text index only
Human-zebrafish non-coding conserved elements act in vivo to regulate transcription.
PMID 16179648 · PMC1236720 · Nucleic acids research · 2005 · 8 claims · 4 setups
Deeply conserved human-zebrafish non-coding elements are enriched for in vivo cis-acting transcriptional regulatory activity.
-
Full-text index only
Retropseudogenes derived from the human Ro/SS-A autoantigen-associated hY RNAs.
PMID 15817567 · PMC1074747 · Nucleic acids research · 2005 · 8 claims · 8 setups
966 pseudogenes derived from the four human Y (hY) RNAs were characterized in the human genome
-
Full-text index only
Genome-wide analyses of retrogenes derived from the human box H/ACA snoRNAs.
PMID 17175533 · PMC1802619 · Nucleic acids research · 2007 · 8 claims · 6 setups
202 novel box H/ACA RNA-related sequences were identified in the human genome
-
Full-text index only
Systematic identification of pseudogenes through whole genome expression evidence profiling.
PMID 16945953 · PMC1636364 · Nucleic acids research · 2006 · 8 claims · 8 setups
Developed a novel bioinformatics method that identifies pseudogenes by profiling whole-genome transcript and protein expression evidence
-
Full-text index only
Human SNPs resulting in premature stop codons and protein truncation.
PMID 16595072 · PMC3500177 · Human genomics · 2006 · 8 claims · 6 setups
Genome-wide screening of dbSNP identified 28 validated X-SNPs from 28 genes with known minor allele frequencies.
-
Full-text index only
Evolution of the NANOG pseudogene family in the human and chimpanzee genomes.
PMID 16469101 · PMC1457002 · BMC evolutionary biology · 2006 · 7 claims · 5 setups
The NANOG gene and all pseudogenes except NANOGP8 occupy orthologous chromosomal positions in the chimpanzee genome, indicating they originated before the human-chimpanzee divergence.
-
Full-text index only
CONTRAST: a discriminative, phylogeny-free approach to multiple informant de novo gene prediction.
PMID 18096039 · PMC2246271 · Genome biology · 2007 · 8 claims · 5 setups
CONTRAST predicts exact coding region structures for 65% more human genes than the previous state-of-the-art de novo predictor (N-SCAN)
-
Full-text index only
F-SNP: computationally predicted functional SNPs for disease association studies.
PMID 17986460 · PMC2238878 · Nucleic acids research · 2008 · 6 claims · 8 setups
F-SNP is a database integrating functional effect predictions for SNPs from 16 bioinformatics tools/databases across four categories: splicing, transcription, translation, and post-translation
-
Full-text index only
Efficacy assessment of SNP sets for genome-wide disease association studies.
PMID 17726055 · PMC2034459 · Nucleic acids research · 2007 · 6 claims · 4 setups
τ, derived from Shannon entropy and swept radius ɛ, approximates the relative sample size efficiency of a marker set for mapping a causal variant at a given map position compared to a maximally polymorphic SNP
-
Full-text index only
Genomic analysis of the chromosome 15q11-q13 Prader-Willi syndrome region and characterization of transcripts for GOLGA8E and WHCD1L1 from the proximal breakpoint region.
PMID 18226259 · PMC2268926 · BMC genomics · 2008 · 8 claims · 7 setups
GOLGA8E and WHDC1L1 are characterized for the first time as protein-coding transcripts from the PWS proximal breakpoint region.
-
Full-text index only
Integrative analysis of the human cis-antisense gene pairs, miRNAs and their transcription regulation patterns.
PMID 19906709 · PMC2811022 · Nucleic acids research · 2010 · 8 claims · 5 setups
A genome-wide catalog of up to ~9000 overlapping antisense loci (23,782 non-redundant SAT pairs, clustered into 8894) was compiled and stored in the USAGP database
-
Full-text index only
Ensembl 2008.
PMID 18000006 · PMC2238821 · Nucleic acids research · 2008 · 8 claims · 6 setups
The Ensembl regulatory build integrates multiple genome-wide functional genomics datasets to automatically annotate regulatory regions and assign putative functions across the genome.
-
Full-text index only
Ensembl 2009.
PMID 19033362 · PMC2686571 · Nucleic acids research · 2009 · 8 claims · 6 setups
Ensembl provides comprehensive, consistently annotated genome information for chordate genomes with automatically generated genesets and comparative genomics data
-
Full-text index only
Ensembl 2005.
PMID 15608235 · PMC540092 · Nucleic acids research · 2005 · 8 claims · 4 setups
Ensembl's automatic gene build system can flexibly and reliably annotate a wide variety of genomes with limited species-specific evidence.
-
Full-text index only
Statistical Viewer: a tool to upload and integrate linkage and association data as plots displayed within the Ensembl genome browser.
PMID 15826305 · PMC1087836 · BMC bioinformatics · 2005 · 8 claims · 3 setups
Statistical Viewer is a plug-in package for Ensembl that displays disease study-specific linkage and/or association data as 2D plots within Ensembl's Contig View and Cyto View pages.
-
Has reproduction · 89
Chemist: A Domain-Specific Language by Chemists for Chemists.
PMID 40815845 · PMC12400401 · The journal of physical chemistry. A · 2025 · 8 claims · 2 setups
Interpackage modules are rare for the most computationally expensive QC algorithms (integral transformations, Fock builds, sigma vector formation) because their APIs are difficult to define using general-purpose programming language (GPPL) types alone.