Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Human epigenome project--up and running.
PMID 14691553 · PMC300691 · PLoS biology · 2003 · 7 claims · 4 setups
Epigenetic modifications (e.g., DNA methylation) rather than DNA sequence differences explain phenotypic differences between genetically identical individuals, such as monozygotic twins or inbred mice.
-
Full-text index only
High-throughput chromatin information enables accurate tissue-specific prediction of transcription factor binding sites.
PMID 18988630 · PMC2662491 · Nucleic acids research · 2009 · 8 claims · 8 setups
Incorporating H3K4me3 chromatin modification estimates greatly improves the accuracy of in silico prediction of in vivo TF binding for a wide range of TFs in human and mouse
-
Full-text index only
Using structural bioinformatics to investigate the impact of non synonymous SNPs and disease mutations: scope and limitations.
PMID 19758473 · PMC2745591 · BMC bioinformatics · 2009 · 8 claims · 8 setups
None of 39 tested structural properties can be used as a sole classification criterion to separate neutral SNPs from disease mutations.
-
Has reproduction · 83
Hobbes: optimized gram-based methods for efficient read alignment.
PMID 22199254 · PMC3315303 · Nucleic acids research · 2012 · 8 claims · 4 setups
Hobbes, a gram-based short-read mapper supporting Hamming and edit distance, is faster than all other read-mapping programs tested while maintaining high mapping quality.
-
Full-text index only
A parsimony approach to biological pathway reconstruction/inference for genomes and metagenomes.
PMID 19680427 · PMC2714467 · PLoS computational biology · 2009 · 8 claims · 6 setups
The naïve mapping approach (present if ≥1 associated function is found) leads to an inflated estimate of biological pathways and overestimates functional diversity of a sample.
-
Full-text index only
SNP-VISTA: an interactive SNP visualization tool.
PMID 16336665 · PMC1325058 · BMC bioinformatics · 2005 · 7 claims · 3 setups
SNP-VISTA is an interactive Java-based visualization tool with two versions, GeneSNP-VISTA and EcoSNP-VISTA, for exploring large-scale SNP datasets
-
Has reproduction · 66
HTSstation: a web application and open-access libraries for high-throughput sequencing data analysis.
PMID 24475057 · PMC3903476 · PloS one · 2014 · 8 claims · 5 setups
HTSstation is a web application suite coupling simple web forms to modular analysis pipelines for ChIP-seq, RNA-seq, 4C-seq and re-sequencing HTS applications, accessible at http://htsstation.epfl.ch.
-
Has reproduction · 80
PanglaoDB: a web server for exploration of mouse and human single-cell RNA sequencing data.
PMID 30951143 · PMC6450036 · Database : the journal of biological databases and curation · 2019 · 7 claims · 7 setups
PanglaoDB is a web server providing pre-processed and pre-computed analyses of >1054 single-cell experiments (>4 million cells) from mouse and human across many tissues and platforms.
-
Full-text index only
The distribution of SNPs in human gene regulatory regions.
PMID 16209714 · PMC1260019 · BMC genomics · 2005 · 8 claims · 6 setups
SNPs occur with higher density closer to the transcriptional start site within gene promoter regions than in further upstream regions
-
Full-text index only
PPC: an algorithm for accurate estimation of SNP allele frequencies in small equimolar pools of DNA using data from high density microarrays.
PMID 16199750 · PMC1240117 · Nucleic acids research · 2005 · 7 claims · 6 setups
The PPC algorithm, which applies a probe-pair-specific second-degree polynomial correction, increases the accuracy of allele frequency estimates from pooled DNA compared with previously described algorithms
-
Full-text index only
Animal models of gene-nutrient interactions.
PMID 19037208 · PMC2703433 · Obesity (Silver Spring, Md.) · 2008 · 8 claims · 5 setups
Mice and rats are well-suited models for human food selection because they share food preferences with humans and are supported by extensive genetic tools (sequenced genome, inbred strains, gene targeting, Cre-lox).
-
Full-text index only
Effect of read-mapping biases on detecting allele-specific expression from RNA-sequencing data.
PMID 19808877 · PMC2788925 · Bioinformatics (Oxford, England) · 2009 · 8 claims · 6 setups
Reads mapped to the reference genome show a significant bias toward the reference allele at heterozygous SNPs
-
Has reproduction · 86
Multi-INTACT: integrative analysis of the genome, transcriptome, and proteome identifies causal mechanisms of complex traits.
PMID 39901160 · PMC11789355 · Genome biology · 2025 · 8 claims · 2 setups
Multi-INTACT achieves higher power than existing single-gene-product methods while maintaining calibrated false discovery rates in simulations.
-
Full-text index only
Identification of disease causing loci using an array-based genotyping approach on pooled DNA.
PMID 16197552 · PMC1262713 · BMC genomics · 2005 · 8 claims · 5 setups
Pooling genomic DNA and genotyping on SNP microarrays accurately predicts allelic frequencies relative to individual genotyping
-
Has reproduction · 67
Sequencing mRNA from cryo-sliced Drosophila embryos to determine genome-wide spatial patterns of gene expression.
PMID 23951250 · PMC3741199 · PloS one · 2013 · 8 claims · 8 setups
Cryosectioning single blastoderm-stage D. melanogaster embryos along the A–P axis and sequencing mRNA from each slice yields reliable genome-wide spatial expression patterns.
-
Full-text index only
Correlation of microsynteny conservation and disease gene distribution in mammalian genomes.
PMID 19909546 · PMC2779822 · BMC genomics · 2009 · 7 claims · 8 setups
Density of mouse orthologs of human disease genes correlates with regions of conserved microsynteny in the mouse genome
-
Full-text index only
Long-range regulation is a major driving force in maintaining genome integrity.
PMID 19682388 · PMC2741452 · BMC evolutionary biology · 2009 · 7 claims · 5 setups
Long-range transcriptional regulation is a major driving force in maintaining genome integrity by constraining where chromosomal breakpoints can become fixed.
-
Full-text index only
Exploring the immunome: A brave new world for human vaccine development.
PMID 20009527 · PMC2919815 · Human vaccines · 2009 · 7 claims · 7 setups
Screening the Mtb proteome in silico for epitopes ('fishing for antigens using epitopes as bait') revealed a remarkable diversity of human immune responses to Mtb proteins without an ascribed function, suggesting human immune response to Mtb is omnivorous rather than focused on single immunodominant proteins.
-
Has reproduction · 86
RNASEQR--a streamlined and accurate RNA-seq sequence analysis program.
PMID 22199257 · PMC3315322 · Nucleic acids research · 2012 · 8 claims · 7 setups
RNASEQR is a new RNA-seq mapper/aligner that combines a BWT-based (Bowtie) transcriptomic/genomic alignment with hash-based BLAT local alignment in three sequential steps: transcriptome mapping, novel exon detection, and anchor-and-align novel splice junction identification.
-
Full-text index only
Assessing the genomic evidence for conserved transcribed pseudogenes under selection.
PMID 19754956 · PMC2753554 · BMC genomics · 2009 · 8 claims · 8 setups
1750 transcribed pseudogene annotations (TPAs) were identified in the human genome, ~11.5% of all human pseudogene annotations.