Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Analyses of deep mammalian sequence alignments and constraint predictions for 1% of the human genome.
PMID 17567995 · PMC1891336 · Genome research · 2007 · 7 claims · 3 setups
Four different alignment methods show large-scale consistency but substantial differences in small-scale rearrangements, sensitivity, and specificity.
-
Full-text index only
Analysis of sequence conservation at nucleotide resolution.
PMID 18166073 · PMC2230682 · PLoS computational biology · 2007 · 8 claims · 4 setups
SCONE (Sequence CONservation Evaluation) is a novel method that estimates evolutionary rate and a neutrality p-value for individual nucleotide positions in a multiple sequence alignment.
-
Full-text index only
Candidate vaccine sequences to represent intra- and inter-clade HIV-1 variation.
PMID 19812689 · PMC2753653 · PloS one · 2009 · 7 claims · 5 setups
Natural CTL immunodominance toward variable proteome regions increases epitope mismatch with challenge strains and recapitulates the escape-driven CTL failure seen in natural infection, contributing to HIV vaccine failure
-
Has reproduction · 83
Public Omics Explorer (POE): Enabling integrative semantic search across GEO omics datasets based on PubMed publications.
PMID 41282419 · PMC12636342 · Computational and structural biotechnology journal · 2025 · 6 claims · 4 setups
POE is a web platform that semantically links GEO datasets and ENA records through their associated PubMed publications for literature-informed dataset retrieval
-
Full-text index only
BioHealthBase: informatics support in the elucidation of influenza virus host pathogen interactions and virulence.
PMID 17965094 · PMC2238987 · Nucleic acids research · 2008 · 7 claims · 5 setups
BioHealthBase BRC is a public integrated bioinformatics database and analysis resource for influenza virus, Francisella tularensis, Mycobacterium tuberculosis, Microsporidia species and ricin toxin.
-
Full-text index only
The specificity and polymorphism of the MHC class I prevents the global adaptation of HIV-1 to the monomorphic proteasome and TAP.
PMID 18949050 · PMC2569417 · PloS one · 2008 · 6 claims · 5 setups
Within individual hosts, proteasome and TAP escape mutations in HIV-1 occur frequently
-
Full-text index only
Tandem repeats modify the structure of human genes hosted in segmental duplications.
PMID 19954527 · PMC2812944 · Genome biology · 2009 · 8 claims · 6 setups
Around 7% of primate-specific genes located within segmental duplications contain variable internal tandem repeats (ITRs).
-
Full-text index only
In silico meets in vivo.
PMID 18304380 · PMC2374716 · Genome biology · 2008 · 8 claims · 8 setups
About 10% of positions in multiple sequence alignments of the human genome with other vertebrate genomes are likely incorrect.
-
Has reproduction · 96
A bioinformatic pipeline for simulating viral integration data.
PMID 35496474 · PMC9046613 · Data in brief · 2022 · 7 claims · 3 setups
A snakemake-based pipeline was developed to simulate integration of a viral or vector genome into a host genome, including sub-genomic fragment integration, structural variation, and host-site deletions.
-
Has reproduction · 77
Comparison of RNA-Seq by poly (A) capture, ribosomal RNA depletion, and DNA microarray for expression profiling.
PMID 24888378 · PMC4070569 · BMC genomics · 2014 · 8 claims · 8 setups
Ribo-Zero-Seq removes rRNA with efficiency comparable to poly(A)-based mRNA-Seq in both FF and FFPE RNA, whereas DSN-Seq leaves significantly more rRNA and shows greater variation.
-
Full-text index only
Designating eukaryotic orthology via processed transcription units.
PMID 18445630 · PMC2425467 · Nucleic acids research · 2008 · 8 claims · 5 setups
Existing ortholog databases discard/ignore alternative splicing via all-against-all protein comparisons, causing ambiguous ortholog calls and misclassification of AS isoforms as in-paralogs