Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 82
Landscape of allele-specific transcription factor binding in the human genome.
PMID 33980847 · PMC8115691 · Nature communications · 2021 · 8 claims · 6 setups
A novel statistical framework (ADASTRA) calls allele-specific TF binding from existing ChIP-Seq alignments by jointly correcting for background allelic dosage (BAD, from aneuploidy/CNVs) and reference mapping bias.
-
Full-text index only
Empirical codon substitution matrix.
PMID 15927081 · PMC1173088 · BMC bioinformatics · 2005 · 8 claims · 5 setups
The authors present the first empirical codon substitution matrix built entirely from alignments of vertebrate coding DNA sequences.
-
Full-text index only
snoSeeker: an advanced computational package for screening of guide and orphan snoRNA genes in the human genome.
PMID 16990247 · PMC1636440 · Nucleic acids research · 2006 · 8 claims · 5 setups
snoSeeker (comprising CDseeker and ACAseeker) is a computational package that can screen for both guide and orphan snoRNA genes, unlike prior programs limited to guide snoRNAs
-
Full-text index only
Using ESTs to improve the accuracy of de novo gene prediction.
PMID 16817966 · PMC1534067 · BMC bioinformatics · 2006 · 8 claims · 8 setups
TWINSCAN_EST combines EST alignments with TWINSCAN via a trainable 'ESTseq' representation and improves exact gene structure prediction accuracy on the whole C. elegans genome
-
Full-text index only
MACSIMS: multiple alignment of complete sequences information management system.
PMID 16792820 · PMC1539025 · BMC bioinformatics · 2006 · 8 claims · 5 setups
MACSIMS is a multiple alignment-based information management system combining knowledge-based database mining with ab initio sequence predictions
-
Full-text index only
Multiple whole genome alignments and novel biomedical applications at the VISTA portal.
PMID 17488840 · PMC1933192 · Nucleic acids research · 2007 · 8 claims · 4 setups
A novel multiple whole-genome alignment algorithm treats all genomes symmetrically, avoiding dependence on a single base/reference genome
-
Full-text index only
Patterns of evolutionary constraints on genes in humans.
PMID 18840274 · PMC2587479 · BMC evolutionary biology · 2008 · 7 claims · 6 setups
BaseDiver, a novel framework integrating GERP score and derived allele frequency (DAF) at nonsynonymous coding SNPs, can classify GO functional categories by patterns of evolutionary constraint
-
Full-text index only
Discovery of novel human transcript variants by analysis of intronic single-block EST with polyadenylation site.
PMID 19906316 · PMC2784480 · BMC genomics · 2009 · 8 claims · 7 setups
Intronic single-block ESTs with poly(A/T) tails reveal previously unidentified novel transcript variants missed by existing databases.
-
Full-text index only
CapsID: a web-based tool for developing parsimonious sets of CAPS molecular markers for genotyping.
PMID 16686952 · PMC1471797 · BMC genetics · 2006 · 7 claims · 1 setups
CapsID identifies snip-SNPs (SNPs that alter restriction endonuclease recognition sites) within reference sequence alignments and designs PCR primers around them
-
Full-text index only
Discovery of human inversion polymorphisms by comparative analysis of human and chimpanzee DNA sequence assemblies.
PMID 16254605 · PMC1270012 · PLoS genetics · 2005 · 8 claims · 6 setups
Comparative net alignment of human and chimpanzee genome assemblies identifies 1,576 putative inverted regions covering more than 154 Mb of DNA
-
Full-text index only
Heterogeneous genomic molecular clocks in primates.
PMID 17029560 · PMC1592237 · PLoS genetics · 2006 · 7 claims · 7 setups
Non-CpG site substitutions show clear generation-time dependency, consistent with a replication-error origin
-
Full-text index only
Vertebrate gene finding from multiple-species alignments using a two-level strategy.
PMID 16925840 · PMC1810555 · Genome biology · 2006 · 8 claims · 5 setups
DOGFISH cleanly separates a multi-species alignment classifier (RVM cascade) from an HMM-based structure predictor, avoiding tight coupling of alignment complexity with HMM formalism
-
Full-text index only
Analysis of chimpanzee history based on genome sequence alignments.
PMID 18421364 · PMC2278377 · PLoS genetics · 2008 · 8 claims · 6 setups
Bonobos and common chimpanzees separated approximately 1.29 million years ago
-
Full-text index only
PLANdbAffy: probe-level annotation database for Affymetrix expression microarrays.
PMID 19906711 · PMC2808952 · Nucleic acids research · 2010 · 6 claims · 4 setups
PLANdbAffy is a database of Affymetrix probe alignments to the human genome for five widely used arrays (HG-U133A, HG-U133B, HG-U133 Plus 2.0, Human Exon 1.0, Human Gene 1.0)
-
Full-text index only
Developments in CORG: a gene-centric comparative genomics resource.
PMID 17135197 · PMC1751536 · Nucleic acids research · 2007 · 7 claims · 4 setups
CORG provides pairwise and multiple sequence alignments of upstream promoter regions and whole gene loci across 10 vertebrate species.
-
Full-text index only
Pairagon+N-SCAN_EST: a model-based gene annotation pipeline.
PMID 16925839 · PMC1810554 · Genome biology · 2006 · 7 claims · 5 setups
Pairagon+N-SCAN_EST, using only native alignments, was as accurate as ENSEMBL and ExoGean in the EGASP mRNA/EST evidence assessment
-
Full-text index only
Analysis of sequence conservation at nucleotide resolution.
PMID 18166073 · PMC2230682 · PLoS computational biology · 2007 · 8 claims · 4 setups
SCONE (Sequence CONservation Evaluation) is a novel method that estimates evolutionary rate and a neutrality p-value for individual nucleotide positions in a multiple sequence alignment.
-
Full-text index only
How accurately is ncRNA aligned within whole-genome multiple alignments?
PMID 17963514 · PMC2206062 · BMC bioinformatics · 2007 · 7 claims · 4 setups
MULTIZ does a fairly accurate job of aligning ncRNA regions across 17 vertebrate genomes, but better alignments exist in some regions.
-
Full-text index only
Much ado about nothing: modeling amino acid replacement with predicted protein structures.
PMID 42036821 · PMC13171170 · Bioinformatics (Oxford, England) · 2026 · 8 claims · 7 setups
AFSM was constructed from over 660,000 structural alignments across ~21,000 proteins (297 InterPro families), following the BLOSUM log-odds methodology.
-
Full-text index only
Lift&Add-rapid and robust addition of new species to alignments of conserved non-coding sequences.
PMID 42203687 · PMC13224966 · Bioinformatics (Oxford, England) · 2026 · 7 claims · 5 setups
Lift&Add, a Snakemake/bash workflow combining UCSC liftOver, Liftoff, and MAFFT, enables rapid addition of new genome sequences to existing multi-species alignments of conserved elements without requiring new whole-genome alignments.