Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
The DAVID Gene Functional Classification Tool: a novel biological module-centric algorithm to functionally analyze large gene lists.
PMID 17784955 · PMC2375021 · Genome biology · 2007 · 8 claims · 6 setups
Gene-gene functional similarity can be measured using kappa statistics applied to a binary gene-annotation-term matrix built from 14 annotation categories.
-
Full-text index only
The sequence and de novo assembly of the giant panda genome.
PMID 20010809 · PMC3951497 · Nature · 2010 · 8 claims · 8 setups
A draft giant panda genome was successfully generated and assembled de novo using only Illumina Genome Analyser short-read sequencing
-
Full-text index only
Genome-wide prioritization of disease genes and identification of disease-disease associations from an integrated human functional linkage network.
PMID 19728866 · PMC2768980 · Genome biology · 2009 · 6 claims · 6 setups
Integrating 16 genomic features (32 sub-features) via a naïve Bayes classifier produces a genome-scale FLN of 21,657 human genes and 22,388,609 weighted links that outperforms any individual data source for inferring functional linkages.
-
Full-text index only
POCUS: mining genomic sequence annotation to predict disease genes.
PMID 14611661 · PMC329128 · Genome biology · 2003 · 8 claims · 6 setups
Genes predisposing to the same disease tend to share functional annotation IDs (GO/InterPro) more than expected by chance
-
Full-text index only
FeatureScan: revealing property-dependent similarity of nucleotide sequences.
PMID 16845077 · PMC1538849 · Nucleic acids research · 2006 · 6 claims · 5 setups
FeatureScan transforms nucleotide sequences into numerical signals of physico-chemical/conformational properties and compares them via a convolution/correlation (Fourier transform) method rather than comparing letters
-
Has reproduction · 87
High-resolution mapping of transcriptional dynamics across tissue development reveals a stable mRNA-tRNA interface.
PMID 25122613 · PMC4216921 · Genome research · 2014 · 8 claims · 7 setups
mRNA codon and amino acid pools are highly stable across mouse development and across tissues, simply reflecting the genomic background distribution of any possible transcriptome.
-
Full-text index only
MatchMiner: a tool for batch navigation among gene and gene product identifiers.
PMID 12702208 · PMC154578 · Genome biology · 2003 · 8 claims · 3 setups
MatchMiner's LookUp function automates batch translation of an input list of gene identifiers into a matching list of a different identifier type.
-
Has reproduction · 23
Analysis of whole-genome re-sequencing data of ducks reveals a diverse demographic history and extensive gene flow between Southeast/South Asian and Chinese populations.
PMID 33849442 · PMC8042899 · Genetics, selection, evolution : GSE · 2021 · 8 claims · 8 setups
Whole-genome resequencing reveals three geographically distinct genetic groups: local Chinese, wild, and local Southeast/South Asian duck populations
-
Full-text index only
INDELSCAN: a web server for comparative identification of species-specific and non-species-specific insertion/deletion events.
PMID 17517762 · PMC1933116 · Nucleic acids research · 2007 · 8 claims · 3 setups
Pair-wise sequence alignment-based indel identification lacks discrimination of species specificity and cannot distinguish insertions from deletions.
-
Has reproduction · 76
What the Phage: a scalable workflow for the identification and analysis of phage sequences.
PMID 36399058 · PMC9673492 · GigaScience · 2022 · 8 claims · 7 setups
WtP combines 11 tools (14 approaches) for phage prediction in a parallel, containerized Nextflow workflow
-
Full-text index only
In vitro and in silico analysis reveals an efficient algorithm to predict the splicing consequences of mutations at the 5' splice sites.
PMID 17726045 · PMC2094079 · Nucleic acids research · 2007 · 8 claims · 6 setups
Two exonic mutations, PINK1 E417G and PARK7 E64D, disrupt binding to U1 snRNA and cause skipping of the mutation-harboring exon
-
Full-text index only
Sequence analysis of p53 response-elements suggests multiple binding modes of the p53 tetramer to DNA targets.
PMID 17439973 · PMC1888811 · Nucleic acids research · 2007 · 8 claims · 5 setups
p53REs are not simple direct repeats of half-sites; the two half-sites couple to form a higher-order 20-bp full-site palindrome
-
Full-text index only
Comparative genomic study reveals a transition from TA richness in invertebrates to GC richness in vertebrates at CpG flanking sites: an indication for context-dependent mutagenicity of methylated CpG sites.
PMID 19329065 · PMC5054122 · Genomics, proteomics & bioinformatics · 2008 · 8 claims · 8 setups
Nucleotide preference at CpG flanking sites transitions from 5' T (invertebrates) to 5' A (vertebrates) at the invertebrate-vertebrate boundary
-
Has reproduction · 75
A step forward for Shiga toxin-producing Escherichia coli identification and characterization in raw milk using long-read metagenomics.
PMID 36748417 · PMC9836091 · Microbial genomics · 2022 · 8 claims · 6 setups
Long-read metagenomics enables isolation-independent identification and characterization of eae-positive STEC directly from raw milk.
-
Full-text index only
Expression profiling of drug response--from genes to pathways.
PMID 17117610 · PMC3181826 · Dialogues in clinical neuroscience · 2006 · 8 claims · 8 setups
Understanding individual response to a drug (efficacy/tolerability) is the major bottleneck in current drug development and clinical trials.
-
Has reproduction · 44
Detecting DNA modifications from SMRT sequencing data by modeling sequence context dependence of polymerase kinetic.
PMID 23516341 · PMC3597545 · PLoS computational biology · 2013 · 8 claims · 7 setups
Local sequence context strongly determines position-specific polymerase kinetic rate: roughly 80% of IPD variation is explained by a 10 bp context (7 bases upstream, 2 bases downstream of the incorporation site), saturating at 7 bases upstream.
-
Full-text index only
Speeding disease gene discovery by sequence based candidate prioritization.
PMID 15766383 · PMC1274252 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Disease genes (OMIM) differ significantly from non-disease genes in sequence-based features including gene/cDNA/protein size, exon number, homolog conservation, secretion signal, 3' UTR length, CpG islands, and distance to nearest gene.