Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Versatile and open software for comparing large genomes.
PMID 14759262 · PMC395750 · Genome biology · 2004 · 8 claims · 8 setups
MUMmer 3.0 efficiently handles comparisons of large eukaryotic genomes at varying evolutionary distances
-
Has reproduction · 71
Hyb: a bioinformatics pipeline for the analysis of CLASH (crosslinking, ligation and sequencing of hybrids) data.
PMID 24211736 · PMC3969109 · Methods (San Diego, Calif.) · 2014 · 8 claims · 6 setups
The 'hyb' pipeline detects, calls, folds and annotates chimeric reads from CLASH high-throughput sequencing data.
-
Full-text index only
PPC: an algorithm for accurate estimation of SNP allele frequencies in small equimolar pools of DNA using data from high density microarrays.
PMID 16199750 · PMC1240117 · Nucleic acids research · 2005 · 7 claims · 6 setups
The PPC algorithm, which applies a probe-pair-specific second-degree polynomial correction, increases the accuracy of allele frequency estimates from pooled DNA compared with previously described algorithms
-
Full-text index only
L2L: a simple tool for discovering the hidden significance in microarray expression data.
PMID 16168088 · PMC1242216 · Genome biology · 2005 · 8 claims · 4 setups
L2L systematically compares a user's differentially expressed gene list against a database of published differentially expressed gene lists to find statistically significant overlaps and generate hypotheses about shared mechanisms
-
Full-text index only
Evaluation of multiple displacement amplification in a 5 cM STR genome-wide scan.
PMID 16055919 · PMC1182175 · Nucleic acids research · 2005 · 7 claims · 5 setups
MDA genotyping call rates and accuracy are only marginally lower than for genomic DNA
-
Full-text index only
Genome-wide scans for loci under selection in humans.
PMID 16004726 · PMC3525256 · Human genomics · 2005 · 8 claims · 4 setups
Natural selection and population demographic history both distort patterns of genetic variation relative to the standard neutral model, so single-locus tests cannot unambiguously distinguish selection from demography.
-
Full-text index only
Satellog: a database for the identification and prioritization of satellite repeats in disease association studies.
PMID 15949044 · PMC1181805 · BMC bioinformatics · 2005 · 7 claims · 6 setups
Satellog is a database cataloging all pure 1-16 unit satellite repeats in the human genome with supplementary polymorphism, gene-location, and expression data for prioritizing repeats in disease-association studies.
-
Full-text index only
Genome comparison without alignment using shortest unique substrings.
PMID 15910684 · PMC1166540 · BMC bioinformatics · 2005 · 8 claims · 8 setups
A number of sequence comparison tasks, including detection of unique genomic regions, can be accomplished efficiently without an alignment step using shortest unique substrings.
-
Full-text index only
Predictive screening for regulators of conserved functional gene modules (gene batteries) in mammals.
PMID 15882449 · PMC1134656 · BMC genomics · 2005 · 8 claims · 4 setups
A predictive computational screen covering ~40% of annotated protein-coding genes identified 21 co-expressed gene clusters with statistically supported sharing of cis-regulatory motifs.
-
Full-text index only
Investigating hookworm genomes by comparative analysis of two Ancylostoma species.
PMID 15854223 · PMC1112591 · BMC genomics · 2005 · 8 claims · 8 setups
Nearly 20,000 ESTs from 7 cDNA libraries define nearly 7,000 hookworm genes across A. caninum and A. ceylanicum
-
Full-text index only
Quadratic regression analysis for gene discovery and pattern recognition for non-cyclic short time-course microarray experiments.
PMID 15850479 · PMC1127068 · BMC bioinformatics · 2005 · 8 claims · 8 setups
A step-down quadratic regression method (fitting quadratic, then linear, then null models per gene) identifies differentially expressed genes and classifies them into 9 temporal expression patterns using continuous time information.
-
Full-text index only
Silhouette scores for assessment of SNP genotype clusters.
PMID 15760469 · PMC555759 · BMC genomics · 2005 · 7 claims · 5 setups
Silhouette scores provide a relevant, objective numeric measure of SNP genotype cluster quality, condensing tightness and separation into a single value from -1.0 to 1.0.
-
Full-text index only
Low density DNA microarray for detection of most frequent TP53 missense point mutations.
PMID 15713227 · PMC553977 · BMC biotechnology · 2005 · 7 claims · 6 setups
A double tandem hybridization microarray with 7-nt gap targets and 7-mer capture probes can search for 9 point mutations in TP53 codons 248, 249, and 273
-
Full-text index only
Sample preparation for serum/plasma profiling and biomarker identification by mass spectrometry.
PMID 17166507 · PMC7094463 · Journal of chromatography. A · 2007 · 8 claims · 8 setups
Standardizing sample preparation procedures for serum/plasma profiling is critical for obtaining reliable biomarkers, since slight procedural changes can produce very different protein profiles.
-
Full-text index only
PeroxisomeDB: a database for the peroxisomal proteome, functional genomics and disease.
PMID 17135190 · PMC1747181 · Nucleic acids research · 2007 · 8 claims · 6 setups
PeroxisomeDB integrates the complete peroxisomal proteome of Homo sapiens and Saccharomyces cerevisiae into interrelated 'Genes', 'Functions', 'Metabolic pathways' and 'Diseases' sections with links to NCBI, ENSEMBL and UCSC
-
Full-text index only
CancerGenes: a gene selection resource for cancer genome projects.
PMID 17088289 · PMC1781153 · Nucleic acids research · 2007 · 6 claims · 4 setups
CancerGenes is a gene list-centric web resource that combines expert-annotated gene lists with data from public databases (Entrez Gene, Ensembl BioMart, Kim et al. promoter data, Sanger COSMIC) to support gene selection for cancer re-sequencing projects.
-
Has reproduction · 73
Vespucci: a system for building annotated databases of nascent transcripts.
PMID 24304890 · PMC3936758 · Nucleic acids research · 2014 · 8 claims · 7 setups
Existing ChIP-seq and RNA-seq analysis platforms (e.g. Cufflinks, peak callers) are unsuited to GRO-seq because they assume spliced/exonic reads, uniform density and paired-end data, and cannot identify transcriptional units de novo across the whole genome.
-
Full-text index only
V-MitoSNP: visualization of human mitochondrial SNPs.
PMID 16907992 · PMC1564046 · BMC bioinformatics · 2006 · 6 claims · 4 setups
V-MitoSNP integrates RFLP genotyping information with mitochondria-related cancer/disease data in a user-friendly, interactive, color-coded visual web interface
-
Full-text index only
Predicting candidate genes for human deafness disorders: a bioinformatics approach.
PMID 16854223 · PMC1564145 · BMC genomics · 2006 · 8 claims · 4 setups
A bioinformatic approach combining expression databases and protein interaction data narrows ~2400 candidate genes across deafness loci to a manageable set of candidates.
-
Full-text index only
TFBScluster web server for the identification of mammalian composite regulatory elements.
PMID 16845063 · PMC1538905 · Nucleic acids research · 2006 · 7 claims · 5 setups
TFBScluster is a web server that identifies genome-wide clusters of TFBSs conserved in multiple mammalian species using human or mouse as the reference genome.