Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Genome-wide copy number profiling on high-density bacterial artificial chromosomes, single-nucleotide polymorphisms, and oligonucleotide microarrays: a platform comparison based on statistical power analysis.
PMID 17363414 · PMC2779891 · DNA research : an international journal for rapid publication of reports on genes and genomes · 2007 · 8 claims · 6 setups
High-density oligonucleotide/SNP platforms are superior to the BAC platform for genome-wide detection of copy-number variations smaller than 1 Mb
-
Has reproduction · 79
Epigenetic loss of heterogeneity from low to high grade localized prostate tumours.
PMID 34911933 · PMC8674326 · Nature communications · 2021 · 7 claims · 4 setups
Shared chromatin accessibility features among low-grade (Gleason pattern 3) prostate cancer cells are lost in high-grade (Gleason pattern 4) tumours.
-
Full-text index only
Comparing whole genomes using DNA microarrays.
PMID 18347592 · PMC7097741 · Nature reviews. Genetics · 2008 · 8 claims · 6 setups
DNA microarrays offer a relatively inexpensive and efficient alternative to genome sequencing for comparing all known classes of genomic diversity between closely related genomes.
-
Full-text index only
Highly cost-efficient genome-wide association studies using DNA pools and dense SNP arrays.
PMID 18276640 · PMC2346606 · Nucleic acids research · 2008 · 8 claims · 5 setups
Illumina HumanHap300 arrays are substantially more efficient than Affymetrix Genechip HindIII arrays for DNA-pooling based GWAS
-
Full-text index only
CGHPRO -- a comprehensive data analysis tool for array CGH.
PMID 15807904 · PMC1274268 · BMC bioinformatics · 2005 · 8 claims · 3 setups
CGHPRO is a user-friendly, versatile, stand-alone Java tool for normalization, visualization, breakpoint detection and comparative analysis of array-CGH data
-
Full-text index only
A compatible exon-exon junction database for the identification of exon skipping events using tandem mass spectrum data.
PMID 19087293 · PMC2636810 · BMC bioinformatics · 2008 · 6 claims · 6 setups
A theoretical exon-exon junction protein database accounting for all in-phase (frame-preserving) exon combinations can be built from the Ensembl Core Database using Perl/Bioperl/MySQL/Ensembl API.
-
Full-text index only
Functional copy-number alterations in cancer.
PMID 18784837 · PMC2527508 · PloS one · 2008 · 8 claims · 3 setups
RAE is a comprehensive computational framework that robustly maps chromosomal alterations in tumor samples and statistically assesses their functional importance in cancer.
-
Has reproduction · 63
hgtseq: A Standard Pipeline to Study Horizontal Gene Transfer.
PMID 36498841 · PMC9738810 · International journal of molecular sciences · 2022 · 8 claims · 8 setups
hgtseq is a fully automated, portable, and scalable Nextflow/nf-core pipeline for detecting horizontal gene transfer signatures from unmapped sequencing reads.
-
Full-text index only
A re-annotation pipeline for Illumina BeadArrays: improving the interpretation of gene expression data.
PMID 19923232 · PMC2817484 · Nucleic acids research · 2010 · 8 claims · 7 setups
A Perl-based pipeline that BLASTs/BLATs Illumina probe sequences against genomes and transcript databases (RefSeq, UCSC Known Genes, UniGene/GenBank, Ensembl) can classify probes by quality grade (Perfect/Good/Bad/No match) and is applicable across 8 BeadArray platforms and other array types
-
Full-text index only
Pathway analysis for intracellular Porphyromonas gingivalis using a strain ATCC 33277 specific database.
PMID 19723305 · PMC2753363 · BMC microbiology · 2009 · 8 claims · 5 setups
Using the ATCC 33277-specific genome annotation improves proteome coverage (more proteins identified and more abundance ratios calculated) compared to the W83 annotation
-
Full-text index only
A novel optineurin genetic mutation associated with open-angle glaucoma in a Chinese family.
PMID 19710941 · PMC2730747 · Molecular vision · 2009 · 8 claims · 3 setups
A novel missense mutation A1274G (Lys322Glu) in exon 10 of OPTN was identified in affected members of the family
-
Has reproduction · 68
Cell-type annotation with accurate unseen cell-type identification using multiple references.
PMID 37379341 · PMC10335708 · PLoS computational biology · 2023 · 8 claims · 4 setups
mtANN integrates multiple reference datasets and eight gene selection methods via ensemble learning (multiple deep classification models + majority voting) to improve cell-type annotation accuracy
-
Full-text index only
Benchmarking ortholog identification methods using functional genomics data.
PMID 16613613 · PMC1557999 · Genome biology · 2006 · 8 claims · 7 setups
InParanoid is the best overall ortholog identification method for identifying functionally equivalent proteins when sensitivity and selectivity are combined into an overall score.
-
Has reproduction · 63
Creation of a Single Cell RNASeq Meta-Atlas to Define Human Liver Immune Homeostasis.
PMID 34335581 · PMC8322955 · Frontiers in immunology · 2021 · 7 claims · 7 setups
Independent human liver immune scRNA-seq datasets can be combined into an integrated meta-atlas in which all datasets co-cluster, despite differing cell-type proportions between studies.
-
Has reproduction · 83
Hobbes: optimized gram-based methods for efficient read alignment.
PMID 22199254 · PMC3315303 · Nucleic acids research · 2012 · 8 claims · 4 setups
Hobbes, a gram-based short-read mapper supporting Hamming and edit distance, is faster than all other read-mapping programs tested while maintaining high mapping quality.
-
Full-text index only
Codon usage comparison of novel genes in clinical isolates of Haemophilus influenzae.
PMID 15983137 · PMC1160521 · Nucleic acids research · 2005 · 8 claims · 4 setups
A codon usage similarity statistic (ε, based on squared/absolute differences of codon frequencies with an optimized amino acid usage factor) was developed to compare ORFs against a set of 80 reference genomes.
-
Full-text index only
Variation analysis and gene annotation of eight MHC haplotypes: the MHC Haplotype Project.
PMID 18193213 · PMC2206249 · Immunogenetics · 2008 · 8 claims · 6 setups
Comparison of eight HLA-homozygous MHC haplotype sequences identified >44,000 variations (substitutions and indels), submitted to dbSNP
-
Full-text index only
Recent additions and improvements to the Onto-Tools.
PMID 15980579 · PMC1160233 · Nucleic acids research · 2005 · 7 claims · 3 setups
The Onto-Tools back-end database was redesigned around the Entrez Gene data model after NCBI phased out LocusLink in February 2005.
-
Full-text index only
Ensembl 2006.
PMID 16381931 · PMC1347495 · Nucleic acids research · 2006 · 8 claims · 5 setups
Ensembl now provides annotation for 19 genomes, up from 4 the previous year, including new mammalian (Rhesus macaque, Opossum), chordate (Ciona intestinalis), and yeast genomes.
-
Has reproduction · 69
A comparison across non-model animals suggests an optimal sequencing depth for de novo transcriptome assembly.
PMID 23496952 · PMC3655071 · BMC genomics · 2013 · 8 claims · 8 setups
Representative de novo transcriptome assemblies are generated with as few as ~20 million reads for single-tissue samples and ~30 million reads for whole animals at the mRNA-coverage level.