Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
A statistical model to identify differentially expressed proteins in 2D PAGE gels.
PMID 19763172 · PMC2734266 · PLoS computational biology · 2009 · 7 claims · 5 setups
A mixture likelihood model incorporating both detected and non-detected proteins has higher statistical power to detect differential expression than standard approaches like the Student's t-test.
-
Full-text index only
Construction and analysis of tag single nucleotide polymorphism maps for six human-mouse orthologous candidate genes in type 1 diabetes.
PMID 15720714 · PMC551616 · BMC genetics · 2005 · 7 claims · 5 setups
None of the six candidate gene regions showed evidence of association with type 1 diabetes (all multi-locus/single-locus test P values > 0.2)
-
Has reproduction · 65
SPEAQeasy: a scalable pipeline for expression analysis and quantification for R/bioconductor-powered RNA-seq analyses.
PMID 33932985 · PMC8088074 · BMC bioinformatics · 2021 · 8 claims · 5 setups
SPEAQeasy is a portable, easy-to-install, Nextflow-powered RNA-seq processing pipeline that lowers the computational entry barrier for biologists/clinicians
-
Full-text index only
Development of proteomic patterns for detecting lung cancer.
PMID 14757945 · PMC3851077 · Disease markers · 2003 · 8 claims · 3 setups
A decision tree classification algorithm built on three serum protein mass peaks (8122Da, 1452Da, 1610Da) can discriminate lung cancer patients from healthy controls
-
Full-text index only
PDA: Pooled DNA analyzer.
PMID 16643673 · PMC1539032 · BMC bioinformatics · 2006 · 8 claims · 4 setups
No software existed prior to PDA for complete pooled-DNA analysis including data standardization, allele frequency estimation, and single/multipoint association tests
-
Has reproduction · 85
ScLRTC: imputation for single-cell RNA-seq data via low-rank tensor completion.
PMID 34844559 · PMC8628418 · BMC genomics · 2021 · 8 claims · 8 setups
scLRTC imputes dropout entries closest to the original expression values on simulated datasets, outperforming other state-of-the-art methods by SSE and PCC.
-
Full-text index only
Development of animal models to test the fundamental basis of gene-environment interactions.
PMID 19037209 · PMC2703424 · Obesity (Silver Spring, Md.) · 2008 · 8 claims · 8 setups
Selective breeding for low and high intrinsic aerobic treadmill running capacity produced divergent rat lines (LCR and HCR) that contrast in propensity for complex disease
-
Full-text index only
BRCA1 and BRCA2 mutation carriers in the Breast Cancer Family Registry: an open resource for collaborative research.
PMID 18704680 · PMC2775077 · Breast cancer research and treatment · 2009 · 8 claims · 8 setups
The Breast Cancer Family Registry is an open resource for collaborative interdisciplinary and translational studies of the genetic epidemiology of breast cancer.
-
Full-text index only
Needles in the haystack: identifying individuals present in pooled genomic data.
PMID 19798441 · PMC2747273 · PLoS genetics · 2009 · 8 claims · 7 setups
The distribution of T for null samples (individuals not in F or G) deviates strongly from the assumed standard normal, in both location and width.
-
Full-text index only
Onto-Tools: new additions and improvements in 2006.
PMID 17584796 · PMC1933142 · Nucleic acids research · 2007 · 8 claims · 3 setups
OE2GO enables functional profiling for organisms lacking public-domain annotations by allowing users to supply custom GO-format annotation files and OBO-format ontology files
-
Full-text index only
ORFer--retrieval of protein sequences and open reading frames from GenBank and storage into relational databases or text files.
PMID 12493080 · PMC139979 · BMC bioinformatics · 2002 · 6 claims · 6 setups
ORFer retrieves protein and nucleic acid sequences and annotations from NCBI GenBank using the XML sequence format
-
Full-text index only
GoMiner: a resource for biological interpretation of genomic and proteomic data.
PMID 12702209 · PMC154579 · Genome biology · 2003 · 8 claims · 4 setups
GoMiner organizes 'interesting' gene lists (e.g., differentially expressed genes) into the Gene Ontology hierarchy for biological interpretation, displaying results as both a tree and a directed acyclic graph (DAG).
-
Full-text index only
Clustering by neurocognition for fine mapping of the schizophrenia susceptibility loci on chromosome 6p.
PMID 19694819 · PMC4286260 · Genes, brain, and behavior · 2009 · 6 claims · 6 setups
A family-based clustering strategy using neurocognitive test scores (CPT, WCST) can identify more homogeneous subgroups of schizophrenia families for genetic association analysis
-
Full-text index only
Assessing individual differences in genome-wide gene expression in human whole blood: reliability over four hours and stability over 10 months.
PMID 19653838 · PMC3819565 · Twin research and human genetics : the official journal of the International Society for Twin Studies · 2009 · 8 claims · 5 setups
A subset of probesets (3,414) shows 4-hour test-retest reliability exceeding r=0.70 for detecting individual differences in gene expression.
-
Full-text index only
A modified T-test feature selection method and its application on the HapMap genotype data.
PMID 18267305 · PMC5054219 · Genomics, proteomics & bioinformatics · 2007 · 7 claims · 4 setups
A modified t-test ranking measure, extended to handle nominal SNP genotype data via vector transformation, can effectively rank SNPs by their discriminative capability for population classification.
-
Full-text index only
Detection of venous thromboembolism by proteomic serum biomarkers.
PMID 17579716 · PMC1891085 · PloS one · 2007 · 5 claims · 8 setups
A neural network-based classifier built from direct MALDI-TOF MS serum protein expression profiles can diagnose VTE with sensitivity/specificity that exceeds D-dimer assays
-
Full-text index only
PRESTO: rapid calculation of order statistic distributions and multiple-testing adjusted P-values via permutation for one and two-stage genetic association studies.
PMID 18620604 · PMC2483288 · BMC bioinformatics · 2008 · 8 claims · 4 setups
PRESTO is an order of magnitude faster than other existing permutation testing software for genetic association studies.
-
Has reproduction · 58
HGA: de novo genome assembly method for bacterial genomes using high coverage short sequencing reads.
PMID 26945881 · PMC4779561 · BMC genomics · 2016 · 8 claims · 7 setups
HGA leads to significant improvement in assembly quality (N50 and corrected N50) for all 7 evaluated GAGE-B bacterial datasets using most of the 8 evaluated assemblers
-
Full-text index only
Mining expressed sequence tags identifies cancer markers of clinical interest.
PMID 17078886 · PMC1635568 · BMC bioinformatics · 2006 · 8 claims · 6 setups
An EST-mining approach (Fisher Exact Test on tumor vs. non-tumor library hit counts) identifies differentially expressed transcripts with an estimated false discovery rate below 22% when human and mouse screens are combined.
-
Full-text index only
Applications for protein sequence-function evolution data: mRNA/protein expression analysis and coding SNP scoring tools.
PMID 16912992 · PMC1538848 · Nucleic acids research · 2006 · 7 claims · 8 setups
PANTHER HMMs built from family/subfamily multiple sequence alignments can classify novel protein sequences into functional groups based on statistically significant HMM match scores