Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Software for tag single nucleotide polymorphism selection.
PMID 16004730 · PMC3525260 · Human genomics · 2005 · 8 claims · 3 setups
Pairwise R2 methods tend to pick more tagging SNPs than strictly needed because they miss redundancy where two or more tag SNPs jointly predict an untagged SNP with no single direct surrogate.
-
Full-text index only
Differential protein expression in human corneal endothelial cells cultured from young and older donors.
PMID 18852868 · PMC2565687 · Molecular vision · 2008 · 8 claims · 5 setups
Cultured HCEC protein profiles show age-related differences between young and older donors
-
Full-text index only
VarDetect: a nucleotide sequence variation exploratory tool.
PMID 19091032 · PMC2638149 · BMC bioinformatics · 2008 · 8 claims · 2 setups
VarDetect is a stand-alone software tool that automatically detects nucleotide variation (SNPs) from fluorescence-based chromatogram traces using pre-calculated peak content ratios and artifact-handling rules.
-
Has reproduction · 74
Software JimenaE allows efficient dynamic simulations of Boolean networks, centrality and system state analysis.
PMID 36725967 · PMC9892028 · Scientific reports · 2023 · 8 claims · 4 setups
JimenaE simulates Boolean networks dynamically and systematically calculates all system states rather than heuristically as SQUAD does.
-
Full-text index only
SNP-RFLPing: restriction enzyme mining for SNPs in genomes.
PMID 16503968 · PMC1386656 · BMC genomics · 2006 · 8 claims · 2 setups
SNP-RFLPing accepts three flexible input types (dbSNP rs#/ss# IDs, HUGO gene name/Entrez gene ID, or free-form SNP-in-sequence including IUPAC or [dNTP1/dNTP2] formats) for human, rat, and mouse genomes
-
Has reproduction · 67
Cyrface: An interface from Cytoscape to R that provides a user interface to R packages.
PMID 24715956 · PMC3962008 · F1000Research · 2013 · 8 claims · 6 setups
Cyrface is a Cytoscape app/Java library providing a general interface from Cytoscape (Java) to any R function or package.
-
Full-text index only
Serum diagnosis of diffuse large B-cell lymphomas and further identification of response to therapy using SELDI-TOF-MS and tree analysis patterning.
PMID 18163913 · PMC2242801 · BMC cancer · 2007 · 8 claims · 8 setups
SELDI-TOF-MS serum proteomic patterns analyzed by decision tree classification (Biomarker Pattern Software) can discriminate DLBCL patients from healthy controls with high sensitivity and specificity.
-
Has reproduction · 86
The selection of software and database for metagenomics sequence analysis impacts the outcome of microbial profiling and pathogen detection.
PMID 37027361 · PMC10081788 · PloS one · 2023 · 7 claims · 7 setups
Obtaining an accurate species-level microbial profile using current direct-read metagenomics profiling software is still a challenging task.
-
Has reproduction · 90
PrimerSeq: Design and visualization of RT-PCR primers for alternative splicing using RNA-seq data.
PMID 24747190 · PMC4411361 · Genomics, proteomics & bioinformatics · 2014 · 8 claims · 3 setups
PrimerSeq is a user-friendly stand-alone software with a GUI for systematic design and visualization of RT-PCR primers for alternative splicing analysis using user-provided RNA-seq data.
-
Full-text index only
Grammar-based distance in progressive multiple sequence alignment.
PMID 18616828 · PMC2478692 · BMC bioinformatics · 2008 · 7 claims · 3 setups
A grammar-based (LZ complexity) distance metric can be used to determine the order in which sequences are progressively pairwise aligned
-
Has reproduction · 71
Sustainable data analysis with Snakemake.
PMID 34035898 · PMC8114187 · F1000Research · 2021 · 8 claims · 4 setups
Reproducibility alone is insufficient for sustainable data analysis; transparency and adaptability are equally important additional properties.
-
Full-text index only
PRESTO: rapid calculation of order statistic distributions and multiple-testing adjusted P-values via permutation for one and two-stage genetic association studies.
PMID 18620604 · PMC2483288 · BMC bioinformatics · 2008 · 8 claims · 4 setups
PRESTO is an order of magnitude faster than other existing permutation testing software for genetic association studies.
-
Has reproduction · 93
aCLImatise: automated generation of tool definitions for bioinformatics workflows.
PMID 33325479 · PMC8016486 · Bioinformatics (Oxford, England) · 2021 · 6 claims · 3 setups
aCLImatise automatically generates workflow-language tool definitions by parsing a command-line tool's help output
-
Full-text index only
Genome annotation errors in pathway databases due to semantic ambiguity in partial EC numbers.
PMID 16034025 · PMC1179732 · Nucleic acids research · 2005 · 7 claims · 4 setups
Partial EC numbers are semantically ambiguous, and databases that assign a gene to all reactions sharing the same partial EC number make a faulty inference, causing systematic misannotation.
-
Full-text index only
Genome-wide copy number profiling on high-density bacterial artificial chromosomes, single-nucleotide polymorphisms, and oligonucleotide microarrays: a platform comparison based on statistical power analysis.
PMID 17363414 · PMC2779891 · DNA research : an international journal for rapid publication of reports on genes and genomes · 2007 · 8 claims · 6 setups
High-density oligonucleotide/SNP platforms are superior to the BAC platform for genome-wide detection of copy-number variations smaller than 1 Mb
-
Has reproduction · 43
Compression of structured high-throughput sequencing data.
PMID 24260313 · PMC3832420 · PloS one · 2013 · 8 claims · 7 setups
Leveraging an explicit data schema (separate field encoding, field modeling, template compression, domain modeling) enables stronger compression of HTS alignment data than general-purpose compression of serialized bytes.
-
Full-text index only
A novel wavelet-based thresholding method for the pre-processing of mass spectrometry data that accounts for heterogeneous noise.
PMID 18615428 · PMC2855839 · Proteomics · 2008 · 6 claims · 4 setups
Noise in SELDI-TOF/MALDI-TOF mass spectrometry data is heteroscedastic across the m/z range, with larger variance at lower m/z values, contrary to the homogeneous noise assumption of existing wavelet denoising methods.