Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Efficacy assessment of SNP sets for genome-wide disease association studies.
PMID 17726055 · PMC2034459 · Nucleic acids research · 2007 · 6 claims · 4 setups
τ, derived from Shannon entropy and swept radius ɛ, approximates the relative sample size efficiency of a marker set for mapping a causal variant at a given map position compared to a maximally polymorphic SNP
-
Full-text index only
Assessing the gene space in draft genomes.
PMID 19042974 · PMC2615622 · Nucleic acids research · 2009 · 6 claims · 7 setups
The proportion of mapped CEGs in a draft genome assembly is a useful metric for describing gene space completeness, complementing N50 and x-fold coverage.
-
Full-text index only
Using structural bioinformatics to investigate the impact of non synonymous SNPs and disease mutations: scope and limitations.
PMID 19758473 · PMC2745591 · BMC bioinformatics · 2009 · 8 claims · 8 setups
None of 39 tested structural properties can be used as a sole classification criterion to separate neutral SNPs from disease mutations.
-
Has reproduction · 94
Systematic assessment of pathway databases, based on a diverse collection of user-submitted experiments.
PMID 36088548 · PMC9487593 · Briefings in bioinformatics · 2022 · 8 claims · 6 setups
Well-established, hierarchically organized pathway annotation systems (e.g. GO, Reactome, KEGG) yield the best overall enrichment performance despite covering much of the human genome only in general terms.
-
Full-text index only
Ensembl 2007.
PMID 17148474 · PMC1761443 · Nucleic acids research · 2007 · 8 claims · 7 setups
Ensembl added 18 new chordate genomes this year, increasing total genomes available from 15 to 33, the largest yearly increase to date.
-
Full-text index only
A genetic variation map for chicken with 2.8 million single-nucleotide polymorphisms.
PMID 15592405 · PMC2263125 · Nature · 2004 · 8 claims · 8 setups
A genetic variation map of 2.8 million SNPs was constructed for chicken by comparing 3 domestic breeds to Red Jungle Fowl
-
Has reproduction · 84
Genome of the Asian longhorned beetle (Anoplophora glabripennis), a globally significant invasive species, reveals key functional and evolutionary innovations at the beetle-plant interface.
PMID 27832824 · PMC5105290 · Genome biology · 2016 · 7 claims · 7 setups
The A. glabripennis genome encodes a uniquely diverse arsenal of enzymes that degrade the main plant cell wall polysaccharide networks (cellulose, hemicellulose, pectin) and detoxify plant allelochemicals.
-
Full-text index only
Automatic discovery of cross-family sequence features associated with protein function.
PMID 16409628 · PMC1395344 · BMC bioinformatics · 2006 · 8 claims · 6 setups
A self-supervised data mining approach can find relationships between sequence features and functional annotations without preconceived functional categories.
-
Full-text index only
DAVID Knowledgebase: a gene-centered database integrating heterogeneous gene annotation resources to facilitate high-throughput gene functional analysis.
PMID 17980028 · PMC2186358 · BMC bioinformatics · 2007 · 7 claims · 3 setups
The DAVID Gene Concept, a single-linkage algorithm, merges gene clusters from Entrez Gene, UniRef100, and PIR-NREF100 that share protein IDs and species into unified DAVID gene clusters, improving cross-referencing between NCBI and UniProt systems
-
Full-text index only
SpliceMiner: a high-throughput database implementation of the NCBI Evidence Viewer for microarray splice variant analysis.
PMID 17338820 · PMC1839109 · BMC bioinformatics · 2007 · 6 claims · 4 setups
EVDB is a comprehensive, non-redundant relational database of known human splice variants built from NCBI Entrez Gene and Evidence Viewer data
-
Full-text index only
Sequence similarity network reveals common ancestry of multidomain proteins.
PMID 18475320 · PMC2377100 · PLoS computational biology · 2008 · 8 claims · 6 setups
Traditional homology definitions do not capture multidomain evolution; the authors extend the definition to include domain insertion via a common ancestral locus model.
-
Full-text index only
Discovery and identification of potential biomarkers of papillary thyroid carcinoma.
PMID 19785722 · PMC2761863 · Molecular cancer · 2009 · 8 claims · 7 setups
A 3-peak (m/z 9190, 6631, 8697 Da) SVM classification model discriminates PTC from non-cancer controls with high sensitivity and specificity
-
Has reproduction · 67
binny: an automated binning algorithm to recover high-quality genomes from complex metagenomic datasets.
PMID 36239393 · PMC9677464 · Briefings in bioinformatics · 2022 · 8 claims · 8 setups
binny outperforms or is highly competitive with commonly used and state-of-the-art binning methods (MetaBAT2, MaxBin2, CONCOCT, VAMB, SemiBin, MetaDecoder)
-
Full-text index only
Babelomics: advanced functional profiling of transcriptomics, proteomics and genomics experiments.
PMID 18515841 · PMC2447758 · Nucleic acids research · 2008 · 8 claims · 5 setups
Babelomics is a web suite offering both conventional functional enrichment methods and more advanced gene set analysis (GSA) methods, a combination offered by only one other tool (FuncAssociate) among competitors.
-
Full-text index only
BioDrugScreen: a computational drug design resource for ranking molecules docked to the human proteome.
PMID 19923229 · PMC2808957 · Nucleic acids research · 2010 · 6 claims · 5 setups
BioDrugScreen is a web resource providing pre-docked and pre-scored receptor-ligand complexes for ranking molecules against human proteome targets
-
Has reproduction · 75
Sequencing of human genomes with nanopore technology.
PMID 31015479 · PMC6478738 · Nature communications · 2019 · 8 claims · 7 setups
A novel single-sample, reference panel-free, read-based phasing algorithm built on the STITCH model improves nanopore SNV calling from modest baseline levels.
-
Has reproduction · 51
Population Genomic Analyses Suggest a Hybrid Origin, Cryptic Sexuality, and Decay of Genes Regulating Seed Development for the Putatively Strictly Asexual Kingdonia uniflora (Circaeasteraceae, Ranunculales).
PMID 36674965 · PMC9866071 · International journal of molecular sciences · 2023 · 8 claims · 8 setups
K. uniflora shows high allelic heterozygosity (negative F_IS) consistent with theoretical expectations under asexual evolution
-
Has reproduction · 51
Evaluation of the Available Variant Calling Tools for Oxford Nanopore Sequencing in Breast Cancer.
PMID 36140751 · PMC9498802 · Genes · 2022 · 7 claims · 6 setups
Clair3 and Human-SNP-wf (which incorporates Clair3) achieved the highest performance among the six variant callers tested.
-
Has reproduction · 44
Detecting DNA modifications from SMRT sequencing data by modeling sequence context dependence of polymerase kinetic.
PMID 23516341 · PMC3597545 · PLoS computational biology · 2013 · 8 claims · 7 setups
Local sequence context strongly determines position-specific polymerase kinetic rate: roughly 80% of IPD variation is explained by a 10 bp context (7 bases upstream, 2 bases downstream of the incorporation site), saturating at 7 bases upstream.
-
Has reproduction · 69
A comparison across non-model animals suggests an optimal sequencing depth for de novo transcriptome assembly.
PMID 23496952 · PMC3655071 · BMC genomics · 2013 · 8 claims · 8 setups
Representative de novo transcriptome assemblies are generated with as few as ~20 million reads for single-tissue samples and ~30 million reads for whole animals at the mRNA-coverage level.