Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
LMPD: LIPID MAPS proteome database.
PMID 16381922 · PMC1347484 · Nucleic acids research · 2006 · 8 claims · 5 setups
LMPD is an object-relational database of lipid-associated protein sequences and annotations, publicly available from the LIPID MAPS Consortium website.
-
Full-text index only
MODBASE: a database of annotated comparative protein structure models and associated resources.
PMID 16381869 · PMC1347422 · Nucleic acids research · 2006 · 8 claims · 7 setups
MODBASE is a database of automatically calculated comparative protein structure models covering all UniProt sequences matchable to a known structure
-
Full-text index only
The Edinburgh human metabolic network reconstruction and its functional analysis.
PMID 17882155 · PMC2013923 · Molecular systems biology · 2007 · 8 claims · 7 setups
EHMN is a high-quality, manually curated human metabolic network combining genome-based and literature-based (EMP) reconstruction, containing nearly 3000 reactions and over 2000 metabolic genes.
-
Full-text index only
DAVID Bioinformatics Resources: expanded annotation database and novel algorithms to better extract biology from large gene lists.
PMID 17576678 · PMC1933169 · Nucleic acids research · 2007 · 8 claims · 4 setups
The DAVID Gene Concept uses a single-linkage method to agglomerate tens of millions of gene/protein identifiers from NCBI, PIR, UniProt and other resources into unified DAVID genes.
-
Full-text index only
BRENDA, AMENDA and FRENDA: the enzyme information system in 2007.
PMID 17202167 · PMC1899097 · Nucleic acids research · 2007 · 7 claims · 6 setups
BRENDA is the largest publicly available enzyme information system worldwide, manually curated from primary literature and covering all identified enzymes regardless of source.
-
Full-text index only
Expansion of the BioCyc collection of pathway/genome databases to 160 genomes.
PMID 16246909 · PMC1266070 · Nucleic acids research · 2005 · 8 claims · 6 setups
The BioCyc collection has been expanded to 160 pathway/genome databases (PGDBs) organized into three curation tiers.
-
Full-text index only
'Genome design' model and multicellular complexity: golden middle.
PMID 17062620 · PMC1635334 · Nucleic acids research · 2006 · 8 claims · 8 setups
Intermediately expressed human genes are the longest genes genome-wide, in both coding and intronic sequence, longer than housekeeping or tissue-specific genes.
-
Full-text index only
MACSIMS: multiple alignment of complete sequences information management system.
PMID 16792820 · PMC1539025 · BMC bioinformatics · 2006 · 8 claims · 5 setups
MACSIMS is a multiple alignment-based information management system combining knowledge-based database mining with ab initio sequence predictions
-
Full-text index only
A survey of integral alpha-helical membrane proteins.
PMID 19760129 · PMC2780624 · Journal of structural and functional genomics · 2009 · 8 claims · 8 setups
An automated annotation pipeline defines the integral membrane genome and family associations for 21,379 proteins from 34 genomes, most belonging to 598 Pfam-derived membrane protein families.
-
Full-text index only
DAVID Knowledgebase: a gene-centered database integrating heterogeneous gene annotation resources to facilitate high-throughput gene functional analysis.
PMID 17980028 · PMC2186358 · BMC bioinformatics · 2007 · 7 claims · 3 setups
The DAVID Gene Concept, a single-linkage algorithm, merges gene clusters from Entrez Gene, UniRef100, and PIR-NREF100 that share protein IDs and species into unified DAVID gene clusters, improving cross-referencing between NCBI and UniProt systems
-
Full-text index only
Recent additions and improvements to the Onto-Tools.
PMID 15980579 · PMC1160233 · Nucleic acids research · 2005 · 7 claims · 3 setups
The Onto-Tools back-end database was redesigned around the Entrez Gene data model after NCBI phased out LocusLink in February 2005.
-
Has reproduction · 83
Metabolite-Centric Reporter Pathway and Tripartite Network Analysis of Arabidopsis Under Cold Stress.
PMID 30258841 · PMC6143811 · Frontiers in bioengineering and biotechnology · 2018 · 8 claims · 8 setups
Metabolite-centric reporter pathway analysis (RPAm) computes reporter metabolites and reporter pathways from transcriptome P-values by aggregating Z-scores of neighboring genes in a genome-scale metabolic network
-
Has reproduction · 69
A comparison across non-model animals suggests an optimal sequencing depth for de novo transcriptome assembly.
PMID 23496952 · PMC3655071 · BMC genomics · 2013 · 8 claims · 8 setups
Representative de novo transcriptome assemblies are generated with as few as ~20 million reads for single-tissue samples and ~30 million reads for whole animals at the mRNA-coverage level.
-
Full-text index only
Optimization of protein solubilization for the analysis of the CD14 human monocyte membrane proteome using LC-MS/MS.
PMID 19709643 · PMC3159575 · Journal of proteomics · 2009 · 7 claims · 5 setups
Methanol-based solubilization, alone or combined with PPS, yields significantly higher membrane protein identification/enrichment than PPS alone in monocyte membrane proteomics
-
Has reproduction · 79
pyrpipe: a Python package for RNA-Seq workflows.
PMID 34085037 · PMC8168212 · NAR genomics and bioinformatics · 2021 · 8 claims · 3 setups
pyrpipe enables development of flexible, reproducible, and easy-to-debug RNA-Seq computational pipelines purely in Python, in an object-oriented manner