Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Evolution of the NANOG pseudogene family in the human and chimpanzee genomes.
PMID 16469101 · PMC1457002 · BMC evolutionary biology · 2006 · 7 claims · 5 setups
The NANOG gene and all pseudogenes except NANOGP8 occupy orthologous chromosomal positions in the chimpanzee genome, indicating they originated before the human-chimpanzee divergence.
-
Full-text index only
iRefIndex: a consolidated protein interaction database with provenance.
PMID 18823568 · PMC2573892 · BMC bioinformatics · 2008 · 6 claims · 3 setups
A reproducible key (ROG) for each protein interactor and a corresponding key (RIG) for each interaction record can be generated by anyone using only primary sequence, taxonomy identifier, and the SHA-1 algorithm (SEGUID).
-
Full-text index only
GenBank.
PMID 16381837 · PMC1347519 · Nucleic acids research · 2006 · 8 claims · 8 setups
GenBank is a comprehensive public database of nucleotide sequences with supporting bibliographic and biological annotation, built and distributed by NCBI.
-
Full-text index only
pSTIING: a 'systems' approach towards integrating signalling pathways, interaction and transcriptional regulatory networks in inflammation and cancer.
PMID 16381926 · PMC1347407 · Nucleic acids research · 2006 · 8 claims · 3 setups
pSTIING is a publicly accessible web-based knowledgebase integrating protein-protein, protein-lipid, protein-small molecule interactions, transcriptional regulatory associations, ligand-receptor-cell type information, and signal transduction modules, with a focus on inflammation, cell migration and cancer.
-
Full-text index only
DAVID Knowledgebase: a gene-centered database integrating heterogeneous gene annotation resources to facilitate high-throughput gene functional analysis.
PMID 17980028 · PMC2186358 · BMC bioinformatics · 2007 · 7 claims · 3 setups
The DAVID Gene Concept, a single-linkage algorithm, merges gene clusters from Entrez Gene, UniRef100, and PIR-NREF100 that share protein IDs and species into unified DAVID gene clusters, improving cross-referencing between NCBI and UniProt systems
-
Full-text index only
Molecular characterization of Campylobacter jejuni clones: a basis for epidemiologic investigation.
PMID 12194772 · PMC2732546 · Emerging infectious diseases · 2002 · 8 claims · 5 setups
Clonal complex, as defined by MLST, is an epidemiologically relevant unit for long- and short-term investigation of C. jejuni epidemiology.
-
Full-text index only
Molecular evolution and multilocus sequence typing of 145 strains of SARS-CoV.
PMID 16112670 · PMC7118731 · FEBS letters · 2005 · 8 claims · 7 setups
145 SARS-CoV genomes can be divided into three groups: animal-origin viruses, first-epidemic clinical viruses, and GD03T0013
-
Full-text index only
MPromDb: an integrated resource for annotation and visualization of mammalian gene promoters and ChIP-chip experimental data.
PMID 16381984 · PMC1347458 · Nucleic acids research · 2006 · 8 claims · 5 setups
MPromDb is a novel database integrating experimentally supported gene promoters, TSS annotation, cis-regulatory elements, CpG islands, and ChIP-chip data with an integrated visualization interface.
-
Full-text index only
A proteomics grade electron transfer dissociation-enabled hybrid linear ion trap-orbitrap mass spectrometer.
PMID 18613715 · PMC2601597 · Journal of proteome research · 2008 · 8 claims · 5 setups
A NCI source coupled via an added octopole and the c-trap to a QLT-orbitrap enables fast, efficient ETD reagent anion injection (4-8 ms)
-
Has reproduction · 83
Macrel: antimicrobial peptide screening in genomes and metagenomes.
PMID 33384902 · PMC7751412 · PeerJ · 2020 · 8 claims · 8 setups
Macrel is an end-to-end pipeline that predicts high-quality AMP candidates from peptides, contigs, or reads of (meta)genomes
-
Full-text index only
Biodefense versus bioterrorism.
PMID 18771576 · PMC2575524 · Genome biology · 2008 · 5 claims · 4 setups
Whole-genome sequencing and comparative genomics of the attack strain were used to trace the anthrax letters to a specific laboratory flask
-
Has reproduction · 83
Current status of use of high throughput nucleotide sequencing in rheumatology.
PMID 33408124 · PMC7789458 · RMD open · 2021 · 8 claims · 4 setups
RNA-Seq is the most represented HTS assay in rheumatology research (n=457, 65%), used for biomarker identification in blood or synovial tissue
-
Full-text index only
The use of coded PCR primers enables high-throughput sequencing of multiple homolog amplification products by 454 parallel sequencing.
PMID 17299583 · PMC1797623 · PloS one · 2007 · 6 claims · 4 setups
5′-tagged PCR primers enable pooling of homologous PCR products from multiple sources into a single GS20 run with accurate post-hoc assignment of sequences to source
-
Full-text index only
How bacterial communities expand functional repertoires.
PMID 17238278 · PMC1750926 · PLoS biology · 2006 · 8 claims · 7 setups
The human microbiome contains roughly 100 times as many genes as does the human genome
-
Has reproduction · 87
De Novo Transcriptome Meta-Assembly of the Mixotrophic Freshwater Microalga Euglena gracilis.
PMID 34072576 · PMC8227486 · Genes · 2021 · 6 claims · 8 setups
A consensus transcriptome assembled by combining reads from five independent studies is the most complete E. gracilis transcriptome released to date, outperforming the two previously available transcriptomes (GEFR01 and GDJR01).
-
Full-text index only
Molecular epidemiology of measles viruses in the United States, 1997-2001.
PMID 12194764 · PMC2732556 · Emerging infectious diseases · 2002 · 8 claims · 6 setups
The diversity of measles virus genotypes observed in the US from 1997–2001 reflects multiple imported sources of virus, indicating no strain of measles is endemic in the United States.
-
Full-text index only
Exogean: a framework for annotating protein-coding genes in eukaryotic genomic DNA.
PMID 16925841 · PMC1810556 · Genome biology · 2006 · 8 claims · 5 setups
Exogean is a framework using directed acyclic coloured multigraphs (DACMs) to represent biological objects (mRNA, ESTs, protein alignments, exons) and iteratively combine them into complex protein-coding transcript models.
-
Full-text index only
G2Cdb: the Genes to Cognition database.
PMID 18984621 · PMC2686544 · Nucleic acids research · 2009 · 7 claims · 7 setups
G2Cdb integrates experimentally validated synapse proteome datasets with mouse/human genomic annotation, phenotype, and human disease data in a gene-centric database.
-
Full-text index only
Comparative analysis of the tear protein profile in mycotic keratitis patients.
PMID 18385783 · PMC2268856 · Molecular vision · 2008 · 8 claims · 5 setups
A glutaredoxin-related protein is expressed only in the tears of fungal keratitis patients and is absent in control tears.
-
Has reproduction · 84
Genome of the Asian longhorned beetle (Anoplophora glabripennis), a globally significant invasive species, reveals key functional and evolutionary innovations at the beetle-plant interface.
PMID 27832824 · PMC5105290 · Genome biology · 2016 · 7 claims · 7 setups
The A. glabripennis genome encodes a uniquely diverse arsenal of enzymes that degrade the main plant cell wall polysaccharide networks (cellulose, hemicellulose, pectin) and detoxify plant allelochemicals.