Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 68
Mining the equine gut metagenome: poorly-characterized taxa associated with cardiovascular fitness in endurance athletes.
PMID 36192523 · PMC9529974 · Communications biology · 2022 · 8 claims · 8 setups
Built an integrated horse gut microbiome gene catalog (~25 million unique genes) and 372 metagenome-assembled genomes (MAGs) spanning 4179 genera and 95 phyla
-
Full-text index only
Oligomeric protein structure networks: insights into protein-protein interactions.
PMID 16336694 · PMC1326230 · BMC bioinformatics · 2005 · 8 claims · 6 setups
Interface amino acid clusters identified at Imin=6% correlate well with residues losing accessible surface area (δASA) upon oligomerization
-
Full-text index only
GeneSeer: a sage for gene names and genomic resources.
PMID 16176584 · PMC1266031 · BMC genomics · 2005 · 7 claims · 4 setups
GeneSeer aggregates gene name synonyms from GenBank, FlyBase, ExPASy, HUGO, ENSEMBL, UCSC and Gene Ontology into a name-translation database that maps any familiar name to a reference (SOFAR) identifier.
-
Full-text index only
Genome comparison without alignment using shortest unique substrings.
PMID 15910684 · PMC1166540 · BMC bioinformatics · 2005 · 8 claims · 8 setups
A number of sequence comparison tasks, including detection of unique genomic regions, can be accomplished efficiently without an alignment step using shortest unique substrings.
-
Full-text index only
SysPIMP: the web-based systematical platform for identifying human disease-related mutated sequences from mass spectrometry.
PMID 19036792 · PMC2686442 · Nucleic acids research · 2009 · 8 claims · 7 setups
SysPIMP is a web-based platform integrating disease mutation databases with X!Tandem and BLAST to identify disease-related mutated proteins from MS results
-
Full-text index only
EPD in its twentieth year: towards complete promoter coverage of selected model organisms.
PMID 16381980 · PMC1347508 · Nucleic acids research · 2006 · 7 claims · 4 setups
EPD is an annotated, non-redundant collection of experimentally defined eukaryotic POL II promoters accessed via genome position pointers.
-
Full-text index only
Human genome research in China.
PMID 15168679 · PMC7079922 · Journal of molecular medicine (Berlin, Germany) · 2004 · 8 claims · 8 setups
China completed its assigned 1% share of the international Human Genome Project sequencing effort and contributed ~10% of the HapMap effort
-
Full-text index only
MODBASE: a database of annotated comparative protein structure models and associated resources.
PMID 16381869 · PMC1347422 · Nucleic acids research · 2006 · 8 claims · 7 setups
MODBASE is a database of automatically calculated comparative protein structure models covering all UniProt sequences matchable to a known structure
-
Has reproduction · 98
Large-scale quality assessment of prokaryotic genomes with metashot/prok-quality.
PMID 35136576 · PMC8804904 · F1000Research · 2021 · 8 claims · 6 setups
metashot/prok-quality is a container-enabled Nextflow pipeline for quality assessment and dereplication of draft prokaryotic genomes
-
Full-text index only
LMPD: LIPID MAPS proteome database.
PMID 16381922 · PMC1347484 · Nucleic acids research · 2006 · 8 claims · 5 setups
LMPD is an object-relational database of lipid-associated protein sequences and annotations, publicly available from the LIPID MAPS Consortium website.
-
Full-text index only
piRNABank: a web resource on classified and clustered Piwi-interacting RNAs.
PMID 17881367 · PMC2238943 · Nucleic acids research · 2008 · 6 claims · 4 setups
piRNABank is a web-accessible database storing empirically known piRNA sequences and annotations for human, mouse and rat.
-
Has reproduction · 79
Species-Wide Phylogenomics of the Staphylococcus aureus Agr Operon Revealed Convergent Evolution of Frameshift Mutations.
PMID 35044202 · PMC8768832 · Microbiology spectrum · 2022 · 8 claims · 7 setups
AgrVATE, a novel kmer-based BLASTn and in silico PCR/Snippy pipeline, enables fast, standardized agr group typing and frameshift/null mutation detection from genome assemblies
-
Full-text index only
The global landscape of sequence diversity.
PMID 17996061 · PMC2258180 · Genome biology · 2007 · 7 claims · 5 setups
Eukaryotic sequence datasets show substantially greater genetic diversity (higher sequence/gene family discovery rates) than bacterial datasets, likely related to differences in modes of genetic inheritance.
-
Has reproduction · 63
Comparative transcriptome analysis of tomato (Solanum lycopersicum) in response to exogenous abscisic acid.
PMID 24289302 · PMC4046761 · BMC genomics · 2013 · 8 claims · 7 setups
Exogenous ABA alters the expression of a majority (54.73%) of expressed tomato leaf transcripts, with 2,787 significantly differentially expressed genes, predominantly up-regulated.
-
Full-text index only
Proteomics studies reveal important information on small molecule therapeutics: a case study on plasma proteins.
PMID 18973825 · PMC7185545 · Drug discovery today · 2008 · 8 claims · 8 setups
Abundant plasma proteins (albumin, IgG, transferrin) act as 'molecular sponges' that bind and transport low molecular weight proteins/peptides and drugs, extending their half-life by preventing rapid renal clearance.
-
Has reproduction · 67
A consensus approach to vertebrate de novo transcriptome assembly from RNA-seq data: assembly of the duck (Anas platyrhynchos) transcriptome.
PMID 25009556 · PMC4070175 · Frontiers in genetics · 2014 · 8 claims · 8 setups
Multiple k-mer (MK) assemblies are more complete than single k-mer (SK) assemblies, showing higher reads-mapped-back-to-transcripts (RMBT) and higher CEGMA complete-gene percentages for all three tools.