Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Genome reannotation of Escherichia coli CFT073 with new insights into virulence.
PMID 19930606 · PMC2785843 · BMC genomics · 2009 · 8 claims · 7 setups
Reannotation excluded 608 CDSs from the original RefSeq annotation, mostly unfunctional 'hypothetical'/'putative' genes
-
Full-text index only
The Universal Protein Resource (UniProt) in 2010.
PMID 19843607 · PMC2808944 · Nucleic acids research · 2010 · 8 claims · 5 setups
UniProt is a centralized, freely accessible, comprehensive knowledgebase of protein sequence and functional annotation maintained by the EBI, SIB and PIR consortium.
-
Full-text index only
Columba: an integrated database of proteins, structures, and annotations.
PMID 15801979 · PMC1087474 · BMC bioinformatics · 2005 · 8 claims · 6 setups
COLUMBA physically integrates data from twelve protein structure-related databases (PDB, KEGG, Swiss-Prot, CATH, SCOP, Gene Ontology, ENZYME, etc.) into a single PostgreSQL data warehouse.
-
Has reproduction · 77
Representing and querying disease networks using graph databases.
PMID 27462371 · PMC4960687 · BioData mining · 2016 · 7 claims · 8 setups
Graph databases are well suited for representing biological information because it is typically highly connected, semi-structured and unpredictable, unlike relational databases which require rigid schemas.
-
Full-text index only
MutDB: update on development of tools for the biochemical analysis of genetic variation.
PMID 17827212 · PMC2238958 · Nucleic acids research · 2008 · 7 claims · 5 setups
MutDB integrates dbSNP and Swiss-Prot genetic variation data with protein structural information, functional disruption prediction scores, and clinical phenotype links (OMIM, dbGAP)
-
Full-text index only
Towards alignment independent quantitative assessment of homology detection.
PMID 17205117 · PMC1762415 · PloS one · 2006 · 8 claims · 6 setups
The Fhom Estimator uses the prevalence of a conserved protein feature (X) in two protein sets to estimate the fraction of true homologs among paired proteins, independent of alignment quality.
-
Full-text index only
Ensembl's 10th year.
PMID 19906699 · PMC2808936 · Nucleic acids research · 2010 · 8 claims · 8 setups
Ensembl provides comprehensive gene annotation and integrated genomic resources (variation, regulation, comparative genomics) across a growing set of chordate genomes
-
Full-text index only
Expansion of the BioCyc collection of pathway/genome databases to 160 genomes.
PMID 16246909 · PMC1266070 · Nucleic acids research · 2005 · 8 claims · 6 setups
The BioCyc collection has been expanded to 160 pathway/genome databases (PGDBs) organized into three curation tiers.
-
Full-text index only
RAId_DbS: mass-spectrometry based peptide identification web server with knowledge integration.
PMID 18954448 · PMC2605478 · BMC genomics · 2008 · 7 claims · 4 setups
Constructed enhanced protein databases integrating annotated SAPs, PTMs, and disease associations for 17 organisms.
-
Has reproduction · 50
MEDUSA: A Pipeline for Sensitive Taxonomic Classification and Flexible Functional Annotation of Metagenomic Shotgun Sequences.
PMID 35330728 · PMC8940201 · Frontiers in genetics · 2022 · 6 claims · 6 setups
MEDUSA is an automated, Conda-installable and Snakemake-managed pipeline performing preprocessing, assembly, alignment, taxonomic classification, and functional annotation on shotgun data.
-
Full-text index only
LMPD: LIPID MAPS proteome database.
PMID 16381922 · PMC1347484 · Nucleic acids research · 2006 · 8 claims · 5 setups
LMPD is an object-relational database of lipid-associated protein sequences and annotations, publicly available from the LIPID MAPS Consortium website.
-
Full-text index only
Systematic identification of pseudogenes through whole genome expression evidence profiling.
PMID 16945953 · PMC1636364 · Nucleic acids research · 2006 · 8 claims · 8 setups
Developed a novel bioinformatics method that identifies pseudogenes by profiling whole-genome transcript and protein expression evidence
-
Has reproduction · 100
The genome of the ant Tetramorium bicarinatum reveals a tandem organization of venom peptides genes allowing the prediction of their regulatory and evolutionary profiles.
PMID 38245722 · PMC10800049 · BMC genomics · 2024 · 8 claims · 8 setups
44 venom peptide genes were identified, distributed across four of the eleven chromosomes and organized in tandem repeat clusters.
-
Has reproduction · 46
De novo transcriptome assembly and comprehensive assessment provide insight into fruiting body formation of Sparassis latifolia.
PMID 35773379 · PMC9247108 · Scientific reports · 2022 · 6 claims · 7 setups
De novo transcriptome assembly of S. latifolia produced 48,549 unigenes, 71.53% (34,728) of which were annotated against KEGG, GO, and/or KOG databases
-
Full-text index only
A genome-wide survey demonstrates widespread non-linear mRNA in expressed sequences from multiple species.
PMID 16237125 · PMC1258171 · Nucleic acids research · 2005 · 8 claims · 6 setups
A genome-wide computational survey identifies 245 genes in mammals (264 across six species) that produce RREO events in expressed sequences
-
Full-text index only
Protein coding potential of retroviruses and other transposable elements in vertebrate genomes.
PMID 15716312 · PMC549403 · Nucleic acids research · 2005 · 8 claims · 5 setups
About 1000 genes across four vertebrate gene sets analyzed contain at least one RETRA marker protein domain
-
Has reproduction · 78
Transcriptomic and physiological analysis of atractylodes chinensis in response to drought stress reveals the putative genes related to sesquiterpenoid biosynthesis.
PMID 38317086 · PMC10845750 · BMC plant biology · 2024 · 8 claims · 6 setups
Drought stress significantly increases MDA, proline, soluble sugar, and crude protein content and antioxidative enzyme (SOD, POD, CAT) activity in A. chinensis seedlings