Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 94
BaRTv2: a highly resolved barley reference transcriptome for accurate transcript-specific RNA-seq quantification.
PMID 35704392 · PMC9546494 · The Plant journal : for cell and molecular biology · 2022 · 8 claims · 6 setups
BaRTv2.18 is the most comprehensive and resolved reference transcriptome in barley to date, containing 39,434 genes and 148,260 transcripts
-
Full-text index only
LMPD: LIPID MAPS proteome database.
PMID 16381922 · PMC1347484 · Nucleic acids research · 2006 · 8 claims · 5 setups
LMPD is an object-relational database of lipid-associated protein sequences and annotations, publicly available from the LIPID MAPS Consortium website.
-
Full-text index only
ChimerDB--a knowledgebase for fusion sequences.
PMID 16381848 · PMC1347382 · Nucleic acids research · 2006 · 8 claims · 6 setups
ChimerDB integrates bioinformatics analysis of mRNA/EST sequences, manually collected literature data, and OMIM translocation data into a single fusion sequence knowledgebase
-
Full-text index only
ARED 3.0: the large and diverse AU-rich transcriptome.
PMID 16381826 · PMC1347415 · Nucleic acids research · 2006 · 7 claims · 6 setups
ARED 3.0 computationally mapped more than 4000 ARE-mRNAs to the human genome, representing 5-8% of human genes.
-
Full-text index only
Dyneins across eukaryotes: a comparative genomic analysis.
PMID 17897317 · PMC2239267 · Traffic (Copenhagen, Denmark) · 2007 · 8 claims · 6 setups
Phylogenetic inference identified nine DHC families (two cytoplasmic, seven axonemal) and six IC families (one cytoplasmic)
-
Full-text index only
piRNABank: a web resource on classified and clustered Piwi-interacting RNAs.
PMID 17881367 · PMC2238943 · Nucleic acids research · 2008 · 6 claims · 4 setups
piRNABank is a web-accessible database storing empirically known piRNA sequences and annotations for human, mouse and rat.
-
Full-text index only
G2Cdb: the Genes to Cognition database.
PMID 18984621 · PMC2686544 · Nucleic acids research · 2009 · 7 claims · 7 setups
G2Cdb integrates experimentally validated synapse proteome datasets with mouse/human genomic annotation, phenotype, and human disease data in a gene-centric database.
-
Full-text index only
Genomic analysis of the TRIM family reveals two groups of genes with distinct evolutionary properties.
PMID 18673550 · PMC2533329 · BMC evolutionary biology · 2008 · 8 claims · 6 setups
The human TRIM family is split into two groups (group 1 and group 2) that differ in domain structure, genomic organization, and evolutionary properties.
-
Full-text index only
Structural evolution of the protein kinase-like superfamily.
PMID 16244704 · PMC1261164 · PLoS computational biology · 2005 · 8 claims · 5 setups
All kinases in the superfamily share a 'universal core' domain consisting only of the regions required for ATP binding and the phosphotransfer reaction.
-
Has reproduction · 62
E3RC: A step-by-step computational protocol for exploring enhancer RNA expression and regulation using conventional RNA-seq data.
PMID 40716058 · PMC12318280 · STAR protocols · 2025 · 6 claims · 3 setups
E3RC is a computational framework for identifying and quantifying eRNAs and characterizing their expression and transcriptional regulation using conventional RNA-seq data.
-
Has reproduction · 79
Species-Wide Phylogenomics of the Staphylococcus aureus Agr Operon Revealed Convergent Evolution of Frameshift Mutations.
PMID 35044202 · PMC8768832 · Microbiology spectrum · 2022 · 8 claims · 7 setups
AgrVATE, a novel kmer-based BLASTn and in silico PCR/Snippy pipeline, enables fast, standardized agr group typing and frameshift/null mutation detection from genome assemblies
-
Full-text index only
Integration with the human genome of peptide sequences obtained by high-throughput mass spectrometry.
PMID 15642101 · PMC549070 · Genome biology · 2005 · 8 claims · 4 setups
PeptideAtlas, a public database integrating MS/MS-derived peptide identifications with the human genome, was built as an expandable resource for proteomic data.
-
Full-text index only
Classification of real and pseudo microRNA precursors using local structure-sequence features and support vector machine.
PMID 16381612 · PMC1360673 · BMC bioinformatics · 2005 · 7 claims · 7 setups
A 32-dimensional triplet structure-sequence feature vector combined with SVM (triplet-SVM) can distinguish real human pre-miRNAs from pseudo pre-miRNA hairpins with ~90% accuracy.
-
Full-text index only
Genome comparison without alignment using shortest unique substrings.
PMID 15910684 · PMC1166540 · BMC bioinformatics · 2005 · 8 claims · 8 setups
A number of sequence comparison tasks, including detection of unique genomic regions, can be accomplished efficiently without an alignment step using shortest unique substrings.
-
Full-text index only
Inparanoid: a comprehensive database of eukaryotic orthologs.
PMID 15608241 · PMC540061 · Nucleic acids research · 2005 · 8 claims · 4 setups
The Inparanoid algorithm identifies true ortholog clusters by seeding on reciprocal best-matching pairs, gathering inparalogs (post-speciation duplicates) while excluding outparalogs (pre-speciation duplicates)
-
Full-text index only
The global landscape of sequence diversity.
PMID 17996061 · PMC2258180 · Genome biology · 2007 · 7 claims · 5 setups
Eukaryotic sequence datasets show substantially greater genetic diversity (higher sequence/gene family discovery rates) than bacterial datasets, likely related to differences in modes of genetic inheritance.
-
Full-text index only
An integrated database-pipeline system for studying single nucleotide polymorphisms and diseases.
PMID 19091018 · PMC2638159 · BMC bioinformatics · 2008 · 6 claims · 5 setups
Existing SNP/disease databases are fragmented; no combined resource widely supports gene-, SNP-, and disease-related information together
-
Full-text index only
Proteomics studies reveal important information on small molecule therapeutics: a case study on plasma proteins.
PMID 18973825 · PMC7185545 · Drug discovery today · 2008 · 8 claims · 8 setups
Abundant plasma proteins (albumin, IgG, transferrin) act as 'molecular sponges' that bind and transport low molecular weight proteins/peptides and drugs, extending their half-life by preventing rapid renal clearance.
-
Full-text index only
Eukan: a fully automated nuclear genome annotation pipeline for less studied and divergent eukaryotes.
PMID 41567515 · PMC12817076 · NAR genomics and bioinformatics · 2026 · 8 claims · 7 setups
Eukan automatically leverages RNA-Seq coverage to inform generalized Hidden Markov Model gene prediction and intron lengths to inform protein sequence alignments
-
Full-text index only
SGCEdb: a flexible database and web interface integrating experimental results and analysis for structural genomics focusing on Caenorhabditis elegans.
PMID 16381914 · PMC1347399 · Nucleic acids research · 2006 · 8 claims · 8 setups
SGCEdb is a flexible, reusable database and web interface for reporting and analyzing structural genomics experiment results, focused on C. elegans