Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
The use of coded PCR primers enables high-throughput sequencing of multiple homolog amplification products by 454 parallel sequencing.
PMID 17299583 · PMC1797623 · PloS one · 2007 · 6 claims · 4 setups
5′-tagged PCR primers enable pooling of homologous PCR products from multiple sources into a single GS20 run with accurate post-hoc assignment of sequences to source
-
Full-text index only
Evolution of the NANOG pseudogene family in the human and chimpanzee genomes.
PMID 16469101 · PMC1457002 · BMC evolutionary biology · 2006 · 7 claims · 5 setups
The NANOG gene and all pseudogenes except NANOGP8 occupy orthologous chromosomal positions in the chimpanzee genome, indicating they originated before the human-chimpanzee divergence.
-
Full-text index only
iRefIndex: a consolidated protein interaction database with provenance.
PMID 18823568 · PMC2573892 · BMC bioinformatics · 2008 · 6 claims · 3 setups
A reproducible key (ROG) for each protein interactor and a corresponding key (RIG) for each interaction record can be generated by anyone using only primary sequence, taxonomy identifier, and the SHA-1 algorithm (SEGUID).
-
Full-text index only
ECgene: genome annotation for alternative splicing.
PMID 15608289 · PMC540072 · Nucleic acids research · 2005 · 8 claims · 5 setups
ECgene combines genome-based EST clustering with a graph-theoretic transcript assembly procedure to predict gene models including alternative splicing events.
-
Full-text index only
GenBank.
PMID 16381837 · PMC1347519 · Nucleic acids research · 2006 · 8 claims · 8 setups
GenBank is a comprehensive public database of nucleotide sequences with supporting bibliographic and biological annotation, built and distributed by NCBI.
-
Full-text index only
BRENDA, AMENDA and FRENDA: the enzyme information system in 2007.
PMID 17202167 · PMC1899097 · Nucleic acids research · 2007 · 7 claims · 6 setups
BRENDA is the largest publicly available enzyme information system worldwide, manually curated from primary literature and covering all identified enzymes regardless of source.
-
Has reproduction · 87
De Novo Transcriptome Meta-Assembly of the Mixotrophic Freshwater Microalga Euglena gracilis.
PMID 34072576 · PMC8227486 · Genes · 2021 · 6 claims · 8 setups
A consensus transcriptome assembled by combining reads from five independent studies is the most complete E. gracilis transcriptome released to date, outperforming the two previously available transcriptomes (GEFR01 and GDJR01).
-
Full-text index only
Molecular epidemiology of measles viruses in the United States, 1997-2001.
PMID 12194764 · PMC2732556 · Emerging infectious diseases · 2002 · 8 claims · 6 setups
The diversity of measles virus genotypes observed in the US from 1997–2001 reflects multiple imported sources of virus, indicating no strain of measles is endemic in the United States.
-
Full-text index only
Molecular evolution and multilocus sequence typing of 145 strains of SARS-CoV.
PMID 16112670 · PMC7118731 · FEBS letters · 2005 · 8 claims · 7 setups
145 SARS-CoV genomes can be divided into three groups: animal-origin viruses, first-epidemic clinical viruses, and GD03T0013
-
Full-text index only
Exogean: a framework for annotating protein-coding genes in eukaryotic genomic DNA.
PMID 16925841 · PMC1810556 · Genome biology · 2006 · 8 claims · 5 setups
Exogean is a framework using directed acyclic coloured multigraphs (DACMs) to represent biological objects (mRNA, ESTs, protein alignments, exons) and iteratively combine them into complex protein-coding transcript models.
-
Has reproduction · 88
Chromosome-Scale Assembly of the Complete Genome Sequence of Porcisia hertigi, Isolate C119, Strain LV43.
PMID 34647802 · PMC8515887 · Microbiology resource announcements · 2021 · 6 claims · 8 setups
The complete, chromosome-scale genome sequence of Porcisia hertigi (isolate C119, strain LV43) was assembled using combined short- and long-read sequencing technologies.
-
Has reproduction · 83
Macrel: antimicrobial peptide screening in genomes and metagenomes.
PMID 33384902 · PMC7751412 · PeerJ · 2020 · 8 claims · 8 setups
Macrel is an end-to-end pipeline that predicts high-quality AMP candidates from peptides, contigs, or reads of (meta)genomes
-
Full-text index only
The role of genomics in the identification, prediction, and prevention of biological threats.
PMID 19855827 · PMC2757898 · PLoS biology · 2009 · 8 claims · 5 setups
Genomics should be used proactively, not just reactively, to build biopreparedness against biological threats
-
Has reproduction · 88
Comprehensive benchmarking of large language models for RNA secondary structure prediction.
PMID 40205851 · PMC11982019 · Briefings in bioinformatics · 2025 · 7 claims · 4 setups
Existing RNA-LLMs had not previously been evaluated for secondary structure prediction in a unified, fair experimental setup with the same datasets and prediction model.
-
Has reproduction · 91
Genomic Description of 'Candidatus Abyssubacteria,' a Novel Subsurface Lineage Within the Candidate Phylum Hydrogenedentes.
PMID 30210471 · PMC6121073 · Frontiers in microbiology · 2018 · 8 claims · 7 setups
SURF_5 and SURF_17 are the first full genomes of a novel bacterial lineage, 'Candidatus Abyssubacteria,' within the candidate phylum Hydrogenedentes
-
Has reproduction · 76
Analysis of the Hypoxic Response in a Mouse Cortical Collecting Duct-Derived Cell Line Suggests That Esrra Is Partially Involved in Hif1α-Mediated Hypoxia-Inducible Gene Expression in mCCD(cl1) Cells.
PMID 35806266 · PMC9267015 · International journal of molecular sciences · 2022 · 8 claims · 7 setups
mCCD cl1 cells mount a broad transcriptional response to 24 h hypoxia (0.2% O2), with 3086 genes differentially expressed