Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Advancement of biomarker discovery and validation through the HUPO plasma proteome project.
PMID 15502245 · PMC3839274 · Disease markers · 2004 · 7 claims · 3 setups
Standardization of specimen collection, handling, storage, and choice of serum vs. plasma/anticoagulant is essential for comparable proteomic biomarker discovery.
-
Full-text index only
EPD in its twentieth year: towards complete promoter coverage of selected model organisms.
PMID 16381980 · PMC1347508 · Nucleic acids research · 2006 · 7 claims · 4 setups
EPD is an annotated, non-redundant collection of experimentally defined eukaryotic POL II promoters accessed via genome position pointers.
-
Full-text index only
Automated recognition of retroviral sequences in genomic data--RetroTector.
PMID 17636050 · PMC1976444 · Nucleic acids research · 2007 · 8 claims · 8 setups
RetroTector uses 'fragment threading' (detection of chains of conserved retroviral motifs satisfying distance constraints) combined with LTR detection and protein reconstruction to identify ERVs in genomic sequences
-
Has reproduction · 68
Cell-type annotation with accurate unseen cell-type identification using multiple references.
PMID 37379341 · PMC10335708 · PLoS computational biology · 2023 · 8 claims · 4 setups
mtANN integrates multiple reference datasets and eight gene selection methods via ensemble learning (multiple deep classification models + majority voting) to improve cell-type annotation accuracy
-
Has reproduction · 45
Identifying and classifying trait linked polymorphisms in non-reference species by walking coloured de bruijn graphs.
PMID 23536903 · PMC3607606 · PloS one · 2013 · 8 claims · 9 setups
Bubbleparse detects sequence variants directly from NGS reads without a reference genome, using the coloured de Bruijn graph implementation of Cortex plus a new depth-first bubble-finding module.
-
Has reproduction · 89
Statistical framework for calling allelic imbalance in high-throughput sequencing data.
PMID 39966391 · PMC11836314 · Nature communications · 2025 · 8 claims · 6 setups
MIXALIME is a versatile computational framework for calling allele-specific variants (ASVs) from diverse high-throughput omics data
-
Full-text index only
Genetic diversity of clinical isolates of Bacillus cereus using multilocus sequence typing.
PMID 18990211 · PMC2585095 · BMC microbiology · 2008 · 8 claims · 7 setups
The 55 clinical B. cereus isolates were phylogenetically diverse, comprising 38 sequence types (STs) distributed across two of three previously described clades.
-
Full-text index only
Genomic diversity and evolution of Mycobacterium ulcerans revealed by next-generation sequencing.
PMID 19806175 · PMC2736377 · PLoS pathogens · 2009 · 8 claims · 6 setups
Genome sequencing of three M. ulcerans strains (NM20/02, NM31/04, Jp8756) identified thousands of SNPs relative to reference strain Agy99
-
Full-text index only
Functional copy-number alterations in cancer.
PMID 18784837 · PMC2527508 · PloS one · 2008 · 8 claims · 3 setups
RAE is a comprehensive computational framework that robustly maps chromosomal alterations in tumor samples and statistically assesses their functional importance in cancer.
-
Full-text index only
Challenges and standards in integrating surveys of structural variation.
PMID 17597783 · PMC2698291 · Nature genetics · 2007 · 7 claims · 5 setups
There is no standard approach to collecting, assessing the quality of, or describing structural variants, risking the entire genome eventually being labeled 'structurally variant' based on uncurated nondisease-sample data.
-
Full-text index only
High-throughput sequencing provides insights into genome variation and evolution in Salmonella Typhi.
PMID 18660809 · PMC2652037 · Nature genetics · 2008 · 7 claims · 8 setups
Evolution in the Typhi population is characterized by ongoing loss of gene function (pseudogene accumulation) rather than gain of function or diversifying selection.
-
Full-text index only
Disease-specific proteins from rheumatoid arthritis patients.
PMID 16778393 · PMC2729955 · Journal of Korean medical science · 2006 · 8 claims · 6 setups
Fibronectin, semaphorin 7A precursor, growth factor receptor-bound protein 7 (GRB7), and immunoglobulin µ chain specifically associate with antibodies purified from RA synovial fluid
-
Full-text index only
Multilocus sequence typing supports the hypothesis that Ochrobactrum anthropi displays a human-associated subpopulation.
PMID 20021660 · PMC2810298 · BMC microbiology · 2009 · 8 claims · 6 setups
A novel Multi-Locus Sequence Typing (MLST) scheme for O. anthropi was developed for the first time, based on 7 genes (3490 nucleotides) evolving mostly by neutral mutations
-
Full-text index only
Multilocus sequence typing of Cronobacter sakazakii and Cronobacter malonaticus reveals stable clonal structures with clinical significance which do not correlate with biotypes.
PMID 19852808 · PMC2770063 · BMC microbiology · 2009 · 8 claims · 6 setups
A seven-locus MLST scheme (atpD, fusA, glnS, gltB, gyrB, infB, pps) reliably identifies and discriminates C. sakazakii and C. malonaticus strains
-
Full-text index only
Complete genome sequence and comparative analysis of the wild-type commensal Escherichia coli strain SE11 isolated from a healthy adult.
PMID 18931093 · PMC2608844 · DNA research : an international journal for rapid publication of reports on genes and genomes · 2008 · 8 claims · 6 setups
The SE11 genome comprises a 4.8 Mb chromosome encoding 4679 protein-coding genes and six plasmids encoding 323 protein-coding genes
-
Has reproduction · 100
Viral Diagnostics in Plants Using Next Generation Sequencing: Computational Analysis in Practice.
PMID 29123534 · PMC5662881 · Frontiers in plant science · 2017 · 8 claims · 8 setups
NGS/RNA-seq enables unbiased, hypothesis-free detection of multiple known and emergent plant viruses, unlike RT-PCR which only detects one or a few known viruses per test.
-
Full-text index only
High resolution array-CGH analysis of single cells.
PMID 17178751 · PMC1807964 · Nucleic acids research · 2007 · 7 claims · 7 setups
Single copy number changes as small as 8.3 Mb can be detected reliably in single cells using GenomePlex WGA combined with high-resolution tiling-path array-CGH.