Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
A re-annotation pipeline for Illumina BeadArrays: improving the interpretation of gene expression data.
PMID 19923232 · PMC2817484 · Nucleic acids research · 2010 · 8 claims · 7 setups
A Perl-based pipeline that BLASTs/BLATs Illumina probe sequences against genomes and transcript databases (RefSeq, UCSC Known Genes, UniGene/GenBank, Ensembl) can classify probes by quality grade (Perfect/Good/Bad/No match) and is applicable across 8 BeadArray platforms and other array types
-
Has reproduction · 89
A near complete genome for goat genetic and genomic research.
PMID 34507524 · PMC8434745 · Genetics, selection, evolution : GSE · 2021 · 8 claims · 8 setups
Saanen_v1 is a high-quality de novo goat genome assembly from a male Saanen buck, including the first goat Y chromosome scaffold.
-
Has reproduction · 54
Population structure analysis of Salmonella serovar Muenchen to redefine geno-serotyping using genome indexing approaches.
PMID 41743541 · PMC12929376 · Frontiers in microbiology · 2025 · 6 claims · 6 setups
Integrating genome-indexing (bettercallsal, DNA sketching + genome proximity) with SeqSero2 yields complementary serovar calls that improve discrimination of genomically distinct but antigenically similar serovars while retaining historical nomenclature
-
Has reproduction · 67
Optimal scaling of digital transcriptomes.
PMID 24223126 · PMC3819321 · PloS one · 2013 · 8 claims · 8 setups
Fifteen existing and novel transcript-count normalization algorithms can be compared with two novel, mutually independent metrics: the number of "uniform" genes (sufficiently low coefficient of variation after normalization) and low average Spearman correlation between normalized expression profiles of gene pairs.
-
Has reproduction · 99
A platinum standard pan-genome resource that represents the population structure of Asian rice.
PMID 32265447 · PMC7138821 · Scientific data · 2020 · 6 claims · 6 setups
The 3,000 Rice Genomes (3K-RG) dataset can be subdivided into 15 subpopulations (K=15), refining the previous K=9 population structure.
-
Has reproduction · 80
Curation of over 10 000 transcriptomic studies to enable data reuse.
PMID 33599246 · PMC7904053 · Database : the journal of biological databases and curation · 2021 · 8 claims · 6 setups
Gemma is a curated database and bioinformatics system that addresses metadata, probe annotation, and expression data inconsistencies in GEO to enable transcriptomic data reuse