Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Benchmarking ortholog identification methods using functional genomics data.
PMID 16613613 · PMC1557999 · Genome biology · 2006 · 8 claims · 7 setups
InParanoid is the best overall ortholog identification method for identifying functionally equivalent proteins when sensitivity and selectivity are combined into an overall score.
-
Full-text index only
Short tandem repeats in human exons: a target for disease mutations.
PMID 18789129 · PMC2543027 · BMC genomics · 2008 · 8 claims · 6 setups
STRs are present in exons of 92% of known human genes, unlike longer tandem repeats which are rare in exons
-
Full-text index only
Effect of read-mapping biases on detecting allele-specific expression from RNA-sequencing data.
PMID 19808877 · PMC2788925 · Bioinformatics (Oxford, England) · 2009 · 8 claims · 6 setups
Reads mapped to the reference genome show a significant bias toward the reference allele at heterozygous SNPs
-
Full-text index only
MBGD update 2010: toward a comprehensive resource for exploring microbial genome diversity.
PMID 19906735 · PMC2808943 · Nucleic acids research · 2010 · 8 claims · 6 setups
MBGD allows users to create ortholog groups using a specified subgroup of organisms, distinguishing it from other comparative genomics resources
-
Has reproduction · 89
Statistical framework for calling allelic imbalance in high-throughput sequencing data.
PMID 39966391 · PMC11836314 · Nature communications · 2025 · 8 claims · 6 setups
MIXALIME is a versatile computational framework for calling allele-specific variants (ASVs) from diverse high-throughput omics data
-
Full-text index only
Designating eukaryotic orthology via processed transcription units.
PMID 18445630 · PMC2425467 · Nucleic acids research · 2008 · 8 claims · 5 setups
Existing ortholog databases discard/ignore alternative splicing via all-against-all protein comparisons, causing ambiguous ortholog calls and misclassification of AS isoforms as in-paralogs
-
Full-text index only
HUPO Highlights.
PMID 19862759 · PMC4594800 · Proteomics · 2009 · 8 claims · 8 setups
Mass spectrometry analysis of human liver reference samples (French Reference liver + Huh7 hepatoma cells) achieves substantial human genome coverage via PeptideAtlas processing
-
Has reproduction · 50
Estimating and Correcting for Off-Target Cellular Contamination in Brain Cell Type Specific RNA-Seq Data.
PMID 33746712 · PMC7966716 · Frontiers in molecular neuroscience · 2021 · 6 claims · 7 setups
A computational method using high-quality scRNA-seq reference data can estimate per-sample, per-cell-type off-target contamination coefficients in sctRNA-seq datasets.
-
Has reproduction · 74
SpaGene: A Deep Adversarial Framework for Spatial Gene Imputation.
PMID 42146899 · PMC13176606 · Computational and structural biotechnology journal · 2026 · 8 claims · 6 setups
SpaGene improves average PCC and SSIM and reduces RMSE compared to 6 baseline methods (SpaGE, gimVI, Tangram, VISTA, spRefine, stDiff) across 8 diverse ST-SC dataset pairs under gene-holdout evaluation.
-
Has reproduction · 96
A bioinformatic pipeline for simulating viral integration data.
PMID 35496474 · PMC9046613 · Data in brief · 2022 · 7 claims · 3 setups
A snakemake-based pipeline was developed to simulate integration of a viral or vector genome into a host genome, including sub-genomic fragment integration, structural variation, and host-site deletions.
-
Has reproduction · 90
Assessment of genotyping array performance for genome-wide association studies and imputation in African cattle.
PMID 36057548 · PMC9441065 · Genetics, selection, evolution : GSE · 2022 · 7 claims · 6 setups
Commercially available bovine arrays are ineffective at capturing variants segregating among African indicine animals, with only 6% of high-LD (r2>0.8) variants captured by the best arrays versus 17% in African taurine and 25% in European taurine.
-
Has reproduction · 75
Genomic regions and candidate genes selected during the breeding of rice in Vietnam.
PMID 35899250 · PMC9309459 · Evolutionary applications · 2022 · 8 claims · 7 setups
XP-CLR and FST scans identify genomic regions with distorted allele frequency/differentiation patterns resulting from differential selective pressures between Vietnamese rice subpopulations
-
Full-text index only
Genome assembly comparison identifies structural variants in the human genome.
PMID 17115057 · PMC2674632 · Nature genetics · 2006 · 7 claims · 7 setups
Genome assembly comparison is a robust approach for identifying all classes of genetic variation, with no lower size limit.
-
Has reproduction · 63
Creation of a Single Cell RNASeq Meta-Atlas to Define Human Liver Immune Homeostasis.
PMID 34335581 · PMC8322955 · Frontiers in immunology · 2021 · 7 claims · 7 setups
Independent human liver immune scRNA-seq datasets can be combined into an integrated meta-atlas in which all datasets co-cluster, despite differing cell-type proportions between studies.
-
Has reproduction · 85
High performance imputation of structural and single nucleotide variants using low-coverage whole genome sequencing.
PMID 40155798 · PMC11951665 · Genetics, selection, evolution : GSE · 2025 · 7 claims · 6 setups
SNVs are imputed with high accuracy and recall across all tested WGS depths (1-4x), including in samples external to the reference panel.
-
Full-text index only
Comparative genomics comes of age.
PMID 12186641 · PMC139393 · Genome biology · 2002 · 8 claims · 8 setups
Only about 50% of conserved sequence elements (exons+introns) in orthologous human-mouse genes correspond to exons, implying substantial non-exonic conservation
-
Full-text index only
Identifying related L1 retrotransposons by analyzing 3' transduced sequences.
PMID 12734010 · PMC156586 · Genome biology · 2003 · 8 claims · 6 setups
L1 elements with transduction-derived 3' sequence (L1-TDs) can be computationally identified using RepeatMasker/TSDfinder and grouped into families sharing a common progenitor via BLAST comparison of downstream sequences.
-
Full-text index only
Fast and systematic genome-wide discovery of conserved regulatory elements using a non-alignment based approach.
PMID 15693947 · PMC551538 · Genome biology · 2005 · 7 claims · 8 setups
FastCompare, a non-alignment-based, linear-time algorithm, computes a genome-wide conservation score for all k-mers (7-9 nt) between two genomes to identify conserved regulatory elements
-
Full-text index only
Automatic annotation of eukaryotic genes, pseudogenes and promoters.
PMID 16925832 · PMC1810547 · Genome biology · 2006 · 8 claims · 6 setups
Fgenesh++ gene prediction pipeline identifies 91% of coding nucleotides with 90% specificity
-
Full-text index only
In silico and in vitro comparative analysis to select, validate and test SNPs for human identification.
PMID 18076761 · PMC2222643 · BMC genomics · 2007 · 8 claims · 7 setups
A panel of 24 SNPs was selected and validated for human identification using 1,040 unrelated samples from three populations (Italian, Benin Gulf, Mongolian)