Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
BHD mutations, clinical and molecular genetic investigations of Birt-Hogg-Dubé syndrome: a new series of 50 families and a review of published reports.
PMID 18234728 · PMC2564862 · Journal of medical genetics · 2008 · 8 claims · 7 setups
BHD germline mutation detection rate was 88% (51/58 families) using direct DNA sequencing
-
Full-text index only
Skittle: a 2-dimensional genome visualization tool.
PMID 20042093 · PMC2817707 · BMC bioinformatics · 2009 · 7 claims · 6 setups
Skittle is a 2D genome visualization tool combining a color-coded Nucleotide Display, a Repeat Map, a Repeat Overview, and an Alignment Cylinder to reveal genomic patterns at multiple scales
-
Full-text index only
High throughput sequencing and proteomics to identify immunogenic proteins of a new pathogen: the dirty genome approach.
PMID 20037647 · PMC2793016 · PloS one · 2009 · 7 claims · 7 setups
A dirty genome approach using unfinished, unclosed genome sequences combined with proteomics can rapidly identify immunogenic proteins useful for diagnostic tool development
-
Full-text index only
Large-scale discovery of insertion hotspots and preferential integration sites of human transposed elements.
PMID 20008508 · PMC2836564 · Nucleic acids research · 2010 · 8 claims · 6 setups
Most TEs insert within specific 'hotspots' along the targeted TE rather than uniformly.
-
Full-text index only
Effects of HIV type-1 immune selection on susceptability to integrase inhibitor resistance.
PMID 19918099 · PMC4155129 · Antiviral therapy · 2009 · 8 claims · 6 setups
Primary integrase inhibitor resistance mutations (T66I, E92Q, G140S, Y143C/H/R, Q148H/R/K, N155S/H) were absent in 342 drug-naive individuals, indicating these sites are highly constrained.
-
Full-text index only
Correlation of microsynteny conservation and disease gene distribution in mammalian genomes.
PMID 19909546 · PMC2779822 · BMC genomics · 2009 · 7 claims · 8 setups
Density of mouse orthologs of human disease genes correlates with regions of conserved microsynteny in the mouse genome
-
Full-text index only
Identification and characterization of HLA-A*0301 epitopes in HIV-1 gag proteins using a novel approach.
PMID 19903485 · PMC2836169 · Journal of immunological methods · 2010 · 7 claims · 7 setups
PS mutations V7I and I34L (p17) and K403R (p7) in HIV-1 gag significantly correlate with HLA-A*0301
-
Full-text index only
Isolated eyelid closure myotonia in two families with sodium channel myotonia.
PMID 19876661 · PMC2854355 · Neurogenetics · 2010 · 6 claims · 5 setups
The L250P mutation in SCN4A is associated with a strictly isolated eyelid closure myotonia phenotype
-
Full-text index only
Metagenomic study of the oral microbiota by Illumina high-throughput sequencing.
PMID 19796657 · PMC3568755 · Journal of microbiological methods · 2009 · 8 claims · 6 setups
The 16S rRNA V5 hypervariable region, amplified as a short ~82-base segment, provides reliable taxonomic identification of oral bacteria against public databases like HOMD.
-
Full-text index only
Improved mutation tagging with gene identifiers applied to membrane protein stability prediction.
PMID 19758467 · PMC2745585 · BMC bioinformatics · 2009 · 8 claims · 4 setups
MutationTagger achieves 87% F-measure for the mutation retrieval task on a benchmark dataset
-
Full-text index only
The promise and reality of personal genomics.
PMID 19723346 · PMC2768970 · Genome biology · 2009 · 7 claims · 6 setups
Despite being the most complete and accurate individually sequenced human genome to date, AK1 sequencing still misses a substantial fraction of variants, showing sequencing technology remains far from complete/reliable.
-
Full-text index only
Targeted capture and massively parallel sequencing of 12 human exomes.
PMID 19684571 · PMC2844771 · Nature · 2009 · 8 claims · 8 setups
Targeted exome capture combined with massively parallel sequencing sensitively and specifically identifies rare and common variants across >300 Mb of coding sequence
-
Full-text index only
Constitutive RB1 mutation in a child conceived by in vitro fertilization: implications for genetic counseling.
PMID 19640284 · PMC2726130 · BMC medical genetics · 2009 · 7 claims · 4 setups
The retinoblastoma proband carries a novel constitutive RB1 mutation (g.2056C>G) at position -4 of the 5'UTR Kozak consensus sequence, absent in her father and unaffected sisters
-
Full-text index only
Assembly, Characterization and Comparative Analysis of the Complete Mitogenome of Small-Leaved Eriobotrya seguinii (Maleae, Rosaceae).
PMID 41595526 · PMC12841229 · Genes · 2026 · 8 claims · 7 setups
The E. seguinii mitogenome is the smallest and has the highest GC content of any known Eriobotrya species
-
Full-text index only
Exploring the regulatory potential of RNA structures in 202 cyanobacterial genomes.
PMID 41641705 · PMC12873609 · Nucleic acids research · 2026 · 7 claims · 8 setups
Screening 202 cyanobacterial genomes identified 402 CRSs matching known RNA families (Rfam and Rho-independent terminators) and 409 novel CRSs.
-
Full-text index only
Protocol to perform cell-type-specific transcriptome-wide association study using scPrediXcan framework.
PMID 41689808 · PMC12925207 · STAR protocols · 2026 · 6 claims · 6 setups
scPrediXcan enables cell-type-specific transcriptome-wide association studies (TWAS) by integrating deep learning-based prediction of gene expression from DNA sequence and epigenetic features.
-
Full-text index only
Lift&Add-rapid and robust addition of new species to alignments of conserved non-coding sequences.
PMID 42203687 · PMC13224966 · Bioinformatics (Oxford, England) · 2026 · 7 claims · 5 setups
Lift&Add, a Snakemake/bash workflow combining UCSC liftOver, Liftoff, and MAFFT, enables rapid addition of new genome sequences to existing multi-species alignments of conserved elements without requiring new whole-genome alignments.
-
Full-text index only
EpiXFormer: a cross-attention neural network for predicting cell type-specific transcription factor binding sites.
PMID 41527854 · PMC12796812 · Briefings in bioinformatics · 2026 · 8 claims · 8 setups
EpiXFormer achieves high accuracy (mean AUROC ~0.99) predicting binding sites of both TFs and non-sequence-specific DBPs across 199 DBP-cell type pairs
-
Full-text index only
CLAMP: predicting specific protein-mediated chromatin loops in diverse species with a chromatin accessibility language model.
PMID 41555433 · PMC12903630 · Genome biology · 2026 · 8 claims · 8 setups
CLAMP, a chromatin-accessibility language model, predicts protein-mediated chromatin loops across 10 species, 18 proteins, and 24 cell types with superior performance versus existing methods.
-
Full-text index only
Eukan: a fully automated nuclear genome annotation pipeline for less studied and divergent eukaryotes.
PMID 41567515 · PMC12817076 · NAR genomics and bioinformatics · 2026 · 8 claims · 7 setups
Eukan automatically leverages RNA-Seq coverage to inform generalized Hidden Markov Model gene prediction and intron lengths to inform protein sequence alignments