Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Benchmarking tools for the alignment of functional noncoding DNA.
PMID 14736341 · PMC344529 · BMC bioinformatics · 2004 · 8 claims · 4 setups
Global alignment tools (Avid, ClustalW, Lagan, Needle, DiAlign-G) typically have higher sensitivity over entire noncoding sequences and within constrained blocks than local tools
-
Full-text index only
Divergence of exonic splicing elements after gene duplication and the impact on gene structures.
PMID 19883501 · PMC3091315 · Genome biology · 2009 · 8 claims · 7 setups
ESEs and ESSs diverge especially fast shortly after gene duplication, correlating with time since duplication (Ks)
-
Full-text index only
The genomic distribution of intraspecific and interspecific sequence divergence of human segmental duplications relative to human/chimpanzee chromosomal rearrangements.
PMID 18699995 · PMC2542386 · BMC genomics · 2008 · 8 claims · 5 setups
Some relatively recent (young) SDs accumulate in regions homologous to chromosomal inversions that occurred in the sister lineage
-
Full-text index only
Genome-wide prediction of functional gene-gene interactions inferred from patterns of genetic differentiation in mice and men.
PMID 18270580 · PMC2217631 · PloS one · 2008 · 8 claims · 6 setups
Pairs of unlinked SNPs showing excess genetic differentiation (LD in mouse RILs, Fst in human populations) beyond what simulations/coalescent models predict by chance represent candidate functionally interacting (epistatic) gene pairs.
-
Has reproduction · 74
SpaGene: A Deep Adversarial Framework for Spatial Gene Imputation.
PMID 42146899 · PMC13176606 · Computational and structural biotechnology journal · 2026 · 8 claims · 6 setups
SpaGene improves average PCC and SSIM and reduces RMSE compared to 6 baseline methods (SpaGE, gimVI, Tangram, VISTA, spRefine, stDiff) across 8 diverse ST-SC dataset pairs under gene-holdout evaluation.
-
Full-text index only
Genomewide pattern of synonymous nucleotide substitution in two complete genomes of Mycobacterium tuberculosis.
PMID 12453367 · PMC2738538 · Emerging infectious diseases · 2002 · 8 claims · 6 setups
Genomewide comparison of two complete M. tuberculosis genomes reveals substantially more nucleotide diversity than prior studies based on few loci suggested
-
Full-text index only
Evolutionary distance estimation and fidelity of pair wise sequence alignment.
PMID 15840174 · PMC1087827 · BMC bioinformatics · 2005 · 8 claims · 8 setups
Evolutionary distance estimation is relatively unaffected by alignment error as long as 50% or more of homologous sites remain identical between sequences
-
Full-text index only
Computing Ka and Ks with a consideration of unequal transitional substitutions.
PMID 16740169 · PMC1552089 · BMC evolutionary biology · 2006 · 7 claims · 7 setups
MYN, a modified version of the Yang-Nielsen (YN) algorithm based on the Tamura-Nei Model, allows unequal transitional substitution rates between purines (κR) and pyrimidines (κY) plus codon frequency bias
-
Full-text index only
Sushi gets serious: the draft genome sequence of the pufferfish Fugu rubripes.
PMID 12225591 · PMC139409 · Genome biology · 2002 · 8 claims · 7 setups
The Fugu rubripes draft genome sequence was generated by whole-genome shotgun sequencing assembled to ~5.6x coverage using the JAZZ pipeline.
-
Full-text index only
Gene duplication: the genomic trade in spare parts.
PMID 15252449 · PMC449868 · PLoS biology · 2004 · 8 claims · 7 setups
Gene duplication relaxes selective constraint on one copy, allowing exploration of evolutionary space that is otherwise forbidden in single-copy genes, making duplication the major opportunity for new gene function evolution.
-
Full-text index only
Heterogeneous genomic molecular clocks in primates.
PMID 17029560 · PMC1592237 · PLoS genetics · 2006 · 7 claims · 7 setups
Non-CpG site substitutions show clear generation-time dependency, consistent with a replication-error origin
-
Full-text index only
Identification, characterization and comparative genomics of chimpanzee endogenous retroviruses.
PMID 16805923 · PMC1779541 · Genome biology · 2006 · 8 claims · 6 setups
The chimpanzee genome contains at least 42 separate families of endogenous retroviruses, 9 newly identified
-
Full-text index only
Evidence for limited genetic compartmentalization of HIV-1 between lung and blood.
PMID 19759830 · PMC2736399 · PloS one · 2009 · 8 claims · 7 setups
Statistical evidence of genetic compartmentalization between lung and blood HIV-1 env sequences was found in 10 of 18 subjects.
-
Full-text index only
Ultraconserved coding regions outside the homeobox of mammalian Hox genes.
PMID 18816392 · PMC2566984 · BMC evolutionary biology · 2008 · 7 claims · 7 setups
Ultraconserved coding regions (UCRs, ≥120 nt with no synonymous or nonsynonymous substitutions) exist outside the homeobox in mammalian Hox genes
-
Full-text index only
The many uses of a genome sequence.
PMID 11423005 · PMC138940 · Genome biology · 2001 · 8 claims · 8 setups
Solved protein structures from structural genomics efforts can be used to model many other proteins by homology, aiding function prediction
-
Has reproduction · 94
Topological signatures in regulatory network enable phenotypic heterogeneity in small cell lung cancer.
PMID 33729159 · PMC8012062 · eLife · 2021 · 8 claims · 7 setups
The SCLC regulatory network is multistable and its steady states map onto four experimentally observed phenotypes (ASCL1high/NEUROD1low, ASCL1low/NEUROD1high, ASCL1high/NEUROD1high, ASCL1low/NEUROD1low)
-
Full-text index only
Prioritization of candidate cancer genes--an aid to oncogenomic studies.
PMID 18710882 · PMC2566894 · Nucleic acids research · 2008 · 8 claims · 8 setups
Computational classifiers using combinations of protein conservation, gene structure, protein domains, protein interactions, and regulatory data can distinguish known cancer genes (CD/CR) from unlabelled human genes