Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 62
Gbdmr: identifying differentially methylated CpG regions in the human genome via generalized beta regressions.
PMID 38443825 · PMC10916021 · BMC bioinformatics · 2024 · 8 claims · 4 setups
gbdmr models DNA methylation levels of CpG sites using a generalized beta distribution instead of assuming normality as in linear-regression-based methods
-
Has reproduction · 83
MetaGT: A pipeline for de novo assembly of metatranscriptomes with the aid of metagenomic data.
PMID 36386613 · PMC9651917 · Frontiers in microbiology · 2022 · 7 claims · 4 setups
MetaGT is a pipeline that combines metatranscriptomic and metagenomic data from the same sample to assemble complete transcript sequences
-
Full-text index only
Benchmarking tools for the alignment of functional noncoding DNA.
PMID 14736341 · PMC344529 · BMC bioinformatics · 2004 · 8 claims · 4 setups
Global alignment tools (Avid, ClustalW, Lagan, Needle, DiAlign-G) typically have higher sensitivity over entire noncoding sequences and within constrained blocks than local tools
-
Full-text index only
Whole genome association mapping by incompatibilities and local perfect phylogenies.
PMID 17042942 · PMC1624851 · BMC bioinformatics · 2006 · 8 claims · 8 setups
Blossoc scores the perfect phylogenetic tree spanning the largest compatible region around each marker as a decision tree for case/control status to detect association
-
Full-text index only
Analyses and comparison of accuracy of different genotype imputation methods.
PMID 18958166 · PMC2569208 · PloS one · 2008 · 8 claims · 3 setups
Stronger LD produces higher imputation accuracy rates for all five methods
-
Has reproduction · 76
Tracing human genetic histories and natural selection with precise local ancestry inference.
PMID 40379651 · PMC12084304 · Nature communications · 2025 · 7 claims · 7 setups
Orchestra, a two-stage LAI method combining a recombination-distance base layer with a deep learning (convolutional + attention) smoothing module, outperforms RFmix, FLARE and Gnomix in precision and recall across simulated admixture generations.
-
Full-text index only
Detecting transcriptionally active regions using genomic tiling arrays.
PMID 16859498 · PMC1779562 · Genome biology · 2006 · 8 claims · 4 setups
A non-parametric method (TranscriptionDetector) integrates single-channel p-values from multiple replicate arrays into a multi-channel p-value (MCPV) to identify transcribed probed loci without assumptions about intensity distributions.
-
Full-text index only
Direct maximum parsimony phylogeny reconstruction from genotype data.
PMID 18053244 · PMC2222657 · BMC bioinformatics · 2007 · 6 claims · 4 setups
The paper presents the first practical method for computing maximum parsimony phylogenies directly from genotype data, using integer linear programming.
-
Full-text index only
Calculating expected DNA remnants from ancient founding events in human population genetics.
PMID 18928554 · PMC2588638 · BMC genetics · 2008 · 8 claims · 3 setups
Genetic parameters (native/migrant population size, mutation rate, generations since admixture) strongly determine the final frequency of migrant alleles detectable today.
-
Full-text index only
The distribution of SNPs in human gene regulatory regions.
PMID 16209714 · PMC1260019 · BMC genomics · 2005 · 8 claims · 6 setups
SNPs occur with higher density closer to the transcriptional start site within gene promoter regions than in further upstream regions
-
Full-text index only
A third approach to gene prediction suggests thousands of additional human transcribed regions.
PMID 16543943 · PMC1391917 · PLoS computational biology · 2006 · 8 claims · 7 setups
A third basic concept for gene prediction exists, based on detecting strand-specific 'transcription footprints' (mutational and selectional biases) rather than gene structure or sequence similarity.
-
Full-text index only
Comparative genomic analysis of the gut bacterium Bifidobacterium longum reveals loci susceptible to deletion during pure culture growth.
PMID 18505588 · PMC2430713 · BMC genomics · 2008 · 8 claims · 8 setups
Comparative genomics of B. longum DJO10A (minimally cultured) and NCC2705 (culture collection strain) reveals 17 unique DNA regions in DJO10A and 6 in NCC2705 despite otherwise high genome collinearity and identity
-
Full-text index only
The genomic distribution of intraspecific and interspecific sequence divergence of human segmental duplications relative to human/chimpanzee chromosomal rearrangements.
PMID 18699995 · PMC2542386 · BMC genomics · 2008 · 8 claims · 5 setups
Some relatively recent (young) SDs accumulate in regions homologous to chromosomal inversions that occurred in the sister lineage
-
Full-text index only
Detecting natural selection by empirical comparison to random regions of the genome.
PMID 19783549 · PMC2778377 · Human molecular genetics · 2009 · 8 claims · 5 setups
Comparing candidate loci to empirically matched random genomic regions (ENCODE data) avoids the strong demographic/mutation assumptions required by theoretical neutral models and provides a robust test for selection
-
Full-text index only
SNP@Evolution: a hierarchical database of positive selection on the human genome.
PMID 19732458 · PMC2755008 · BMC evolutionary biology · 2009 · 7 claims · 6 setups
SNP@Evolution is a hierarchical database integrating HET, FST, and iHS from HapMap Phase II and III to identify genome-wide positive selection signals
-
Has reproduction · 50
MoDLE: high-performance stochastic modeling of DNA loop extrusion interactions.
PMID 36451166 · PMC9710047 · Genome biology · 2022 · 7 claims · 6 setups
MoDLE is a high-performance stochastic model that simulates DNA-DNA contacts from loop extrusion genome-wide in minutes using less than 1 GB of RAM
-
Full-text index only
Commonality of functional annotation: a method for prioritization of candidate genes from genome-wide linkage studies.
PMID 18263617 · PMC2275105 · Nucleic acids research · 2008 · 8 claims · 7 setups
Genes correlated with a common complex trait are more likely to share GO functional annotations than genes not correlated with that trait
-
Full-text index only
POCUS: mining genomic sequence annotation to predict disease genes.
PMID 14611661 · PMC329128 · Genome biology · 2003 · 8 claims · 6 setups
Genes predisposing to the same disease tend to share functional annotation IDs (GO/InterPro) more than expected by chance
-
Full-text index only
Positive selection for the male functionality of a co-retroposed gene in the hominoids.
PMID 19832993 · PMC2773790 · BMC evolutionary biology · 2009 · 8 claims · 8 setups
PIPSL is an extraordinary co-retroposed protein-coding gene that may participate in male-specific functions of humans and close relatives
-
Full-text index only
Assessing the genomic evidence for conserved transcribed pseudogenes under selection.
PMID 19754956 · PMC2753554 · BMC genomics · 2009 · 8 claims · 8 setups
1750 transcribed pseudogene annotations (TPAs) were identified in the human genome, ~11.5% of all human pseudogene annotations.