Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 82
Reusable building blocks in biological systems.
PMID 30958230 · PMC6303794 · Journal of the Royal Society, Interface · 2018 · 8 claims · 4 setups
Biological systems can be decomposed into phenotypic building blocks (PBBs) via k-maximally reusable decompositions (k-MRD) that maximize average reusability across conditions.
-
Full-text index only
Sequence determinants of human microsatellite variability.
PMID 20015383 · PMC2806349 · BMC genomics · 2009 · 6 claims · 4 setups
Mean and maximum number of repeats across individuals are positively correlated with heterozygosity
-
Full-text index only
Calculating expected DNA remnants from ancient founding events in human population genetics.
PMID 18928554 · PMC2588638 · BMC genetics · 2008 · 8 claims · 3 setups
Genetic parameters (native/migrant population size, mutation rate, generations since admixture) strongly determine the final frequency of migrant alleles detectable today.
-
Full-text index only
Direct inference of SNP heterozygosity rates and resolution of LOH detection.
PMID 18052545 · PMC2098867 · PLoS computational biology · 2007 · 6 claims · 7 setups
A large proportion of SNPs in dbSNP have high-variance HET rate estimates, limiting their reliability for LOH study design.
-
Has reproduction · 57
Genome-wide kinetic properties of transcriptional bursting in mouse embryonic stem cells.
PMID 32596448 · PMC7299619 · Science advances · 2020 · 8 claims · 8 setups
Genome-wide transcriptional bursting kinetics (intrinsic noise, burst size, frequency) can be estimated from allele-specific scRNA-seq of hybrid mESCs
-
Has reproduction · 72
Analysis of the genome of the New Zealand giant collembolan (Holacanthella duospinosa) sheds light on hexapod evolution.
PMID 29041914 · PMC5644144 · BMC genomics · 2017 · 8 claims · 8 setups
A high-quality ~375 Mbp draft genome and transcriptome of Holacanthella duospinosa was assembled and annotated, providing a genomic resource for hexapod evolution.
-
Has reproduction · 83
Accurate prediction of metagenome-assembled genome completeness by MAGISTA, a random forest model built on alignment-free intra-bin statistics.
PMID 35248155 · PMC8898458 · Environmental microbiome · 2022 · 7 claims · 7 setups
MAGISTA, a random forest model built on alignment-free intra-bin distance-distribution statistics, can estimate MAG completeness and purity without relying on reference marker genes.
-
Has reproduction · 83
SIRE 2.0: a novel method for estimating polygenic host effects underlying infectious disease transmission, and analytical expressions for prediction accuracies.
PMID 40169992 · PMC11963337 · Genetics, selection, evolution : GSE · 2025 · 8 claims · 2 setups
SIRE 2.0 is a novel Bayesian methodology and software tool for estimating polygenic contributions (variance components and additive genetic effects) to host susceptibility, infectivity and recoverability from temporal epidemic data using pedigree/genomic relationship matrices.
-
Has reproduction · 99
Verrucomicrobia are prevalent in north-temperate freshwater lakes and display class-level preferences between lake habitats.
PMID 29590198 · PMC5874073 · PloS one · 2018 · 8 claims · 8 setups
Verrucomicrobia is highly prevalent in north-temperate freshwater lakes, on average the 4th most abundant phylum (range 1.7–41.7%).
-
Full-text index only
Prioritization of candidate cancer genes--an aid to oncogenomic studies.
PMID 18710882 · PMC2566894 · Nucleic acids research · 2008 · 8 claims · 8 setups
Computational classifiers using combinations of protein conservation, gene structure, protein domains, protein interactions, and regulatory data can distinguish known cancer genes (CD/CR) from unlabelled human genes
-
Full-text index only
Modeling the amplification dynamics of human Alu retrotransposons.
PMID 16201008 · PMC1239904 · PLoS computational biology · 2005 · 8 claims · 4 setups
Combining sequence diversity (π) and insertion polymorphism level (IPL) statistics can statistically exclude implausible Alu amplification scenarios and narrow the range of plausible ones for individual subfamilies.
-
Full-text index only
Using comparative genomics to reorder the human genome sequence into a virtual sheep genome.
PMID 17663790 · PMC2323240 · Genome biology · 2007 · 8 claims · 6 setups
A sheep BAC library (CHORI-243) with ~13.5-fold genome coverage was constructed and end-sequenced.
-
Full-text index only
Analysis of concordance of different haplotype block partitioning algorithms.
PMID 16356172 · PMC1343594 · BMC bioinformatics · 2005 · 7 claims · 7 setups
Each block partitioning algorithm infers blocks differing in number, size, and coverage under different SNP density and allele frequency conditions.
-
Has reproduction · 71
Prospects of telomere-to-telomere assembly in barley: Analysis of sequence gaps in the MorexV3 reference genome.
PMID 35338551 · PMC9241371 · Plant biotechnology journal · 2022 · 7 claims · 8 setups
Almost all centromeric sequences and 45S ribosomal DNA repeat arrays are absent from the MorexV3 pseudomolecules
-
Full-text index only
Evolutionary distance estimation and fidelity of pair wise sequence alignment.
PMID 15840174 · PMC1087827 · BMC bioinformatics · 2005 · 8 claims · 8 setups
Evolutionary distance estimation is relatively unaffected by alignment error as long as 50% or more of homologous sites remain identical between sequences
-
Full-text index only
Simple models of genomic variation in human SNP density.
PMID 17553150 · PMC1919371 · BMC genomics · 2007 · 6 claims · 4 setups
Hierarchical Poisson model B, which allows both the mutation-rate proxy (Beta-distributed Λ) and the ARG-size proxy (Gamma-distributed T) to vary, fits the observed SNP density distribution significantly better than models with only one or neither varying.
-
Full-text index only
Genome-wide identification of specific oligonucleotides using artificial neural network and computational genomic analysis.
PMID 17518996 · PMC1892811 · BMC bioinformatics · 2007 · 7 claims · 4 setups
The IAB algorithm (integration of ANN and BLAST) identifies genome-wide specific oligos much faster than pure BLAST search while maintaining comparable success rate and cross homology
-
Has reproduction · 94
Systematic assessment of pathway databases, based on a diverse collection of user-submitted experiments.
PMID 36088548 · PMC9487593 · Briefings in bioinformatics · 2022 · 8 claims · 6 setups
Well-established, hierarchically organized pathway annotation systems (e.g. GO, Reactome, KEGG) yield the best overall enrichment performance despite covering much of the human genome only in general terms.
-
Has reproduction · 93
Population genomics of the Wolbachia endosymbiont in Drosophila melanogaster.
PMID 23284297 · PMC3527207 · PLoS genetics · 2012 · 8 claims · 8 setups
Wolbachia infection status can be accurately predicted in silico from whole-genome shotgun sequence of individual host strains, showing 99% concordance with diagnostic PCR.
-
Has reproduction · 69
A comparison across non-model animals suggests an optimal sequencing depth for de novo transcriptome assembly.
PMID 23496952 · PMC3655071 · BMC genomics · 2013 · 8 claims · 8 setups
Representative de novo transcriptome assemblies are generated with as few as ~20 million reads for single-tissue samples and ~30 million reads for whole animals at the mRNA-coverage level.