Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Identification of disease causing loci using an array-based genotyping approach on pooled DNA.
PMID 16197552 · PMC1262713 · BMC genomics · 2005 · 8 claims · 5 setups
Pooling genomic DNA and genotyping on SNP microarrays accurately predicts allelic frequencies relative to individual genotyping
-
Full-text index only
Unusual linkage patterns of ligands and their cognate receptors indicate a novel reason for non-random gene order in the human genome.
PMID 16277660 · PMC1309615 · BMC evolutionary biology · 2005 · 8 claims · 5 setups
Ligands are not more closely linked (shorter physical distance) to their cognate receptors than expected by chance
-
Full-text index only
Direct inference of SNP heterozygosity rates and resolution of LOH detection.
PMID 18052545 · PMC2098867 · PLoS computational biology · 2007 · 6 claims · 7 setups
A large proportion of SNPs in dbSNP have high-variance HET rate estimates, limiting their reliability for LOH study design.
-
Has reproduction · 83
Accurate prediction of metagenome-assembled genome completeness by MAGISTA, a random forest model built on alignment-free intra-bin statistics.
PMID 35248155 · PMC8898458 · Environmental microbiome · 2022 · 7 claims · 7 setups
MAGISTA, a random forest model built on alignment-free intra-bin distance-distribution statistics, can estimate MAG completeness and purity without relying on reference marker genes.
-
Full-text index only
A modified T-test feature selection method and its application on the HapMap genotype data.
PMID 18267305 · PMC5054219 · Genomics, proteomics & bioinformatics · 2007 · 7 claims · 4 setups
A modified t-test ranking measure, extended to handle nominal SNP genotype data via vector transformation, can effectively rank SNPs by their discriminative capability for population classification.
-
Full-text index only
"Reverse ecology" and the power of population genomics.
PMID 18752601 · PMC2626434 · Evolution; international journal of organic evolution · 2008 · 8 claims · 7 setups
Population genomic data can be used to rapidly identify genes targeted by adaptive natural selection, an approach termed 'reverse ecology'.
-
Has reproduction · 75
Genomic regions and candidate genes selected during the breeding of rice in Vietnam.
PMID 35899250 · PMC9309459 · Evolutionary applications · 2022 · 8 claims · 7 setups
XP-CLR and FST scans identify genomic regions with distorted allele frequency/differentiation patterns resulting from differential selective pressures between Vietnamese rice subpopulations
-
Full-text index only
Association between telomere length and V(H) gene mutation status in chronic lymphocytic leukaemia: clinical and biological implications.
PMID 12592375 · PMC2377180 · British journal of cancer · 2003 · 7 claims · 5 setups
Unmutated VH gene CLL cases have significantly shorter telomeres than mutated VH gene CLL cases
-
Full-text index only
Expoldb: expression linked polymorphism database with inbuilt tools for analysis of expression and simple repeats.
PMID 17038195 · PMC1618849 · BMC genomics · 2006 · 8 claims · 6 setups
EXPOLDB is a novel database integrating human gene expression variability data (including monozygotic twin comparisons) with (TG/CA)n repeat polymorphism information
-
Full-text index only
Development of animal models to test the fundamental basis of gene-environment interactions.
PMID 19037209 · PMC2703424 · Obesity (Silver Spring, Md.) · 2008 · 8 claims · 8 setups
Selective breeding for low and high intrinsic aerobic treadmill running capacity produced divergent rat lines (LCR and HCR) that contrast in propensity for complex disease
-
Full-text index only
Correlation between pre-treatment quasispecies complexity and treatment outcome in chronic HCV genotype 3a.
PMID 18613968 · PMC2483966 · Virology journal · 2008 · 7 claims · 7 setups
Quasispecies complexity and diversity within HVR1 are lower in the SVR group than in the TF group
-
Full-text index only
A statistical model to identify differentially expressed proteins in 2D PAGE gels.
PMID 19763172 · PMC2734266 · PLoS computational biology · 2009 · 7 claims · 5 setups
A mixture likelihood model incorporating both detected and non-detected proteins has higher statistical power to detect differential expression than standard approaches like the Student's t-test.
-
Has reproduction · 71
Protein structure quality assessment based on the distance profiles of consecutive backbone Cα atoms.
PMID 24555103 · PMC3892923 · F1000Research · 2013 · 8 claims · 8 setups
The distance between consecutive backbone Cα atoms in high-quality structures is normally distributed with mean 3.8 Å and standard deviation 0.04 Å, justifying a reference state in which all consecutive Cα atoms are 3.8 Å apart.
-
Full-text index only
Prevalence of variations in melanoma susceptibility genes among Slovenian melanoma families.
PMID 18803811 · PMC2556318 · BMC medical genetics · 2008 · 8 claims · 7 setups
CDKN2A germline mutations were found in 7 of 25 (28.0%) Slovenian melanoma families
-
Full-text index only
MiPred: classification of real and pseudo microRNA precursors using random forest prediction model with combined features.
PMID 17553836 · PMC1933124 · Nucleic acids research · 2007 · 8 claims · 8 setups
A hybrid feature combining local contiguous triplet structure-sequence composition, MFE of the secondary structure, and P-value of a randomization test improves classification of real vs pseudo pre-miRNAs
-
Full-text index only
Rapid detection of genomic imbalances using micro-arrays consisting of pooled BACs covering all human chromosome arms.
PMID 16221972 · PMC1253841 · Nucleic acids research · 2005 · 8 claims · 6 setups
Reducing array complexity by pooling five BACs per spot (covering a chromosome arm) increases robustness to amplification-related ratio variation compared with single-BAC spotting
-
Full-text index only
Genome-wide identification of specific oligonucleotides using artificial neural network and computational genomic analysis.
PMID 17518996 · PMC1892811 · BMC bioinformatics · 2007 · 7 claims · 4 setups
The IAB algorithm (integration of ANN and BLAST) identifies genome-wide specific oligos much faster than pure BLAST search while maintaining comparable success rate and cross homology
-
Full-text index only
Integrative analysis of RUNX1 downstream pathways and target genes.
PMID 18671852 · PMC2529319 · BMC genomics · 2008 · 7 claims · 8 setups
Integrating gene expression profiles from three independent RUNX1 perturbation platforms (FPD-AML patient cell lines, RUNX1/CBFβ overexpression in HeLa cells, Runx1 knockout mouse embryos) identifies RUNX1-regulated genes and downstream pathways
-
Has reproduction · 50
Exploiting convergent phenotypes to derive a pan-cancer cisplatin response gene expression signature.
PMID 37076665 · PMC10115855 · NPJ precision oncology · 2023 · 8 claims · 8 setups
A convergent-phenotype-based seed gene/co-expression method can extract consensus gene expression signatures predictive of response to chemotherapeutic drugs in the GDSC database
-
Full-text index only
Species-specific protein sequence and fold optimizations.
PMID 12487631 · PMC139977 · BMC bioinformatics · 2002 · 7 claims · 7 setups
Environmental niche is a significant factor explaining variability in amino acid composition across 100 complete genomes