Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Computational disease gene identification: a concert of methods prioritizes type 2 diabetes and obesity candidate genes.
PMID 16757574 · PMC1475747 · Nucleic acids research · 2006 · 6 claims · 8 setups
Applying seven independent computational disease-gene prioritization methods in concert to 9556 positional candidate genes identifies a prioritized set of likely T2D and obesity candidate genes
-
Full-text index only
Genome-wide identification of human functional DNA using a neutral indel model.
PMID 16410828 · PMC1326222 · PLoS computational biology · 2006 · 8 claims · 8 setups
A neutral indel model predicting a geometric distribution of intergap segment (IGS) lengths fits human-mouse ancestral repeat (AR) alignment data excellently
-
Full-text index only
Computation of haplotypes on SNPs subsets: advantage of the "global method".
PMID 17067372 · PMC1636337 · BMC genetics · 2006 · 6 claims · 4 setups
The global method for subhaplotyping always yields a lower error rate than the direct method across datasets and SNP subset sizes
-
Full-text index only
In silico whole-genome screening for cancer-related single-nucleotide polymorphisms located in human mRNA untranslated regions.
PMID 17201911 · PMC1774567 · BMC genomics · 2007 · 8 claims · 5 setups
A computational EST-based pipeline can identify UTR-SNPs that are statistically over-represented in cancerous versus normal tissue libraries
-
Full-text index only
Estimation of relevant variables on high-dimensional biological patterns using iterated weighted kernel functions.
PMID 18509521 · PMC2396875 · PloS one · 2008 · 7 claims · 6 setups
wKIERA combines a weighted-kernel discriminant (kernel perceptron) with an iterative stochastic probability estimation-of-distribution algorithm to estimate a relevance distribution over variables
-
Full-text index only
PRDM1 Is Associated with Chemoradiotherapy-Associated Enrichment of Adaptive NK Cells in Cervical Cancer.
PMID 42110212 · PMC13150072 · Computational and structural biotechnology journal · 2026 · 8 claims · 8 setups
CRT enriches adaptive NK (aNK) cells within the cervical cancer immune microenvironment, with enrichment apparent after the first fraction and persisting after the second
-
Has reproduction · 32
Developing prognostic gene panel of survival time in lung adenocarcinoma patients using machine learning.
PMID 35117753 · PMC8799101 · Translational cancer research · 2020 · 8 claims · 5 setups
Naïve Bayes using a 22-gene panel is the best-performing and most stable machine learning model for predicting LUAD survival time (>3 vs <3 years)
-
Full-text index only
Identifying synonymous regulatory elements in vertebrate genomes.
PMID 15980499 · PMC1160227 · Nucleic acids research · 2005 · 7 claims · 4 setups
SynoR is a tool that performs de novo genome-wide identification of synonymous regulatory elements (SREs) using evolutionarily conserved TFBS modules as seeds
-
Full-text index only
Biocomputing enters its adolescence.
PMID 15960815 · PMC1175967 · Genome biology · 2005 · 8 claims · 8 setups
A 'match augmentation' algorithm efficiently matches structural motifs by prioritizing functionally significant residues, enabling function prediction between evolutionarily unrelated proteins
-
Full-text index only
SW-ARRAY: a dynamic programming solution for the identification of copy-number changes in genomic DNA using array comparative genome hybridization data.
PMID 15961730 · PMC1151590 · Nucleic acids research · 2005 · 7 claims · 5 setups
SW-ARRAY, an adaptation of the Smith-Waterman dynamic programming algorithm, provides a sensitive and robust method for identifying copy-number changes in array CGH data
-
Has reproduction · 59
The rates of adult neurogenesis and oligodendrogenesis are linked to cell cycle regulation through p27-dependent gene repression of SOX2.
PMID 36627412 · PMC9832098 · Cellular and molecular life sciences : CMLS · 2023 · 8 claims · 8 setups
p27 restricts residual CDK activity after mitogen withdrawal to antagonize cell cycling, but is not essential for cell cycle exit per se
-
Full-text index only
Exploration of the omics evidence landscape: adding qualitative labels to predicted protein-protein interactions.
PMID 17880677 · PMC2375035 · Genome biology · 2007 · 7 claims · 8 setups
Combining pairs of omics evidence types into two-dimensional 'evidence landscapes' allows regions to be identified that specifically and purely predict either physical or metabolic protein interactions
-
Full-text index only
An SVM-based system for predicting protein subnuclear localizations.
PMID 16336650 · PMC1325059 · BMC bioinformatics · 2005 · 7 claims · 3 setups
New kernels defined on k-peptide vectors mapped by BLOSUM62-based high-scored pair matrices (D1, D2, D3) improve SVM discrimination of protein subnuclear localization compared to conventional k-peptide encodings.
-
Full-text index only
Inferring combinatorial regulation of transcription in silico.
PMID 15647509 · PMC546154 · Nucleic acids research · 2005 · 8 claims · 5 setups
Combining Cluster-Buster (TFBS cluster prediction) with GOSSIP (rigorous GO enrichment statistics with multiple-testing/FDR correction) predicts biological functions controlled by combinatorial transcription factor action, without prior knowledge of factor targets
-
Full-text index only
CTCFBSDB: a CTCF-binding site database for characterization of vertebrate genomic insulators.
PMID 17981843 · PMC2238977 · Nucleic acids research · 2008 · 7 claims · 8 setups
CTCF is the only identified trans-acting factor in vertebrates that confers enhancer-blocking insulator activity
-
Full-text index only
Global sequencing of proteolytic cleavage sites in apoptosis by specific labeling of protein N termini.
PMID 18722006 · PMC2566540 · Cell · 2008 · 7 claims · 8 setups
A subtiligase-based N-terminal biotinylation and enrichment method enables global identification and sequencing of protease cleavage sites in complex mixtures
-
Full-text index only
Spatial transcriptomics maps host-gut microbiome biogeography at high resolution.
PMID 41792309 · PMC13171632 · Nature microbiology · 2026 · 7 claims · 6 setups
Enzymatic in situ polyadenylation increases bacterial RNA recovery in oligo(dT)-based spatial transcriptomics arrays by up to ~100-fold while preserving host gene capture
-
Full-text index only
Features affecting Cas9-induced editing efficiency and patterns in tomato: evidence from a large CRISPR dataset.
PMID 41877594 · PMC13014117 · The Plant journal : for cell and molecular biology · 2026 · 8 claims · 5 setups
Chromatin accessibility significantly increases editing efficiency, with higher editing at targets in accessible versus inaccessible chromatin.
-
Has reproduction · 67
Adaptive learning embedding features to improve the predictive performance of SARS-CoV-2 phosphorylation sites.
PMID 37847658 · PMC10628388 · Bioinformatics (Oxford, England) · 2023 · 8 claims · 6 setups
PSPred-ALE outperforms state-of-the-art SARS-CoV-2 phosphorylation site predictors (e.g. DeepIPs) and handcrafted feature-based methods in benchmarking comparisons
-
Has reproduction · 71
Transcriptome and machine learning analysis of the impact of COVID-19 on mitochondria and multiorgan damage.
PMID 38295140 · PMC10830027 · PloS one · 2024 · 6 claims · 7 setups
Potential cardiac, hepatic, and renal impairments in COVID-19 are associated with ACE2, inflammatory cytokine storms, and mitochondrial pathways.