Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 87
Translation affects mRNA stability in a codon-dependent manner in human cells.
PMID 31012849 · PMC6529216 · eLife · 2019 · 8 claims · 8 setups
Translation strongly affects mRNA stability in a codon-dependent manner in human cells, with specific codons stabilizing or destabilizing mRNAs.
-
Has reproduction · 94
An archaeal histone-like protein regulates gene expression in response to salt stress.
PMID 34883507 · PMC8682779 · Nucleic acids research · 2021 · 8 claims · 4 setups
HpyA is important for maintaining wild-type growth rate under reduced salinity
-
Full-text index only
Multilocus analysis of SNP and metabolic data within a given pathway.
PMID 16412218 · PMC1382210 · BMC genomics · 2006 · 8 claims · 7 setups
The combinatorial partitioning method (CPM) with optimal thresholds can identify SNPs associated with quantitative metabolite levels rather than only categorical traits.
-
Has reproduction · 43
Integration of Dual Stress Transcriptomes and Major QTLs from a Pair of Genotypes Contrasting for Drought and Chronic Nitrogen Starvation Identifies Key Stress Responsive Genes in Rice.
PMID 34089405 · PMC8179884 · Rice (New York, N.Y.) · 2021 · 8 claims · 7 setups
N22 performs better under dual (low N + low water) stress owing to better root architecture, chlorophyll/porphyrin synthesis and oxidative stress management
-
Full-text index only
Testing whether genetic variation explains correlation of quantitative measures of gene expression, and application to genetic network analysis.
PMID 18444230 · PMC2729096 · Statistics in medicine · 2008 · 8 claims · 3 setups
A statistical test (delta method and Steiger-Browne optimal linear composites) is developed to test equality of the marginal correlation and the partial correlation of two gene expression traits conditional on a set of covariates.
-
Full-text index only
Prediction of catalytic residues using Support Vector Machine with selected protein sequence and structural properties.
PMID 16790052 · PMC1534064 · BMC bioinformatics · 2006 · 8 claims · 7 setups
The Sequential Minimal Optimization (SMO) SVM algorithm was the best-performing classifier among 26 WEKA classifiers for predicting catalytic residues
-
Full-text index only
Constructing support vector machine ensembles for cancer classification based on proteomic profiling.
PMID 16689692 · PMC5173238 · Genomics, proteomics & bioinformatics · 2005 · 7 claims · 4 setups
CSVME, built by selecting a subset of base SVMs via SVM-RFE ranking and fusing them with a trained upper-layer SVM, achieves better classification performance than an ensemble of all base SVMs.
-
Full-text index only
Optimality driven nearest centroid classification from genomic data.
PMID 17912341 · PMC1991588 · PloS one · 2007 · 7 claims · 5 setups
A theoretical result determines the subset of features of a given size that minimizes the misclassification rate for a nearest-centroid (LDA) classifier, based on equation (4).
-
Full-text index only
iMapper: a web application for the automated analysis and mapping of insertional mutagenesis sequence data against Ensembl genomes.
PMID 18974167 · PMC2639305 · Bioinformatics (Oxford, England) · 2008 · 6 claims · 3 setups
iMapper is a web application for automated analysis and mapping of insertional mutagenesis sequence data against vertebrate and invertebrate Ensembl genomes (human, mouse, rat, zebrafish, Drosophila, S. cerevisiae).
-
Has reproduction · 83
Hobbes: optimized gram-based methods for efficient read alignment.
PMID 22199254 · PMC3315303 · Nucleic acids research · 2012 · 8 claims · 4 setups
Hobbes, a gram-based short-read mapper supporting Hamming and edit distance, is faster than all other read-mapping programs tested while maintaining high mapping quality.
-
Full-text index only
A statistical approach designed for finding mathematically defined repeats in shotgun data and determining the length distribution of clone-inserts.
PMID 15626332 · PMC5172250 · Genomics, proteomics & bioinformatics · 2003 · 8 claims · 6 setups
Repeats of different copy number have distinct probabilities of appearance in shotgun data, which can be modeled statistically to define recognition thresholds (MDRs) at different shotgun coverages.
-
Full-text index only
BTW: a web server for Boltzmann time warping of gene expression time series.
PMID 16845055 · PMC1538860 · Nucleic acids research · 2006 · 5 claims · 4 setups
Symmetric time warping distance is more flexible than Euclidean distance or correlation coefficient for identifying genes with similar temporal expression profiles, especially across sequences of different length.
-
Full-text index only
From genomics to chemical genomics: new developments in KEGG.
PMID 16381885 · PMC1347464 · Nucleic acids research · 2006 · 8 claims · 5 setups
KEGG BRITE has been formally added as a fourth main KEGG database to establish a logical foundation for functional interpretation and pathway reconstruction.
-
Full-text index only
Efficacy assessment of SNP sets for genome-wide disease association studies.
PMID 17726055 · PMC2034459 · Nucleic acids research · 2007 · 6 claims · 4 setups
τ, derived from Shannon entropy and swept radius ɛ, approximates the relative sample size efficiency of a marker set for mapping a causal variant at a given map position compared to a maximally polymorphic SNP
-
Full-text index only
Clustering of phosphorylation site recognition motifs can be exploited to predict the targets of cyclin-dependent kinase.
PMID 17316440 · PMC1852407 · Genome biology · 2007 · 8 claims · 6 setups
CDK consensus motifs are frequently clustered (closely spaced) in known CDK substrate proteins rather than uniformly distributed
-
Full-text index only
Sequence similarity network reveals common ancestry of multidomain proteins.
PMID 18475320 · PMC2377100 · PLoS computational biology · 2008 · 8 claims · 6 setups
Traditional homology definitions do not capture multidomain evolution; the authors extend the definition to include domain insertion via a common ancestral locus model.
-
Full-text index only
An efficient method for the prediction of deleterious multiple-point mutations in the secondary structure of RNAs using suboptimal folding solutions.
PMID 18445289 · PMC2386494 · BMC bioinformatics · 2008 · 8 claims · 6 setups
Using RNAsubopt suboptimal solutions computed once for the wild-type sequence, specific multiple-point mutations likely to cause conformational rearrangement can be selected without brute-force enumeration.
-
Full-text index only
Evolutionary algorithms for the selection of single nucleotide polymorphisms.
PMID 12875658 · PMC183839 · BMC bioinformatics · 2003 · 8 claims · 3 setups
Evolutionary algorithms are well suited to multiobjective optimization problems with large, intractable search spaces such as SNP selection, unlike exact methods (exhaustive enumeration) or single-objective search techniques (tabu search, simulated annealing).
-
Has reproduction · 84
Strong population differentiation in lingcod (Ophiodon elongatus) is driven by a small portion of the genome.
PMID 33294007 · PMC7691466 · Evolutionary applications · 2020 · 7 claims · 8 setups
Lingcod comprise two distinct genetic clusters separated latitudinally at a break near Point Reyes off Northern California, with a high frequency of admixed individuals near the break.
-
Has reproduction · 64
Environmental Drivers of Genetic Divergence in Two Corals From the Florida Keys.
PMID 40589611 · PMC12206660 · Evolutionary applications · 2025 · 8 claims · 8 setups
Temperature was not the most important driver of coral genetic divergence in the Florida Keys; depth and water chemistry were more important