Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Computational analysis of splicing errors and mutations in human transcripts.
PMID 18194514 · PMC2234086 · BMC genomics · 2008 · 8 claims · 4 setups
Retained introns are significantly shorter than constitutively spliced introns
-
Full-text index only
A unique, consistent identifier for alternatively spliced transcript variants.
PMID 19865484 · PMC2765725 · PloS one · 2009 · 6 claims · 1 setups
Existing transcript identifiers (NM_ accessions, ENST identifiers) are unsuitable for uniquely identifying isoform structure across databases, methods, or organisms
-
Full-text index only
Twin peaks: the draft human genome sequence.
PMID 11276423 · PMC138909 · Genome biology · 2001 · 8 claims · 8 setups
The predicted number of human genes (~26,000-40,000) is far lower than the widely assumed ~100,000, though downstream RNA/protein complexity can still generate substantial biological complexity.
-
Full-text index only
High fidelity of whole-genome amplified DNA on high-density single nucleotide polymorphism arrays.
PMID 18786630 · PMC2659594 · Genomics · 2008 · 8 claims · 7 setups
WGA product performs well on the Affymetrix 250K SNP array compared to genomic DNA, especially with the BRLMM calling algorithm.
-
Full-text index only
Quantitative serum proteomics using dual stable isotope coding and nano LC-MS/MSMS.
PMID 19817497 · PMC4684172 · Journal of proteome research · 2009 · 7 claims · 6 setups
DSIC labeling achieves high efficiency: 100% for Cysteine (acrylamide) and 98% for Lysine (succinic anhydride)
-
Has reproduction · 92
Large-scale integration of single-cell transcriptomic data captures transitional progenitor states in mouse skeletal muscle regeneration.
PMID 34773081 · PMC8589952 · Communications biology · 2021 · 8 claims · 7 setups
Large-scale integration of 111 sc/snRNAseq datasets captures rare, transitional myogenic progenitor states (commitment and fusion) that are poorly represented in individual datasets.
-
Full-text index only
A molecular signature of epithelial host defense: comparative gene expression analysis of cultured bronchial epithelial cells and keratinocytes.
PMID 16420688 · PMC1382211 · BMC genomics · 2006 · 7 claims · 4 setups
PBEC and KC SAGE libraries show approximately 80% overlap in expressed tags, indicating high similarity in gene repertoire
-
Has reproduction · 73
GREIN: An Interactive Web Platform for Re-analyzing GEO RNA-seq Data.
PMID 31110304 · PMC6527554 · Scientific reports · 2019 · 8 claims · 7 setups
GREIN is a web application providing user-friendly interfaces to manipulate, visualize, and analyze GEO RNA-seq data.
-
Full-text index only
Genomic rearrangements by LINE-1 insertion-mediated deletion in the human and chimpanzee lineages.
PMID 16034026 · PMC1179734 · Nucleic acids research · 2005 · 8 claims · 6 setups
L1 insertions are directly responsible for genomic deletions (L1IMDs) confirmed in both human and chimpanzee genomes
-
Has reproduction · 89
Spatial information matters: are traditional imputation methods effective for spatial transcriptomics data?
PMID 41627342 · PMC12862982 · Briefings in bioinformatics · 2026 · 7 claims · 3 setups
No single existing SOTA imputation method consistently performs well across newer SRT platforms/datasets
-
Has reproduction · 62
Bayesian prediction of RNA translation from ribosome profiling.
PMID 28126919 · PMC5389577 · Nucleic acids research · 2017 · 8 claims · 4 setups
Rp-Bp is an unsupervised Bayesian approach that uses a two-component 'high-low-low' mixture model to predict translated ORFs from ribosome profiles
-
Full-text index only
Flanking p10 contribution and sequence bias in matrix based epitope prediction: revisiting the assumption of independent binding pockets.
PMID 18925947 · PMC2600787 · BMC structural biology · 2008 · 8 claims · 3 setups
The extended matrix PP10 (built from a proline-containing peptide library) shows significant improvement in binding prediction over the original nine-residue matrix P9
-
Full-text index only
A general definition and nomenclature for alternative splicing events.
PMID 18688268 · PMC2467475 · PLoS computational biology · 2008 · 6 claims · 4 setups
Existing AS nomenclatures (Malko et al.'s 5-letter strings, Nagasaki et al.'s bit matrices, and the ASD/ATD/AEdb system) are redundant, ambiguous, or incapable of representing complex or large splicing variations.
-
Full-text index only
An SVM-based system for predicting protein subnuclear localizations.
PMID 16336650 · PMC1325059 · BMC bioinformatics · 2005 · 7 claims · 3 setups
New kernels defined on k-peptide vectors mapped by BLOSUM62-based high-scored pair matrices (D1, D2, D3) improve SVM discrimination of protein subnuclear localization compared to conventional k-peptide encodings.
-
Full-text index only
A novel wavelet-based thresholding method for the pre-processing of mass spectrometry data that accounts for heterogeneous noise.
PMID 18615428 · PMC2855839 · Proteomics · 2008 · 6 claims · 4 setups
Noise in SELDI-TOF/MALDI-TOF mass spectrometry data is heteroscedastic across the m/z range, with larger variance at lower m/z values, contrary to the homogeneous noise assumption of existing wavelet denoising methods.
-
Full-text index only
Importance sampling for the infinite sites model.
PMID 18976228 · PMC2832804 · Statistical applications in genetics and molecular biology · 2008 · 7 claims · 2 setups
A new importance sampling proposal distribution for the ISM, derived from a new result on exact sampling from a single segregating site, generally shows greater efficiency than the GT and SD proposals.
-
Full-text index only
Computation of haplotypes on SNPs subsets: advantage of the "global method".
PMID 17067372 · PMC1636337 · BMC genetics · 2006 · 6 claims · 4 setups
The global method for subhaplotyping always yields a lower error rate than the direct method across datasets and SNP subset sizes
-
Full-text index only
Prediction-based approaches to characterize bidirectional promoters in the mammalian genome.
PMID 18366609 · PMC2386062 · BMC genomics · 2008 · 8 claims · 7 setups
The mapping algorithm identified 5,647 candidate bidirectional promoter regions in the mouse genome, similar in number to those previously found in human.
-
Full-text index only
KEGG for linking genomes to life and the environment.
PMID 18077471 · PMC2238879 · Nucleic acids research · 2008 · 8 claims · 4 setups
KEGG provides a reference knowledge base for linking genomes to life via PATHWAY mapping and to the environment via BRITE mapping.
-
Full-text index only
A genome-wide survey demonstrates widespread non-linear mRNA in expressed sequences from multiple species.
PMID 16237125 · PMC1258171 · Nucleic acids research · 2005 · 8 claims · 6 setups
A genome-wide computational survey identifies 245 genes in mammals (264 across six species) that produce RREO events in expressed sequences