Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Applicability of DNA pools on 500 K SNP microarrays for cost-effective initial screens in genomewide association studies.
PMID 17610740 · PMC1925094 · BMC genomics · 2007 · 8 claims · 5 setups
SNP-MaP can be effectively applied to the Affymetrix 500K GeneChip, providing a cost-effective, reliable and valid initial genomewide screen
-
Full-text index only
MiPred: classification of real and pseudo microRNA precursors using random forest prediction model with combined features.
PMID 17553836 · PMC1933124 · Nucleic acids research · 2007 · 8 claims · 8 setups
A hybrid feature combining local contiguous triplet structure-sequence composition, MFE of the secondary structure, and P-value of a randomization test improves classification of real vs pseudo pre-miRNAs
-
Full-text index only
Cross species genomic analysis identifies a mouse model as undifferentiated pleomorphic sarcoma/malignant fibrous histiocytoma.
PMID 19956606 · PMC2779485 · PloS one · 2009 · 8 claims · 7 setups
A 100-gene signature from LSL-KrasG12D;Trp53Flox/Flox mouse sarcomas (vs normal muscle) is specifically and significantly enriched in human MFH but not other soft tissue sarcoma subtypes across three independent human datasets.
-
Full-text index only
Cross-ancestry genome-wide association studies of liver function biomarkers uncover pleiotropic variants, systemic disease links and therapeutic targets.
PMID 41689074 · PMC13005531 · Genome medicine · 2026 · 8 claims · 8 setups
5,507 lead signals (P<5x10^-9) were identified for seven LFQBs across ancestries, including 210 novel loci
-
Full-text index only
Identification of deleterious non-synonymous single nucleotide polymorphisms using sequence-derived information.
PMID 18588693 · PMC2446391 · BMC bioinformatics · 2008 · 8 claims · 5 setups
A decision tree built on 10 selected sequence-derived features classifies SAPs as Disease or Polymorphism with 82.6% accuracy and 0.607 MCC in cross-validation.
-
Full-text index only
Canine tumor cross-species genomics uncovers targets linked to osteosarcoma progression.
PMID 20028558 · PMC2803201 · BMC genomics · 2009 · 8 claims · 7 setups
High expression of IL-8 and SLC1A3, identified via cross-species mining, is associated with poor outcome in an independent population of human osteosarcoma patients
-
Full-text index only
Predicting the phenotypic effects of non-synonymous single nucleotide polymorphisms based on support vector machines.
PMID 18005451 · PMC2216041 · BMC bioinformatics · 2007 · 8 claims · 5 setups
Parepro, an SVM-based method integrating three attribute sets (RD, MI, IE) derived from evolutionary and residue-property information, predicts whether an nsSNP is deleterious or neutral.
-
Has reproduction · 90
A Decentralized Kidney Transplant Biopsy Classifier for Transplant Rejection Developed Using Genes of the Banff-Human Organ Transplant Panel.
PMID 35619722 · PMC9128066 · Frontiers in immunology · 2022 · 6 claims · 6 setups
A random forest model trained solely on B-HOT panel genes (B-HOT Model) accurately classifies kidney transplant biopsies as NR, ABMR, or TCMR.
-
Full-text index only
SNAP: predict effect of non-synonymous polymorphisms on function.
PMID 17526529 · PMC1920242 · Nucleic acids research · 2007 · 7 claims · 8 setups
SNAP, a neural network-based method using sequence-derived information, predicts whether a non-synonymous SNP is neutral or non-neutral for protein function
-
Has reproduction · 88
AuPairWise: A Method to Estimate RNA-Seq Replicability through Co-expression.
PMID 27082953 · PMC4833304 · PLoS computational biology · 2016 · 7 claims · 6 setups
Sample-sample correlation of transcript abundances is a misleading measure of replicability for assessing differential expression, because it is dominated by gene-specific dynamic ranges rather than condition-dependent variation.
-
Has reproduction · 78
Enhancing chemotherapy response prediction via matched colorectal tumor-organoid gene expression analysis and network-based biomarker selection.
PMID 39754813 · PMC11754497 · Translational oncology · 2025 · 6 claims · 8 setups
A consensus WGCNA approach combining matched tumor-organoid and independent organoid drug-response expression data identifies gene modules and hub genes predictive of 5-FU chemotherapy response
-
Full-text index only
RAPseq enables large-scale identification of RBP-RNA interactions and reveals essentials of post-transcriptional gene regulation.
PMID 41755635 · PMC12956339 · Nucleic acids research · 2026 · 7 claims · 8 setups
RAPseq is a novel in vitro, antibody-free and cross-linking-free method that profiles RBP binding to native cellular RNA transcriptome-wide using recombinant Halo-tagged RBPs and affinity purification followed by sequencing.
-
Full-text index only
Bioinformatic Analysis of Differentially Expressed Long Non-Coding RNAs in Skeletal Muscle Following Aerobic and Resistance Exercise.
PMID 41595529 · PMC12840796 · Genes · 2026 · 8 claims · 3 setups
Distinct lncRNA expression profiles exist in skeletal muscle between acute AE and RE, suggesting modality-specific lncRNA roles in exercise adaptation.
-
Full-text index only
Classification of real and pseudo microRNA precursors using local structure-sequence features and support vector machine.
PMID 16381612 · PMC1360673 · BMC bioinformatics · 2005 · 7 claims · 7 setups
A 32-dimensional triplet structure-sequence feature vector combined with SVM (triplet-SVM) can distinguish real human pre-miRNAs from pseudo pre-miRNA hairpins with ~90% accuracy.
-
Full-text index only
Speeding disease gene discovery by sequence based candidate prioritization.
PMID 15766383 · PMC1274252 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Disease genes (OMIM) differ significantly from non-disease genes in sequence-based features including gene/cDNA/protein size, exon number, homolog conservation, secretion signal, 3' UTR length, CpG islands, and distance to nearest gene.
-
Full-text index only
Comparative phosphoproteomics reveals evolutionary and functional conservation of phosphorylation across eukaryotes.
PMID 18828897 · PMC2760871 · Genome biology · 2008 · 8 claims · 8 setups
The overlap between phosphoproteomes of six eukaryotes (human, mouse, fly, yeast, plant, zebrafish) is significantly greater than expected by chance.
-
Full-text index only
Conserved elements with potential to form polymorphic G-quadruplex structures in the first intron of human genes.
PMID 18187510 · PMC2275096 · Nucleic acids research · 2008 · 8 claims · 6 setups
G-richness downstream of the TSS is strand-biased, concentrated on the nontemplate strand, with a peak at +200 to +300 bp
-
Full-text index only
Retentive Network promotes efficient RNA language modeling of long sequences.
PMID 41814064 · PMC13111708 · Communications biology · 2026 · 8 claims · 6 setups
RNAret, a RetNet-based RNA language model with O(n) complexity, achieves training parallelism and low computational overhead while processing long RNA sequences
-
Full-text index only
An SVM-based system for predicting protein subnuclear localizations.
PMID 16336650 · PMC1325059 · BMC bioinformatics · 2005 · 7 claims · 3 setups
New kernels defined on k-peptide vectors mapped by BLOSUM62-based high-scored pair matrices (D1, D2, D3) improve SVM discrimination of protein subnuclear localization compared to conventional k-peptide encodings.
-
Has reproduction · 76
Topologically inferring pathway activity toward precise cancer classification via integrating genomic and metabolomic data: prostate cancer as a case.
PMID 26286638 · PMC4541321 · Scientific reports · 2015 · 6 claims · 6 setups
DRW-GM evaluates gene topological importance by directed random walk on a global gene–metabolite pathway graph integrating gene expression and metabolomic profiles to infer pathway activities