Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 77
SurvConvMixer: robust and interpretable cancer survival prediction based on ConvMixer using pathway-level gene expression images.
PMID 38539106 · PMC10967213 · BMC bioinformatics · 2024 · 7 claims · 6 setups
SurvConvMixer reformats KEGG Pathways-in-Cancer gene expression values into pathway-level 2D gene expression images and applies a ConvMixer-based model for overall survival prediction
-
Full-text index only
Inferring combinatorial regulation of transcription in silico.
PMID 15647509 · PMC546154 · Nucleic acids research · 2005 · 8 claims · 5 setups
Combining Cluster-Buster (TFBS cluster prediction) with GOSSIP (rigorous GO enrichment statistics with multiple-testing/FDR correction) predicts biological functions controlled by combinatorial transcription factor action, without prior knowledge of factor targets
-
Full-text index only
PlasmoDraft: a database of Plasmodium falciparum gene function predictions based on postgenomic data.
PMID 18925948 · PMC2605471 · BMC bioinformatics · 2008 · 8 claims · 4 setups
Gonna, a supervised k-nearest-neighbor Guilt-By-Association predictor, proposes GO annotations for a gene based on similarity of its transcriptome, proteome, or interactome profile to genes already annotated by GeneDB
-
Full-text index only
Reconstruction of human protein interolog network using evolutionary conserved network.
PMID 17493278 · PMC1885812 · BMC bioinformatics · 2007 · 8 claims · 7 setups
A relative conservation score derived from maximal quasi-cliques in protein interaction networks, combined with other interaction features, can score and rank predicted human interologs for confidence.
-
Full-text index only
Non-EST-based prediction of novel alternatively spliced cassette exons with cell signaling function in Caenorhabditis elegans and human.
PMID 17452356 · PMC1904267 · Nucleic acids research · 2007 · 8 claims · 7 setups
PASE (Prediction of Alternative Signaling Exons) is a computational algorithm combining Markov splice-site models, a Bayesian classifier, species conservation, and Scansite motif scoring to identify novel alternative cassette exons involved in cell signaling.
-
Full-text index only
The Princeton Protein Orthology Database (P-POD): a comparative genomics analysis tool for biologists.
PMID 17712414 · PMC1942082 · PloS one · 2007 · 8 claims · 5 setups
P-POD is the first comparative genomics database to combine results from multiple computational ortholog/homolog prediction methods with manually curated literature-derived experimental evidence of functional conservation.
-
Full-text index only
Protein function assignment through mining cross-species protein-protein interactions.
PMID 18253506 · PMC2216687 · PloS one · 2008 · 8 claims · 6 setups
CSIDOP predicts protein molecular function with 95.42% accuracy using 2,972 GO functional categories in H. sapiens
-
Has reproduction · 69
A hybrid empirical and parametric approach for managing ecosystem complexity: Water quality in Lake Geneva under nonstationary futures.
PMID 35733249 · PMC9245694 · Proceedings of the National Academy of Sciences of the United States of America · 2022 · 7 claims · 6 setups
A hybrid model combining empirical dynamic modeling (EDM/S-map) for biogeochemical sink terms with the equation-based Simstrat hydrodynamic model for physical source terms produces substantially better historical DOB forecasts than conventional parametric models.
-
Full-text index only
Functional annotation and identification of candidate disease genes by computational analysis of normal tissue gene expression data.
PMID 18560577 · PMC2409962 · PloS one · 2008 · 7 claims · 5 setups
Ranked Coexpression Groups (RCG) built from k=6 nearest coexpressed genes, combined with a majority-rule functional characterization, integrate multiple datasets/coexpression measures to generate high-confidence functional annotation predictions
-
Has reproduction · 50
Genetic parallels in biomineralization of the calcareous sponge Sycon ciliatum and stony corals.
PMID 40922549 · PMC12419799 · eLife · 2025 · 8 claims · 8 setups
829 genes are overexpressed in the oscular region of increased calcite spicule formation in S. ciliatum
-
Full-text index only
Human and mouse introns are linked to the same processes and functions through each genome's most frequent non-conserved motifs.
PMID 18450818 · PMC2425492 · Nucleic acids research · 2008 · 8 claims · 5 setups
Pyknons (recurrent, genome-specific, ≥16nt motifs with ≥30 intact intergenic/intronic copies and ≥1 exonic copy) span a substantial fraction of previously uncharacterized intronic space (7.4% human, 4.4% mouse)
-
Full-text index only
The other side of comparative genomics: genes with no orthologs between the cow and other mammalian species.
PMID 20003425 · PMC2808326 · BMC genomics · 2009 · 7 claims · 4 setups
3,801 bovine genes have no orthologs in human, mouse and dog, and 1,010 human genes have no orthologs in cow despite having orthologs in mouse and dog
-
Has reproduction · 87
Next-generation phenotyping integrated in a national framework for patients with ultrarare disorders improves genetic diagnostics and yields new molecular findings.
PMID 39039281 · PMC11319204 · Nature genetics · 2024 · 6 claims · 5 setups
A structured multidisciplinary exome sequencing framework established molecular genetic diagnoses in 32% of patients with suspected ultrarare disorders, comprising 370 distinct molecular causes.
-
Has reproduction · 81
Macrophages on the run: Exercise balances macrophage polarization for improved health.
PMID 39476967 · PMC11585839 · Molecular metabolism · 2024 · 8 claims · 7 setups
Immediate/acute exercise triggers an M1 (pro-inflammatory) macrophage polarization surge.
-
Full-text index only
Large-scale identification and characterization of alternative splicing variants of human gene transcripts using 56,419 completely sequenced and manually annotated full-length cDNAs.
PMID 16914452 · PMC1557807 · Nucleic acids research · 2006 · 8 claims · 8 setups
Analysis of 56,419 full-length cDNAs identified 6877 alternative splicing genes encoding 18,297 alternative splicing variants made of 37,670 exons.
-
Has reproduction · 90
A2TEA: Identifying trait-specific evolutionary adaptations.
PMID 37224329 · PMC10186066 · F1000Research · 2022 · 8 claims · 7 setups
A2TEA integrates gene family expansion analysis with differential expression data across species to identify genes that were targets of evolutionary adaptation to a given stress/treatment
-
Has reproduction · 50
Comparative analysis of circular RNAs between soybean cytoplasmic male-sterile line NJCMS1A and its maintainer NJCMS1B by high-throughput sequencing.
PMID 30208848 · PMC6134632 · BMC genomics · 2018 · 8 claims · 7 setups
2867 circRNAs were identified in soybean flower buds via high-throughput sequencing with RNase R enrichment, of which 1009 were differentially expressed between NJCMS1A and NJCMS1B
-
Has reproduction · 60
TRAPID 2.0: a web application for taxonomic and functional analysis of de novo transcriptomes.
PMID 34197621 · PMC8464036 · Nucleic acids research · 2021 · 8 claims · 8 setups
TRAPID 2.0 is a web application performing global characterization of de novo transcriptomes via structural, functional, and taxonomic annotation in an initial processing phase, followed by an exploratory phase of downstream analyses.
-
Has reproduction · 76
The genome and development-dependent transcriptomes of Pyronema confluens: a window into fungal evolution.
PMID 24068976 · PMC3778014 · PLoS genetics · 2013 · 8 claims · 8 setups
The 50 Mb P. confluens genome with 13,369 predicted protein-coding genes is more characteristic of higher filamentous ascomycetes than of the large, repeat-rich Tuber melanosporum genome, showing that the truffle's expanded genome is not typical of the Pezizales.
-
Full-text index only
Computational disease gene identification: a concert of methods prioritizes type 2 diabetes and obesity candidate genes.
PMID 16757574 · PMC1475747 · Nucleic acids research · 2006 · 6 claims · 8 setups
Applying seven independent computational disease-gene prioritization methods in concert to 9556 positional candidate genes identifies a prioritized set of likely T2D and obesity candidate genes