Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 51
SGCP: a spectral self-learning method for clustering genes in co-expression networks.
PMID 38956463 · PMC11221046 · BMC bioinformatics · 2024 · 7 claims · 4 setups
SGCP, a spectral self-learning method, yields gene co-expression modules with higher GO enrichment than WGCNA, CoExpNets, and CEMiTool across 12 real gene expression datasets.
-
Has reproduction · 50
Viewing RNA-seq data on the entire human genome.
PMID 28979763 · PMC5605993 · F1000Research · 2017 · 6 claims · 3 setups
RNA-Seq Viewer is a web application that visualizes genome-wide expression data from NCBI's SRA and GEO databases using an ideogram across the entire human genome.
-
Full-text index only
Meeting highlights: beyond the genome 2000: the 18th International Congress of Biochemistry and Molecular Biology.
PMID 11119309 · PMC2448388 · Yeast (Chichester, England) · 2000 · 8 claims · 8 setups
Celera sequenced a human genome to ~45-fold coverage from one donor and used high-quality sequence stretches to define ~6 million SNPs
-
Has reproduction · 59
Advances in genomic and pharmacokinetic profiling for clinical stratification of metastatic breast cancer.
PMID 41369820 · PMC12799884 · Discover oncology · 2025 · 7 claims · 8 setups
Eight gene modules linked to metastasis were identified via scored network analysis and validated through pathway databases.
-
Full-text index only
Similarities and differences in genome-wide expression data of six organisms.
PMID 14737187 · PMC300882 · PLoS biology · 2004 · 8 claims · 8 setups
Coexpression of functionally related genes is frequently conserved across evolutionarily distant organisms
-
Full-text index only
Computational verification of protein-protein interactions by orthologous co-expression.
PMID 15740634 · PMC555590 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Co-expression of orthologous protein pairs across multiple species can verify/predict S. cerevisiae PPIs with better performance than S. cerevisiae co-expression alone.
-
Full-text index only
VIRGO: computational prediction of gene functions.
PMID 16845022 · PMC1538839 · Nucleic acids research · 2006 · 8 claims · 6 setups
VIRGO constructs a functional linkage network (FLN) from gene expression and molecular interaction data, labels genes with GO annotations, and propagates these labels to predict functions of unlabelled genes
-
Full-text index only
A global definition of expression context is conserved between orthologs, but does not correlate with sequence conservation.
PMID 16423292 · PMC1382217 · BMC genomics · 2006 · 7 claims · 6 setups
Expression context is largely conserved between orthologs across four eukaryote species.
-
Full-text index only
Recent additions and improvements to the Onto-Tools.
PMID 15980579 · PMC1160233 · Nucleic acids research · 2005 · 7 claims · 3 setups
The Onto-Tools back-end database was redesigned around the Entrez Gene data model after NCBI phased out LocusLink in February 2005.
-
Full-text index only
RiboSubstrates: a web application addressing the cleavage specificities of ribozymes in designated genomes.
PMID 17076887 · PMC1634876 · BMC bioinformatics · 2006 · 7 claims · 4 setups
RiboSubstrates is a web-based Perl application that scans a cDNA database for all potential substrates of a given ribozyme, including perfect matches, Wobble base-pair matches, and mismatch-containing matches.
-
Full-text index only
BTW: a web server for Boltzmann time warping of gene expression time series.
PMID 16845055 · PMC1538860 · Nucleic acids research · 2006 · 5 claims · 4 setups
Symmetric time warping distance is more flexible than Euclidean distance or correlation coefficient for identifying genes with similar temporal expression profiles, especially across sequences of different length.
-
Full-text index only
A rigorous method for multigenic families' functional annotation: the peptidyl arginine deiminase (PADs) proteins family example.
PMID 16271148 · PMC1310624 · BMC genomics · 2005 · 8 claims · 5 setups
Integrating EST-based expression data with phylogenetic analysis is a valid new method for functionally annotating multigenic protein families
-
Full-text index only
A combined approach exploring gene function based on worm-human orthology.
PMID 15877817 · PMC1112593 · BMC genomics · 2005 · 8 claims · 6 setups
Strict phylogenetic criteria (concordant Neighbor Joining and Maximum Parsimony tree topology) can select single most-likely human orthologs for C. elegans genes despite the large phylogenetic distance between worm and human sequences.
-
Full-text index only
CROPPER: a metagene creator resource for cross-platform and cross-species compendium studies.
PMID 16995941 · PMC1592126 · BMC bioinformatics · 2006 · 7 claims · 5 setups
CROPPER is a web-based software resource that combines genomic data from heterogeneous sources using identifier and orthologous gene information from the Ensembl database.
-
Full-text index only
The protein-phosphatome of the human malaria parasite Plasmodium falciparum.
PMID 18793411 · PMC2559854 · BMC genomics · 2008 · 8 claims · 8 setups
P. falciparum possesses 27 putative protein phosphatase sequences across the four major PP families (PPP, PPM, PTP, NIF), plus 7 additional sequences predicted to dephosphorylate non-protein substrates, totaling 34.
-
Full-text index only
L1Base: from functional annotation to prediction of active LINE-1 elements.
PMID 15608246 · PMC539998 · Nucleic acids research · 2005 · 7 claims · 6 setups
L1Base is a database of putatively active LINE-1 insertions in human, mouse and rat genomes, containing FLI-L1s (intact in both ORFs), ORF2-L1s (intact ORF2, disrupted ORF1), and FLnI-L1s (full-length, >6000 bp, non-intact)
-
Full-text index only
Comprehensive splice-site analysis using comparative genomics.
PMID 16914448 · PMC1557818 · Nucleic acids research · 2006 · 8 claims · 6 setups
Over half a million splice sites were collected from five species (H. sapiens, M. musculus, D. melanogaster, C. elegans, A. thaliana) and classified into four main subtypes: U2-type GT-AG and GC-AG, and U12-type GT-AG and AT-AC.
-
Has reproduction · 85
Exploring microproteins from various model organisms using the mip-mining database.
PMID 37919660 · PMC10623795 · BMC genomics · 2023 · 5 claims · 4 setups
Mip-mining is a database of 336 curated RNA-seq datasets from 8626 samples across nine species, built specifically to explore microprotein functions under stress and disease conditions
-
Full-text index only
Phylogenetic profiling of the Arabidopsis thaliana proteome: what proteins distinguish plants from other organisms?
PMID 15287975 · PMC507878 · Genome biology · 2004 · 8 claims · 6 setups
3,848 Arabidopsis proteins were identified as likely plant-specific based on phylogenetic profiling and EST confirmation in multiple plant species
-
Has reproduction · 86
Screening of core genes prognostic for sepsis and construction of a ceRNA regulatory network.
PMID 36855106 · PMC9976425 · BMC medical genomics · 2023 · 8 claims · 7 setups
RNA-seq of peripheral blood from 23 sepsis patients and 10 healthy controls identifies 1,044 DEmRNAs, 66 DEmiRNAs and 155 DElncRNAs.