Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Genome-wide identification of human functional DNA using a neutral indel model.
PMID 16410828 · PMC1326222 · PLoS computational biology · 2006 · 8 claims · 8 setups
A neutral indel model predicting a geometric distribution of intergap segment (IGS) lengths fits human-mouse ancestral repeat (AR) alignment data excellently
-
Full-text index only
Evola: Ortholog database of all human genes in H-InvDB with manual curation of phylogenetic trees.
PMID 17982176 · PMC2238928 · Nucleic acids research · 2008 · 6 claims · 7 setups
Evola combines genome synteny-based computational ortholog detection with manual curation of phylogenetic trees by experts to yield more reliable orthologs than automated pairwise methods
-
Full-text index only
TRED: a Transcriptional Regulatory Element Database and a platform for in silico gene regulation studies.
PMID 15608156 · PMC539958 · Nucleic acids research · 2005 · 8 claims · 5 setups
TRED is a database collecting both cis-regulatory elements (promoters) and trans-regulatory elements (transcription factor binding/regulation data) with linked access.
-
Full-text index only
Identification of the proliferation/differentiation switch in the cellular network of multicellular organisms.
PMID 17166053 · PMC1664705 · PLoS computational biology · 2006 · 8 claims · 8 setups
Integrating interactome and transcriptome data reveals a pair of transcriptionally anticorrelated network modules (P and D) each comprising hundreds of genes, present across individuals and species.
-
Full-text index only
InSite: a computational method for identifying protein-protein interaction binding sites on a proteome-wide scale.
PMID 17868464 · PMC2375030 · Genome biology · 2007 · 8 claims · 8 setups
InSite predicts protein-pair-specific binding motifs ('Motif M on protein A binds to protein B') by integrating heterogeneous PPI and motif-motif interaction evidence within a Bayesian network trained by EM
-
Full-text index only
The Princeton Protein Orthology Database (P-POD): a comparative genomics analysis tool for biologists.
PMID 17712414 · PMC1942082 · PloS one · 2007 · 8 claims · 5 setups
P-POD is the first comparative genomics database to combine results from multiple computational ortholog/homolog prediction methods with manually curated literature-derived experimental evidence of functional conservation.
-
Full-text index only
The biological function of some human transcription factor binding motifs varies with position relative to the transcription start site.
PMID 18367472 · PMC2377430 · Nucleic acids research · 2008 · 8 claims · 5 setups
1226 eight-letter DNA words show statistically significant positional preferences relative to the TSS across 7914 human promoter regions
-
Full-text index only
A methodological framework for the reconstruction of contiguous regions of ancestral genomes and its application to mammalian genomes.
PMID 19043541 · PMC2580819 · PLoS computational biology · 2008 · 8 claims · 5 setups
A general model-free methodological framework is proposed for reconstructing Contiguous Ancestral Regions (CARs) from conserved syntenies, generalizing prior computational and cytogenetic approaches
-
Full-text index only
The stem cell population of the human colon crypt: analysis via methylation patterns.
PMID 17335343 · PMC1808490 · PLoS computational biology · 2007 · 8 claims · 3 setups
A coalescent-based, full probabilistic model with MCMC Bayesian inference provides a more powerful alternative to prior forward-simulation approaches for analyzing methylation pattern data from crypts.
-
Full-text index only
Optimized mixed Markov models for motif identification.
PMID 16749929 · PMC1534070 · BMC bioinformatics · 2006 · 8 claims · 4 setups
OMiMa can incorporate more than NNSplice's pairwise dependencies
-
Full-text index only
MODBASE, a database of annotated comparative protein structure models and associated resources.
PMID 18948282 · PMC2686492 · Nucleic acids research · 2009 · 8 claims · 8 setups
MODBASE contains 5,152,695 reliable comparative protein structure models for 1,593,209 unique protein sequences.
-
Has reproduction · 84
An integrated in silico-in vitro approach for identifying therapeutic targets against osteoarthritis.
PMID 36352408 · PMC9648005 · BMC biology · 2022 · 7 claims · 5 setups
A signal transduction/gene regulatory network model of the articular chondrocyte was built combining knowledge-based curation and data-driven (machine learning) network inference
-
Full-text index only
KEGG for linking genomes to life and the environment.
PMID 18077471 · PMC2238879 · Nucleic acids research · 2008 · 8 claims · 4 setups
KEGG provides a reference knowledge base for linking genomes to life via PATHWAY mapping and to the environment via BRITE mapping.
-
Full-text index only
The distribution of SNPs in human gene regulatory regions.
PMID 16209714 · PMC1260019 · BMC genomics · 2005 · 8 claims · 6 setups
SNPs occur with higher density closer to the transcriptional start site within gene promoter regions than in further upstream regions
-
Full-text index only
Information-based methods for predicting gene function from systematic gene knock-downs.
PMID 18959798 · PMC2596148 · BMC bioinformatics · 2008 · 8 claims · 4 setups
Information-based metrics, which incorporate a phenotype's genomic frequency, outperform non-information-based metrics for detecting gene-gene functional similarity from phenotypic knock-down profiles.
-
Full-text index only
CTCF binding site classes exhibit distinct evolutionary, genomic, epigenomic and transcriptomic features.
PMID 19922652 · PMC3091324 · Genome biology · 2009 · 8 claims · 8 setups
CTCF binding sites can be classified into three occupancy-based classes (LowOc, MedOc, HighOc) based on similarity to the CTCF PWM motif
-
Full-text index only
Computer identification of snoRNA genes using a Mammalian Orthologous Intron Database.
PMID 16093549 · PMC1184218 · Nucleic acids research · 2005 · 8 claims · 5 setups
Created the Mammalian Orthologous Intron Database (MOID) containing orthologous introns of human, mouse and rat identified via conserved reading-frame position
-
Full-text index only
ABS: a database of Annotated regulatory Binding Sites from orthologous promoters.
PMID 16381947 · PMC1347478 · Nucleic acids research · 2006 · 7 claims · 6 setups
ABS is a public database of experimentally identified TF binding sites conserved in orthologous vertebrate gene promoters, manually curated from the literature.
-
Has reproduction · 94
Hierarchical cell-type identifier accurately distinguishes immune-cell subtypes enabling precise profiling of tissue microenvironment with single-cell RNA-sequencing.
PMID 36681937 · PMC10025442 · Briefings in bioinformatics · 2023 · 8 claims · 8 setups
HiCAT is a hierarchical, marker-based cell-type identifier that uses gene set analysis (GSA) scoring with markers structured in a three-level taxonomy tree (major-type, minor-type, subset)
-
Has reproduction · 67
Optimal scaling of digital transcriptomes.
PMID 24223126 · PMC3819321 · PloS one · 2013 · 8 claims · 8 setups
Fifteen existing and novel transcript-count normalization algorithms can be compared with two novel, mutually independent metrics: the number of "uniform" genes (sufficiently low coefficient of variation after normalization) and low average Spearman correlation between normalized expression profiles of gene pairs.