Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Short tandem repeats in human exons: a target for disease mutations.
PMID 18789129 · PMC2543027 · BMC genomics · 2008 · 8 claims · 6 setups
STRs are present in exons of 92% of known human genes, unlike longer tandem repeats which are rare in exons
-
Full-text index only
Speeding disease gene discovery by sequence based candidate prioritization.
PMID 15766383 · PMC1274252 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Disease genes (OMIM) differ significantly from non-disease genes in sequence-based features including gene/cDNA/protein size, exon number, homolog conservation, secretion signal, 3' UTR length, CpG islands, and distance to nearest gene.
-
Full-text index only
Comparative metagenomics revealed commonly enriched gene sets in human gut microbiomes.
PMID 17916580 · PMC2533590 · DNA research : an international journal for rapid publication of reports on genes and genomes · 2007 · 7 claims · 7 setups
Adult and weaned-children gut microbiota show high functional (gene-content) uniformity despite taxonomic differences, while unweaned infant microbiota show high inter-individual variation in both taxonomic and gene composition.
-
Full-text index only
Direct in-gel fluorescence detection and cellular imaging of O-GlcNAc-modified proteins.
PMID 18683930 · PMC2649877 · Journal of the American Chemical Society · 2008 · 7 claims · 5 setups
Y289L GalT-mediated transfer of UDP-GalNAz followed by Cu(I)-catalyzed azide-alkyne cycloaddition enables selective, highly sensitive fluorescent/biotin labeling of O-GlcNAc-modified proteins
-
Full-text index only
Fast and systematic genome-wide discovery of conserved regulatory elements using a non-alignment based approach.
PMID 15693947 · PMC551538 · Genome biology · 2005 · 7 claims · 8 setups
FastCompare, a non-alignment-based, linear-time algorithm, computes a genome-wide conservation score for all k-mers (7-9 nt) between two genomes to identify conserved regulatory elements
-
Full-text index only
Analysis of protein sequence and interaction data for candidate disease gene prediction.
PMID 17020920 · PMC1636487 · Nucleic acids research · 2006 · 8 claims · 7 setups
Combining CPS and CMP using known disease genes as input achieves sensitivity 0.52 and specificity 0.97, reducing candidate lists 13-fold
-
Full-text index only
Clustering of phosphorylation site recognition motifs can be exploited to predict the targets of cyclin-dependent kinase.
PMID 17316440 · PMC1852407 · Genome biology · 2007 · 8 claims · 6 setups
CDK consensus motifs are frequently clustered (closely spaced) in known CDK substrate proteins rather than uniformly distributed
-
Full-text index only
Prediction-based approaches to characterize bidirectional promoters in the mammalian genome.
PMID 18366609 · PMC2386062 · BMC genomics · 2008 · 8 claims · 7 setups
The mapping algorithm identified 5,647 candidate bidirectional promoter regions in the mouse genome, similar in number to those previously found in human.
-
Full-text index only
G-quadruplexes in promoters throughout the human genome.
PMID 17169996 · PMC1802602 · Nucleic acids research · 2007 · 8 claims · 6 setups
Promoter regions (1 kb upstream of TSS) are significantly enriched in quadruplex motifs (PQS) relative to the rest of the genome
-
Full-text index only
Identification of candidate disease genes by integrating Gene Ontologies and protein-interaction networks: case study of primary immunodeficiencies.
PMID 19073697 · PMC2632920 · Nucleic acids research · 2009 · 8 claims · 5 setups
Combining high protein-interaction network scores with significant PID-related GO terms identifies novel PID candidate genes
-
Full-text index only
G-quadruplexes: the beginning and end of UTRs.
PMID 18832370 · PMC2577360 · Nucleic acids research · 2008 · 8 claims · 5 setups
UTRs show significant strand asymmetry with C-PQS more common than G-PQS, consistent with general depletion of G-quadruplex-forming RNA
-
Has reproduction · 63
RummaGEO: Automatic mining of human and mouse gene sets from GEO.
PMID 39569206 · PMC11573963 · Patterns (New York, N.Y.) · 2024 · 8 claims · 7 setups
RummaGEO is a gene expression signature search engine built from automatically mined human and mouse RNA-seq perturbation studies in GEO
-
Has reproduction · 81
Identification of Proteins Deregulated by Platinum-Based Chemotherapy as Novel Biomarkers and Therapeutic Targets in Non-Small Cell Lung Cancer.
PMID 33777753 · PMC7991912 · Frontiers in oncology · 2021 · 7 claims · 8 setups
Cisplatin exposure induces significant deregulation of protein expression networks in NSCLC cells
-
Full-text index only
Dcode.org anthology of comparative genomic tools.
PMID 15980535 · PMC1160116 · Nucleic acids research · 2005 · 8 claims · 7 setups
The dcode.org suite (zPicture, Mulan, eShadow, rVista 2.0, multiTF, Creme 2.0, ECR Browser) provides integrated tools for comparative genomic analysis and non-coding regulatory element discovery.
-
Full-text index only
The use of edge-betweenness clustering to investigate biological function in protein interaction networks.
PMID 15740614 · PMC555937 · BMC bioinformatics · 2005 · 8 claims · 7 setups
Edge-Betweenness clustering separates protein interaction graphs into subgraphs whose GO term distributions show significant correlations, revealing biologically meaningful functional modules.
-
Full-text index only
Comparative proteomics of clathrin-coated vesicles.
PMID 17116749 · PMC2064594 · The Journal of cell biology · 2006 · 8 claims · 4 setups
A comparative proteomics strategy contrasting CCV fractions from control and clathrin-depleted (CHC siRNA knockdown) HeLa cells can distinguish genuine CCV proteins from copurifying contaminants
-
Full-text index only
Computational disease gene identification: a concert of methods prioritizes type 2 diabetes and obesity candidate genes.
PMID 16757574 · PMC1475747 · Nucleic acids research · 2006 · 6 claims · 8 setups
Applying seven independent computational disease-gene prioritization methods in concert to 9556 positional candidate genes identifies a prioritized set of likely T2D and obesity candidate genes
-
Full-text index only
A comprehensive modular map of molecular interactions in RB/E2F pathway.
PMID 18319725 · PMC2290939 · Molecular systems biology · 2008 · 8 claims · 4 setups
A comprehensive, curated map of RB/E2F pathway molecular interactions was built using SBGN notation in CellDesigner and converted to BioPAX 2.0 format
-
Full-text index only
Genome-wide analysis of small RNA and novel MicroRNA discovery in human acute lymphoblastic leukemia based on extensive sequencing approach.
PMID 19724645 · PMC2731166 · PloS one · 2009 · 7 claims · 5 setups
159 novel miRNAs and 116 novel miRNA*s were identified from ALL patient and normal donor small RNA libraries
-
Has reproduction · 85
Digital sorting of complex tissues for cell type-specific gene expression profiles.
PMID 23497278 · PMC3626856 · BMC bioinformatics · 2013 · 8 claims · 8 setups
The Digital Sorting Algorithm (DSA) deconvolves mixed tissue expression into cell type-specific profiles using only marker genes, without requiring prior knowledge of cell type frequencies or in vitro pure-cell profiles.