Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Clustering by neurocognition for fine mapping of the schizophrenia susceptibility loci on chromosome 6p.
PMID 19694819 · PMC4286260 · Genes, brain, and behavior · 2009 · 6 claims · 6 setups
A family-based clustering strategy using neurocognitive test scores (CPT, WCST) can identify more homogeneous subgroups of schizophrenia families for genetic association analysis
-
Full-text index only
Polymorphism discovery and association analyses of the interferon genes in type 1 diabetes.
PMID 16504056 · PMC1402321 · BMC genetics · 2006 · 7 claims · 8 setups
No statistical evidence of a major association between T1D and any of the interferon or interferon-related genes tested (IFNA cluster, IFNB1, IFNW1, IFNG, ICSBP1)
-
Full-text index only
Comparative genomic analysis of Campylobacter jejuni associated with Guillain-Barré and Miller Fisher syndromes: neuropathogenic and enteritis-associated isolates can share high levels of genomic similarity.
PMID 17919333 · PMC2174954 · BMC genomics · 2007 · 8 claims · 4 setups
GBS/MFS strains are genomically heterogeneous, falling into about six major lineages rather than a single clonal group
-
Full-text index only
A taxonomy of epithelial human cancer and their metastases.
PMID 20017941 · PMC2806369 · BMC medical genomics · 2009 · 8 claims · 6 setups
Unsupervised hierarchical clustering of 1566 primary epithelial tumors yields large tissue-enriched clusters (breast, colon/GI, lung, ovary, kidney) plus smaller prostate, thyroid-kidney, and mixed clusters
-
Full-text index only
CLEAN: CLustering Enrichment ANalysis.
PMID 19640299 · PMC2734555 · BMC bioinformatics · 2009 · 8 claims · 4 setups
The gene-specific CLEAN score improves reproducibility of cluster analysis conclusions across independent datasets compared to the traditional cluster-wide score (cwCLEAN).
-
Full-text index only
Proteome analysis enables separate clustering of normal breast, benign breast and breast cancer tissues.
PMID 12865921 · PMC2394238 · British journal of cancer · 2003 · 6 claims · 3 setups
Hierarchical cluster analysis of 2-DE proteome data can distinguish normal breast, benign breast, and breast cancer tissues based on protein expression profiles
-
Full-text index only
SPRINT: a new parallel framework for R.
PMID 19114001 · PMC2628907 · BMC bioinformatics · 2008 · 8 claims · 1 setups
SPRINT is a prototype R framework that wraps parallelised functions, requiring minimal modification to existing sequential R scripts and no parallel programming expertise from the user
-
Has reproduction · 95
A role for ColV plasmids in the evolution of pathogenic Escherichia coli ST58.
PMID 35115531 · PMC8813906 · Nature communications · 2022 · 8 claims · 8 setups
ST58 contains a major sub-lineage (BAP2, n=363) characterized by near-ubiquitous carriage of ColV plasmids
-
Full-text index only
Automatic discovery of cross-family sequence features associated with protein function.
PMID 16409628 · PMC1395344 · BMC bioinformatics · 2006 · 8 claims · 6 setups
A self-supervised data mining approach can find relationships between sequence features and functional annotations without preconceived functional categories.
-
Full-text index only
An analysis of human microRNA and disease associations.
PMID 18923704 · PMC2559869 · PloS one · 2008 · 8 claims · 8 setups
MicroRNAs tend to show similar dysfunctional evidence (both up- or both down-regulated) for diseases within the same disease cluster, and different dysfunctional evidence between different disease clusters.
-
Has reproduction · 83
Functional module detection through integration of single-cell RNA sequencing data with protein-protein interaction networks.
PMID 33138772 · PMC7607865 · BMC genomics · 2020 · 8 claims · 6 setups
scPPIN integrates scRNA-seq-derived p-values with PPINs to detect maximum-weight connected subgraphs (active/functional modules) via an exact Steiner-tree approach
-
Has reproduction · 83
Hierarchical classification-based pan-cancer methylation analysis to classify primary cancer.
PMID 38066424 · PMC10709847 · BMC bioinformatics · 2023 · 8 claims · 5 setups
CHCT, a hierarchical classification tool, splits classification of 30 cancer types into ten smaller subproblems using a two-tier architecture to classify primary cancer by methylation profile
-
Full-text index only
The use of edge-betweenness clustering to investigate biological function in protein interaction networks.
PMID 15740614 · PMC555937 · BMC bioinformatics · 2005 · 8 claims · 7 setups
Edge-Betweenness clustering separates protein interaction graphs into subgraphs whose GO term distributions show significant correlations, revealing biologically meaningful functional modules.
-
Full-text index only
ARED Organism: expansion of ARED reveals AU-rich element cluster variations between human and mouse.
PMID 17984078 · PMC2238997 · Nucleic acids research · 2008 · 6 claims · 4 setups
ARED Organism and ARED-Integrated are new/updated public databases cataloguing ARE-containing mRNAs/genes in human, mouse and rat
-
Full-text index only
EPD in its twentieth year: towards complete promoter coverage of selected model organisms.
PMID 16381980 · PMC1347508 · Nucleic acids research · 2006 · 7 claims · 4 setups
EPD is an annotated, non-redundant collection of experimentally defined eukaryotic POL II promoters accessed via genome position pointers.
-
Full-text index only
Diversity of preferred nucleotide sequences around the translation initiation codon in eukaryote genomes.
PMID 18086709 · PMC2241899 · Nucleic acids research · 2008 · 8 claims · 5 setups
Preferred nucleotide sequences around the initiation codon are diverse among eukaryote species, but differences roughly reflect evolutionary relationships between species
-
Full-text index only
Clustering of phosphorylation site recognition motifs can be exploited to predict the targets of cyclin-dependent kinase.
PMID 17316440 · PMC1852407 · Genome biology · 2007 · 8 claims · 6 setups
CDK consensus motifs are frequently clustered (closely spaced) in known CDK substrate proteins rather than uniformly distributed
-
Full-text index only
EGenBio: a data management system for evolutionary genomics and biodiversity.
PMID 17118150 · PMC1683573 · BMC bioinformatics · 2006 · 7 claims · 7 setups
EGenBio is a web-based system for integrated management, filtering, curation, and visualization of large-scale genomic sequences, alignments, and phylogenetic trees for evolutionary genomics and biodiversity research.
-
Has reproduction · 76
A Deeper Insight into Evolutionary Patterns and Phylogenetic History of ASFV Epidemics in Sardinia (Italy) through Extensive Genomic Sequencing.
PMID 34696424 · PMC8539718 · Viruses · 2021 · 7 claims · 8 setups
58 new whole genomes of Sardinian ASFV isolates were sequenced, the largest ASFV whole-genome sequencing effort to date
-
Full-text index only
Satellog: a database for the identification and prioritization of satellite repeats in disease association studies.
PMID 15949044 · PMC1181805 · BMC bioinformatics · 2005 · 7 claims · 6 setups
Satellog is a database cataloging all pure 1-16 unit satellite repeats in the human genome with supplementary polymorphism, gene-location, and expression data for prioritizing repeats in disease-association studies.