Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 79
Enriched domain detector: a program for detection of wide genomic enrichment domains robust against local variations.
PMID 24782521 · PMC4066758 · Nucleic acids research · 2014 · 8 claims · 5 setups
EDD is a new algorithm that detects broad (megabase-size) enrichment domains from ChIP-seq data of widely distributed chromatin proteins such as A- and B-type lamins.
-
Full-text index only
WebGestalt: an integrated system for exploring gene sets in various biological contexts.
PMID 15980575 · PMC1160236 · Nucleic acids research · 2005 · 8 claims · 6 setups
WebGestalt is an integrated web-based system composed of four modules: gene set management, information retrieval, organization/visualization, and statistics.
-
Full-text index only
Discovering cancer genes by integrating network and functional properties.
PMID 19765316 · PMC2758898 · BMC medical genomics · 2009 · 8 claims · 6 setups
Cancer genes have distinct PPI network topology (higher connectivity, higher clustering coefficient, shorter path length to known cancer genes) compared to non-cancer genes
-
Full-text index only
Proteomic-based identification of maternal proteins in mature mouse oocytes.
PMID 19646285 · PMC2730056 · BMC genomics · 2009 · 8 claims · 6 setups
625 different proteins were identified from 2700 zona pellucida-free mature mouse MII oocytes, the largest oocyte proteome catalog to date
-
Full-text index only
The DNA sequence of the human X chromosome.
PMID 15772651 · PMC2665286 · Nature · 2005 · 8 claims · 8 setups
The euchromatic sequence of the human X chromosome was determined to 99.3% completeness (~155 Mb total)
-
Has reproduction · 75
Genomic regions and candidate genes selected during the breeding of rice in Vietnam.
PMID 35899250 · PMC9309459 · Evolutionary applications · 2022 · 8 claims · 7 setups
XP-CLR and FST scans identify genomic regions with distorted allele frequency/differentiation patterns resulting from differential selective pressures between Vietnamese rice subpopulations
-
Has reproduction · 55
N6-methyladenosine (m6A) reader Pho92 is recruited co-transcriptionally and couples translation to mRNA decay to promote meiotic fitness in yeast.
PMID 36422864 · PMC9731578 · eLife · 2022 · 8 claims · 8 setups
Pho92 specifically binds m6A-modified RNA via its YTH domain, both in vitro and in vivo
-
Full-text index only
Advances in the study of SR protein family.
PMID 15626328 · PMC5172405 · Genomics, proteomics & bioinformatics · 2003 · 8 claims · 8 setups
SR proteins promote assembly of the early splicesome via protein-protein interactions in their RS-domain that recruit components of the splicing machinery.
-
Full-text index only
Long-term trends in evolution of indels in protein sequences.
PMID 17298668 · PMC1805498 · BMC evolutionary biology · 2007 · 8 claims · 5 setups
More than one third of protein domains show a statistically significant tendency to increase or decrease in size over evolutionary distance.
-
Full-text index only
POCUS: mining genomic sequence annotation to predict disease genes.
PMID 14611661 · PMC329128 · Genome biology · 2003 · 8 claims · 6 setups
Genes predisposing to the same disease tend to share functional annotation IDs (GO/InterPro) more than expected by chance
-
Full-text index only
Phosphoproteomics: unraveling the signaling web.
PMID 18922462 · PMC2754874 · Molecular cell · 2008 · 8 claims · 8 setups
Integrating discovery-based MS phosphoproteomics with targeted protein microarray technologies yields a more complete picture of signaling networks and improves clinical translation of findings.
-
Full-text index only
The dystrobrevin-binding protein 1 gene: features and networks.
PMID 18663367 · PMC2859304 · Molecular psychiatry · 2009 · 8 claims · 6 setups
DTNBP1 gene structure, protein-coding sequence, and dysbindin domain are conserved across 13 vertebrate species, while noncoding sequence is diverse.
-
Has reproduction · 67
Essential Genes of Vibrio anguillarum and Other Vibrio spp. Guide the Development of New Drugs and Vaccines.
PMID 34745063 · PMC8564382 · Frontiers in microbiology · 2021 · 7 claims · 7 setups
Tn-seq using the TnSC189 mariner transposon identified 329 essential genes in V. anguillarum NB10Sm from a library of 52,662 insertion mutants.
-
Full-text index only
Analysis of protein sequence and interaction data for candidate disease gene prediction.
PMID 17020920 · PMC1636487 · Nucleic acids research · 2006 · 8 claims · 7 setups
Combining CPS and CMP using known disease genes as input achieves sensitivity 0.52 and specificity 0.97, reducing candidate lists 13-fold
-
Full-text index only
Mutation of ERBB2 provides a novel alternative mechanism for the ubiquitous activation of RAS-MAPK in ovarian serous low malignant potential tumors.
PMID 19010816 · PMC6953412 · Molecular cancer research : MCR · 2008 · 8 claims · 8 setups
Activating RAS-MAPK pathway mutations are present in >70% of serous LMP tumors versus ~12.5% of serous ovarian carcinomas
-
Full-text index only
Sequence and structure signatures of cancer mutation hotspots in protein kinases.
PMID 19834613 · PMC2759519 · PloS one · 2009 · 8 claims · 6 setups
Developed CKMD (Composite Kinase Mutation Database), an integrated bioinformatics resource mapping genetic variation in protein kinase genes to sequence, structural, and functional data
-
Full-text index only
Enthoprotin: a novel clathrin-associated protein identified through subcellular proteomics.
PMID 12213833 · PMC2173151 · The Journal of cell biology · 2002 · 8 claims · 8 setups
Subcellular proteomics of purified CCVs identifies enthoprotin (encoded by KIAA0171), a novel ENTH domain-containing protein not previously detected at the protein level.
-
Full-text index only
'Genome design' model and multicellular complexity: golden middle.
PMID 17062620 · PMC1635334 · Nucleic acids research · 2006 · 8 claims · 8 setups
Intermediately expressed human genes are the longest genes genome-wide, in both coding and intronic sequence, longer than housekeeping or tissue-specific genes.
-
Has reproduction · 76
Bayesian prediction of microbial oxygen requirement.
PMID 26913185 · PMC4743139 · F1000Research · 2013 · 7 claims · 8 setups
A naive Bayesian classifier based on presence/absence of class-associated Pfam-A domains can distinguish three oxygen requirement classes (aerobe, anaerobe, facultative anaerobe) from genome sequence, unlike prior studies that only made pairwise distinctions.
-
Full-text index only
Pseudofam: the pseudogene families database.
PMID 18957444 · PMC2686518 · Nucleic acids research · 2009 · 8 claims · 7 setups
Pseudofam is an online database of pseudogene families built by mapping pseudogenes to Pfam protein families, providing query tools, statistics, and sequence alignments