Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
RAId_DbS: mass-spectrometry based peptide identification web server with knowledge integration.
PMID 18954448 · PMC2605478 · BMC genomics · 2008 · 7 claims · 4 setups
Constructed enhanced protein databases integrating annotated SAPs, PTMs, and disease associations for 17 organisms.
-
Has reproduction · 86
Molecular Classification Models for Triple Negative Breast Cancer Subtype Using Machine Learning.
PMID 34575658 · PMC8472680 · Journal of personalized medicine · 2021 · 6 claims · 4 setups
A training gene set of 719 unique upregulated DEGs (subtype-specific) can be used to build ML models that classify TNBC into BLIA, BLIS, MES, and LAR subtypes.
-
Full-text index only
Pseudofam: the pseudogene families database.
PMID 18957444 · PMC2686518 · Nucleic acids research · 2009 · 8 claims · 7 setups
Pseudofam is an online database of pseudogene families built by mapping pseudogenes to Pfam protein families, providing query tools, statistics, and sequence alignments
-
Full-text index only
Proteomics analysis reveals novel components in the detergent-insoluble subproteome in Alzheimer's disease.
PMID 19746990 · PMC2784247 · Journal of proteome research · 2009 · 8 claims · 5 setups
A label-free XIC-based LC-MS/MS quantitation strategy, combined with an FTLD-U comparator cohort, can be used to identify AD-specific changes in the detergent-insoluble brain subproteome.
-
Has reproduction · 50
Comparative analysis of circular RNAs between soybean cytoplasmic male-sterile line NJCMS1A and its maintainer NJCMS1B by high-throughput sequencing.
PMID 30208848 · PMC6134632 · BMC genomics · 2018 · 8 claims · 7 setups
2867 circRNAs were identified in soybean flower buds via high-throughput sequencing with RNase R enrichment, of which 1009 were differentially expressed between NJCMS1A and NJCMS1B
-
Has reproduction · 50
DeeReCT-APA: Prediction of Alternative Polyadenylation Site Usage Through Deep Learning.
PMID 33662629 · PMC9801043 · Genomics, proteomics & bioinformatics · 2022 · 7 claims · 3 setups
DeeReCT-APA, a CNN-LSTM deep learning architecture, quantitatively predicts the usage level of all alternative PASs within a gene regardless of PAS number, treating it as a variable-length regression task.
-
Has reproduction · 89
A Meta-Analysis of Wolbachia Transcriptomics Reveals a Stage-Specific Wolbachia Transcriptional Response Shared Across Different Hosts.
PMID 32718933 · PMC7467002 · G3 (Bethesda, Md.) · 2020 · 7 claims · 8 setups
Across datasets re-analyzed with a unified workflow, there is a general lack of global Wolbachia gene regulation.
-
Has reproduction · 71
Microbial diversity of plant pathogens and insect endosymbionts in Reptalus artemisiae.
PMID 41826827 · PMC13202766 · BMC microbiology · 2026 · 8 claims · 8 setups
R. artemisiae harbors six prokaryotic taxa: two plant pathogens ('Ca. P. solani', 'Ca. A. phytopathogenicus') and four insect endosymbionts ('Ca. Vidania', 'Ca. Purcelliella', 'Ca. Karelsulcia', and Wolbachia).
-
Full-text index only
Limited copy number-high resolution melting (LCN-HRM) enables the detection and identification by sequencing of low level mutations in cancer biopsies.
PMID 19811662 · PMC2766370 · Molecular cancer · 2009 · 7 claims · 6 setups
LCN-HRM enables detection and sequencing-based characterisation of low-level mutations that are undetectable by direct sequencing alone
-
Full-text index only
Tracing the origin of functional and conserved domains in the human proteome: implications for protein evolution at the modular level.
PMID 17090320 · PMC1654190 · BMC evolutionary biology · 2006 · 8 claims · 5 setups
HHpred (HMM-HMM comparison) detects remote homologs in the human proteome with higher sensitivity than hmmpfam (HMMER), giving 10% more functional domain coverage and 20% higher residue coverage against Pfam-A families.
-
Full-text index only
Comparative metagenomics revealed commonly enriched gene sets in human gut microbiomes.
PMID 17916580 · PMC2533590 · DNA research : an international journal for rapid publication of reports on genes and genomes · 2007 · 7 claims · 7 setups
Adult and weaned-children gut microbiota show high functional (gene-content) uniformity despite taxonomic differences, while unweaned infant microbiota show high inter-individual variation in both taxonomic and gene composition.
-
Full-text index only
Genomic organization and recombinational unit duplication-driven evolution of ovine and bovine T cell receptor gamma loci.
PMID 18282289 · PMC2270265 · BMC genomics · 2008 · 8 claims · 6 setups
The sheep TRG1 and TRG2 loci evolved through a series of duplication events involving either entire V-J-J-C recombinational cassettes or single V genes
-
Full-text index only
Proteomics identifies multipotent and low oncogenic risk stem cells of the spleen.
PMID 20005973 · PMC2891339 · The international journal of biochemistry & cell biology · 2010 · 8 claims · 5 setups
CD45- splenic stem cell-specific proteins are identical to core iPS/ES markers OCT3/4, SOX2, KLF4, c-MYC and NANOG.
-
Full-text index only
Integrated proteomic analysis of human cancer cells and plasma from tumor bearing mice for ovarian cancer biomarker discovery.
PMID 19936259 · PMC2775948 · PloS one · 2009 · 8 claims · 8 setups
Integrated proteomic analysis of a cancer mouse model and human cancer cell populations provides an effective approach to identify potential circulating protein biomarkers.
-
Full-text index only
A European focus on proteomics.
PMID 15128441 · PMC416463 · Genome biology · 2004 · 8 claims · 8 setups
MALDI-MS and ESI-MS are complementary techniques that identify overlapping but distinct subsets of proteins
-
Has reproduction · 67
Heterogeneity and Differentiation Trajectories of Infiltrating CD8+ T Cells in Lung Adenocarcinoma.
PMID 36358600 · PMC9658355 · Cancers · 2022 · 7 claims · 8 setups
Infiltrating CD8+ T cells in LUAD can be divided into ten transcriptionally distinct subsets: eight cytotoxic (CTL) subsets, one naive-like (NTL) subset, and one exhausted (ETL) subset.
-
Full-text index only
Polymorphix: a sequence polymorphism database.
PMID 15608242 · PMC540030 · Nucleic acids research · 2005 · 8 claims · 5 setups
Polymorphix is an ACNUC-structured database that organizes EMBL/GenBank sequences into within-species homologous sequence families using similarity and bibliographic criteria, with alignments, outgroups and phylogenetic trees provided.
-
Full-text index only
ARED 3.0: the large and diverse AU-rich transcriptome.
PMID 16381826 · PMC1347415 · Nucleic acids research · 2006 · 7 claims · 6 setups
ARED 3.0 computationally mapped more than 4000 ARE-mRNAs to the human genome, representing 5-8% of human genes.
-
Full-text index only
EPGD: a comprehensive web resource for integrating and displaying eukaryotic paralog/paralogon information.
PMID 17984073 · PMC2238967 · Nucleic acids research · 2008 · 8 claims · 8 setups
EPGD is a gene-centered, internet-accessible database integrating paralog family and paralogon information for 26 eukaryotic genomes.
-
Full-text index only
Power analysis for genome-wide association studies.
PMID 17725844 · PMC2042984 · BMC genetics · 2007 · 8 claims · 6 setups
Developed a method to compute genome-wide association study power using tag SNPs and representative population genotype data (HapMap), equivalent to the cumulative r2-adjusted power of Jorgenson and Witte.