Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Genome sequences and great expectations.
PMID 11178275 · PMC150431 · Genome biology · 2001 · 8 claims · 3 setups
Function is known or can be predicted for an average of 62% of proteins across 31 analyzed genomes.
-
Full-text index only
Identification of "pathologs" (disease-related genes) from the RIKEN mouse cDNA dataset using human curation plus FACTS, a new biological information extraction system.
PMID 15115540 · PMC420239 · BMC genomics · 2004 · 6 claims · 3 setups
Bioinformatic sequence comparison of 60,770 RIKEN FANTOM2 mouse cDNA clones identified 2,578 sequences with 70-85% identity to known human disease genes/proteins
-
Full-text index only
SECIS elements in the coding regions of selenoprotein transcripts are functional in higher eukaryotes.
PMID 17169995 · PMC1802603 · Nucleic acids research · 2007 · 8 claims · 5 setups
SECIS elements located within coding regions of selenoprotein mRNAs support functional Sec insertion in mammalian cells
-
Full-text index only
Generation of a restriction minus enteropathogenic Escherichia coli E2348/69 strain that is efficiently transformed with large, low copy plasmids.
PMID 18681975 · PMC2518929 · BMC microbiology · 2008 · 8 claims · 7 setups
E2348/69 possesses a type I restriction-modification system encoded by an hsdMSR-like operon identified by homology to known Hsd proteins.
-
Full-text index only
SelenoDB 1.0 : a database of selenoprotein genes, proteins and SECIS elements.
PMID 18174224 · PMC2238826 · Nucleic acids research · 2008 · 6 claims · 5 setups
Standard genome annotation pipelines misannotate selenoprotein genes because they rely on UGA as a universal stop codon, failing to recognize its dual role as the selenocysteine-recoding codon.
-
Full-text index only
Sequence similarity network reveals common ancestry of multidomain proteins.
PMID 18475320 · PMC2377100 · PLoS computational biology · 2008 · 8 claims · 6 setups
Traditional homology definitions do not capture multidomain evolution; the authors extend the definition to include domain insertion via a common ancestral locus model.
-
Full-text index only
Genome-wide census and expression profiling of chicken neuropeptide and prohormone convertase genes.
PMID 20006904 · PMC2814002 · Neuropeptides · 2010 · 8 claims · 5 setups
Bioinformatic survey of chicken genome/EST/HTGS databases identifies previously unreported chicken neuropeptide genes
-
Full-text index only
Update of the G2D tool for prioritization of gene candidates to inherited diseases.
PMID 17478516 · PMC1933178 · Nucleic acids research · 2007 · 8 claims · 4 setups
G2D is a web server that prioritizes candidate genes for inherited diseases using three distinct algorithms based on different input information.
-
Full-text index only
Comparative analysis of genome tiling array data reveals many novel primate-specific functional RNAs in human.
PMID 17288572 · PMC1796608 · BMC evolutionary biology · 2007 · 8 claims · 6 setups
Widespread transcription occurs across the human genome outside known gene annotations, and the bulk of TARs represent genuine transcripts rather than experimental artifacts
-
Full-text index only
The cohesin complex: sequence homologies, interaction networks and shared motifs.
PMID 11276426 · PMC30708 · Genome biology · 2001 · 8 claims · 8 setups
Mouse Mmip1 and Smc3 (SMCD) share 99% sequence identity and are products of the same gene
-
Full-text index only
Annotation and analysis of 10,000 expressed sequence tags from developing mouse eye and adult retina.
PMID 14519200 · PMC328454 · Genome biology · 2003 · 8 claims · 5 setups
Annotation of 8,633 high-quality non-mitochondrial/non-ribosomal ESTs shows 57% represent known genes and 43% are unknown or novel, with M15E having the highest proportion of novel ESTs
-
Full-text index only
TPRpred: a tool for prediction of TPR-, PPR- and SEL1-like repeats from protein sequences.
PMID 17199898 · PMC1774580 · BMC bioinformatics · 2007 · 7 claims · 8 setups
TPRpred detects divergent/remote-homolog TPR repeat units that existing resources (Pfam, SMART, REP) fail to detect
-
Full-text index only
Large genomic rearrangements in the CFTR gene contribute to CBAVD.
PMID 17448246 · PMC1876208 · BMC medical genetics · 2007 · 7 claims · 6 setups
Large genomic rearrangements in CFTR contribute to CBAVD and should be systematically investigated alongside point mutation screening
-
Full-text index only
Comparative genomics supports a deep evolutionary origin for the large, four-module transcriptional mediator complex.
PMID 18515835 · PMC2475620 · Nucleic acids research · 2008 · 8 claims · 6 setups
Yeast Med2, Med3/Pgd1 and Med5/Nut1 (Tail module) are homologs of human Med29, Med27 and Med24, respectively
-
Has reproduction · 83
Macrel: antimicrobial peptide screening in genomes and metagenomes.
PMID 33384902 · PMC7751412 · PeerJ · 2020 · 8 claims · 8 setups
Macrel introduces a novel set of 22 peptide features (6 local, 16 global), including a new Free Energy Transition (FET) feature group, for AMP and hemolytic activity classification
-
Has reproduction · 67
Satellitome Analysis and Transposable Elements Comparison in Geographically Distant Populations of Spodoptera frugiperda.
PMID 35455012 · PMC9026859 · Life (Basel, Switzerland) · 2022 · 8 claims · 5 setups
Most transposable elements are commonly shared across all eight geographically distant S. frugiperda samples, except Maverick and PIF/Harbinger elements which show divergent repeat copies
-
Full-text index only
The UCSC Proteome Browser.
PMID 15608236 · PMC540054 · Nucleic acids research · 2005 · 8 claims · 5 setups
The UCSC Proteome Browser is tightly integrated with the UCSC Genome Browser, giving users simultaneous access to genome and proteome data.
-
Has reproduction · 71
Systematic and computational identification of Androctonus crassicauda long non-coding RNAs.
PMID 33633149 · PMC7907363 · Scientific reports · 2021 · 8 claims · 6 setups
13,401 lncRNAs were identified in the A. crassicauda transcriptome using the ECF pipeline
-
Full-text index only
A candidate metastasis-associated DNA marker for ductal mammary carcinoma.
PMID 12631399 · PMC154149 · Breast cancer research : BCR · 2003 · 8 claims · 8 setups
RDA comparing normal and metastatic ductal breast carcinoma cell DNA identified 10 unique metastasis-associated DNA sequences (MADS) apparently lost in metastatic cells
-
Full-text index only
Comparative genomics of cyclin-dependent kinases suggest co-evolution of the RNAP II C-terminal domain and CTD-directed CDKs.
PMID 15380029 · PMC521075 · BMC genomics · 2004 · 8 claims · 6 setups
Cell-cycle related CDKs (orthologs of CDK1-6) are present in all sampled eukaryotic organisms, including the most ancestral protists.