Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Novel gene and gene model detection using a whole genome open reading frame analysis in proteomics.
PMID 16646984 · PMC1557991 · Genome biology · 2006 · 8 claims · 4 setups
A six-frame genomic ORF translation used as an MS search database can detect novel peptides absent from standard protein databases, revealing incomplete genome annotation.
-
Has reproduction · 78
Identification of the Wheat (Triticum aestivum) IQD Gene Family and an Expression Analysis of Candidate Genes Associated with Seed Dormancy and Germination.
PMID 35456910 · PMC9025732 · International journal of molecular sciences · 2022 · 8 claims · 8 setups
73 IQD gene family members (TaIQD1-73) were identified in the wheat genome and phylogenetically divided into six major groups.
-
Has reproduction · 61
lncEvo: automated identification and conservation study of long noncoding RNAs.
PMID 33563213 · PMC7871587 · BMC bioinformatics · 2021 · 8 claims · 5 setups
lncEvo is an integrated Nextflow/Docker pipeline combining transcriptome assembly, lncRNA identification, and cross-species conservation analysis into a single workflow.
-
Full-text index only
Comparative analysis of cancer genes in the human and chimpanzee genomes.
PMID 16438707 · PMC1382208 · BMC genomics · 2006 · 7 claims · 6 setups
All 333 examined human cancer genes have intact, highly conserved orthologs in the chimpanzee genome (99.38% protein identity).
-
Full-text index only
Microbial genomics: from sequence to function.
PMID 10998380 · PMC2627950 · Emerging infectious diseases · 2000 · 8 claims · 4 setups
Whole-genome shotgun sequencing (sequencing and assembly of random genome fragments), first demonstrated with Haemophilus influenzae in 1995, is now the method of choice for sequencing most genomes, including the human genome.
-
Full-text index only
Genetic diversity among five T4-like bacteriophages.
PMID 16716236 · PMC1524935 · Virology journal · 2006 · 8 claims · 8 setups
A core set of 82 conserved genes (T4-like genes) is present in all five genomes analyzed, clustered in large collinear blocks.
-
Full-text index only
Report of the 9th HLPP Workshop October 2007, Seoul, Korea.
PMID 18683817 · PMC4601560 · Proteomics · 2008 · 8 claims · 8 setups
An integrated separating-identifying platform identified 6788 proteins (≥2 peptides, 95% confidence) in Chinese human liver samples, including 3721 new to liver and 977 hypothetical proteins
-
Full-text index only
Proteomics data repositories.
PMID 19795424 · PMC2908408 · Proteomics · 2009 · 5 claims · 5 setups
The YRC Public Data Repository (YRC PDR) provides a single unified interface disseminating multi-technology proteomics data (mass spectrometry, yeast two-hybrid, fluorescence microscopy, structure prediction) linked to protein annotations from many source databases.
-
Full-text index only
The human L-threonine 3-dehydrogenase gene is an expressed pseudogene.
PMID 12361482 · PMC131051 · BMC genetics · 2002 · 8 claims · 7 setups
The human TDH gene is located at chromosome 8p23-22, spans 10 kb, and has 8 exons that would be expected to encode a 369-residue ORF.
-
Full-text index only
Gene losses during human origins.
PMID 16464126 · PMC1361800 · PLoS biology · 2006 · 7 claims · 7 setups
A comparative genomic screen identified 67 new human-specific nonprocessed pseudogenes, bringing the total (with 13 from prior literature) to 80 human-specific pseudogenes.
-
Has reproduction · 50
Comparative analysis of circular RNAs between soybean cytoplasmic male-sterile line NJCMS1A and its maintainer NJCMS1B by high-throughput sequencing.
PMID 30208848 · PMC6134632 · BMC genomics · 2018 · 8 claims · 7 setups
2867 circRNAs were identified in soybean flower buds via high-throughput sequencing with RNase R enrichment, of which 1009 were differentially expressed between NJCMS1A and NJCMS1B
-
Has reproduction · 60
TRAPID 2.0: a web application for taxonomic and functional analysis of de novo transcriptomes.
PMID 34197621 · PMC8464036 · Nucleic acids research · 2021 · 8 claims · 8 setups
TRAPID 2.0 is a web application performing global characterization of de novo transcriptomes via structural, functional, and taxonomic annotation in an initial processing phase, followed by an exploratory phase of downstream analyses.
-
Full-text index only
Characterization of 954 bovine full-CDS cDNA sequences.
PMID 16305752 · PMC1314900 · BMC genomics · 2005 · 7 claims · 8 setups
954 bovine full-length insert cDNA (bFLIC) clones representing 762 distinct loci were sequenced and characterized
-
Full-text index only
An emerging cyberinfrastructure for biodefense pathogen and pathogen-host data.
PMID 17984082 · PMC2239001 · Nucleic acids research · 2008 · 8 claims · 7 setups
The Biodefense Proteomics Resource Center (RC) is a public cyberinfrastructure that stores, integrates, and disseminates experimental data from seven Proteomics Research Centers (PRCs) on biodefense pathogens and host interactions
-
Has reproduction · 57
Analysis and comprehensive comparison of PacBio and nanopore-based RNA sequencing of the Arabidopsis transcriptome.
PMID 32536962 · PMC7291481 · Plant methods · 2020 · 8 claims · 8 setups
ONT Pc produces higher raw data quality (higher alignment rate, lower error rate) than ONT Dc, while PacBio generates the longest reads
-
Has reproduction · 93
Characterization of protein isoform diversity in human umbilical vein endothelial cells via long-read proteogenomics.
PMID 36457147 · PMC9721438 · RNA biology · 2022 · 8 claims · 7 setups
Long-read RNA-seq detected 53,863 transcript isoforms from 10,426 genes in HUVECs, of which 22,195 were novel
-
Full-text index only
Comprehensive search for intra- and inter-specific sequence polymorphisms among coding envelope genes of retroviral origin found in the human genome: genes and pseudogenes.
PMID 16150157 · PMC1236922 · BMC genomics · 2005 · 8 claims · 5 setups
HERV-W (envW) and HERV-FRD (envFRD) envelope genes, both specifically expressed in placenta, show strong sequence conservation with only two nonsynonymous SNPs identified across 91 individuals
-
Full-text index only
Comparative genomics of multidrug resistance in Acinetobacter baumannii.
PMID 16415984 · PMC1326220 · PLoS genetics · 2006 · 7 claims · 5 setups
A. baumannii strain AYE contains an 86-kb genomic resistance island (AbaR1) clustering 45 of its 52 resistance genes, the largest resistance island described to date
-
Full-text index only
Evolutionary genomics reveals lineage-specific gene loss and rapid evolution of a sperm-specific ion channel complex: CatSpers and CatSperbeta.
PMID 18974790 · PMC2572835 · PloS one · 2008 · 8 claims · 6 setups
The CatSper channel complex (four CatSpers plus CatSperβ) originated as early as primitive metazoans such as the Cnidarian Nematostella vectensis
-
Full-text index only
High accuracy mass spectrometry analysis as a tool to verify and improve gene annotation using Mycobacterium tuberculosis as an example.
PMID 18597682 · PMC2483986 · BMC genomics · 2008 · 8 claims · 5 setups
High-accuracy MS proteomics (LTQ-Orbitrap) can be used to verify and improve gene annotation by identifying peptides specific to one of two competing annotation datasets.