Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Exonic remnants of whole-genome duplication reveal cis-regulatory function of coding exons.
PMID 19969543 · PMC2831330 · Nucleic acids research · 2010 · 8 claims · 8 setups
38 candidate cis-regulatory coding exons (RCEs) with predicted target genes were identified genome-wide
-
Full-text index only
Exogean: a framework for annotating protein-coding genes in eukaryotic genomic DNA.
PMID 16925841 · PMC1810556 · Genome biology · 2006 · 8 claims · 5 setups
Exogean is a framework using directed acyclic coloured multigraphs (DACMs) to represent biological objects (mRNA, ESTs, protein alignments, exons) and iteratively combine them into complex protein-coding transcript models.
-
Full-text index only
Novel gene and gene model detection using a whole genome open reading frame analysis in proteomics.
PMID 16646984 · PMC1557991 · Genome biology · 2006 · 8 claims · 4 setups
A six-frame genomic ORF translation used as an MS search database can detect novel peptides absent from standard protein databases, revealing incomplete genome annotation.
-
Has reproduction · 77
Spatially clustered loci with multiple enhancers are frequent targets of HIV-1 integration.
PMID 31492853 · PMC6731298 · Nature communications · 2019 · 8 claims · 7 setups
HIV-1 recurrently integrates into genes that are proximal to super-enhancer (SE) genomic elements in both patients and in vitro T cell cultures.
-
Full-text index only
Assessing the genomic evidence for conserved transcribed pseudogenes under selection.
PMID 19754956 · PMC2753554 · BMC genomics · 2009 · 8 claims · 8 setups
1750 transcribed pseudogene annotations (TPAs) were identified in the human genome, ~11.5% of all human pseudogene annotations.
-
Has reproduction · 91
Chromosome-level genome assembly of agar-producing red seaweed Gracilaria vermiculophylla.
PMID 41629338 · PMC12966425 · Scientific data · 2026 · 8 claims · 8 setups
A chromosome-level genome assembly of G. vermiculophylla was generated by combining DNBSeq short reads, Nanopore long reads, and Hi-C data.
-
Has reproduction · 75
geneshot: gene-level metagenomics identifies genome islands associated with immunotherapy response.
PMID 33952321 · PMC8097837 · Genome biology · 2021 · 8 claims · 4 setups
geneshot is a gene-level metagenomic bioinformatics tool that clusters de novo assembled protein-coding genes into co-abundant gene groups (CAGs) to reduce dimensionality and generate testable hypotheses from WGS microbiome data
-
Full-text index only
The promoter and the enhancer region of the KLK 3 (prostate specific antigen) gene is frequently mutated in breast tumours and in breast carcinoma cell lines.
PMID 10188912 · PMC2362704 · British journal of cancer · 1999 · 7 claims · 5 setups
No mutations were found in the protein-coding exons of the PSA gene in breast tumours or cell lines
-
Full-text index only
AUGUSTUS at EGASP: using EST, protein and genomic alignments for improved gene prediction in the human genome.
PMID 16925833 · PMC1810548 · Genome biology · 2006 · 8 claims · 5 setups
AUGUSTUS predicted significantly more genes correctly than any other ab initio program in EGASP
-
Full-text index only
Human PAML browser: a database of positive selection on human genes using phylogenetic methods.
PMID 17962310 · PMC2238824 · Nucleic acids research · 2008 · 8 claims · 5 setups
The Human PAML Browser is a web-accessible database of codeml-based positive selection test results for 13,721 human genes with orthologs in UCSC multispecies alignments.
-
Full-text index only
Paired-end mapping reveals extensive structural variation in the human genome.
PMID 17901297 · PMC2674581 · Science (New York, N.Y.) · 2007 · 8 claims · 8 setups
Paired-end mapping (PEM) combining 3-kb fragment paired-end capture, massive 454 sequencing, and computational mapping detects SVs ~3 kb or larger with an average breakpoint resolution of 644 bp
-
Full-text index only
The DNA sequence and analysis of human chromosome 13.
PMID 15057823 · PMC2665288 · Nature · 2004 · 8 claims · 8 setups
95.5 Mb of finished sequence from chromosome 13 was completed, containing 633 genes and 296 pseudogenes.
-
Has reproduction · 76
The genome and development-dependent transcriptomes of Pyronema confluens: a window into fungal evolution.
PMID 24068976 · PMC3778014 · PLoS genetics · 2013 · 8 claims · 8 setups
The 50 Mb P. confluens genome with 13,369 predicted protein-coding genes is more characteristic of higher filamentous ascomycetes than of the large, repeat-rich Tuber melanosporum genome, showing that the truffle's expanded genome is not typical of the Pezizales.
-
Full-text index only
Retropseudogenes derived from the human Ro/SS-A autoantigen-associated hY RNAs.
PMID 15817567 · PMC1074747 · Nucleic acids research · 2005 · 8 claims · 8 setups
966 pseudogenes derived from the four human Y (hY) RNAs were characterized in the human genome
-
Full-text index only
Complete genome sequence and comparative analysis of the wild-type commensal Escherichia coli strain SE11 isolated from a healthy adult.
PMID 18931093 · PMC2608844 · DNA research : an international journal for rapid publication of reports on genes and genomes · 2008 · 8 claims · 6 setups
The SE11 genome comprises a 4.8 Mb chromosome encoding 4679 protein-coding genes and six plasmids encoding 323 protein-coding genes
-
Full-text index only
Next-generation sequencing.
PMID 20030863 · PMC2797692 · Breast cancer research : BCR · 2009 · 8 claims · 7 setups
Massively parallel sequencing can simultaneously capture base-pair mutations, copy number aberrations and somatic rearrangements of a cancer genome in a single experiment
-
Full-text index only
F-SNP: computationally predicted functional SNPs for disease association studies.
PMID 17986460 · PMC2238878 · Nucleic acids research · 2008 · 6 claims · 8 setups
F-SNP is a database integrating functional effect predictions for SNPs from 16 bioinformatics tools/databases across four categories: splicing, transcription, translation, and post-translation
-
Full-text index only
Web-based resources for comparative genomics.
PMID 16197736 · PMC3525128 · Human genomics · 2005 · 8 claims · 8 setups
Comparative genomics is an indispensable tool for identifying functional genome elements and exploring evolutionary genome dynamics
-
Has reproduction · 94
A Deluge of Complex Repeats: The Solanum Genome.
PMID 26241045 · PMC4524691 · PloS one · 2015 · 8 claims · 7 setups
~50–60% of the S. tuberosum and S. lycopersicum genomes are composed of repetitive elements
-
Full-text index only
Integrative annotation of 21,037 human genes validated by full-length cDNA clones.
PMID 15103394 · PMC393292 · PLoS biology · 2004 · 8 claims · 5 setups
41,118 full-length human cDNAs from six high-throughput sequencing projects were exhaustively integratively characterized