Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
With the finished human genome in hand, what next?
PMID 12844356 · PMC193627 · Genome biology · 2003 · 8 claims · 8 setups
Gene Ontology (GO) provides a syntax/query framework for functional classification of genes, expanding beyond E. coli origins into anatomy, pathology, and phenotype data.
-
Has reproduction · 96
GC-biased gene conversion conceals the prediction of the nearly neutral theory in avian genomes.
PMID 30616647 · PMC6322265 · Genome biology · 2019 · 8 claims · 6 setups
gBGC conceals the correlation between life-history traits and dN/dS in birds; accounting for it reveals correlations consistent with nearly neutral theory
-
Full-text index only
Large-scale analysis of human alternative protein isoforms: pattern classification and correlation with subcellular localization signals.
PMID 15860772 · PMC1087780 · Nucleic acids research · 2005 · 8 claims · 8 setups
Constructed a large-scale dataset of 6876 human alternative protein isoforms from 2624 genes by combining H-Invitational full-length cDNA data and SwissProt VARSPLIC entries
-
Has reproduction · 68
Rfam 15: RNA families database in 2025.
PMID 39526405 · PMC11701678 · Nucleic acids research · 2025 · 8 claims · 6 setups
Rfamseq was expanded to 26 106 genomes, a 76% increase, by incorporating the latest UniProt reference proteomes and additional viral genomes
-
Full-text index only
Analysis of the prostate cancer cell line LNCaP transcriptome using a sequencing-by-synthesis approach.
PMID 17010196 · PMC1592491 · BMC genomics · 2006 · 8 claims · 7 setups
High-throughput 454 sequencing-by-synthesis of LNCaP cDNA can profile transcript abundance across the transcriptome
-
Full-text index only
Mitochondrial diversity within modern human populations.
PMID 17439969 · PMC1888801 · Nucleic acids research · 2007 · 8 claims · 5 setups
Modern humans show extremely low divergence from the mitochondrial consensus sequence, differing on average by only 21.6 nucleotide sites
-
Full-text index only
Characterization of the human DYRK1A promoter and its regulation by the transcription factor E2F1.
PMID 18366763 · PMC2292204 · BMC molecular biology · 2008 · 8 claims · 8 setups
Transcription start sites of human DYRK1A are distributed over an 800 bp region within an unmethylated CpG island
-
Has reproduction · 86
Assessing Bos taurus introgression in the UOA Bos indicus assembly.
PMID 34922445 · PMC8684283 · Genetics, selection, evolution : GSE · 2021 · 7 claims · 6 setups
Aligning divergent (cross-subspecies) sequence data detects substantially more SNVs than aligning to a same-subspecies reference, indicating reference/assembly bias in variant calling.
-
Has reproduction · 57
Analysis and comprehensive comparison of PacBio and nanopore-based RNA sequencing of the Arabidopsis transcriptome.
PMID 32536962 · PMC7291481 · Plant methods · 2020 · 8 claims · 8 setups
ONT Pc produces higher raw data quality (higher alignment rate, lower error rate) than ONT Dc, while PacBio generates the longest reads
-
Full-text index only
Genome informatics: taming the avalanche of genomic data.
PMID 15642109 · PMC549058 · Genome biology · 2005 · 8 claims · 7 setups
Ultraconserved regions (>100 bp, 100% conserved among mammals) exist in the genome and their function remains unknown
-
Full-text index only
ARED 3.0: the large and diverse AU-rich transcriptome.
PMID 16381826 · PMC1347415 · Nucleic acids research · 2006 · 7 claims · 6 setups
ARED 3.0 computationally mapped more than 4000 ARE-mRNAs to the human genome, representing 5-8% of human genes.
-
Full-text index only
Twin peaks: the draft human genome sequence.
PMID 11276423 · PMC138909 · Genome biology · 2001 · 8 claims · 8 setups
The predicted number of human genes (~26,000-40,000) is far lower than the widely assumed ~100,000, though downstream RNA/protein complexity can still generate substantial biological complexity.
-
Full-text index only
Extending Asia Pacific bioinformatics into new realms in the "-omics" era.
PMID 19958472 · PMC2788361 · BMC genomics · 2009 · 8 claims · 6 setups
88 full paper submissions were peer-reviewed for InCoB2009, with 49 shortlisted for oral presentation and 34 accepted into this BMC Genomics supplement, reflecting an overall acceptance rate of 50% across venues.
-
Has reproduction · 69
Discovery and characterization of Alu repeat sequences via precise local read assembly.
PMID 26503250 · PMC4666360 · Nucleic acids research · 2015 · 7 claims · 8 setups
Combining Alu-supporting read detection (RetroSeq) with local de novo assembly (CAP3) reconstructs the full sequence of non-reference Alu insertions from Illumina paired-end WGS reads
-
Full-text index only
Differences in the evolutionary history of disease genes affected by dominant or recessive mutations.
PMID 16817963 · PMC1534034 · BMC genomics · 2006 · 8 claims · 8 setups
Dominant disease genes are more conserved at the protein level (mouse orthologues) than recessive disease genes.