Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Towards alignment independent quantitative assessment of homology detection.
PMID 17205117 · PMC1762415 · PloS one · 2006 · 8 claims · 6 setups
The Fhom Estimator uses the prevalence of a conserved protein feature (X) in two protein sets to estimate the fraction of true homologs among paired proteins, independent of alignment quality.
-
Full-text index only
Proteomic analysis of stage I primary lung adenocarcinoma aimed at individualisation of postoperative therapy.
PMID 18212748 · PMC2243141 · British journal of cancer · 2008 · 5 claims · 6 setups
LC-MS/MS proteomic analysis of stage I lung adenocarcinoma specimens identified myosin IIA and vimentin as candidate biomarker proteins with signal intensities that differed significantly among patient outcome groups
-
Has reproduction · 100
A Bioinformatics Workflow to Identify eccDNA Using ECCFP From Long-Read Nanopore Sequencing Data.
PMID 41924242 · PMC13037781 · Bio-protocol · 2026 · 7 claims · 5 setups
ECCFP significantly improves eccDNA detection sensitivity, accuracy, and runtime efficiency compared to other pipelines
-
Full-text index only
FeatureScan: revealing property-dependent similarity of nucleotide sequences.
PMID 16845077 · PMC1538849 · Nucleic acids research · 2006 · 6 claims · 5 setups
FeatureScan transforms nucleotide sequences into numerical signals of physico-chemical/conformational properties and compares them via a convolution/correlation (Fourier transform) method rather than comparing letters
-
Full-text index only
Discovery of novel human transcript variants by analysis of intronic single-block EST with polyadenylation site.
PMID 19906316 · PMC2784480 · BMC genomics · 2009 · 8 claims · 7 setups
Intronic single-block ESTs with poly(A/T) tails reveal previously unidentified novel transcript variants missed by existing databases.
-
Full-text index only
Emerging genomic and proteomic evidence on relationships among the animal, plant and fungal kingdoms.
PMID 15629046 · PMC5172449 · Genomics, proteomics & bioinformatics · 2004 · 8 claims · 7 setups
Sequence-based molecular phylogenies widely support a sister relationship between animals and fungi, grouped as the Opisthokonta
-
Full-text index only
Large-scale analysis of human alternative protein isoforms: pattern classification and correlation with subcellular localization signals.
PMID 15860772 · PMC1087780 · Nucleic acids research · 2005 · 8 claims · 8 setups
Constructed a large-scale dataset of 6876 human alternative protein isoforms from 2624 genes by combining H-Invitational full-length cDNA data and SwissProt VARSPLIC entries
-
Full-text index only
Phylogenetic analysis of RhoGAP domain-containing proteins.
PMID 17127216 · PMC5054073 · Genomics, proteomics & bioinformatics · 2006 · 7 claims · 6 setups
RhoGAP domain-containing proteins, sharing the conserved arginine residue, form a monophyletic group with a common ancestor.
-
Full-text index only
Genome-wide in silico identification and analysis of cis natural antisense transcripts (cis-NATs) in ten species.
PMID 16849434 · PMC1524920 · Nucleic acids research · 2006 · 8 claims · 7 setups
A fast integrative in silico pipeline combining UniGene mRNA/EST mapping to GoldenPath genomes with CDS, poly(A) signal, poly(A) tail and splicing site evidence can reliably identify cis-NATs genome-wide across multiple species
-
Full-text index only
Large-scale identification and characterization of alternative splicing variants of human gene transcripts using 56,419 completely sequenced and manually annotated full-length cDNAs.
PMID 16914452 · PMC1557807 · Nucleic acids research · 2006 · 8 claims · 8 setups
Analysis of 56,419 full-length cDNAs identified 6877 alternative splicing genes encoding 18,297 alternative splicing variants made of 37,670 exons.
-
Full-text index only
Genome wide identification of recessive cancer genes by combinatorial mutation analysis.
PMID 18846217 · PMC2557123 · PloS one · 2008 · 7 claims · 4 setups
A combinatorial mutation analysis identified 154 candidate recessive cancer genes (pRecessiveCancer<1.5x10-7, FDR=0.39)
-
Has reproduction · 100
ChIP-seq Data Processing and Relative and Quantitative Signal Normalization for Saccharomyces cerevisiae.
PMID 40364978 · PMC12067309 · Bio-protocol · 2025 · 8 claims · 6 setups
siQ-ChIP measures absolute protein–DNA interaction (IP efficiency) genome-wide without relying on exogenous spike-in chromatin, overcoming limitations of spike-in normalization.
-
Full-text index only
Characterization of rabbit myocilin: Implications for human myocilin glycosylation and signal peptide usage.
PMID 12697062 · PMC156599 · BMC genetics · 2003 · 8 claims · 6 setups
Rabbit MYOC encodes a 490 amino acid, 54,882-Da protein that is 84% identical overall to human myocilin
-
Full-text index only
Phylogenetic analysis of mRNA polyadenylation sites reveals a role of transposable elements in evolution of the 3'-end of genes.
PMID 18757892 · PMC2553571 · Nucleic acids research · 2008 · 8 claims · 6 setups
3'-most (L type) poly(A) sites are more conserved than upstream F/M type sites, while intronic (C/H type) sites are the least conserved
-
Full-text index only
Comparative genomic analysis reveals a novel mitochondrial isoform of human rTS protein and unusual phylogenetic distribution of the rTS gene.
PMID 16162288 · PMC1261261 · BMC genomics · 2005 · 8 claims · 5 setups
A novel rTS protein isoform, rTSγ, exists with a 27-residue longer N-terminus generated from an alternative upstream start codon relative to rTSβ.
-
Full-text index only
A third approach to gene prediction suggests thousands of additional human transcribed regions.
PMID 16543943 · PMC1391917 · PLoS computational biology · 2006 · 8 claims · 7 setups
A third basic concept for gene prediction exists, based on detecting strand-specific 'transcription footprints' (mutational and selectional biases) rather than gene structure or sequence similarity.
-
Has reproduction · 98
Massively parallel genomic perturbations with multi-target CRISPR interrogates Cas9 activity and DNA repair at endogenous sites.
PMID 36064968 · PMC9481459 · Nature cell biology · 2022 · 8 claims · 6 setups
Multi-target gRNAs (mgRNAs) can direct Cas9 to over a hundred well-mapped endogenous genomic sites simultaneously, enabling massively parallel, high-throughput interrogation of Cas9 activity via short-read sequencing
-
Full-text index only
SeqDoC: rapid SNP and mutation detection by direct comparison of DNA sequence chromatograms.
PMID 15927052 · PMC1156871 · BMC bioinformatics · 2005 · 8 claims · 6 setups
SeqDoC generates a subtracted difference trace between a reference and test chromatogram that highlights single base changes
-
Full-text index only
The ENCODE Project at UC Santa Cruz.
PMID 17166863 · PMC1781110 · Nucleic acids research · 2007 · 8 claims · 4 setups
The UCSC ENCODE portal serves as the primary repository and access point for sequence-based ENCODE pilot phase data
-
Full-text index only
Identification and characterisation of the angiotensin converting enzyme-3 (ACE3) gene: a novel mammalian homologue of ACE.
PMID 17597519 · PMC1925091 · BMC genomics · 2007 · 7 claims · 7 setups
A novel single-domain ACE-like gene, ACE3, exists in mouse, rat, cow, dog and human genomes, located on the same chromosome downstream of ACE.