Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Retropseudogenes derived from the human Ro/SS-A autoantigen-associated hY RNAs.
PMID 15817567 · PMC1074747 · Nucleic acids research · 2005 · 8 claims · 8 setups
966 pseudogenes derived from the four human Y (hY) RNAs were characterized in the human genome
-
Has reproduction · 30
Genomic and transcriptomic plasticity in treatment-naive ovarian cancer.
PMID 24221193 · PMC3912411 · Genome research · 2014 · 8 claims · 8 setups
Treatment-naïve epithelial ovarian cancers show extensive intra-tumor heterogeneity of genomic rearrangements, with the most substantial differences occurring between omentum/peritoneum metastases and ovarian tumor sites.
-
Full-text index only
Transcription of the human and rodent SPAM1 / PH-20 genes initiates within an ancient endogenous retrovirus.
PMID 15804358 · PMC1079825 · BMC genomics · 2005 · 8 claims · 8 setups
Human, mouse, and rat SPAM1/Spam1 transcripts initiate within an ERV1 pol (internal coding) region rather than within an LTR
-
Full-text index only
Two modes of microsatellite instability in human cancer: differential connection of defective DNA mismatch repair to dinucleotide repeat instability.
PMID 15778432 · PMC1067522 · Nucleic acids research · 2005 · 8 claims · 8 setups
Dinucleotide microsatellite alterations in human cancer fall into two distinct modes: Type A (length changes ≤6 bp) and Type B (changes ≥8 bp)
-
Full-text index only
Large-scale discovery and validation of functional elements in the human genome.
PMID 15774039 · PMC1088940 · Genome biology · 2005 · 8 claims · 8 setups
Genome-wide tiling microarray hybridization reveals large, diverse sets of transcripts, many of which lack existing gene annotations
-
Full-text index only
The DNA sequence of the human X chromosome.
PMID 15772651 · PMC2665286 · Nature · 2005 · 8 claims · 8 setups
The euchromatic sequence of the human X chromosome was determined to 99.3% completeness (~155 Mb total)
-
Full-text index only
Speeding disease gene discovery by sequence based candidate prioritization.
PMID 15766383 · PMC1274252 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Disease genes (OMIM) differ significantly from non-disease genes in sequence-based features including gene/cDNA/protein size, exon number, homolog conservation, secretion signal, 3' UTR length, CpG islands, and distance to nearest gene.
-
Full-text index only
Mice have a transcribed L-threonine aldolase/GLY1 gene, but the human GLY1 gene is a non-processed pseudogene.
PMID 15757516 · PMC555945 · BMC genomics · 2005 · 8 claims · 8 setups
Mouse has a transcribed, 7-exon L-threonine aldolase (GLY1) gene on chromosome 11 encoding a 400-residue protein homologous to bacterial threonine aldolase
-
Full-text index only
Phylogenomic approaches to common problems encountered in the analysis of low copy repeats: the sulfotransferase 1A gene family example.
PMID 15752422 · PMC555591 · BMC evolutionary biology · 2005 · 8 claims · 8 setups
A previously unidentified fourth human SULT1A gene (SULT1A4) exists on chromosome 16 and is transcriptionally active
-
Full-text index only
Computational verification of protein-protein interactions by orthologous co-expression.
PMID 15740634 · PMC555590 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Co-expression of orthologous protein pairs across multiple species can verify/predict S. cerevisiae PPIs with better performance than S. cerevisiae co-expression alone.
-
Has reproduction
Using random walks to identify cancer-associated modules in expression data.
PMID 24128261 · PMC4015830 · BioData mining · 2013 · 8 claims · 8 setups
Walktrap-GM, a random-walk community detection algorithm adapted with stopping criteria (maximum modularity, maximum size, maximum module score), identifies modules significantly enriched with cancer genes in expression-weighted interaction networks.
-
Has reproduction · 95
Quantitative epigenetic co-variation in CpG islands and co-regulation of developmental genes.
PMID 23999385 · PMC6505400 · Scientific reports · 2013 · 8 claims · 8 setups
Four epigenetic modifications (DNA methylation, H3K4me2, H3K4me3, H3K27me3) in mouse CGIs undergo combinatorial variation (co-variation) across ESCs, NPCs and adult brain during neuron differentiation.
-
Has reproduction · 49
Aberration in DNA methylation in B-cell lymphomas has a complex origin and increases with disease severity.
PMID 23326238 · PMC3542081 · PLoS genetics · 2013 · 8 claims · 8 setups
B-cell non-Hodgkin lymphomas display striking intra-tumor (intra-sample) and inter-patient (inter-sample) cytosine methylation heterogeneity that increases progressively with disease aggressiveness (NBC<NGC<FL<GCB<ABC).
-
Full-text index only
Cystic fibrosis in Korean children:a case report identified by a quantitative pilocarpine iontophoresis sweat test and genetic analysis.
PMID 15716623 · PMC2808565 · Journal of Korean medical science · 2005 · 8 claims · 8 setups
CF should be suspected in Korean/Asian children with chronic respiratory symptoms despite its rarity in Asian populations
-
Full-text index only
Ontological visualization of protein-protein interactions.
PMID 15707487 · PMC550656 · BMC bioinformatics · 2005 · 8 claims · 8 setups
Aggregating independently made GO 'protein binding' (IPI) annotations reveals larger, previously undescribed mouse protein-protein interaction networks
-
Full-text index only
Fast and systematic genome-wide discovery of conserved regulatory elements using a non-alignment based approach.
PMID 15693947 · PMC551538 · Genome biology · 2005 · 7 claims · 8 setups
FastCompare, a non-alignment-based, linear-time algorithm, computes a genome-wide conservation score for all k-mers (7-9 nt) between two genomes to identify conserved regulatory elements
-
Full-text index only
Non-linear mapping for exploratory data analysis in functional genomics.
PMID 15661072 · PMC548129 · BMC bioinformatics · 2005 · 8 claims · 8 setups
A relaxation method for non-linear mapping adapts one pair of points per step rather than all points at once, and was originally shown by Chang and Lee to outperform Sammon's mapping in cluster detection effectiveness and computational efficiency.
-
Full-text index only
Bioinformatic mapping of AlkB homology domains in viruses.
PMID 15627404 · PMC544882 · BMC genomics · 2005 · 8 claims · 8 setups
AlkB-like domains are found in at least 22 different single-stranded RNA positive-strand plant viruses, mainly within a subgroup of the Flexiviridae family.
-
Full-text index only
DG-CST (Disease Gene Conserved Sequence Tags), a database of human-mouse conserved elements associated to disease genes.
PMID 15608249 · PMC539965 · Nucleic acids research · 2005 · 5 claims · 8 setups
Comparative human-mouse genome analysis identifies conserved sequence tags (CSTs, >=70% identity over >=100bp) that frequently correspond to non-coding elements with putative regulatory or structural roles
-
Has reproduction · 71
Protein structure quality assessment based on the distance profiles of consecutive backbone Cα atoms.
PMID 24555103 · PMC3892923 · F1000Research · 2013 · 8 claims · 8 setups
The distance between consecutive backbone Cα atoms in high-quality structures is normally distributed with mean 3.8 Å and standard deviation 0.04 Å, justifying a reference state in which all consecutive Cα atoms are 3.8 Å apart.