Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
In vitro identification and in silico utilization of interspecies sequence similarities using GeneChip technology.
PMID 15871745 · PMC1156887 · BMC genomics · 2005 · 7 claims · 6 setups
Only 14±2% of canine transcripts were detected by U133A probe sets versus 49±6% of human transcripts when hybridized to the same chip
-
Full-text index only
Genome-wide survey for biologically functional pseudogenes.
PMID 16680195 · PMC1456316 · PLoS computational biology · 2006 · 8 claims · 6 setups
A subset of ancient, cross-species-conserved pseudogenes (30 of 1,453 candidate quartets) show evidence consistent with retained biological function
-
Full-text index only
Design factors that influence PCR amplification success of cross-species primers among 1147 mammalian primer pairs.
PMID 17029642 · PMC1635982 · BMC genomics · 2006 · 8 claims · 7 setups
The number of index-species (IS) mismatches in a primer pair significantly reduces amplification success, with an estimated 6-8% decrease in success rate per additional mismatch.
-
Full-text index only
Given the complexity of the human genome, can 'personalised medicine' or 'individualised drug therapy' ever be achieved?
PMID 19706359 · PMC3525196 · Human genomics · 2009 · 7 claims · 3 setups
The human genome is far too complex, given current understanding, for personalised medicine or individualised drug therapy to be realised in the near term
-
Full-text index only
Science review: searching for gene candidates in acute lung injury.
PMID 15566614 · PMC1065043 · Critical care (London, England) · 2004 · 8 claims · 8 setups
The candidate gene approach combined with an ortholog gene database and gene ontology analysis identifies ALI candidate genes, with blood coagulation and inflammation ontologies most highly represented
-
Full-text index only
A clustering property of highly-degenerate transcription factor binding sites in the mammalian genome.
PMID 16670430 · PMC1456330 · Nucleic acids research · 2006 · 8 claims · 7 setups
Highly-degenerate RE1 sites are significantly enriched in promoters of validated and putative REST target genes compared to control promoters
-
Full-text index only
DAVID Knowledgebase: a gene-centered database integrating heterogeneous gene annotation resources to facilitate high-throughput gene functional analysis.
PMID 17980028 · PMC2186358 · BMC bioinformatics · 2007 · 7 claims · 3 setups
The DAVID Gene Concept, a single-linkage algorithm, merges gene clusters from Entrez Gene, UniRef100, and PIR-NREF100 that share protein IDs and species into unified DAVID gene clusters, improving cross-referencing between NCBI and UniProt systems
-
Has reproduction · 86
Improving the annotation of the cattle genome by annotating transcription start sites in a diverse set of tissues and populations using Cap Analysis Gene Expression sequencing.
PMID 37216666 · PMC10411599 · G3 (Bethesda, Md.) · 2023 · 7 claims · 8 setups
CAGE sequencing of 24 tissues from 3 cattle populations (dairy, beef-dairy cross, Kinsella composite) defines TSS and coexpressed short-range enhancers in the ARS-UCD1.2 reference genome
-
Full-text index only
Structural organization and interactions of transmembrane domains in tetraspanin proteins.
PMID 15985154 · PMC1190194 · BMC structural biology · 2005 · 8 claims · 5 setups
TM1, TM2 and TM3 of human tetraspanins display a distinct heptad repeat motif (abcdefg)n, while TM4 lacks this motif.
-
Full-text index only
Gene Prospector: an evidence gateway for evaluating potential susceptibility genes and interacting risk factors for human diseases.
PMID 19063745 · PMC2613935 · BMC bioinformatics · 2008 · 8 claims · 5 setups
Gene Prospector is a Web-based application that selects and prioritizes potential disease-related genes using a curated, updated literature database of genetic association studies
-
Full-text index only
Differentiation of core promoter architecture between plants and mammals revealed by LDSS analysis.
PMID 17855401 · PMC2094075 · Nucleic acids research · 2007 · 7 claims · 8 setups
LDSS analysis identifies octamer sequences with localized distribution profiles as promoter constituents, classifiable into groups (REG, TATA, Inr, Kozak, CpG, Y Patch)
-
Full-text index only
Conserved elements with potential to form polymorphic G-quadruplex structures in the first intron of human genes.
PMID 18187510 · PMC2275096 · Nucleic acids research · 2008 · 8 claims · 6 setups
G-richness downstream of the TSS is strand-biased, concentrated on the nontemplate strand, with a peak at +200 to +300 bp
-
Full-text index only
Undergraduate research. Genomics Education Partnership.
PMID 18974335 · PMC2953277 · Science (New York, N.Y.) · 2008 · 6 claims · 4 setups
A course-embedded, multi-institution undergraduate research model (the Genomics Education Partnership) can deliver authentic research experiences during the academic year rather than only in summer programs.
-
Full-text index only
Genomewide pattern of synonymous nucleotide substitution in two complete genomes of Mycobacterium tuberculosis.
PMID 12453367 · PMC2738538 · Emerging infectious diseases · 2002 · 8 claims · 6 setups
Genomewide comparison of two complete M. tuberculosis genomes reveals substantially more nucleotide diversity than prior studies based on few loci suggested
-
Full-text index only
Coiled-coil protein composition of 22 proteomes--differences and common themes in subcellular infrastructure and traffic control.
PMID 16288662 · PMC1322226 · BMC evolutionary biology · 2005 · 7 claims · 5 setups
Proteins with extended coiled-coil domains (>250 amino acids) are largely absent from bacterial genomes but present in archaea and eukaryotes.
-
Full-text index only
Filtering high-throughput protein-protein interaction data using a combination of genomic features.
PMID 15833142 · PMC1127019 · BMC bioinformatics · 2005 · 8 claims · 8 setups
A combination of three genomic features (interacting Pfam domains, GO annotations, sequence homology) using naive Bayesian networks predicts true protein-protein interactions with high sensitivity and good specificity.
-
Full-text index only
The jewels of our genome: the search for the genomic changes underlying the evolutionarily unique capacities of the human brain.
PMID 16733552 · PMC1464830 · PLoS genetics · 2006 · 8 claims · 7 setups
Human and chimp genomes differ by ~35 million single nucleotide substitutions, corresponding to ~1.06% divergence after removing polymorphic sites
-
Full-text index only
Pol II promoter prediction using characteristic 4-mer motifs: a machine learning approach.
PMID 18834544 · PMC2575220 · BMC bioinformatics · 2008 · 8 claims · 8 setups
128 discriminating 4-mer motifs combined with an SVM (RBF kernel, LIBSVM) can distinguish promoter from non-promoter DNA sequences
-
Full-text index only
Meta-analysis of inter-species liver co-expression networks elucidates traits associated with common human diseases.
PMID 20019805 · PMC2787626 · PLoS computational biology · 2009 · 8 claims · 8 setups
A novel semi-parametric meta-analysis method (based on a gene-centric Glass's d effect size) outperforms existing parametric and non-parametric meta-analysis methods at identifying functionally coherent gene pairs across species.
-
Full-text index only
Atlas of nascent RNA transcripts reveals tissue-specific enhancer to gene linkages.
PMID 40281430 · PMC12032694 · BMC genomics · 2025 · 7 claims · 8 setups
A large repository of nascent run-on RNA-seq samples (DBNascent) was assembled and uniformly processed to identify sites of bidirectional transcription genome-wide.