Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Web-based resources for comparative genomics.
PMID 16197736 · PMC3525128 · Human genomics · 2005 · 8 claims · 8 setups
Comparative genomics is an indispensable tool for identifying functional genome elements and exploring evolutionary genome dynamics
-
Full-text index only
Integrative functional genomics.
PMID 15239826 · PMC463286 · Genome biology · 2004 · 8 claims · 8 setups
Ultra-conserved noncoding elements exist across human, mouse and rat genomes at very high sequence identity, often far from genes
-
Full-text index only
Genome-wide identification of human functional DNA using a neutral indel model.
PMID 16410828 · PMC1326222 · PLoS computational biology · 2006 · 8 claims · 8 setups
A neutral indel model predicting a geometric distribution of intergap segment (IGS) lengths fits human-mouse ancestral repeat (AR) alignment data excellently
-
Full-text index only
Analyses of deep mammalian sequence alignments and constraint predictions for 1% of the human genome.
PMID 17567995 · PMC1891336 · Genome research · 2007 · 7 claims · 3 setups
Four different alignment methods show large-scale consistency but substantial differences in small-scale rearrangements, sensitivity, and specificity.
-
Full-text index only
Comparative genomics.
PMID 14624258 · PMC261895 · PLoS biology · 2003 · 8 claims · 7 setups
Conserved DNA between species tends to encode shared functional features, while divergent DNA underlies species differences
-
Full-text index only
Evolutionarily conserved and diverged alternative splicing events show different expression and functional profiles.
PMID 16195578 · PMC1240112 · Nucleic acids research · 2005 · 8 claims · 5 setups
Alternative splices in 10,818 human-mouse gene pairs can be classified as conserved, novel, or diverged based on genomic and transcript-level cross-species comparison.
-
Full-text index only
MutDB: update on development of tools for the biochemical analysis of genetic variation.
PMID 17827212 · PMC2238958 · Nucleic acids research · 2008 · 7 claims · 5 setups
MutDB integrates dbSNP and Swiss-Prot genetic variation data with protein structural information, functional disruption prediction scores, and clinical phenotype links (OMIM, dbGAP)
-
Full-text index only
Analysis of the glutathione S-transferase (GST) gene family.
PMID 15607001 · PMC3500200 · Human genomics · 2004 · 8 claims · 4 setups
The complete human GST gene family comprises 16 genes in six subfamilies: alpha (GSTA), mu (GSTM), omega (GSTO), pi (GSTP), theta (GSTT) and zeta (GSTZ).
-
Full-text index only
L1Base: from functional annotation to prediction of active LINE-1 elements.
PMID 15608246 · PMC539998 · Nucleic acids research · 2005 · 7 claims · 6 setups
L1Base is a database of putatively active LINE-1 insertions in human, mouse and rat genomes, containing FLI-L1s (intact in both ORFs), ORF2-L1s (intact ORF2, disrupted ORF1), and FLnI-L1s (full-length, >6000 bp, non-intact)
-
Full-text index only
Sequence and structure signatures of cancer mutation hotspots in protein kinases.
PMID 19834613 · PMC2759519 · PloS one · 2009 · 8 claims · 6 setups
Developed CKMD (Composite Kinase Mutation Database), an integrated bioinformatics resource mapping genetic variation in protein kinase genes to sequence, structural, and functional data
-
Full-text index only
NCBI Reference Sequences: current status, policy and new initiatives.
PMID 18927115 · PMC2686572 · Nucleic acids research · 2009 · 7 claims · 5 setups
RefSeq is a curated, non-redundant, explicitly linked database of nucleotide and protein sequences spanning genomes, transcripts and proteins across prokaryotes, eukaryotes and viruses
-
Has reproduction · 86
Assessing Bos taurus introgression in the UOA Bos indicus assembly.
PMID 34922445 · PMC8684283 · Genetics, selection, evolution : GSE · 2021 · 7 claims · 6 setups
Aligning divergent (cross-subspecies) sequence data detects substantially more SNVs than aligning to a same-subspecies reference, indicating reference/assembly bias in variant calling.
-
Full-text index only
Linking disease-associated genes to regulatory networks via promoter organization.
PMID 15701758 · PMC549397 · Nucleic acids research · 2005 · 8 claims · 7 setups
Pairs of TFBSs conserved both vertically (orthologous genes) and horizontally (co-regulated genes) can serve as seeds to build promoter models representing potential co-regulation networks
-
Full-text index only
Genome-wide in silico identification and analysis of cis natural antisense transcripts (cis-NATs) in ten species.
PMID 16849434 · PMC1524920 · Nucleic acids research · 2006 · 8 claims · 7 setups
A fast integrative in silico pipeline combining UniGene mRNA/EST mapping to GoldenPath genomes with CDS, poly(A) signal, poly(A) tail and splicing site evidence can reliably identify cis-NATs genome-wide across multiple species
-
Full-text index only
MACSIMS: multiple alignment of complete sequences information management system.
PMID 16792820 · PMC1539025 · BMC bioinformatics · 2006 · 8 claims · 5 setups
MACSIMS is a multiple alignment-based information management system combining knowledge-based database mining with ab initio sequence predictions
-
Full-text index only
Functional analysis of novel SNPs and mutations in human and mouse genomes.
PMID 19091009 · PMC2638150 · BMC bioinformatics · 2008 · 8 claims · 7 setups
FANS streamlines functional analysis of novel SNPs and mutations into a simplified, few-click, four-step procedure.
-
Full-text index only
Genome wide identification of recessive cancer genes by combinatorial mutation analysis.
PMID 18846217 · PMC2557123 · PloS one · 2008 · 7 claims · 4 setups
A combinatorial mutation analysis identified 154 candidate recessive cancer genes (pRecessiveCancer<1.5x10-7, FDR=0.39)
-
Has reproduction · 65
SPEAQeasy: a scalable pipeline for expression analysis and quantification for R/bioconductor-powered RNA-seq analyses.
PMID 33932985 · PMC8088074 · BMC bioinformatics · 2021 · 8 claims · 5 setups
SPEAQeasy is a portable, easy-to-install, Nextflow-powered RNA-seq processing pipeline that lowers the computational entry barrier for biologists/clinicians
-
Full-text index only
A genome-wide survey demonstrates widespread non-linear mRNA in expressed sequences from multiple species.
PMID 16237125 · PMC1258171 · Nucleic acids research · 2005 · 8 claims · 6 setups
A genome-wide computational survey identifies 245 genes in mammals (264 across six species) that produce RREO events in expressed sequences
-
Full-text index only
ECgene: genome annotation for alternative splicing.
PMID 15608289 · PMC540072 · Nucleic acids research · 2005 · 8 claims · 5 setups
ECgene combines genome-based EST clustering with a graph-theoretic transcript assembly procedure to predict gene models including alternative splicing events.