Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
DDBJ in collaboration with mass-sequencing teams on annotation.
PMID 15608189 · PMC539974 · Nucleic acids research · 2005 · 7 claims · 5 setups
DDBJ collected and released 1,066,084 entries (718,072,425 bases) in the past year, including the complete chimpanzee chromosome 22 sequence and silkworm whole-genome shotgun data
-
Full-text index only
The jewels of our genome: the search for the genomic changes underlying the evolutionarily unique capacities of the human brain.
PMID 16733552 · PMC1464830 · PLoS genetics · 2006 · 8 claims · 7 setups
Human and chimp genomes differ by ~35 million single nucleotide substitutions, corresponding to ~1.06% divergence after removing polymorphic sites
-
Full-text index only
KEGG for linking genomes to life and the environment.
PMID 18077471 · PMC2238879 · Nucleic acids research · 2008 · 8 claims · 4 setups
KEGG provides a reference knowledge base for linking genomes to life via PATHWAY mapping and to the environment via BRITE mapping.
-
Full-text index only
Coverage and characteristics of the Affymetrix GeneChip Human Mapping 100K SNP set.
PMID 16680197 · PMC1456318 · PLoS genetics · 2006 · 7 claims · 7 setups
SNPs in the Affymetrix 100K set are undersampled from coding regions (both synonymous and nonsynonymous) and oversampled from regions outside genes, relative to HapMap SNPs
-
Has reproduction · 85
Systematic Analysis of Long Non-Coding RNA Genes in Nonalcoholic Fatty Liver Disease.
PMID 35893239 · PMC9332188 · Non-coding RNA · 2022 · 6 claims · 5 setups
Many lncRNAs, like protein-coding genes, are differentially expressed in NAFLD patients compared to healthy normal-weight and obese individuals.
-
Has reproduction · 69
A comparison across non-model animals suggests an optimal sequencing depth for de novo transcriptome assembly.
PMID 23496952 · PMC3655071 · BMC genomics · 2013 · 8 claims · 8 setups
Representative de novo transcriptome assemblies are generated with as few as ~20 million reads for single-tissue samples and ~30 million reads for whole animals at the mRNA-coverage level.
-
Full-text index only
Optineurin coding variants in Ghanaian patients with primary open-angle glaucoma.
PMID 19096531 · PMC2605106 · Molecular vision · 2008 · 8 claims · 4 setups
OPTN coding variant allele frequencies do not differ significantly between POAG cases and controls in the Ghanaian population.
-
Full-text index only
Mutation and deletion analysis of GFR alpha-1, encoding the co-receptor for the GDNF/RET complex, in human brain tumours.
PMID 10408842 · PMC2362327 · British journal of cancer · 1999 · 8 claims · 3 setups
No mutations were found in the coding region of GDNF in any of the 36 brain tumours analysed
-
Full-text index only
Genome assembly comparison identifies structural variants in the human genome.
PMID 17115057 · PMC2674632 · Nature genetics · 2006 · 7 claims · 7 setups
Genome assembly comparison is a robust approach for identifying all classes of genetic variation, with no lower size limit.
-
Full-text index only
GeneAlign: a coding exon prediction tool based on phylogenetical comparisons.
PMID 16845010 · PMC1538901 · Nucleic acids research · 2006 · 8 claims · 5 setups
GeneAlign predicts coding exons by using signal detection (GeneSplicer/WMM) combined with CORAL, a heuristic linear-time alignment tool, to align candidate signal-flanked regions against annotated exons of a homologous organism's genes
-
Full-text index only
Genetic variation in an individual human exome.
PMID 18704161 · PMC2493042 · PLoS genetics · 2008 · 8 claims · 7 setups
The ~12,500 nonsilent coding variants in the HuRef exome can be reduced ~8-fold to a set of ~1,600 variants most likely to affect protein function.
-
Full-text index only
The evolution of human influenza A viruses from 1999 to 2006: a complete genome study.
PMID 18325125 · PMC2311284 · Virology journal · 2008 · 8 claims · 6 setups
H3N2 was the prevalent influenza A strain in Denmark from 1999 to 2006, except the 2000–2001 season when H1N1 dominated
-
Full-text index only
Worldwide distribution of NAT2 diversity: implications for NAT2 evolutionary history.
PMID 18304320 · PMC2292740 · BMC genetics · 2008 · 8 claims · 8 setups
NAT2 coding region sequence variation in the Mandenka and other sub-Saharan African populations is consistent with selective neutrality and constant population size.
-
Full-text index only
In silico discovery of gene-coding variants in murine quantitative trait loci using strain-specific genome sequence databases.
PMID 12537567 · PMC151180 · Genome biology · 2002 · 6 claims · 4 setups
Strain-specific mouse genome sequence databases can be used in a high-throughput in silico pipeline to discover gene-coding variants within murine QTLs, without de novo sequencing.
-
Full-text index only
SVC: structured visualization of evolutionary sequence conservation.
PMID 15991338 · PMC1160265 · Nucleic acids research · 2005 · 7 claims · 5 setups
SVC aligns protein-coding sequences of orthologous gene pairs and maps them back onto their encoding exons/introns to generate a scaffold of conserved gene structure.
-
Full-text index only
Ab initio identification of putative human transcription factor binding sites by comparative genomics.
PMID 15865625 · PMC1097714 · BMC bioinformatics · 2005 · 8 claims · 5 setups
An integrated algorithm combining human-mouse genomic comparison, motif overrepresentation, and coregulation filters (GO annotation and microarray coexpression) can identify candidate transcription factor binding sites genome-wide
-
Full-text index only
Exogean: a framework for annotating protein-coding genes in eukaryotic genomic DNA.
PMID 16925841 · PMC1810556 · Genome biology · 2006 · 8 claims · 5 setups
Exogean is a framework using directed acyclic coloured multigraphs (DACMs) to represent biological objects (mRNA, ESTs, protein alignments, exons) and iteratively combine them into complex protein-coding transcript models.
-
Full-text index only
Automatic annotation of eukaryotic genes, pseudogenes and promoters.
PMID 16925832 · PMC1810547 · Genome biology · 2006 · 8 claims · 6 setups
Fgenesh++ gene prediction pipeline identifies 91% of coding nucleotides with 90% specificity
-
Full-text index only
CONTRAST: a discriminative, phylogeny-free approach to multiple informant de novo gene prediction.
PMID 18096039 · PMC2246271 · Genome biology · 2007 · 8 claims · 5 setups
CONTRAST predicts exact coding region structures for 65% more human genes than the previous state-of-the-art de novo predictor (N-SCAN)
-
Full-text index only
The role of positive selection in determining the molecular cause of species differences in disease.
PMID 18837980 · PMC2576240 · BMC evolutionary biology · 2008 · 8 claims · 6 setups
Genes predicted to be under positive selection during human evolution are implicated in diseases (epithelial cancers, schizophrenia, autoimmune diseases, Alzheimer's disease) that differ in prevalence and symptomatology between humans and other mammals