Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Manual annotation and analysis of the defensin gene cluster in the C57BL/6J mouse reference genome.
PMID 20003482 · PMC2807441 · BMC genomics · 2009 · 8 claims · 6 setups
Manual annotation of the mouse Chromosome 8 defensin region identifies 98 gene loci: 54 in the alpha-defensin cluster and 44 in the beta-defensin cluster
-
Full-text index only
Gene- and evidence-based candidate gene selection for schizophrenia and gene feature analysis.
PMID 19944577 · PMC2826526 · Artificial intelligence in medicine · 2010 · 8 claims · 5 setups
The SCOR method outperforms the CCOR method for prioritizing schizophrenia candidate genes
-
Full-text index only
Comparative analysis of plant genomes allows the definition of the "Phytolongins": a novel non-SNARE longin domain protein family.
PMID 19889231 · PMC2779197 · BMC genomics · 2009 · 8 claims · 6 setups
A novel, plant-specific family of longin-related proteins, the 'Phytolongins', was identified in land plant genomes.
-
Full-text index only
Helminth genomics: The implications for human health.
PMID 19855829 · PMC2757907 · PLoS neglected tropical diseases · 2009 · 8 claims · 7 setups
More than two billion people (one-third of humanity) are infected with helminth parasites, causing major morbidity, mortality, and poverty maintenance
-
Full-text index only
Network-assisted protein identification and data interpretation in shotgun proteomics.
PMID 19690572 · PMC2736651 · Molecular systems biology · 2009 · 7 claims · 7 setups
Confidently identified proteins in a sample form tightly connected sub-networks in the protein interaction network, with significantly higher clustering coefficients than random or topology-matched random sub-networks.
-
Full-text index only
Bases and spaces: resources on the web for accessing the draft human genome.
PMID 11178254 · PMC138875 · Genome biology · 2000 · 8 claims · 8 setups
By combining currently available genomic databases and mapping resources (GenBank/Entrez, UniGene, RH maps, BAC fingerprint maps, Ensembl, NIX), it is possible to devise strategies that fully exploit the fragmentary draft human genome sequence.
-
Full-text index only
Twin peaks: the draft human genome sequence.
PMID 11276423 · PMC138909 · Genome biology · 2001 · 8 claims · 8 setups
The predicted number of human genes (~26,000-40,000) is far lower than the widely assumed ~100,000, though downstream RNA/protein complexity can still generate substantial biological complexity.
-
Full-text index only
With the finished human genome in hand, what next?
PMID 12844356 · PMC193627 · Genome biology · 2003 · 8 claims · 8 setups
Gene Ontology (GO) provides a syntax/query framework for functional classification of genes, expanding beyond E. coli origins into anatomy, pathology, and phenotype data.
-
Full-text index only
T1DBase, a community web-based resource for type 1 diabetes research.
PMID 15608258 · PMC540049 · Nucleic acids research · 2005 · 8 claims · 6 setups
T1DBase is an integrated, open-access web resource that unifies genetic, genomic, and biological data to support type 1 diabetes (T1D) research
-
Full-text index only
Database resources of the National Center for Biotechnology Information.
PMID 17170002 · PMC1781113 · Nucleic acids research · 2007 · 8 claims · 8 setups
NCBI maintains an integrated suite of database resources (Entrez, PubMed, RefSeq, dbSNP, BLAST, etc.) for molecular biology data retrieval and analysis
-
Full-text index only
DNA sequencing: bench to bedside and beyond.
PMID 17855400 · PMC2094077 · Nucleic acids research · 2007 · 8 claims · 7 setups
DNA sequencing methods derived from Sanger's 1977 dideoxy method have dominated sequencing for 30 years despite being only incrementally refined.
-
Full-text index only
Phenotypic variation meets systems biology.
PMID 19664197 · PMC2745761 · Genome biology · 2009 · 8 claims · 8 setups
Cellular differentiation states are constrained by complex networks with substantial positive and negative regulation, challenging the concept of single 'master regulators'
-
Full-text index only
Ensembl 2005.
PMID 15608235 · PMC540092 · Nucleic acids research · 2005 · 8 claims · 4 setups
Ensembl's automatic gene build system can flexibly and reliably annotate a wide variety of genomes with limited species-specific evidence.
-
Full-text index only
A cell biological perspective on genome research.
PMID 8522596 · PMC2120688 · The Journal of cell biology · 1995 · 7 claims · 7 setups
Genome sequencing represents a sixth stage in the historical progression of structural biology (comparative anatomy through crystallography), and will be similarly valuable once related to function.
-
Has reproduction · 53
Combining evidence of preferential gene-tissue relationships from multiple sources.
PMID 23950964 · PMC3741196 · PloS one · 2013 · 8 claims · 8 setups
A high-level integration approach combining three methods across four human microarray datasets, merged by consensus voting and a rule-based inner/total score, predicts preferentially expressed genes while reducing method- and study-specific bias.
-
Has reproduction · 82
Ordinal-level phylogenomics of the arthropod class Diplopoda (millipedes) based on an analysis of 221 nuclear protein-coding loci generated using next-generation sequence analyses.
PMID 24236165 · PMC3827447 · PloS one · 2013 · 8 claims · 8 setups
An ordinal-level phylogeny of Diplopoda reconstructed from 221 nuclear protein-coding loci (61,641 aligned amino acid columns) differs from existing classifications in fundamental ways.
-
Full-text index only
DG-CST (Disease Gene Conserved Sequence Tags), a database of human-mouse conserved elements associated to disease genes.
PMID 15608249 · PMC539965 · Nucleic acids research · 2005 · 5 claims · 8 setups
Comparative human-mouse genome analysis identifies conserved sequence tags (CSTs, >=70% identity over >=100bp) that frequently correspond to non-coding elements with putative regulatory or structural roles
-
Full-text index only
GENCODE: producing a reference annotation for ENCODE.
PMID 16925838 · PMC1810553 · Genome biology · 2006 · 8 claims · 8 setups
GENCODE annotation combines initial manual annotation by HAVANA, experimental validation, and refinement based on results to identify protein-coding genes in ENCODE regions
-
Full-text index only
EGASP: the human ENCODE Genome Annotation Assessment Project.
PMID 16925836 · PMC1810551 · Genome biology · 2006 · 8 claims · 6 setups
Best-performing computational gene prediction methods correctly predict at least one transcript for close to 70% of annotated genes in the ENCODE regions.
-
Full-text index only
Detecting transcriptionally active regions using genomic tiling arrays.
PMID 16859498 · PMC1779562 · Genome biology · 2006 · 8 claims · 4 setups
A non-parametric method (TranscriptionDetector) integrates single-channel p-values from multiple replicate arrays into a multi-channel p-value (MCPV) to identify transcribed probed loci without assumptions about intensity distributions.