Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
pmid-42131214
PMID 42131214 · PMC13161053 · 8 claims · 7 setups
Cyanobacteria dominate the microbial communities of Hawaiian steam vent habitats, with community composition varying by microhabitat
-
Full-text index only
'Chumanzee' evolution: the urge to diverge and merge.
PMID 17129363 · PMC1794591 · Genome biology · 2006 · 8 claims · 3 setups
Human-chimpanzee divergence was not a simple clean split; evidence suggests hybridization continued after an initial split.
-
Full-text index only
Comparative genomics of vertebrate Fox cluster loci.
PMID 17062144 · PMC1634998 · BMC genomics · 2006 · 8 claims · 3 setups
Two additional human paralogous Fox cluster regions exist, on chromosomes 14 and 20, beyond the previously known chromosome 6 and 16 loci
-
Full-text index only
DBD--taxonomically broad transcription factor predictions: new content and functionality.
PMID 18073188 · PMC2238844 · Nucleic acids research · 2008 · 8 claims · 3 setups
DBD is a database of predicted sequence-specific DNA-binding transcription factors covering over 700 publicly available proteomes, up from 150 in the initial version.
-
Full-text index only
QuadBase: genome-wide database of G4 DNA--occurrence and conservation in human, chimpanzee, mouse and rat promoters and 146 microbes.
PMID 17962308 · PMC2238983 · Nucleic acids research · 2008 · 8 claims · 3 setups
QuadBase is a compendium of G4 DNA (quadruplex) motifs focused on their occurrence and conservation in promoters, composed of EuQuad and ProQuad
-
Full-text index only
Protein co-evolution, co-adaptation and interactions.
PMID 18818697 · PMC2556093 · The EMBO journal · 2008 · 8 claims · 6 setups
The mirrortree method predicts protein-protein interactions by detecting pairs of protein families with similar phylogenetic trees (quantified as Pearson correlation of sequence similarity matrices).
-
Full-text index only
Using multiple alignments to improve seeded local alignment algorithms.
PMID 16100379 · PMC1185574 · Nucleic acids research · 2005 · 8 claims · 2 setups
Using information implicit in a multiple alignment to dynamically build a spaced-seed index weighted toward promising regions increases sensitivity of local alignment search compared to indexing a sequence alone
-
Full-text index only
HaploSNPer: a web-based allele and SNP detection tool.
PMID 18307806 · PMC2288614 · BMC genetics · 2008 · 6 claims · 2 setups
HaploSNPer is a web-based tool integrating BLASTN, CAP3/PHRAP, and QualitySNP into a single pipeline for allele and SNP detection from diploid and polyploid species
-
Has reproduction · 88
nf-core/isoseq: simple gene and isoform annotation with PacBio Iso-Seq long-read sequencing.
PMID 36961337 · PMC10199315 · Bioinformatics (Oxford, England) · 2023 · 7 claims · 4 setups
nf-core/isoseq is a new automated Nextflow-based pipeline that processes raw Iso-Seq subreads through to genome annotation (BED format) without requiring transcriptome assembly.
-
Has reproduction · 81
Chromosome-scale Elaeis guineensis and E. oleifera assemblies: comparative genomics of oil palm and other Arecaceae.
PMID 38918881 · PMC11373658 · G3 (Bethesda, Md.) · 2024 · 8 claims · 8 setups
Improved E. guineensis genome assembly achieved with substantially increased continuity and completeness compared to prior assemblies
-
Full-text index only
The integrated world of functional genomics.
PMID 12537543 · PMC151279 · Genome biology · 2003 · 8 claims · 8 setups
Integrating chromatin immunoprecipitation (promoter-binding) data with expression data reveals the yeast cell-cycle transcriptional regulatory network, including network motifs such as autoregulation, multi-component loops, and feedforward loops.
-
Full-text index only
Genomics, proteomics and bioinformatics: all in the same boat.
PMID 12374575 · PMC244909 · Genome biology · 2002 · 8 claims · 8 setups
Microarray sensitivity has improved enough to analyze clinical samples as small as 10-50 nanograms, and many early microarray experiments may need to be repeated due to data quality issues
-
Full-text index only
Of rats and men.
PMID 15003114 · PMC395761 · Genome biology · 2004 · 8 claims · 10 setups
The rat genome has been sequenced to draft level, with over 90% of the genome sampled using more than 36 million sequence reads (assembly version 3.1)
-
Full-text index only
From single cells to whole organisms.
PMID 16420683 · PMC1414103 · Genome biology · 2005 · 8 claims · 8 setups
The genetic-interaction map in S. cerevisiae is roughly four times as complex as the protein-protein interaction map, and genetic interactions do not overlap with physical interactions but instead predict functional neighborhoods
-
Full-text index only
The European Bioinformatics Institute's data resources: towards systems biology.
PMID 15608238 · PMC539980 · Nucleic acids research · 2005 · 8 claims · 5 setups
Since 2003 the EBI has launched new databases covering protein-protein interactions (IntAct), pathways (Reactome) and small molecules (ChEBI)
-
Full-text index only
Variation resources at UC Santa Cruz.
PMID 17151077 · PMC1781230 · Nucleic acids research · 2007 · 8 claims · 8 setups
The UCSC Genome Browser variation resources integrate polymorphism data from public collections (dbSNP, HapMap, Affymetrix, Perlegen, SeattleSNPs) into a common format with additional annotations and genomic context.
-
Full-text index only
The PeptideAtlas project.
PMID 16381952 · PMC1347403 · Nucleic acids research · 2006 · 8 claims · 5 setups
PeptideAtlas provides an automated repository that identifies peptides by MS/MS, statistically validates identifications, and maps them to eukaryotic genomes to enable data exchange and integration with genomic data.
-
Full-text index only
The UCSC Genome Browser Database: update 2006.
PMID 16381938 · PMC1347506 · Nucleic acids research · 2006 · 8 claims · 8 setups
The UCSC Genome Browser Database (GBD) provides integrated sequence and annotation data, with web tools (Genome Browser, Table Browser, Proteome Browser, Gene Sorter, BLAT, In Silico PCR) for visualizing and querying genomes of about a dozen vertebrate species and several model organisms.
-
Full-text index only
Ensembl 2008.
PMID 18000006 · PMC2238821 · Nucleic acids research · 2008 · 8 claims · 6 setups
The Ensembl regulatory build integrates multiple genome-wide functional genomics datasets to automatically annotate regulatory regions and assign putative functions across the genome.
-
Full-text index only
BRENDA, AMENDA and FRENDA: the enzyme information system in 2007.
PMID 17202167 · PMC1899097 · Nucleic acids research · 2007 · 7 claims · 6 setups
BRENDA is the largest publicly available enzyme information system worldwide, manually curated from primary literature and covering all identified enzymes regardless of source.