Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Human and mouse introns are linked to the same processes and functions through each genome's most frequent non-conserved motifs.
PMID 18450818 · PMC2425492 · Nucleic acids research · 2008 · 8 claims · 5 setups
Pyknons (recurrent, genome-specific, ≥16nt motifs with ≥30 intact intergenic/intronic copies and ≥1 exonic copy) span a substantial fraction of previously uncharacterized intronic space (7.4% human, 4.4% mouse)
-
Full-text index only
A new method for 2D gel spot alignment: application to the analysis of large sample sets in clinical proteomics.
PMID 18957120 · PMC2628390 · BMC bioinformatics · 2008 · 8 claims · 2 setups
Sili2DGel represents recursive gel matching results as a weighted undirected graph and identifies SAP by finding cliques and pseudocliques (dense subgraphs) after edge-weight filtering, strength-metric-based graph reduction, and cluster refinement.
-
Full-text index only
BIPASS: BioInformatics Pipeline Alternative Splicing Services.
PMID 17584795 · PMC1933140 · Nucleic acids research · 2007 · 8 claims · 4 setups
BIPASS offers two complementary services for alternative splicing (AS) research: BIPAS-SpliceDB, a queryable pre-computed AS data warehouse, and BIPAS-Align&Splice, an online pipeline for user-submitted sequences.
-
Full-text index only
Towards alignment independent quantitative assessment of homology detection.
PMID 17205117 · PMC1762415 · PloS one · 2006 · 8 claims · 6 setups
The Fhom Estimator uses the prevalence of a conserved protein feature (X) in two protein sets to estimate the fraction of true homologs among paired proteins, independent of alignment quality.
-
Full-text index only
Sequence context affects the rate of short insertions and deletions in flies and primates.
PMID 18291026 · PMC2374710 · Genome biology · 2008 · 8 claims · 6 setups
The rate of insertion or deletion of specific lengths can vary by more than 100-fold depending on the surrounding sequence context
-
Full-text index only
Versatile and open software for comparing large genomes.
PMID 14759262 · PMC395750 · Genome biology · 2004 · 8 claims · 8 setups
MUMmer 3.0 efficiently handles comparisons of large eukaryotic genomes at varying evolutionary distances
-
Full-text index only
Association of poly-purine/poly-pyrimidine sequences with meiotic recombination hot spots.
PMID 16846522 · PMC1543642 · BMC genomics · 2006 · 7 claims · 6 setups
PPT frequency is significantly elevated in yeast meiotic recombination hot spots compared with cold spots
-
Full-text index only
Improving the specificity of exon prediction using comparative genomics.
PMID 18831778 · PMC2559877 · BMC genomics · 2008 · 8 claims · 6 setups
A log-odds ratio scoring method based on codon conservation across human-mouse/human-dog alignments and adjacent-codon dependency can classify putative exons as coding vs non-coding.
-
Full-text index only
Linking disease-associated genes to regulatory networks via promoter organization.
PMID 15701758 · PMC549397 · Nucleic acids research · 2005 · 8 claims · 7 setups
Pairs of TFBSs conserved both vertically (orthologous genes) and horizontally (co-regulated genes) can serve as seeds to build promoter models representing potential co-regulation networks
-
Has reproduction · 83
Accurate prediction of metagenome-assembled genome completeness by MAGISTA, a random forest model built on alignment-free intra-bin statistics.
PMID 35248155 · PMC8898458 · Environmental microbiome · 2022 · 7 claims · 7 setups
MAGISTA, a random forest model built on alignment-free intra-bin distance-distribution statistics, can estimate MAG completeness and purity without relying on reference marker genes.
-
Full-text index only
Database resources of the National Center for Biotechnology Information.
PMID 17170002 · PMC1781113 · Nucleic acids research · 2007 · 8 claims · 8 setups
NCBI maintains an integrated suite of database resources (Entrez, PubMed, RefSeq, dbSNP, BLAST, etc.) for molecular biology data retrieval and analysis
-
Full-text index only
Endonuclease-independent insertion provides an alternative pathway for L1 retrotransposition in the human genome.
PMID 17517773 · PMC1920257 · Nucleic acids research · 2007 · 8 claims · 5 setups
An endonuclease-independent pathway (NCLI) for L1 insertion has been active in recent human genome evolution
-
Full-text index only
Ensembl 2006.
PMID 16381931 · PMC1347495 · Nucleic acids research · 2006 · 8 claims · 5 setups
Ensembl now provides annotation for 19 genomes, up from 4 the previous year, including new mammalian (Rhesus macaque, Opossum), chordate (Ciona intestinalis), and yeast genomes.
-
Has reproduction · 95
Determining virus-host interactions and glycerol metabolism profiles in geographically diverse solar salterns with metagenomics.
PMID 28097058 · PMC5228507 · PeerJ · 2017 · 8 claims · 8 setups
Similar virus-host interactions and glycerol metabolism gene associations (notably dihydroxyacetone kinase with Haloquadratum/Halorubrum) exist across geographically diverse solar salterns
-
Full-text index only
TRED: a Transcriptional Regulatory Element Database and a platform for in silico gene regulation studies.
PMID 15608156 · PMC539958 · Nucleic acids research · 2005 · 8 claims · 5 setups
TRED is a database collecting both cis-regulatory elements (promoters) and trans-regulatory elements (transcription factor binding/regulation data) with linked access.
-
Full-text index only
Analysis of the prostate cancer cell line LNCaP transcriptome using a sequencing-by-synthesis approach.
PMID 17010196 · PMC1592491 · BMC genomics · 2006 · 8 claims · 7 setups
High-throughput 454 sequencing-by-synthesis of LNCaP cDNA can profile transcript abundance across the transcriptome
-
Full-text index only
TFBScluster web server for the identification of mammalian composite regulatory elements.
PMID 16845063 · PMC1538905 · Nucleic acids research · 2006 · 7 claims · 5 setups
TFBScluster is a web server that identifies genome-wide clusters of TFBSs conserved in multiple mammalian species using human or mouse as the reference genome.
-
Full-text index only
Sequence occurrence and structural uniqueness of a G-quadruplex in the human c-kit promoter.
PMID 17720713 · PMC2034477 · Nucleic acids research · 2007 · 8 claims · 4 setups
The native 22-nt c-kit87 sequence occurs only once in the entire human genome.
-
Full-text index only
CompMoby: comparative MobyDick for detection of cis-regulatory motifs.
PMID 18950538 · PMC2605473 · BMC bioinformatics · 2008 · 7 claims · 4 setups
CompMoby identifies cis-regulatory binding sites at both transcriptional and post-transcriptional levels in metazoans without prior knowledge of the trans-acting factor
-
Full-text index only
CorGen--measuring and generating long-range correlations for DNA sequence analysis.
PMID 16845099 · PMC1538783 · Nucleic acids research · 2006 · 8 claims · 3 setups
CorGen is a web server that measures long-range correlations in DNA sequences and generates random sequences with the same (or user-specified) correlation and composition parameters