Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Recent additions and improvements to the Onto-Tools.
PMID 15980579 · PMC1160233 · Nucleic acids research · 2005 · 7 claims · 3 setups
The Onto-Tools back-end database was redesigned around the Entrez Gene data model after NCBI phased out LocusLink in February 2005.
-
Full-text index only
The biological function of some human transcription factor binding motifs varies with position relative to the transcription start site.
PMID 18367472 · PMC2377430 · Nucleic acids research · 2008 · 8 claims · 5 setups
1226 eight-letter DNA words show statistically significant positional preferences relative to the TSS across 7914 human promoter regions
-
Has reproduction · 51
SGCP: a spectral self-learning method for clustering genes in co-expression networks.
PMID 38956463 · PMC11221046 · BMC bioinformatics · 2024 · 7 claims · 4 setups
SGCP, a spectral self-learning method, yields gene co-expression modules with higher GO enrichment than WGCNA, CoExpNets, and CEMiTool across 12 real gene expression datasets.
-
Full-text index only
Meeting highlights: beyond the genome 2000: the 18th International Congress of Biochemistry and Molecular Biology.
PMID 11119309 · PMC2448388 · Yeast (Chichester, England) · 2000 · 8 claims · 8 setups
Celera sequenced a human genome to ~45-fold coverage from one donor and used high-quality sequence stretches to define ~6 million SNPs
-
Full-text index only
The truth about mouse, human, worms and yeast.
PMID 15601543 · PMC3525071 · Human genomics · 2004 · 8 claims · 8 setups
Comparing genomes in pairs or larger sets (mouse-human, C. elegans-C. briggsae, multiple Saccharomyces, human-pufferfish, etc.) reveals unsuspected genes and helps eliminate false-positive gene predictions
-
Full-text index only
Continued colonization of the human genome by mitochondrial DNA.
PMID 15361937 · PMC515365 · PLoS biology · 2004 · 7 claims · 6 setups
NUMT insertion into nuclear chromosomes is an ongoing process shaped by double-strand-break repair (as shown in yeast) and continuing in humans.
-
Full-text index only
Assessing the gene space in draft genomes.
PMID 19042974 · PMC2615622 · Nucleic acids research · 2009 · 6 claims · 7 setups
The proportion of mapped CEGs in a draft genome assembly is a useful metric for describing gene space completeness, complementing N50 and x-fold coverage.
-
Has reproduction · 85
An extensive evaluation of read trimming effects on Illumina NGS data analysis.
PMID 24376861 · PMC3871669 · PloS one · 2013 · 8 claims · 8 setups
Read trimming increases the quality and reliability of downstream NGS analyses (RNA-Seq mapping, SNP identification, genome assembly) while reducing execution time and computational resources.
-
Full-text index only
Ensembl 2006.
PMID 16381931 · PMC1347495 · Nucleic acids research · 2006 · 8 claims · 5 setups
Ensembl now provides annotation for 19 genomes, up from 4 the previous year, including new mammalian (Rhesus macaque, Opossum), chordate (Ciona intestinalis), and yeast genomes.
-
Full-text index only
CROPPER: a metagene creator resource for cross-platform and cross-species compendium studies.
PMID 16995941 · PMC1592126 · BMC bioinformatics · 2006 · 7 claims · 5 setups
CROPPER is a web-based software resource that combines genomic data from heterogeneous sources using identifier and orthologous gene information from the Ensembl database.
-
Full-text index only
Gene losses during human origins.
PMID 16464126 · PMC1361800 · PLoS biology · 2006 · 7 claims · 7 setups
A comparative genomic screen identified 67 new human-specific nonprocessed pseudogenes, bringing the total (with 13 from prior literature) to 80 human-specific pseudogenes.
-
Full-text index only
BiSearch: primer-design and search tool for PCR on bisulfite-treated genomes.
PMID 15653630 · PMC546182 · Nucleic acids research · 2005 · 7 claims · 4 setups
BiSearch is a new web-available primer-design software for bisulfite-treated genomes that also analyzes primer pairs for mispriming sites via a novel search algorithm.
-
Full-text index only
N-acetyltransferase 8, a positional candidate for blood pressure and renal regulation: resequencing, association and in silico study.
PMID 18402670 · PMC2330028 · BMC medical genetics · 2008 · 7 claims · 6 setups
NAT8 is a novel positional candidate gene for blood pressure and renal function based on its chromosomal location within a BP linkage region and its expression in embryonic/adult kidney and liver
-
Has reproduction · 86
Screening of core genes prognostic for sepsis and construction of a ceRNA regulatory network.
PMID 36855106 · PMC9976425 · BMC medical genomics · 2023 · 8 claims · 7 setups
RNA-seq of peripheral blood from 23 sepsis patients and 10 healthy controls identifies 1,044 DEmRNAs, 66 DEmiRNAs and 155 DElncRNAs.
-
Full-text index only
A comparison of programmed cell death between species.
PMID 11178240 · PMC138857 · Genome biology · 2000 · 8 claims · 8 setups
The core apoptotic pathway (CED-3/caspases, CED-4/Apaf-1, CED-9/Bcl-2, EGL-1) is conserved across C. elegans, Drosophila, and mammals.
-
Full-text index only
Filtering high-throughput protein-protein interaction data using a combination of genomic features.
PMID 15833142 · PMC1127019 · BMC bioinformatics · 2005 · 8 claims · 8 setups
A combination of three genomic features (interacting Pfam domains, GO annotations, sequence homology) using naive Bayesian networks predicts true protein-protein interactions with high sensitivity and good specificity.
-
Full-text index only
A computational study of off-target effects of RNA interference.
PMID 15800213 · PMC1072799 · Nucleic acids research · 2005 · 8 claims · 5 setups
The chance of RNAi off-target effects is considerable, ranging from 5% to 80% depending on organism and parameters, when using exact sequence identity between siRNA and transcripts.
-
Full-text index only
DDBJ in collaboration with mass-sequencing teams on annotation.
PMID 15608189 · PMC539974 · Nucleic acids research · 2005 · 7 claims · 5 setups
DDBJ collected and released 1,066,084 entries (718,072,425 bases) in the past year, including the complete chimpanzee chromosome 22 sequence and silkworm whole-genome shotgun data
-
Has reproduction · 92
Systematic review of human post-mortem immunohistochemical studies and bioinformatics analyses unveil the complexity of astrocyte reaction in Alzheimer's disease.
PMID 34297416 · PMC8766893 · Neuropathology and applied neurobiology · 2022 · 8 claims · 5 setups
Systematic review of 306 eligible articles identified 196 distinct proteins constituting the ADRA (AD reactive astrocyte) protein set
-
Has reproduction · 78
annotate_my_genomes: an easy-to-use pipeline to improve genome annotation and uncover neglected genes by hybrid RNA sequencing.
PMID 36472574 · PMC9724561 · GigaScience · 2022 · 7 claims · 8 setups
annotate_my_genomes is an easy-to-use genome-guided pipeline that uses hybrid (PacBio+Illumina) assembled transcripts to distinguish coding genes from long non-coding RNAs and reconcile them with prior annotations.