Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Similarities and differences in genome-wide expression data of six organisms.
PMID 14737187 · PMC300882 · PLoS biology · 2004 · 8 claims · 8 setups
Coexpression of functionally related genes is frequently conserved across evolutionarily distant organisms
-
Full-text index only
A computational study of off-target effects of RNA interference.
PMID 15800213 · PMC1072799 · Nucleic acids research · 2005 · 8 claims · 5 setups
The chance of RNAi off-target effects is considerable, ranging from 5% to 80% depending on organism and parameters, when using exact sequence identity between siRNA and transcripts.
-
Full-text index only
BiSearch: primer-design and search tool for PCR on bisulfite-treated genomes.
PMID 15653630 · PMC546182 · Nucleic acids research · 2005 · 7 claims · 4 setups
BiSearch is a new web-available primer-design software for bisulfite-treated genomes that also analyzes primer pairs for mispriming sites via a novel search algorithm.
-
Full-text index only
SNPmasker: automatic masking of SNPs and repeats across eukaryotic genomes.
PMID 16845091 · PMC1538889 · Nucleic acids research · 2006 · 8 claims · 4 setups
SNPmasker is a web service combining SNP masking and repeat masking, supporting both coordinate-defined and homology-search-defined input regions, a combination not offered by prior tools
-
Full-text index only
Mitochondrial DNA mutations in renal cell carcinomas revealed no general impact on energy metabolism.
PMID 16404428 · PMC2361126 · British journal of cancer · 2006 · 6 claims · 5 setups
Somatic mtDNA mutations occur in renal cell carcinoma but are infrequent and frequently present at low (below 25%) heteroplasmy levels
-
Full-text index only
Accurate splice site prediction using support vector machines.
PMID 18269701 · PMC2230508 · BMC bioinformatics · 2007 · 8 claims · 5 setups
Weighted degree (WD) kernel SVMs outperform Markov Chains, GeneSplicer and SpliceMachine for genome-wide splice site recognition
-
Full-text index only
An emerging cyberinfrastructure for biodefense pathogen and pathogen-host data.
PMID 17984082 · PMC2239001 · Nucleic acids research · 2008 · 8 claims · 7 setups
The Biodefense Proteomics Resource Center (RC) is a public cyberinfrastructure that stores, integrates, and disseminates experimental data from seven Proteomics Research Centers (PRCs) on biodefense pathogens and host interactions
-
Full-text index only
InSite: a computational method for identifying protein-protein interaction binding sites on a proteome-wide scale.
PMID 17868464 · PMC2375030 · Genome biology · 2007 · 8 claims · 8 setups
InSite predicts protein-pair-specific binding motifs ('Motif M on protein A binds to protein B') by integrating heterogeneous PPI and motif-motif interaction evidence within a Bayesian network trained by EM
-
Full-text index only
Identification and characterization of insect-specific proteins by genome data analysis.
PMID 17407609 · PMC1852559 · BMC genomics · 2007 · 8 claims · 7 setups
Comparative genome analysis across five holometabolous insects and three non-insect eukaryotes (opisthokonts) identifies 154 insect-specific orthologous groups (refined to 51 proteins) and 466 eukaryote/opisthokont-core orthologous groups
-
Full-text index only
Assessing the gene space in draft genomes.
PMID 19042974 · PMC2615622 · Nucleic acids research · 2009 · 6 claims · 7 setups
The proportion of mapped CEGs in a draft genome assembly is a useful metric for describing gene space completeness, complementing N50 and x-fold coverage.
-
Full-text index only
A novel nonsense mutation in CRYBB1 associated with autosomal dominant congenital cataract.
PMID 18432316 · PMC2324115 · Molecular vision · 2008 · 7 claims · 5 setups
A novel heterozygous nonsense mutation (c.C737T, p.Q223X) in CRYBB1 is responsible for autosomal dominant congenital nuclear cataract in this family.
-
Full-text index only
The biological function of some human transcription factor binding motifs varies with position relative to the transcription start site.
PMID 18367472 · PMC2377430 · Nucleic acids research · 2008 · 8 claims · 5 setups
1226 eight-letter DNA words show statistically significant positional preferences relative to the TSS across 7914 human promoter regions
-
Full-text index only
TEDD: a comprehensive database for translation efficiency dynamics.
PMID 41217970 · PMC12807600 · Nucleic acids research · 2026 · 8 claims · 4 setups
TEDD integrates 1518 RNA-seq, Ribo-seq, and RNC-seq samples from 143 human projects (279 datasets) spanning 24 tissues/cell types, 74 cell lines, and 52 conditions.
-
Has reproduction · 89
Spatial information matters: are traditional imputation methods effective for spatial transcriptomics data?
PMID 41627342 · PMC12862982 · Briefings in bioinformatics · 2026 · 7 claims · 3 setups
No single existing SOTA imputation method consistently performs well across newer SRT platforms/datasets
-
Full-text index only
DDBJ in collaboration with mass-sequencing teams on annotation.
PMID 15608189 · PMC539974 · Nucleic acids research · 2005 · 7 claims · 5 setups
DDBJ collected and released 1,066,084 entries (718,072,425 bases) in the past year, including the complete chimpanzee chromosome 22 sequence and silkworm whole-genome shotgun data
-
Full-text index only
Ensembl 2006.
PMID 16381931 · PMC1347495 · Nucleic acids research · 2006 · 8 claims · 5 setups
Ensembl now provides annotation for 19 genomes, up from 4 the previous year, including new mammalian (Rhesus macaque, Opossum), chordate (Ciona intestinalis), and yeast genomes.
-
Has reproduction · 78
annotate_my_genomes: an easy-to-use pipeline to improve genome annotation and uncover neglected genes by hybrid RNA sequencing.
PMID 36472574 · PMC9724561 · GigaScience · 2022 · 7 claims · 8 setups
annotate_my_genomes is an easy-to-use genome-guided pipeline that uses hybrid (PacBio+Illumina) assembled transcripts to distinguish coding genes from long non-coding RNAs and reconcile them with prior annotations.
-
Has reproduction
Genome-wide signatures of convergent evolution in echolocating mammals.
PMID 24005325 · PMC3836225 · Nature · 2013 · 8 claims · 8 setups
Genome-wide convergent sequence evolution between echolocating lineages is not rare but widespread and continuously distributed, with signatures consistent with convergence in nearly 200 loci out of 2,326 examined.
-
Full-text index only
Chimp genome: branching out.
PMID 16136102 · PMC7420934 · Nature · 2005 · 8 claims · 8 setups
The Chimpanzee Sequencing and Analysis Consortium published the initial draft chimpanzee genome sequence and compared it to the human genome.
-
Full-text index only
Filtering high-throughput protein-protein interaction data using a combination of genomic features.
PMID 15833142 · PMC1127019 · BMC bioinformatics · 2005 · 8 claims · 8 setups
A combination of three genomic features (interacting Pfam domains, GO annotations, sequence homology) using naive Bayesian networks predicts true protein-protein interactions with high sensitivity and good specificity.