Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 87
Enhanced Generalizability of RNA Secondary Structure Prediction via Convolutional Block Attention Network and Ensemble Learning.
PMID 40871599 · PMC12388828 · Molecules (Basel, Switzerland) · 2025 · 8 claims · 8 setups
TrioFold integrates base-pairing clues from thermodynamic- and DL-based methods via ensemble learning and a convolutional block attention mechanism to enhance RSS prediction generalizability.
-
Has reproduction · 84
Strong population differentiation in lingcod (Ophiodon elongatus) is driven by a small portion of the genome.
PMID 33294007 · PMC7691466 · Evolutionary applications · 2020 · 7 claims · 8 setups
Lingcod comprise two distinct genetic clusters separated latitudinally at a break near Point Reyes off Northern California, with a high frequency of admixed individuals near the break.
-
Has reproduction · 44
Population differentiation and epidemic tracking of Bursaphelenchus xylophilus in China based on chromosome-level assembly and whole-genome sequencing data.
PMID 34839581 · PMC9300093 · Pest management science · 2022 · 6 claims · 8 setups
Generated the first chromosome-level genome assembly (AH1) of B. xylophilus using PacBio, Illumina, BioNano, and Hi-C data
-
Full-text index only
52-kD SS-A/Ro: genomic structure and identification of an alternatively spliced transcript encoding a novel leucine zipper-minus autoantigen expressed in fetal and adult heart.
PMID 7561701 · PMC2192297 · The Journal of experimental medicine · 1995 · 7 claims · 7 setups
The 52-kD SS-A/Ro gene spans 10 kb of DNA and is composed of seven exons, with the translation initiation codon in exon 2 and the leucine zipper encoded by exon 4.
-
Full-text index only
The human L-threonine 3-dehydrogenase gene is an expressed pseudogene.
PMID 12361482 · PMC131051 · BMC genetics · 2002 · 8 claims · 7 setups
The human TDH gene is located at chromosome 8p23-22, spans 10 kb, and has 8 exons that would be expected to encode a 369-residue ORF.
-
Full-text index only
Identifying related L1 retrotransposons by analyzing 3' transduced sequences.
PMID 12734010 · PMC156586 · Genome biology · 2003 · 8 claims · 6 setups
L1 elements with transduction-derived 3' sequence (L1-TDs) can be computationally identified using RepeatMasker/TSDfinder and grouped into families sharing a common progenitor via BLAST comparison of downstream sequences.
-
Full-text index only
A screen for proteins that interact with PAX6: C-terminal mutations disrupt interaction with HOMER3, DNCL1 and TRIM11.
PMID 16098226 · PMC1208879 · BMC genetics · 2005 · 8 claims · 7 setups
PAX6 interacts with three novel proteins: HOMER3, DNCL1 and TRIM11
-
Full-text index only
Computer identification of snoRNA genes using a Mammalian Orthologous Intron Database.
PMID 16093549 · PMC1184218 · Nucleic acids research · 2005 · 8 claims · 5 setups
Created the Mammalian Orthologous Intron Database (MOID) containing orthologous introns of human, mouse and rat identified via conserved reading-frame position
-
Full-text index only
Characterization of the linkage disequilibrium structure and identification of tagging-SNPs in five DNA repair genes.
PMID 16091150 · PMC1208870 · BMC cancer · 2005 · 7 claims · 5 setups
Three of the five DNA repair genes (MRE11A, RAD50, XRCC4) do not conform to a contiguous haplotype block structure; instead SNPs in high LD can be non-contiguous, fitting a more flexible LD group paradigm
-
Full-text index only
Columba: an integrated database of proteins, structures, and annotations.
PMID 15801979 · PMC1087474 · BMC bioinformatics · 2005 · 8 claims · 6 setups
COLUMBA physically integrates data from twelve protein structure-related databases (PDB, KEGG, Swiss-Prot, CATH, SCOP, Gene Ontology, ENZYME, etc.) into a single PostgreSQL data warehouse.
-
Full-text index only
Solving structures of protein complexes by molecular replacement with Phaser.
PMID 17164524 · PMC2483468 · Acta crystallographica. Section D, Biological crystallography · 2007 · 7 claims · 4 setups
Maximum-likelihood MR functions enable complex asymmetric units to be built up from individual components using a 'tree search with pruning' approach implemented in Phaser's automated MR mode.
-
Full-text index only
Genomic view of the evolution of the complement system.
PMID 16896831 · PMC2480602 · Immunogenetics · 2006 · 8 claims · 6 setups
Bony fish and higher vertebrates share practically the same set of complement genes, indicating most complement gene duplications occurred by the teleost/mammalian divergence (~500 MYA)
-
Full-text index only
Using ESTs to improve the accuracy of de novo gene prediction.
PMID 16817966 · PMC1534067 · BMC bioinformatics · 2006 · 8 claims · 8 setups
TWINSCAN_EST combines EST alignments with TWINSCAN via a trainable 'ESTseq' representation and improves exact gene structure prediction accuracy on the whole C. elegans genome
-
Full-text index only
Large-scale trends in the evolution of gene structures within 11 animal genomes.
PMID 16518452 · PMC1386723 · PLoS computational biology · 2006 · 8 claims · 5 setups
Change in intron–exon gene structure is gradual, clock-like, and largely independent of coding-sequence (protein) evolution
-
Full-text index only
miRNAMap: genomic maps of microRNA genes and their target genes in mammalian genomes.
PMID 16381831 · PMC1347497 · Nucleic acids research · 2006 · 6 claims · 6 setups
miRNAMap integrates known miRNA genes from miRBase, literature-curated validated targets, and computationally predicted miRNA genes and targets for human, mouse, rat and dog.
-
Full-text index only
GLIDA: GPCR--ligand database for chemical genomics drug discovery--database and tools update.
PMID 17986454 · PMC2238933 · Nucleic acids research · 2008 · 7 claims · 5 setups
GLIDA is a public relational database integrating biological information on GPCRs with chemical information on their ligands and their binding interactions.
-
Full-text index only
The scientific impact of the Structural Genomics Consortium: a protein family and ligand-centered approach to medically-relevant human proteins.
PMID 17932789 · PMC2140095 · Journal of structural and functional genomics · 2007 · 8 claims · 6 setups
A family-based target selection approach (rather than genome-wide or fold-novelty based selection) maximizes cross-member methodological transfer and biological insight
-
Full-text index only
X:Map: annotation and visualization of genome structure for Affymetrix exon array analysis.
PMID 17932061 · PMC2238884 · Nucleic acids research · 2008 · 7 claims · 4 setups
X:Map is a genome annotation database that maps every Affymetrix exon array probeset to Ensembl genome features (genes, ESTs, GenScan predictions) and supports both high-throughput and gene-centric analysis.
-
Full-text index only
Structural genomics and drug discovery.
PMID 17488474 · PMC3822824 · Journal of cellular and molecular medicine · 2007 · 8 claims · 8 setups
Membrane proteins represent ~70% of current drug targets but only just over 100 high-resolution structures exist for them, versus >30,000 total structures in public databases dominated by soluble proteins.
-
Full-text index only
Prediction by graph theoretic measures of structural effects in proteins arising from non-synonymous single nucleotide polymorphisms.
PMID 18654622 · PMC2447880 · PLoS computational biology · 2008 · 8 claims · 5 setups
Bongo identifies mutations causing local and global structural effects with a remarkably low false positive rate