Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Classification of real and pseudo microRNA precursors using local structure-sequence features and support vector machine.
PMID 16381612 · PMC1360673 · BMC bioinformatics · 2005 · 7 claims · 7 setups
A 32-dimensional triplet structure-sequence feature vector combined with SVM (triplet-SVM) can distinguish real human pre-miRNAs from pseudo pre-miRNA hairpins with ~90% accuracy.
-
Full-text index only
Design factors that influence PCR amplification success of cross-species primers among 1147 mammalian primer pairs.
PMID 17029642 · PMC1635982 · BMC genomics · 2006 · 8 claims · 7 setups
The number of index-species (IS) mismatches in a primer pair significantly reduces amplification success, with an estimated 6-8% decrease in success rate per additional mismatch.
-
Full-text index only
Identification and evolutionary analysis of novel exons and alternative splicing events using cross-species EST-to-genome comparisons in human, mouse and rat.
PMID 16536879 · PMC1479377 · BMC bioinformatics · 2006 · 8 claims · 6 setups
ENACE, a cross-species EST-to-genome comparison algorithm, can identify novel cassette-on exons and retained introns for EST-scanty species and distinguish conserved vs lineage-specific exons
-
Full-text index only
Protein function assignment through mining cross-species protein-protein interactions.
PMID 18253506 · PMC2216687 · PloS one · 2008 · 8 claims · 6 setups
CSIDOP predicts protein molecular function with 95.42% accuracy using 2,972 GO functional categories in H. sapiens
-
Full-text index only
Identifying cis-regulatory sequences by word profile similarity.
PMID 19730735 · PMC2731932 · PloS one · 2009 · 8 claims · 8 setups
WPH-finder identifies putative co-regulated CRMs by scanning the genome for sequences with word profiles similar to a known CRM, without explicitly defining binding sites
-
Full-text index only
AUGUSTUS at EGASP: using EST, protein and genomic alignments for improved gene prediction in the human genome.
PMID 16925833 · PMC1810548 · Genome biology · 2006 · 8 claims · 5 setups
AUGUSTUS predicted significantly more genes correctly than any other ab initio program in EGASP
-
Full-text index only
Ab initio identification of human microRNAs based on structure motifs.
PMID 18088431 · PMC2238772 · BMC bioinformatics · 2007 · 8 claims · 7 setups
MiRPred predicts miRNA precursors ab initio using only predicted secondary structure motifs, ignoring nucleotide sequence
-
Has reproduction · 42
CanCellCap: robust cancer cell capture across tissue types on single-cell RNA-seq data by multi-domain learning.
PMID 40739511 · PMC12312500 · BMC biology · 2025 · 8 claims · 8 setups
CanCellCap, a multi-domain learning framework integrating domain adversarial learning and Mixture of Experts, identifies cancer cells across all tissues, cancers, and sequencing platforms by extracting tissue-common and tissue-specific gene expression patterns.
-
Full-text index only
Resequencing PNMT in European hypertensive and normotensive individuals: no common susceptibilily variants for hypertension and purifying selection on intron 1.
PMID 17645789 · PMC1947951 · BMC medical genetics · 2007 · 7 claims · 7 setups
Resequencing of PNMT found no common susceptibility variants that distinguish hypertensive from normotensive individuals
-
Full-text index only
Conserved elements with potential to form polymorphic G-quadruplex structures in the first intron of human genes.
PMID 18187510 · PMC2275096 · Nucleic acids research · 2008 · 8 claims · 6 setups
G-richness downstream of the TSS is strand-biased, concentrated on the nontemplate strand, with a peak at +200 to +300 bp
-
Full-text index only
Meta-analysis of inter-species liver co-expression networks elucidates traits associated with common human diseases.
PMID 20019805 · PMC2787626 · PLoS computational biology · 2009 · 8 claims · 8 setups
A novel semi-parametric meta-analysis method (based on a gene-centric Glass's d effect size) outperforms existing parametric and non-parametric meta-analysis methods at identifying functionally coherent gene pairs across species.
-
Full-text index only
CLC-2 single nucleotide polymorphisms (SNPs) as potential modifiers of cystic fibrosis disease severity.
PMID 15507145 · PMC526769 · BMC medical genetics · 2004 · 8 claims · 7 setups
PCR amplification and sequencing of CLC-2 revealed 1 SNP in the promoter, 4 SNPs in intron 1, and none in exon 20
-
Full-text index only
DNA sequence and analysis of human chromosome 9.
PMID 15164053 · PMC2734081 · Nature · 2004 · 8 claims · 8 setups
The finished euchromatic sequence of chromosome 9 comprises 109,044,351 base pairs, representing >99.6% of the region.
-
Full-text index only
Identification of polymorphisms and balancing selection in the male infertility candidate gene, ornithine decarboxylase antizyme 3.
PMID 16542438 · PMC1526716 · BMC medical genetics · 2006 · 8 claims · 6 setups
Mutations in the OAZ3 gene are not a common cause of male infertility
-
Full-text index only
A space-efficient and accurate method for mapping and aligning cDNA sequences onto genomic sequence.
PMID 18344523 · PMC2377433 · Nucleic acids research · 2008 · 7 claims · 6 setups
Spaln maps and aligns large cDNA sequence sets onto whole mammalian genomes using substantially less memory than comparable existing tools