Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Modeling chromosomes in mouse to explore the function of genes, genomic disorders, and chromosomal organization.
PMID 16839184 · PMC1500809 · PLoS genetics · 2006 · 8 claims · 8 setups
Cre/loxP recombination in ES cells can generate megabase-scale deletions, duplications, and inversions depending on loxP orientation, cis/trans configuration, and cell cycle stage
-
Full-text index only
Female monozygotic twins discordant for hemophilia A due to nonrandom X-chromosome inactivation.
PMID 18645989 · PMC5715470 · American journal of hematology · 2008 · 7 claims · 8 setups
Monozygotic twin A (severe hemophilia A, FVIII:C <1%) shows complete nonrandom X-inactivation skewed toward the paternal (normal factor VIII) X-chromosome
-
Full-text index only
Gene prediction in eukaryotes with a generalized hidden Markov model that uses hints from external sources.
PMID 16469098 · PMC1409804 · BMC bioinformatics · 2006 · 7 claims · 3 setups
AUGUSTUS+ extends the AUGUSTUS GHMM by combining intrinsic sequence information with extrinsic hints via an extended emission alphabet, so the GHMM jointly models the DNA sequence, gene structure, and hint collection.
-
Has reproduction · 44
Population differentiation and epidemic tracking of Bursaphelenchus xylophilus in China based on chromosome-level assembly and whole-genome sequencing data.
PMID 34839581 · PMC9300093 · Pest management science · 2022 · 6 claims · 8 setups
Generated the first chromosome-level genome assembly (AH1) of B. xylophilus using PacBio, Illumina, BioNano, and Hi-C data
-
Has reproduction · 80
Chromosome-level genome of the long-tailed marine-living ornate spiny lobster, Panulirus ornatus.
PMID 38909031 · PMC11193758 · Scientific data · 2024 · 6 claims · 5 setups
A chromosome-level genome of P. ornatus spanning 2.65 Gb was assembled with a contig N50 of 51.05 Mb, anchoring 99.11% of sequences to 73 chromosomes.
-
Full-text index only
Classification of real and pseudo microRNA precursors using local structure-sequence features and support vector machine.
PMID 16381612 · PMC1360673 · BMC bioinformatics · 2005 · 7 claims · 7 setups
A 32-dimensional triplet structure-sequence feature vector combined with SVM (triplet-SVM) can distinguish real human pre-miRNAs from pseudo pre-miRNA hairpins with ~90% accuracy.
-
Full-text index only
Ab initio identification of human microRNAs based on structure motifs.
PMID 18088431 · PMC2238772 · BMC bioinformatics · 2007 · 8 claims · 7 setups
MiRPred predicts miRNA precursors ab initio using only predicted secondary structure motifs, ignoring nucleotide sequence
-
Full-text index only
Lightweight genome viewer: portable software for browsing genomics data in its chromosomal context.
PMID 17877794 · PMC2238324 · BMC bioinformatics · 2007 · 7 claims · 7 setups
lwgv provides a lightweight alternative to large genome browsers for visualizing biological annotations and dynamic analyses without requiring a database or complex software infrastructure
-
Full-text index only
Human chromosome 7: DNA sequence and biology.
PMID 12690205 · PMC2882961 · Science (New York, N.Y.) · 2003 · 6 claims · 2 setups
Presents the DNA sequence and annotation of the entire human chromosome 7
-
Full-text index only
The many uses of a genome sequence.
PMID 11423005 · PMC138940 · Genome biology · 2001 · 8 claims · 8 setups
Solved protein structures from structural genomics efforts can be used to model many other proteins by homology, aiding function prediction
-
Full-text index only
Using ESTs to improve the accuracy of de novo gene prediction.
PMID 16817966 · PMC1534067 · BMC bioinformatics · 2006 · 8 claims · 8 setups
TWINSCAN_EST combines EST alignments with TWINSCAN via a trainable 'ESTseq' representation and improves exact gene structure prediction accuracy on the whole C. elegans genome
-
Full-text index only
MiPred: classification of real and pseudo microRNA precursors using random forest prediction model with combined features.
PMID 17553836 · PMC1933124 · Nucleic acids research · 2007 · 8 claims · 8 setups
A hybrid feature combining local contiguous triplet structure-sequence composition, MFE of the secondary structure, and P-value of a randomization test improves classification of real vs pseudo pre-miRNAs
-
Full-text index only
Manual annotation and analysis of the defensin gene cluster in the C57BL/6J mouse reference genome.
PMID 20003482 · PMC2807441 · BMC genomics · 2009 · 8 claims · 6 setups
Manual annotation of the mouse Chromosome 8 defensin region identifies 98 gene loci: 54 in the alpha-defensin cluster and 44 in the beta-defensin cluster
-
Full-text index only
Reconstructing Indian population history.
PMID 19779445 · PMC2842210 · Nature · 2009 · 8 claims · 8 setups
Most Indian populations descend from a mixture of two ancient, genetically divergent populations: ANI (close to Middle Easterners, Central Asians, Europeans) and ASI (as distinct from ANI and East Asians as those are from each other).
-
Has reproduction · 84
Strong population differentiation in lingcod (Ophiodon elongatus) is driven by a small portion of the genome.
PMID 33294007 · PMC7691466 · Evolutionary applications · 2020 · 7 claims · 8 setups
Lingcod comprise two distinct genetic clusters separated latitudinally at a break near Point Reyes off Northern California, with a high frequency of admixed individuals near the break.
-
Full-text index only
Genomic analysis of the chromosome 15q11-q13 Prader-Willi syndrome region and characterization of transcripts for GOLGA8E and WHCD1L1 from the proximal breakpoint region.
PMID 18226259 · PMC2268926 · BMC genomics · 2008 · 8 claims · 7 setups
GOLGA8E and WHDC1L1 are characterized for the first time as protein-coding transcripts from the PWS proximal breakpoint region.
-
Full-text index only
Genome sequences and great expectations.
PMID 11178275 · PMC150431 · Genome biology · 2001 · 8 claims · 3 setups
Function is known or can be predicted for an average of 62% of proteins across 31 analyzed genomes.
-
Has reproduction · 80
Ancient variation of the AvrPm17 gene in powdery mildew limits the effectiveness of the introgressed rye Pm17 resistance gene in wheat.
PMID 35857869 · PMC9335242 · Proceedings of the National Academy of Sciences of the United States of America · 2022 · 6 claims · 8 setups
AvrPm17 is encoded by a paralogous, tandemly duplicated effector gene pair located in a pericentromeric, mildew sublineage-specific effector cluster (family E003) showing signs of recurring gene conversion.
-
Full-text index only
Mice have a transcribed L-threonine aldolase/GLY1 gene, but the human GLY1 gene is a non-processed pseudogene.
PMID 15757516 · PMC555945 · BMC genomics · 2005 · 8 claims · 8 setups
Mouse has a transcribed, 7-exon L-threonine aldolase (GLY1) gene on chromosome 11 encoding a 400-residue protein homologous to bacterial threonine aldolase
-
Full-text index only
Predicting failure rate of PCR in large genomes.
PMID 18492719 · PMC2441781 · Nucleic acids research · 2008 · 7 claims · 8 setups
The number of predicted primer-binding sites in genomic DNA is the most important factor determining PCR failure.