Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
The TIGR Gene Indices: clustering and assembling EST and known genes and integration with eukaryotic genomes.
PMID 15608288 · PMC540018 · Nucleic acids research · 2005 · 8 claims · 8 setups
The TIGR Gene Indices (TGI) are a collection of 77 species-specific databases that cluster and assemble EST and known gene sequences into tentative consensus (TC) sequences to identify and characterize expressed transcripts.
-
Full-text index only
Direct evidence of extensive diversity of HIV-1 in Kinshasa by 1960.
PMID 18833279 · PMC3682493 · Nature · 2008 · 7 claims · 8 setups
Recovered and characterized HIV-1 sequences (DRC60) from a 1960 Bouin's-fixed paraffin-embedded lymph node biopsy from Léopoldville, Belgian Congo
-
Full-text index only
Metagenomic analysis of human diarrhea: viral detection and discovery.
PMID 18398449 · PMC2290972 · PLoS pathogens · 2008 · 8 claims · 7 setups
Micro-mass sequencing (minimal stool input, minimal purification, ~384 reads/sample) can detect known enteric viruses in diarrhea specimens
-
Full-text index only
Predicting the phenotypic effects of non-synonymous single nucleotide polymorphisms based on support vector machines.
PMID 18005451 · PMC2216041 · BMC bioinformatics · 2007 · 8 claims · 5 setups
Parepro, an SVM-based method integrating three attribute sets (RD, MI, IE) derived from evolutionary and residue-property information, predicts whether an nsSNP is deleterious or neutral.
-
Full-text index only
Conservation, variability and the modeling of active protein kinases.
PMID 17912359 · PMC1989141 · PloS one · 2007 · 7 claims · 5 setups
A novel sequence-order independent (fold-independent) structural alignment algorithm was developed that maximizes side-chain similarity to produce a consensus kinase structure.
-
Full-text index only
Ratiocinative screen of eukaryotic integral membrane protein expression and solubilization for structure determination.
PMID 19031011 · PMC2756966 · Journal of structural and functional genomics · 2009 · 8 claims · 6 setups
A discovery-oriented pipeline using standardized single-condition methods (one expression system, one detergent, one SEC buffer) can efficiently triage large numbers of eukaryotic IMP targets to identify well-behaved candidates for crystallization
-
Full-text index only
Bayesian coestimation of phylogeny and sequence alignment.
PMID 15804354 · PMC1087833 · BMC bioinformatics · 2005 · 7 claims · 3 setups
Alignment and phylogenetic inference are mutually dependent, and treating them as separate sequential steps (align then infer tree) is fundamentally flawed and produces biased, overconfident estimates.
-
Full-text index only
Computer-aided identification of polymorphism sets diagnostic for groups of bacterial and viral genetic variants.
PMID 17672919 · PMC1973086 · BMC bioinformatics · 2007 · 6 claims · 8 setups
The Not-N algorithm, incorporated into the Minimum SNPs program, identifies small marker sets diagnostic for user-defined subgroups of genetic variants with 0% false negatives
-
Full-text index only
Distinctive pattern of sequence polymorphism in the NS3 protein of hepatitis C virus type 1b reflects conflicting evolutionary pressures.
PMID 18632963 · PMC2577380 · The Journal of general virology · 2008 · 7 claims · 6 setups
NS3 shows less evidence of purifying selection acting on its CTL epitopes than the other 9 HCV proteins, while outside the CTL epitopes NS3 is more conserved than the other proteins.
-
Full-text index only
CorGen--measuring and generating long-range correlations for DNA sequence analysis.
PMID 16845099 · PMC1538783 · Nucleic acids research · 2006 · 8 claims · 3 setups
CorGen is a web server that measures long-range correlations in DNA sequences and generates random sequences with the same (or user-specified) correlation and composition parameters
-
Full-text index only
MiPred: classification of real and pseudo microRNA precursors using random forest prediction model with combined features.
PMID 17553836 · PMC1933124 · Nucleic acids research · 2007 · 8 claims · 8 setups
A hybrid feature combining local contiguous triplet structure-sequence composition, MFE of the secondary structure, and P-value of a randomization test improves classification of real vs pseudo pre-miRNAs
-
Full-text index only
Conserved elements with potential to form polymorphic G-quadruplex structures in the first intron of human genes.
PMID 18187510 · PMC2275096 · Nucleic acids research · 2008 · 8 claims · 6 setups
G-richness downstream of the TSS is strand-biased, concentrated on the nontemplate strand, with a peak at +200 to +300 bp
-
Full-text index only
Recent segmental and gene duplications in the mouse genome.
PMID 12914656 · PMC193640 · Genome biology · 2003 · 8 claims · 8 setups
33.6 Mb (1.2%) of the February 2003 mouse genome assembly (2,695 Mb) is involved in recent segmental duplications
-
Full-text index only
Backseat drivers take the wheel.
PMID 18068625 · PMC2705833 · Cancer cell · 2007 · 8 claims · 8 setups
Systematic resequencing combined with functional validation can distinguish rare driver FLT3 mutations from passenger mutations in AML patients negative for known activating mutations
-
Full-text index only
SNAP: predict effect of non-synonymous polymorphisms on function.
PMID 17526529 · PMC1920242 · Nucleic acids research · 2007 · 7 claims · 8 setups
SNAP, a neural network-based method using sequence-derived information, predicts whether a non-synonymous SNP is neutral or non-neutral for protein function
-
Has reproduction · 50
Genome-wide identification of Hfq-regulated small RNAs in the fire blight pathogen Erwinia amylovora discovered small RNAs with virulence regulatory function.
PMID 24885615 · PMC4070566 · BMC genomics · 2014 · 8 claims · 8 setups
A total of 40 candidate Hfq-dependent sRNAs were identified genome-wide in E. amylovora by combining RNA-seq with a Rho-independent terminator search.
-
Full-text index only
Genetic variability of the P120' surface protein gene of Mycoplasma hominis isolates recovered from Tunisian patients with uro-genital and infertility disorders.
PMID 18053243 · PMC2225410 · BMC infectious diseases · 2007 · 7 claims · 5 setups
The P120' surface-exposed N-terminal region undergoes substantial genetic variability among Tunisian M. hominis clinical isolates
-
Full-text index only
Challenges and standards in integrating surveys of structural variation.
PMID 17597783 · PMC2698291 · Nature genetics · 2007 · 7 claims · 5 setups
There is no standard approach to collecting, assessing the quality of, or describing structural variants, risking the entire genome eventually being labeled 'structurally variant' based on uncurated nondisease-sample data.
-
Full-text index only
Variation in the E2-binding domain of HPV 16 is associated with high-grade squamous intraepithelial lesions of the cervix.
PMID 11308254 · PMC2363853 · British journal of cancer · 2001 · 8 claims · 6 setups
E2 gene disruption in the analysed regions is significantly more frequent in high-grade SILs than low-grade SILs