Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Large-scale analysis of Macaca fascicularis transcripts and inference of genetic divergence between M. fascicularis and M. mulatta.
PMID 18294402 · PMC2287170 · BMC genomics · 2008 · 8 claims · 6 setups
Constructed full-length-enriched cDNA libraries and determined 85,721 EST sequences and 9407 full-insert sequences from cynomolgus macaque brain (7 regions), testis, and liver
-
Full-text index only
Exome sequencing of a multigenerational human pedigree.
PMID 20011588 · PMC2788131 · PloS one · 2009 · 8 claims · 6 setups
Microarray-based exome capture combined with 454 GS FLX NGS is an efficient and reliable method to enrich for chromosomal regions of interest, validated on eight individuals from a three-generation pedigree
-
Full-text index only
Prodepth: predict residue depth by support vector regression approach from protein sequences only.
PMID 19759917 · PMC2742725 · PloS one · 2009 · 8 claims · 8 setups
Residue depth can be reliably predicted solely from protein primary sequence using support vector regression on sequence-derived features.
-
Full-text index only
Analysis of virulence factors of Helicobacter pylori isolated from a Vietnamese population.
PMID 19698173 · PMC2739534 · BMC microbiology · 2009 · 8 claims · 5 setups
Three distinct deletion patterns (39-bp, 18-bp, no deletion) exist upstream of the cagA EPIYA repeat region (pre-EPIYA region), providing a novel genotyping marker.
-
Full-text index only
Fully haplotyped genome assemblies of healthy individuals reveal variability in 5'ss strength and support by splicing regulatory proteins.
PMID 40191587 · PMC11970367 · NAR genomics and bioinformatics · 2025 · 8 claims · 5 setups
44 individuals' fully haplotyped diploid genome assemblies (88 haplotypes) from the 1000 Genomes Project were used to comprehensively assess homozygous and heterozygous sequence variations around and within 5'ss
-
Full-text index only
BioAfrica's HIV-1 proteomics resource: combining protein data with bioinformatics tools.
PMID 15757512 · PMC555852 · Retrovirology · 2005 · 8 claims · 3 setups
BioAfrica's HIV-1 Proteomics Resource integrates protein structure, gene expression, post-translational modification, functional activity and protein-macromolecule interaction data with bioinformatics tools in a single website.
-
Full-text index only
Computational identification of transcriptional regulatory elements in DNA sequence.
PMID 16855295 · PMC1524905 · Nucleic acids research · 2006 · 8 claims · 3 setups
Weight matrix (PWM/PSSM) models of TF binding sites are grounded in biophysical theory of protein-DNA interactions, with position weights corresponding to log-odds contributions to binding free energy
-
Full-text index only
CorGen--measuring and generating long-range correlations for DNA sequence analysis.
PMID 16845099 · PMC1538783 · Nucleic acids research · 2006 · 8 claims · 3 setups
CorGen is a web server that measures long-range correlations in DNA sequences and generates random sequences with the same (or user-specified) correlation and composition parameters
-
Full-text index only
Decoding of superimposed traces produced by direct sequencing of heterozygous indels.
PMID 18654614 · PMC2429969 · PLoS computational biology · 2008 · 7 claims · 3 setups
A dynamic programming method (implemented as web app Indelligent) can decode superimposed allelic sequences from a single mixed trace, using only the observed string of ambiguous peak calls, without a reference sequence or reverse trace.
-
Full-text index only
Germline mutations of the INK4a-ARF gene in patients with suspected genetic predisposition to melanoma.
PMID 14735200 · PMC2409576 · British journal of cancer · 2004 · 8 claims · 2 setups
Seven germline INK4a-ARF changes (five novel) were found in 7 of 89 patients (8%) suspected of genetic predisposition to melanoma.
-
Full-text index only
Using multiple alignments to improve seeded local alignment algorithms.
PMID 16100379 · PMC1185574 · Nucleic acids research · 2005 · 8 claims · 2 setups
Using information implicit in a multiple alignment to dynamically build a spaced-seed index weighted toward promising regions increases sensitivity of local alignment search compared to indexing a sequence alone
-
Full-text index only
HaploSNPer: a web-based allele and SNP detection tool.
PMID 18307806 · PMC2288614 · BMC genetics · 2008 · 6 claims · 2 setups
HaploSNPer is a web-based tool integrating BLASTN, CAP3/PHRAP, and QualitySNP into a single pipeline for allele and SNP detection from diploid and polyploid species
-
Full-text index only
Targeted next-generation sequencing of a cancer transcriptome enhances detection of sequence variants and novel fusion transcripts.
PMID 19835606 · PMC2784330 · Genome biology · 2009 · 7 claims · 2 setups
Hybrid selection of cDNA dramatically increases the specificity of sequencing reads mapping to targeted cancer-related transcripts.
-
Full-text index only
Bases and spaces: resources on the web for accessing the draft human genome.
PMID 11178254 · PMC138875 · Genome biology · 2000 · 8 claims · 8 setups
By combining currently available genomic databases and mapping resources (GenBank/Entrez, UniGene, RH maps, BAC fingerprint maps, Ensembl, NIX), it is possible to devise strategies that fully exploit the fragmentary draft human genome sequence.
-
Full-text index only
Charting the map of life.
PMID 11171541 · PMC1242061 · Environmental health perspectives · 2001 · 8 claims · 8 setups
Only 3-5% of the human genome, corresponding to 30,000-100,000 genes, is thought to be biologically functional
-
Full-text index only
The Vertebrate Genome Annotation (Vega) database.
PMID 15608237 · PMC540089 · Nucleic acids research · 2005 · 8 claims · 8 setups
Vega is a community database for browsing manual annotation of finished vertebrate genome sequences, based on an extended Ensembl-style schema.
-
Full-text index only
The UCSC Genome Browser Database: 2008 update.
PMID 18086701 · PMC2238835 · Nucleic acids research · 2008 · 8 claims · 8 setups
The UCSC Genome Browser Database (GBD) provides integrated sequence and annotation data for a large collection of vertebrate and model organism genomes.
-
Full-text index only
The UCSC Genome Browser database: update 2010.
PMID 19906737 · PMC2808870 · Nucleic acids research · 2010 · 8 claims · 5 setups
The UCSC Genome Browser provides a large database of publicly available sequence and annotation data with an integrated tool set for examining, comparing, aligning, and displaying genomes
-
Has reproduction · 61
Tourmaline: A containerized workflow for rapid and iterable amplicon sequence analysis using QIIME 2 and Snakemake.
PMID 35902092 · PMC9334028 · GigaScience · 2022 · 8 claims · 1 setups
Tourmaline is a Python-based workflow that implements QIIME 2 using the Snakemake workflow management system to automate amplicon sequence analysis.
-
Has reproduction · 67
Adaptive learning embedding features to improve the predictive performance of SARS-CoV-2 phosphorylation sites.
PMID 37847658 · PMC10628388 · Bioinformatics (Oxford, England) · 2023 · 8 claims · 6 setups
PSPred-ALE outperforms state-of-the-art SARS-CoV-2 phosphorylation site predictors (e.g. DeepIPs) and handcrafted feature-based methods in benchmarking comparisons