Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 89
TrEMOLO: accurate transposable element allele frequency estimation using long-read sequencing data combining assembly and mapping-based approaches.
PMID 37013657 · PMC10069131 · Genome biology · 2023 · 6 claims · 6 setups
TrEMOLO combines an assembly-based INSIDER module and a mapping-based OUTSIDER module to detect TE insertions/deletions from long-read sequencing data and estimate their allele frequency
-
Full-text index only
A critical reassessment of the role of mitochondria in tumorigenesis.
PMID 16187796 · PMC1240051 · PLoS medicine · 2005 · 8 claims · 8 setups
A significant number of published medical mtDNA cancer studies are based on obviously flawed sequencing results.
-
Full-text index only
WebGestalt: an integrated system for exploring gene sets in various biological contexts.
PMID 15980575 · PMC1160236 · Nucleic acids research · 2005 · 8 claims · 6 setups
WebGestalt is an integrated web-based system composed of four modules: gene set management, information retrieval, organization/visualization, and statistics.
-
Full-text index only
Filtering high-throughput protein-protein interaction data using a combination of genomic features.
PMID 15833142 · PMC1127019 · BMC bioinformatics · 2005 · 8 claims · 8 setups
A combination of three genomic features (interacting Pfam domains, GO annotations, sequence homology) using naive Bayesian networks predicts true protein-protein interactions with high sensitivity and good specificity.
-
Full-text index only
The MAPPER database: a multi-genome catalog of putative transcription factor binding sites.
PMID 15608292 · PMC540057 · Nucleic acids research · 2005 · 8 claims · 6 setups
Built a library of 1134 HMM models (359 matrix-derived, 718 factor-derived, 57 JASPAR-derived), corresponding to 863 distinct TF names, from TRANSFAC and JASPAR binding site data
-
Full-text index only
Genome assembly comparison identifies structural variants in the human genome.
PMID 17115057 · PMC2674632 · Nature genetics · 2006 · 7 claims · 7 setups
Genome assembly comparison is a robust approach for identifying all classes of genetic variation, with no lower size limit.
-
Full-text index only
Analysis of protein sequence and interaction data for candidate disease gene prediction.
PMID 17020920 · PMC1636487 · Nucleic acids research · 2006 · 8 claims · 7 setups
Combining CPS and CMP using known disease genes as input achieves sensitivity 0.52 and specificity 0.97, reducing candidate lists 13-fold
-
Full-text index only
Bias of selection on human copy-number variants.
PMID 16482228 · PMC1366494 · PLoS genetics · 2006 · 8 claims · 8 setups
Human CNVs are significantly overrepresented near telomeres and centromeres and enriched in simple tandem repeats relative to the genome as a whole
-
Has reproduction · 51
Polyploidy and the petal transcriptome of Gossypium.
PMID 24393201 · PMC3890615 · BMC plant biology · 2014 · 8 claims · 8 setups
Most homoeologous gene pairs in polyploid cotton petals are expressed at equal levels, indicating a surprising level of expression homeostasis; only ~20% of expressed genes show significant genome bias.
-
Full-text index only
Dark matter in a deep-sea vent and in human mouth.
PMID 17803764 · PMC2040194 · Environmental microbiology · 2007 · 7 claims · 8 setups
The first genome sequence from the uncultured TM7 phylum was obtained by capturing and sequencing DNA from a single cell using a microfluidic device, yielding a 2.86 Mb assembly with 3245 predicted genes.
-
Full-text index only
A global view of protein expression in human cells, tissues, and organs.
PMID 20029370 · PMC2824494 · Molecular systems biology · 2009 · 7 claims · 6 setups
A high fraction (>65%) of proteins is expressed in most human cells and tissues, while very few proteins (<2%) are detected in any single cell type.
-
Full-text index only
The sequence and de novo assembly of the giant panda genome.
PMID 20010809 · PMC3951497 · Nature · 2010 · 8 claims · 8 setups
A draft giant panda genome was successfully generated and assembled de novo using only Illumina Genome Analyser short-read sequencing
-
Full-text index only
Sequence variation of PfEMP1-DBLalpha in association with rosette formation in Plasmodium falciparum isolates causing severe and uncomplicated malaria.
PMID 19650937 · PMC3224928 · Malaria journal · 2009 · 8 claims · 5 setups
Sequence group 1 (mainly from uncomplicated malaria) is significantly different in distribution from sequence group 3 (mainly from severe malaria)
-
Full-text index only
Mutations associated with HNPCC predisposition -- Update of ICG-HNPCC/INSiGHT mutation database.
PMID 15528792 · PMC3839397 · Disease markers · 2004 · 8 claims · 4 setups
The ICG-HNPCC/INSiGHT mutation database has grown from 126 predisposing mutations (1997) to 448 mutations occurring in 748 families (2003 update)