Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 93
Characterization of protein isoform diversity in human umbilical vein endothelial cells via long-read proteogenomics.
PMID 36457147 · PMC9721438 · RNA biology · 2022 · 8 claims · 7 setups
Long-read RNA-seq detected 53,863 transcript isoforms from 10,426 genes in HUVECs, of which 22,195 were novel
-
Has reproduction
Genome-wide signatures of convergent evolution in echolocating mammals.
PMID 24005325 · PMC3836225 · Nature · 2013 · 8 claims · 8 setups
Genome-wide convergent sequence evolution between echolocating lineages is not rare but widespread and continuously distributed, with signatures consistent with convergence in nearly 200 loci out of 2,326 examined.
-
Full-text index only
The HIV positive selection mutation database.
PMID 17108357 · PMC1669717 · Nucleic acids research · 2007 · 8 claims · 5 setups
The database provides codon-level Ka/Ks selection pressure maps for HIV protease and the first 381 codons of RT, built from a novel ~50,000-sample clinical dataset.
-
Full-text index only
Gene losses during human origins.
PMID 16464126 · PMC1361800 · PLoS biology · 2006 · 7 claims · 7 setups
A comparative genomic screen identified 67 new human-specific nonprocessed pseudogenes, bringing the total (with 13 from prior literature) to 80 human-specific pseudogenes.
-
Full-text index only
Large-scale discovery of insertion hotspots and preferential integration sites of human transposed elements.
PMID 20008508 · PMC2836564 · Nucleic acids research · 2010 · 8 claims · 6 setups
Most TEs insert within specific 'hotspots' along the targeted TE rather than uniformly.
-
Has reproduction · 73
Detecting aberrant DNA methylation in Illumina DNA methylation arrays: a toolbox and recommendations for its use.
PMID 37218167 · PMC10208159 · Epigenetics · 2023 · 8 claims · 7 setups
Probe-specific upper and lower thresholds for flagging aberrant DNA methylation can be derived from a reference database of >2,000 normal and tumour-adjacent normal samples spanning 25 tissue types.
-
Full-text index only
Genome mapping and expression analyses of human intronic noncoding RNAs reveal tissue-specific patterns and enrichment in genes related to regulation of transcription.
PMID 17386095 · PMC1868932 · Genome biology · 2007 · 8 claims · 4 setups
More than 55,000 totally intronic noncoding (TIN) RNAs are transcribed from the introns of 74% of unique RefSeq genes.
-
Has reproduction · 82
Landscape of allele-specific transcription factor binding in the human genome.
PMID 33980847 · PMC8115691 · Nature communications · 2021 · 8 claims · 6 setups
A novel statistical framework (ADASTRA) calls allele-specific TF binding from existing ChIP-Seq alignments by jointly correcting for background allelic dosage (BAD, from aneuploidy/CNVs) and reference mapping bias.
-
Full-text index only
Inference of transcriptional regulation using gene expression data from the bovine and human genomes.
PMID 17683551 · PMC1978505 · BMC genomics · 2007 · 7 claims · 8 setups
Using human reference promoter sequences is a useful approach for studying gene expression regulation in species with limited or non-existing genomic sequence, such as cattle.
-
Full-text index only
Epidemiology of doublet/multiplet mutations in lung cancers: evidence that a subset arises by chronocoordinate events.
PMID 19005564 · PMC2579325 · PloS one · 2008 · 8 claims · 7 setups
Doublet mutations are significantly more frequent in EGFR (6.0%) and TP53 (2.3%) in human lung cancer than spontaneous doublets in mouse lacI (0.7%), about 8-fold and 3-fold higher respectively.
-
Has reproduction · 81
Transcriptional regulation and chromatin architecture maintenance are decoupled functions at the Sox2 locus.
PMID 35710138 · PMC9296009 · Genes & development · 2022 · 8 claims · 7 setups
Sox2 transcriptional activation is traced almost entirely to two key transcription factor-bound regions (SRR107 and SRR111) within the SCR
-
Full-text index only
Widespread A-to-I RNA editing of Alu-containing mRNAs in the human transcriptome.
PMID 15534692 · PMC526178 · PLoS biology · 2004 · 8 claims · 6 setups
Intramolecular pairs of oppositely oriented Alu elements within the same pre-mRNA form dsRNA foldback structures that are major substrates for A-to-I RNA editing
-
Full-text index only
CRSD: a comprehensive web server for composite regulatory signature discovery.
PMID 16845073 · PMC1538777 · Nucleic acids research · 2006 · 7 claims · 5 setups
CRSD is a comprehensive web server integrating six large-scale databases (UniGene, mature microRNAs, putative promoter, TRANSFAC, pathway, GO) plus two newly constructed genome-wide databases (MRS and TRS) for composite regulatory signature discovery
-
Full-text index only
HaploSNPer: a web-based allele and SNP detection tool.
PMID 18307806 · PMC2288614 · BMC genetics · 2008 · 6 claims · 2 setups
HaploSNPer is a web-based tool integrating BLASTN, CAP3/PHRAP, and QualitySNP into a single pipeline for allele and SNP detection from diploid and polyploid species
-
Full-text index only
A surrogate-based approach for post-genomic partner identification.
PMID 11602024 · PMC57814 · BMC biotechnology · 2001 · 8 claims · 5 setups
Peptide surrogates derived from random phage display libraries contain amino acid sequence information that identifies the natural biological partner of the panned target via database searching.
-
Full-text index only
Analysis of human sarcospan as a candidate gene for CFEOM1.
PMID 11180757 · PMC29083 · BMC genetics · 2001 · 7 claims · 5 setups
Sarcospan sequence is unmutated in all six CFEOM1 families studied
-
Full-text index only
Genome annotation errors in pathway databases due to semantic ambiguity in partial EC numbers.
PMID 16034025 · PMC1179732 · Nucleic acids research · 2005 · 7 claims · 4 setups
Partial EC numbers are semantically ambiguous, and databases that assign a gene to all reactions sharing the same partial EC number make a faulty inference, causing systematic misannotation.
-
Full-text index only
GeneKeyDB: a lightweight, gene-centric, relational database to support data mining environments.
PMID 15790402 · PMC1274265 · BMC bioinformatics · 2005 · 8 claims · 6 setups
GeneKeyDB is a lightweight, gene-centric relational database that supports data mining and integration with computational analysis tools.
-
Full-text index only
Natural variation of HIV-1 group M integrase: implications for a new class of antiretroviral inhibitors.
PMID 18687142 · PMC2546438 · Retrovirology · 2008 · 7 claims · 6 setups
Integrase displays significantly less inter- and intra-subtype amino acid diversity and lower Shannon's entropy than protease or RT.
-
Full-text index only
Toward proteome-scale identification and quantification of isoaspartyl residues in biological samples.
PMID 19663459 · PMC2756321 · Journal of proteome research · 2009 · 8 claims · 4 setups
In ECD MS/MS, isoaspartyl (but not aspartyl) residues produce specific fragments c_n•+58.0054 (C2H2O2) and z_(l-n)-56.9976 (C2HO2), which serve as markers for Asp isomerization