Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
A new procedure for determining the genetic basis of a physiological process in a non-model species, illustrated by cold induced angiogenesis in the carp.
PMID 19852815 · PMC2771047 · BMC genomics · 2009 · 8 claims · 5 setups
The Conditional Stepped Reciprocal Best Hit (CSRBH) approach, combining direct RBH and zebrafish-stepped RBH (SRBH), outperformed other ortholog assignment methods and attained 8,726 carp-human functional homolog relationships for 16,650 carp contigs
-
Full-text index only
Tissue proteomics reveals differential and compartment-specific expression of the homologs transgelin and transgelin-2 in lung adenocarcinoma and its stroma.
PMID 19848416 · PMC2789179 · Journal of proteome research · 2009 · 6 claims · 6 setups
Transgelin (TAGLN), transgelin-2 (TAGLN2), and cyclophilin A (PPIA) are overexpressed in lung adenocarcinoma tissue compared to matched normal lung
-
Full-text index only
Pre-operative urinary cathepsin D is associated with survival in patients with renal cell carcinoma.
PMID 19789534 · PMC2768081 · British journal of cancer · 2009 · 8 claims · 7 setups
Cathepsin D, identified via comparative 2D PAGE of conditioned media from RCC cell lines vs normal renal cultures, is a candidate secreted biomarker of RCC
-
Full-text index only
Reconstructing Indian population history.
PMID 19779445 · PMC2842210 · Nature · 2009 · 8 claims · 8 setups
Most Indian populations descend from a mixture of two ancient, genetically divergent populations: ANI (close to Middle Easterners, Central Asians, Europeans) and ASI (as distinct from ANI and East Asians as those are from each other).
-
Full-text index only
Evidence for limited genetic compartmentalization of HIV-1 between lung and blood.
PMID 19759830 · PMC2736399 · PloS one · 2009 · 8 claims · 7 setups
Statistical evidence of genetic compartmentalization between lung and blood HIV-1 env sequences was found in 10 of 18 subjects.
-
Full-text index only
Genome-wide prioritization of disease genes and identification of disease-disease associations from an integrated human functional linkage network.
PMID 19728866 · PMC2768980 · Genome biology · 2009 · 6 claims · 6 setups
Integrating 16 genomic features (32 sub-features) via a naïve Bayes classifier produces a genome-scale FLN of 21,657 human genes and 22,388,609 weighted links that outperforms any individual data source for inferring functional linkages.
-
Full-text index only
MrHAMER yields highly accurate single molecule viral sequences enabling analysis of intra-host evolution.
PMID 33849057 · PMC8266615 · Nucleic acids research · 2021 · 8 claims · 7 setups
MrHAMER yields >1000s of viral genomes per sample at 99.9% accuracy
-
Full-text index only
NeoPrecis: enhancing immunotherapy response prediction through integration of qualified immunogenicity and clonality-aware neoantigen landscapes.
PMID 41577704 · PMC12932759 · Nature communications · 2026 · 8 claims · 8 setups
NeoPrecis-Immuno, a T-cell recognition model incorporating MHC-binding motif enrichment into a cross-reactivity-distance framework, improves neoantigen immunogenicity prediction.
-
Full-text index only
Retentive Network promotes efficient RNA language modeling of long sequences.
PMID 41814064 · PMC13111708 · Communications biology · 2026 · 8 claims · 6 setups
RNAret, a RetNet-based RNA language model with O(n) complexity, achieves training parallelism and low computational overhead while processing long RNA sequences
-
Full-text index only
Separating selection from mutation in antibody language models.
PMID 41944291 · PMC13056363 · eLife · 2026 · 8 claims · 6 setups
Masked antibody language models such as AbLang2 are biased by nucleotide-level mutation processes (germline memorization, codon table, SHM rate variation)
-
Full-text index only
Much ado about nothing: modeling amino acid replacement with predicted protein structures.
PMID 42036821 · PMC13171170 · Bioinformatics (Oxford, England) · 2026 · 8 claims · 7 setups
AFSM was constructed from over 660,000 structural alignments across ~21,000 proteins (297 InterPro families), following the BLOSUM log-odds methodology.
-
Full-text index only
A bioinformatics pipeline for a tick pathogen surveillance multiplex amplicon sequencing assay.
PMID 37247570 · PMC10878300 · Ticks and tick-borne diseases · 2023 · 7 claims · 3 setups
The MPAS pipeline is a portable, reproducible Nextflow-based bioinformatics pipeline that identifies and summarizes amplicon sequences produced by the MPAS assay.
-
Full-text index only
The matrix metalloproteinase-3 (MMP-3) 5A/6A promoter polymorphism is not associated with ischaemic heart disease: analysis employing a family based approach.
PMID 15665388 · PMC3839324 · Disease markers · 2004 · 6 claims · 3 setups
The MMP-3 -1612 5A/6A polymorphism is not associated with ischaemic heart disease in an Irish population.
-
Full-text index only
Alignoth: portable and interactive visualization of read alignments.
PMID 41392197 · PMC12777968 · Bioinformatics (Oxford, England) · 2026 · 8 claims · 3 setups
Alignoth is a command-line tool that generates self-contained, portable HTML reports of DNA sequencing read alignment pileups, with export to PNG, SVG, PDF, and JSON.
-
Full-text index only
PHScaffolding: a hypergraph clustering and dual-weight integration strategy for scaffolding with Pore-C reads.
PMID 41569288 · PMC12825295 · Briefings in bioinformatics · 2026 · 7 claims · 6 setups
PHScaffolding builds a weighted hypergraph from Pore-C read-to-contig alignments, with hyperedges representing multi-way contig interactions and weighted by the geometric mean of per-contig alignment lengths
-
Full-text index only
Evaluating the performance of ancient DNA genetic relatedness estimation methods using high-fidelity pedigree simulations.
PMID 41796349 · PMC13081257 · Genome biology · 2026 · 8 claims · 5 setups
BADGER, an automated snakemake pipeline, was developed to simulate high-fidelity pedigrees and raw ancient DNA sequence data for benchmarking genetic relatedness methods
-
Full-text index only
The next epidemic.
PMID 16737558 · PMC1779519 · Genome biology · 2006 · 8 claims · 1 setups
More than 75% of susceptibility to sporadic Alzheimer's disease may be due to genetic factors
-
Full-text index only
MODBASE: a database of annotated comparative protein structure models and associated resources.
PMID 16381869 · PMC1347422 · Nucleic acids research · 2006 · 8 claims · 7 setups
MODBASE is a database of automatically calculated comparative protein structure models covering all UniProt sequences matchable to a known structure
-
Full-text index only
Report of the 9th HLPP Workshop October 2007, Seoul, Korea.
PMID 18683817 · PMC4601560 · Proteomics · 2008 · 8 claims · 8 setups
An integrated separating-identifying platform identified 6788 proteins (≥2 peptides, 95% confidence) in Chinese human liver samples, including 3721 new to liver and 977 hypothetical proteins
-
Full-text index only
Outlook on Thailand's genomics and computational biology research and development.
PMID 18654621 · PMC2446437 · PLoS computational biology · 2008 · 8 claims · 8 setups
Thai government policy support, infrastructure investment, education programs, and human resource development have substantially advanced genomics and bioinformatics research capacity in Thailand