Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Ensembl 2008.
PMID 18000006 · PMC2238821 · Nucleic acids research · 2008 · 8 claims · 6 setups
The Ensembl regulatory build integrates multiple genome-wide functional genomics datasets to automatically annotate regulatory regions and assign putative functions across the genome.
-
Full-text index only
Ensembl 2006.
PMID 16381931 · PMC1347495 · Nucleic acids research · 2006 · 8 claims · 5 setups
Ensembl now provides annotation for 19 genomes, up from 4 the previous year, including new mammalian (Rhesus macaque, Opossum), chordate (Ciona intestinalis), and yeast genomes.
-
Full-text index only
Molecular archeology of L1 insertions in the human genome.
PMID 12372140 · PMC134481 · Genome biology · 2002 · 8 claims · 4 setups
TSDfinder, a new algorithm, refines RepeatMasker-identified L1 boundaries by locating poly(A) tails, TSDs, and inversion breakpoints
-
Full-text index only
Human-zebrafish non-coding conserved elements act in vivo to regulate transcription.
PMID 16179648 · PMC1236720 · Nucleic acids research · 2005 · 8 claims · 4 setups
Deeply conserved human-zebrafish non-coding elements are enriched for in vivo cis-acting transcriptional regulatory activity.
-
Full-text index only
Manual annotation and analysis of the defensin gene cluster in the C57BL/6J mouse reference genome.
PMID 20003482 · PMC2807441 · BMC genomics · 2009 · 8 claims · 6 setups
Manual annotation of the mouse Chromosome 8 defensin region identifies 98 gene loci: 54 in the alpha-defensin cluster and 44 in the beta-defensin cluster
-
Full-text index only
Integrative annotation of 21,037 human genes validated by full-length cDNA clones.
PMID 15103394 · PMC393292 · PLoS biology · 2004 · 8 claims · 5 setups
41,118 full-length human cDNAs from six high-throughput sequencing projects were exhaustively integratively characterized
-
Full-text index only
Identifying related L1 retrotransposons by analyzing 3' transduced sequences.
PMID 12734010 · PMC156586 · Genome biology · 2003 · 8 claims · 6 setups
L1 elements with transduction-derived 3' sequence (L1-TDs) can be computationally identified using RepeatMasker/TSDfinder and grouped into families sharing a common progenitor via BLAST comparison of downstream sequences.
-
Full-text index only
Retropseudogenes derived from the human Ro/SS-A autoantigen-associated hY RNAs.
PMID 15817567 · PMC1074747 · Nucleic acids research · 2005 · 8 claims · 8 setups
966 pseudogenes derived from the four human Y (hY) RNAs were characterized in the human genome
-
Full-text index only
Systematic identification of pseudogenes through whole genome expression evidence profiling.
PMID 16945953 · PMC1636364 · Nucleic acids research · 2006 · 8 claims · 8 setups
Developed a novel bioinformatics method that identifies pseudogenes by profiling whole-genome transcript and protein expression evidence
-
Full-text index only
Ensembl's 10th year.
PMID 19906699 · PMC2808936 · Nucleic acids research · 2010 · 8 claims · 8 setups
Ensembl provides comprehensive gene annotation and integrated genomic resources (variation, regulation, comparative genomics) across a growing set of chordate genomes
-
Has reproduction · 68
Rfam 15: RNA families database in 2025.
PMID 39526405 · PMC11701678 · Nucleic acids research · 2025 · 8 claims · 6 setups
Rfamseq was expanded to 26 106 genomes, a 76% increase, by incorporating the latest UniProt reference proteomes and additional viral genomes
-
Full-text index only
F-SNP: computationally predicted functional SNPs for disease association studies.
PMID 17986460 · PMC2238878 · Nucleic acids research · 2008 · 6 claims · 8 setups
F-SNP is a database integrating functional effect predictions for SNPs from 16 bioinformatics tools/databases across four categories: splicing, transcription, translation, and post-translation
-
Has reproduction · 85
Ensembl 2013.
PMID 23203987 · PMC3531136 · Nucleic acids research · 2013 · 8 claims · 8 setups
Ensembl (http://www.ensembl.org) provides genome information for sequenced chordate genomes, currently supporting 70 species with a focus on human, mouse, zebrafish and rat.
-
Full-text index only
Genome-wide detection of segmental duplications and potential assembly errors in the human genome sequence.
PMID 12702206 · PMC154576 · Genome biology · 2003 · 8 claims · 6 setups
Segmental duplications comprise 3.53% (107.4/3,043.1 Mb) of the June 2002 human genome assembly
-
Full-text index only
Ensembl 2007.
PMID 17148474 · PMC1761443 · Nucleic acids research · 2007 · 8 claims · 7 setups
Ensembl added 18 new chordate genomes this year, increasing total genomes available from 15 to 33, the largest yearly increase to date.
-
Full-text index only
A space-efficient and accurate method for mapping and aligning cDNA sequences onto genomic sequence.
PMID 18344523 · PMC2377433 · Nucleic acids research · 2008 · 7 claims · 6 setups
Spaln maps and aligns large cDNA sequence sets onto whole mammalian genomes using substantially less memory than comparable existing tools
-
Full-text index only
Performance assessment of promoter predictions on ENCODE regions in the EGASP experiment.
PMID 16925837 · PMC1810552 · Genome biology · 2006 · 6 claims · 3 setups
Promoter predictors that combine promoter prediction with gene prediction (N-SCAN, Fprom) achieve better performance than pure ab initio promoter predictors, mainly by reducing the promoter search space and false positives
-
Full-text index only
Using several pair-wise informant sequences for de novo prediction of alternatively spliced transcripts.
PMID 16925842 · PMC1810557 · Genome biology · 2006 · 8 claims · 4 setups
MARS, an extension of the Twinscan algorithm, uses multiple pairwise informant genomes to predict human alternatively spliced transcripts de novo without expressed sequence information.
-
Full-text index only
GeneTide--Terra Incognita Discovery Endeavor: a new transcriptome focused member of the GeneCards/GeneNote suite of databases.
PMID 15608261 · PMC540076 · Nucleic acids research · 2005 · 8 claims · 7 setups
GeneTide integrates UniGene, DoTS, AceView, BLAT/GeneLoc genomic alignment, and GeneAnnot probe-set data into a unified Consensus/Uniqueness/Score scheme to associate ESTs with GeneCards genes
-
Full-text index only
Expoldb: expression linked polymorphism database with inbuilt tools for analysis of expression and simple repeats.
PMID 17038195 · PMC1618849 · BMC genomics · 2006 · 8 claims · 6 setups
EXPOLDB is a novel database integrating human gene expression variability data (including monozygotic twin comparisons) with (TG/CA)n repeat polymorphism information