Accelerating the Retrieval Efficiency of Biological Genes Based on Cached Trie Tree
preprint
OA: closed
CC-BY-4.0
Abstract
with the rapid development of the third generation sequencing technology and the rapid growth of biological sequence data such as DNA, how to efficiently retrieve a gene cluster from it is a major challenge. This paper mainly introduces the acceleration of DNA gene sequence retrieval based on trie. The use of database to retrieve huge gene sequence clusters has encountered a bottleneck in performance, and the retrieval efficiency is becoming more and more slow. When faced with millions of data, it takes a minute to query a gene cluster. When we build a biological gene database, it takes 1-2 minutes to query the matched genes, which is extremely inefficient for user experience or scientific research retrieval. This paper introduces a reasonable use of memory space to build a trie tree with cache to speed up the retrieval of gene sequences.
My notes (saved in your browser only)
Citation neighborhood (no data yet)
We don't have any in-corpus citations linked to this paper yet. The paper's references may be in our DB but unresolved to ``paper_id`` (resolution happens at ingest when the cited DOI matches a row we already have). Run the cross-source citation reconcile pass to retry.
Source provenance
- europepmc
- last seen: 2026-05-19T01:45:01.086888+00:00
- unpaywall
- last seen: 2026-05-23T02:00:01.238055+00:00
License: CC-BY-4.0