Accelerating the Retrieval Efficiency of Biological Genes Based on Cached Trie Tree

preprint OA: closed CC-BY-4.0
📄 Open PDF View at publisher

Abstract

with the rapid development of the third generation sequencing technology and the rapid growth of biological sequence data such as DNA, how to efficiently retrieve a gene cluster from it is a major challenge. This paper mainly introduces the acceleration of DNA gene sequence retrieval based on trie. The use of database to retrieve huge gene sequence clusters has encountered a bottleneck in performance, and the retrieval efficiency is becoming more and more slow. When faced with millions of data, it takes a minute to query a gene cluster. When we build a biological gene database, it takes 1-2 minutes to query the matched genes, which is extremely inefficient for user experience or scientific research retrieval. This paper introduces a reasonable use of memory space to build a trie tree with cache to speed up the retrieval of gene sequences.

My notes (saved in your browser only)

Citation neighborhood (no data yet)

We don't have any in-corpus citations linked to this paper yet. The paper's references may be in our DB but unresolved to ``paper_id`` (resolution happens at ingest when the cited DOI matches a row we already have). Run the cross-source citation reconcile pass to retry.

Source provenance

europepmc
last seen: 2026-05-19T01:45:01.086888+00:00
unpaywall
last seen: 2026-05-23T02:00:01.238055+00:00
License: CC-BY-4.0