ADAPTIVE MULTI-SCALE GRAPH TRANSFORMER FRAMEWORK FOR HISTOPATHOLOGICAL IMAGES

preprint OA: closed
📄 Open PDF Full text JSON View at publisher

Abstract

ABSTRACT Whole slide images (WSIs) contain hierarchical information from cellular to tissue architecture but their gigapixel scale poses major memory and computational challenges. Existing multi-scale graph and transformer models capture complex WSI features effectively but struggle with efficiency. We propose an Adaptive Multi-Scale Graph Transformer (AMGT) for WSI classification that addresses this limitation through two key modules: a Self-Guided Token Aggregation (SGTA) mechanism that fuses multi-resolution features to reduce redundancy, and a Prototypical Transformer (PT) that groups similar tokens into phenotype-representative prototypes with linear complexity. This design preserves essential spatial and semantic information, substantially lowering memory cost and improving interpretability by prototypical learning. AMGT achieves superior performance and efficiency, outperforming state-of-the-art models by 1.8% and 5.3% AUC on high-grade ovarian cancer and Camelyon16 datasets, respectively. These results demonstrate AMGT’s capacity for scalable, interpretable multi-scale representation learning.
Full text 1,196 characters · extracted from oa-doi-fallback · click to expand
ABSTRACT Whole slide images (WSIs) contain hierarchical information from cellular to tissue architecture but their gigapixel scale poses major memory and computational challenges. Existing multi-scale graph and transformer models capture complex WSI features effectively but struggle with efficiency. We propose an Adaptive Multi-Scale Graph Transformer (AMGT) for WSI classification that addresses this limitation through two key modules: a Self-Guided Token Aggregation (SGTA) mechanism that fuses multi-resolution features to reduce redundancy, and a Prototypical Transformer (PT) that groups similar tokens into phenotype-representative prototypes with linear complexity. This design preserves essential spatial and semantic information, substantially lowering memory cost and improving interpretability by prototypical learning. AMGT achieves superior performance and efficiency, outperforming state-of-the-art models by 1.8% and 5.3% AUC on high-grade ovarian cancer and Camelyon16 datasets, respectively. These results demonstrate AMGT’s capacity for scalable, interpretable multi-scale representation learning. Competing Interest Statement The authors have declared no competing interest.

Text is read by the "Ask this paper" AI Q&A widget below. Extraction quality varies by source — PMC NXML preserves structure cleanly, OA-HTML may include some navigation residue, and OA-PDF can have broken hyphenation. The publisher copy (via DOI) is the canonical version.

My notes (saved in your browser only)

Ask this paper AI returns verbatim quotes from the full text · source: oa-doi-fallback

Answers must be backed by verbatim quotes from this paper's full text. Hallucinated quotes are dropped automatically; if no verbatim passage answers the question, we say so. How this works

Citation neighborhood (no data yet)

We don't have any in-corpus citations linked to this paper yet. This is a recent paper (2025) — citers typically take a year or two to land, and the OpenAlex reference graph may still be filling in.

Source provenance

europepmc
last seen: 2026-05-20T01:45:00.602351+00:00