Lightweight Multiuser Multimodal Semantic Communication System for Multimodal Large Language Model Communication
preprint
OA: closed
AI-generated summary
A Mamba-based system (3M-DeepSC) was developed for efficient, low-latency multimodal LLM communication with a new semantic similarity metric and two-stage training.
One-sentence paraphrase of the abstract; not a substitute for reading it. No clinical advice. How this works
Abstract
Developing multimodal large language models (MLLMs) requires efficient semantic communication systems for diverse data modalities in constrained networks, highlighting the need for lightweight semantic communication models optimized for resource-constrained environments. A Mamba-based multiuser multimodal deep-learning semantic communication (3M-DeepSC) system is developed to serve MLLM communication to address this problem. The proposed framework applies the efficient Mamba architecture to replace traditional Transformerbased designs, improving performance and lowering the latency under diverse channel conditions. Moreover, a new semantic similarity metric is introduced to evaluate the system performance from a semantic perspective. In addition, a two-stage training algorithm is developed that jointly optimizes bit-based metrics and semantic similarity. According to the extensive results, the proposed 3M-DeepSC demonstrates promise as a robust, scalable solution supporting the increasing communication demands of MLLMs in diverse network environments.
My notes (saved in your browser only)
Citation neighborhood (no data yet)
We don't have any in-corpus citations linked to this paper yet. This is a recent paper (2024) — citers typically take a year or two to land, and the OpenAlex reference graph may still be filling in.
Source provenance
- europepmc
- last seen: 2026-05-20T01:45:00.602351+00:00