Semantic Communication based on Large Language Model for Underwater Image Transmission

Chen, Weilong; Xu, Wenxuan; Chen, Haoran; Zhang, Xinran; Qin, Zhijin; Zhang, Yanru; Han, Zhu

Computer Science > Computer Vision and Pattern Recognition

arXiv:2408.12616 (cs)

[Submitted on 8 Aug 2024 (v1), last revised 26 Aug 2024 (this version, v2)]

Title:Semantic Communication based on Large Language Model for Underwater Image Transmission

Authors:Weilong Chen, Wenxuan Xu, Haoran Chen, Xinran Zhang, Zhijin Qin, Yanru Zhang, Zhu Han

View PDF HTML (experimental)

Abstract:Underwater communication is essential for environmental monitoring, marine biology research, and underwater exploration. Traditional underwater communication faces limitations like low bandwidth, high latency, and susceptibility to noise, while semantic communication (SC) offers a promising solution by focusing on the exchange of semantics rather than symbols or bits. However, SC encounters challenges in underwater environments, including semantic information mismatch and difficulties in accurately identifying and transmitting critical information that aligns with the diverse requirements of underwater applications. To address these challenges, we propose a novel Semantic Communication (SC) framework based on Large Language Models (LLMs). Our framework leverages visual LLMs to perform semantic compression and prioritization of underwater image data according to the query from users. By identifying and encoding key semantic elements within the images, the system selectively transmits high-priority information while applying higher compression rates to less critical regions. On the receiver side, an LLM-based recovery mechanism, along with Global Vision ControlNet and Key Region ControlNet networks, aids in reconstructing the images, thereby enhancing communication efficiency and robustness. Our framework reduces the overall data size to 0.8\% of the original. Experimental results demonstrate that our method significantly outperforms existing approaches, ensuring high-quality, semantically accurate image reconstruction.

Subjects:	Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
Cite as:	arXiv:2408.12616 [cs.CV]
	(or arXiv:2408.12616v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2408.12616

Submission history

From: Weilong Chen [view email]
[v1] Thu, 8 Aug 2024 16:46:14 UTC (13,371 KB)
[v2] Mon, 26 Aug 2024 03:47:06 UTC (13,430 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Semantic Communication based on Large Language Model for Underwater Image Transmission

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Semantic Communication based on Large Language Model for Underwater Image Transmission

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators