| Citation: | HU Tianwei, ZHANG Xiangrui, CHEN Jian, DUAN Haodong, JIA Jie. A Structure-Preserving Semantic Transmission Method for Low-Bandwidth Networks[J]. Journal of Electronics & Information Technology. doi: 10.11999/JEIT260525 |
| [1] |
JERNBERG C, SANDIN J, ZIEMKE T, et al. The effect of latency, speed and task on remote operation of vehicles[J]. Transportation Research Interdisciplinary Perspectives, 2024, 26: 101152. doi: 10.1016/j.trip.2024.101152.
|
| [2] |
BOURTSOULATZE E, KURKA D B, and GÜNDÜZ D. Deep joint source-channel coding for wireless image transmission[J]. IEEE Transactions on Cognitive Communications and Networking, 2019, 5(3): 567–579. doi: 10.1109/TCCN.2019.2919300.
|
| [3] |
DONG Chao, DENG Yubin, LOY C C, et al. Compression artifacts reduction by a deep convolutional network[C]. Proceedings of the 2015 IEEE International Conference on Computer Vision (ICCV), Santiago, Chile, 2015: 576–584. doi: 10.1109/ICCV.2015.73.
|
| [4] |
LIU Yating, WANG Xiaojie, NING Zhaolong, et al. A survey on semantic communications: Technologies, solutions, applications and challenges[J]. Digital Communications and Networks, 2024, 10(3): 528–545. doi: 10.1016/j.dcan.2023.05.010.
|
| [5] |
CHAI Jingxuan, XIAO Yong, and SHI Guangming. On the rate-distortion-complexity tradeoff for semantic communication[J]. IEEE Internet of Things Journal, 2026, 13(14): 31768–31781. doi: 10.1109/JIOT.2026.3689652.
|
| [6] |
陈阳, 马欢, 姬智, 等. 面向图像恢复任务的语义通信网络能耗优化[J]. 电子与信息学报, 2026, 48(1): 183–190. doi: 10.11999/JEIT250915.
CHEN Yang, MA Huan, JI Zhi, et al. Optimization of energy consumption in semantic communication networks for image recovery tasks[J]. Journal of Electronics & Information Technology, 2026, 48(1): 183–190. doi: 10.11999/JEIT250915.
|
| [7] |
AITHAL S K, MAINI P, LIPTON Z, et al. Understanding hallucinations in diffusion models through mode interpolation[C]. Proceedings of the 38th International Conference on Neural Information Processing Systems, Vancouver, Canada, 2024: 134614–134644. doi: 10.52202/079017-4278.
|
| [8] |
JIA Zhaoyang, LI Jiahao, LI Bin, et al. Generative latent coding for ultra-low bitrate image compression[C]. Proceedings of the 2024 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Seattle, USA, 2024: 26088–26098. doi: 10.1109/CVPR52733.2024.02465.
|
| [9] |
PEZONE F, MUSA O, CAIRE G, et al. Semantic-preserving image coding based on conditional diffusion models[C]. ICASSP 2024–2024 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), Seoul, Korea, 2024: 13501–13505. doi: 10.1109/ICASSP48485.2024.10447279.
|
| [10] |
ZHANG Chaoning, CHO J, PUSPITASARI F D, et al. A survey on Segment Anything Model (SAM): Vision foundation model meets prompt engineering[EB/OL]. arXiv preprint arXiv: 2306.06211. https://arxiv.org/abs/2306.06211, 2023.
|
| [11] |
KINGMA D P and WELLING M. Auto-encoding variational Bayes[EB/OL]. arXiv preprint arXiv: 1312.6114. https://arxiv.org/abs/1312.6114, 2013.
|
| [12] |
BURGESS C P, HIGGINS I, PAL A, et al. Understanding disentangling in β-VAE[EB/OL]. arXiv preprint arXiv: 1804.03599. https://arxiv.org/abs/1804.03599?context=cs.LG#1, 2018.
|
| [13] |
罗一畅, 齐析屿, 张博锐, 等. 分割一切模型的轻量化研究综述[J]. 电子与信息学报, 2026, 48(2): 713–731. doi: 10.11999/JEIT250894.
LUO Yichang, QI Xiyu, ZHANG Borui, et al. A survey of lightweight techniques for segment anything model[J]. Journal of Electronics & Information Technology, 2026, 48(2): 713–731. doi: 10.11999/JEIT250894.
|
| [14] |
TAPPAREL J, AFISIADIS O, MAYORAZ P, et al. An open-source LoRa physical layer prototype on GNU radio[C]. Proceedings of the 2020 IEEE 21st International Workshop on Signal Processing Advances in Wireless Communications (SPAWC), Atlanta, USA, 2020: 1–5. doi: 10.1109/SPAWC48557.2020.9154273.
|
| [15] |
SUN Xiaorui, LIU Jun, SHEN Hengtao, et al. On efficient variants of Segment Anything Model: A survey[J]. International Journal of Computer Vision, 2025, 133(10): 7406–7436. doi: 10.1007/s11263-025-02539-8.
|
| [16] |
RADFORD A, KIM J W, HALLACY C, et al. Learning transferable visual models from natural language supervision[C]. Proceedings of the 38th International Conference on Machine Learning, 2021: 8748–8763. (查阅网上资料, 未找见本条文献出版地信息, 请确认).
|
| [17] |
PARK T, LIU Mingyu, WANG Tingchun, et al. Semantic image synthesis with spatially-adaptive normalization[C]. Proceedings of the 2019 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Long Beach, USA, 2019: 2332–2341. doi: 10.1109/CVPR.2019.00244.
|
| [18] |
BLAU Y and MICHAELI T. The perception-distortion tradeoff[C]. Proceedings of the 2018 IEEE/CVF Conference on Computer Vision and Pattern Recognition, Salt Lake City, USA, 2018: 6228–6237. doi: 10.1109/CVPR.2018.00652.
|
| [19] |
JIANG Wei, ZHAI Yongqi, LI Hangyu, et al. Learned image compression with ROI-weighted distortion and bit allocation[EB/OL]. CoRR. https://arxiv.org/html/2401.08154v2, 2024.
|
| [20] |
MAO Qi, YANG Tinghan, ZHANG Yinuo, et al. Extreme image compression using fine-tuned VQGANs[C]. Proceedings of the 2024 Data Compression Conference (DCC), Snowbird, USA, 2024: 203–212. doi: 10.1109/DCC58796.2024.00028.
|
| [21] |
CHEN Weilong, XU Wenxuan, CHEN Haoran, et al. Semantic communication based on large language model for underwater image transmission[J]. IEEE Transactions on Mobile Computing, 2026, 25(2): 2060–2075. doi: 10.1109/TMC.2025.3607717.
|