arXiv:2607. 09183v1 Announce Type: cross Abstract: The groundbreaking development of generative artificial intelligence (AI) is rapidly boosting the ability to generate content such as images and videos, reshaping communication paradigms.
By Wenjun Zhang, Zhiyong Chen, Tong Wu, Guo Lu, Li Song, Feng Yang, Meixia Tao
arXiv:2609.10714v1 Announce Type: cross
Abstract: The ambitious requirements of sixth-generation (6G) networks are driving communication systems from reliable bit delivery toward meaning-aware and ta...
By Yu Ma, Zhen Gao, Li Qiao, Xiaoyuan Zhang, Mahdi Boloursaz Mashhadi, Yin Xu, Wenjun Xu, Xiaodong Xu, Kaibin Huang, Jiangzhou Wang, Rahim Tafazolli, Sheng Chen, Tony Q. S. Quek, Ping Zhang
arXiv:2610.01826v1 Announce Type: cross
Abstract: Collaborative embodied artificial intelligence (CEAI) enables multiple physical agents to perceive, reason, and act cooperatively in dynamic environm...
By Peng Yi, Ying-Chang Liang
The article discusses how vision generative AI models, while rapidly advancing, have largely been developed with a focus on output quality, leading to hardware that adapts reactively to increasing model demands. It evaluates the parameter cost and energy efficiency of these models across various accelerator platforms and aligns four generative model families with seven real-world application domains. The authors propose a software‑hardware co‑design strategy that considers deployment constraints from the outset, ensuring that the appropriate model runs on suitable hardware for specific applications, thereby making generative AI deployment more sustainable and widely accessible.
By Eleni Tselepi, Cristian Sestito, Shady Agwa, Themis Prodromakis
arXiv:2608. 14600v1 Announce Type: cross Abstract: We present a demonstration for generative multicasting with on-device, intent-aware semantic decomposition.
By Xinkai Liu, Mahdi Boloursaz Mashhadi, Yi Ma, Rahim Tafazolli
arXiv:2607. 27372v1 Announce Type: new Abstract: The deep learning revolution, kicked off by AlexNet, taught us that end-to-end training beats decomposing a problem into hand-designed stages.
By Alexi Gladstone, Heng Ji, Yilun Du
arXiv:2608. 08101v1 Announce Type: new Abstract: Generative AI has emerged as one of the most transformative forces in modern artificial intelligence, reshaping how we create, imagine, and interact with digital content.
By Jun Lu
The paper introduces UniSandbox, a decoupled evaluation framework with controlled synthetic datasets, to study whether understanding informs generation in Unified Multimodal Models. Results show a notable understanding‑generation gap, especially in reasoning generation and knowledge transfer. Explicit Chain‑of‑Thought (CoT) in the understanding module bridges this gap, and self‑training can internalize CoT for implicit reasoning during generation; query‑based architectures also exhibit latent CoT‑like properties that aid knowledge transfer.
By Yuwei Niu, Weiyang Jin, Jiaqi Liao, Chaoran Feng, Peng Jin, Bin Lin, Zongjian Li, Bin Zhu, Weihao Yu, Li Yuan
arXiv:2407. 09013v2 Announce Type: replace Abstract: The attempt to utilize machine learning in PCG has been made in the past.
By Xinyu Mao, Wanli Yu, Kazunori D Yamada, Michael R. Zielewski
The paper introduces FMSGOC, a framework that leverages visual‑linguistic foundation models to improve semantic and goal‑oriented communication for 6G. By transmitting a sparse set of semantic anchors and using a pretrained diffusion model for masked completion, it reduces overfitting and achieves high rate efficiency, reaching 0.039 BPP while maintaining strong semantic fidelity and robustness on unseen data.
By Boliang Liu, Wint Yi Poe, Riccardo Trivisonno, Giuseppe Caire
arXiv:2608. 14694v1 Announce Type: new Abstract: Foundation models are emerging as a transformative paradigm for AI-native sixth-generation (6G) wireless networks by enabling scalable, transferable, and data-efficient intelligence across diverse communication tasks.
By Naveed Khan, Besan Al Sbeihi, Maryam Alshehhi, Nasir Saeed
arXiv:2604. 20329v3 Announce Type: replace-cross Abstract: Recent works show that image and video generators exhibit zero-shot visual understanding behaviors, in a way reminiscent of how LLMs develop emergent capabilities of language understanding and reasoning from generative pretraining.
By Valentin Gabeur, Shangbang Long, Songyou Peng, Paul Voigtlaender, Shuyang Sun, Yanan Bao, Karen Truong, Zhicheng Wang, Wenlei Zhou, Jonathan T. Barron, Kyle Genova, Nithish Kannen, Sherry Ben, Yandong Li, Mandy Guo, Suhas Yogin, Yiming Gu, Huizhong Chen, Oliver Wang, Saining Xie, Howard Zhou, Kaiming He, Thomas Funkhouser, Jean-Baptiste Alayrac, Radu Soricut