File Download

There are no files associated with this item.

  Links for fulltext
     (May Require Subscription)
Supplementary

Article: Toward Scalable Generative AI via Mixture of Experts in Mobile Edge Networks

TitleToward Scalable Generative AI via Mixture of Experts in Mobile Edge Networks
Authors
Issue Date2024
Citation
IEEE Wireless Communications, 2024 How to Cite?
AbstractThe advancement of generative artificial intelligence (GAI) has driven revolutionary applications like ChatGPT. The widespread use of these applications relies on a mixture of experts (MoE), which contains multiple experts, and selectively engages them for each task to lower operation costs while maintaining performance. Despite MoE, GAI faces challenges in resource consumption when deployed on user devices. Hence, this article proposes mobile edge networks supported MoE-based GAI. We first review the MoE from traditional AI and GAI perspectives, including structure, principles, and applications. We then propose a framework that transfers subtasks to experts in mobile edge networks, aiding GAI model operation on user devices. We discuss challenges in this process and introduce a deep reinforcement learning-based algorithm to select edge experts for subtask execution. Experimental results show that our framework not only facilitates GAI's deployment on resource-limited devices, but also generates higher-quality content compared to methods without edge network support.
Persistent Identifierhttp://hdl.handle.net/10722/353222
ISSN
2023 Impact Factor: 10.9
2023 SCImago Journal Rankings: 5.926
ISI Accession Number ID

 

DC FieldValueLanguage
dc.contributor.authorWang, Jiacheng-
dc.contributor.authorDu, Hongyang-
dc.contributor.authorNiyato, Dusit-
dc.contributor.authorKang, Jiawen-
dc.contributor.authorXiong, Zehui-
dc.contributor.authorKim, Dong In-
dc.contributor.authorLetaief, Khaled B.-
dc.date.accessioned2025-01-13T03:02:43Z-
dc.date.available2025-01-13T03:02:43Z-
dc.date.issued2024-
dc.identifier.citationIEEE Wireless Communications, 2024-
dc.identifier.issn1536-1284-
dc.identifier.urihttp://hdl.handle.net/10722/353222-
dc.description.abstractThe advancement of generative artificial intelligence (GAI) has driven revolutionary applications like ChatGPT. The widespread use of these applications relies on a mixture of experts (MoE), which contains multiple experts, and selectively engages them for each task to lower operation costs while maintaining performance. Despite MoE, GAI faces challenges in resource consumption when deployed on user devices. Hence, this article proposes mobile edge networks supported MoE-based GAI. We first review the MoE from traditional AI and GAI perspectives, including structure, principles, and applications. We then propose a framework that transfers subtasks to experts in mobile edge networks, aiding GAI model operation on user devices. We discuss challenges in this process and introduce a deep reinforcement learning-based algorithm to select edge experts for subtask execution. Experimental results show that our framework not only facilitates GAI's deployment on resource-limited devices, but also generates higher-quality content compared to methods without edge network support.-
dc.languageeng-
dc.relation.ispartofIEEE Wireless Communications-
dc.titleToward Scalable Generative AI via Mixture of Experts in Mobile Edge Networks-
dc.typeArticle-
dc.description.naturelink_to_subscribed_fulltext-
dc.identifier.doi10.1109/MWC.003.2400046-
dc.identifier.scopuseid_2-s2.0-85207009601-
dc.identifier.eissn1558-0687-
dc.identifier.isiWOS:001336046300001-

Export via OAI-PMH Interface in XML Formats


OR


Export to Other Non-XML Formats