Towards an AI-Enabled Metaverse: Architecture, Resource Allocation, and Multimodal Benchmarking

dc.contributor.authorLong, Zijian
dc.contributor.supervisorEl Saddik, Abdulmotaleb
dc.contributor.supervisorDong, Haiwei
dc.date.accessioned2026-07-16T13:14:52Z
dc.date.issued2026-07-16
dc.description.abstractThe metaverse, envisioned as a paradigm for next-generation digital environments, remains fundamentally constrained by isolated, pre-defined, and reactive virtual systems. Artificial intelligence (AI) provides a conceptual and technical foundation for overcoming these limitations by enabling persistent perception, autonomous decision-making, and adaptive coordination. This thesis investigates how AI can be systematically embedded into metaverse systems from a system-level perspective. It begins with a comprehensive analysis of metaverse network traffic characteristics, examining the necessity, performance implications, and trade-offs between remote rendering and local rendering. Building on these empirical insights, the thesis proposes a unified thing–edge–cloud architecture that coordinates heterogeneous intelligent agents operating across network layers. Within this architecture, the thesis addresses two core challenges in AI-enabled metaverse systems: adaptive resource allocation and cognitive scene understanding. At the edge layer, adaptive bandwidth allocation for immersive streaming is formulated as a cooperative multi-agent decision-making problem under dynamic network conditions. A Multi-Agent Soft Actor Critic (MASAC)–based strategy is developed and yields consistent improvements in user Quality of Experience (QoE) of at least 14\% compared with representative streaming baselines. At the cloud layer, the thesis introduces a metaverse-oriented benchmarking framework for evaluating the scene understanding capabilities of Multimodal Large Language Models (MLLMs). Thirteen representative MLLMs are assessed through a pairwise comparison protocol under an MLLM-as-a-judge paradigm guided by metaverse-specific semantic criteria. Human expert validation confirms the robustness of the proposed benchmark, achieving an agreement rate of 87.6\% with human consensus. Overall, this thesis establishes an integrated methodological and architectural foundation for the design, optimization, and evaluation of AI-enabled metaverse systems.
dc.identifier.urihttp://hdl.handle.net/10393/51853
dc.identifier.urihttps://doi.org/10.20381/ruor-32092
dc.language.isoen
dc.publisherUniversité d'Ottawa | University of Ottawa
dc.subjectMetaverse
dc.subjectEdge resource allocation
dc.subjectLarge language model
dc.subjectDeep reinforcement learning
dc.titleTowards an AI-Enabled Metaverse: Architecture, Resource Allocation, and Multimodal Benchmarking
dc.typeThesisen
thesis.degree.disciplineGénie / Engineering
thesis.degree.levelDoctoral
thesis.degree.namePhD
uottawa.departmentScience informatique et génie électrique / Electrical Engineering and Computer Science

Fichiers

Trousse originale

Voici les éléments 1 - 1 sur 1
En cours de chargement...
Vignette d'image
Nom:
Long_Zijian_2026_thesis.pdf
Taille:
5.78 MB
Format:
Adobe Portable Document Format

Trousse de licence

Voici les éléments 1 - 1 sur 1
En cours de chargement...
Vignette d'image
Nom:
license.txt
Taille:
2.51 KB
Format:
Item-specific license agreed upon to submission
Description: