Large Language Models (LLMs) Inference Offloading and Resource Allocation in Cloud-Edge Networks: An Active Inference Approach.
Jingcheng Fang, Ying He, F. Richard Yu, Jianqiang Li, Victor C. M. Leung
Browse the full VTC paper archive.
Jingcheng Fang, Ying He, F. Richard Yu, Jianqiang Li, Victor C. M. Leung
Browse the full VTC paper archive.