MoE-Inference-Bench: Performance Evaluation of Mixture of Expert Large Language and Vision Models.
Krishna Teja Chitty-Venkata, Sylvia Howland, Golara Azar, Daria Soboleva, Natalia Vassilieva, Siddhisanket Raskar, Murali Emani, Venkatram Vishwanath
Browse the full SC paper archive.