| 2024 | Fine-Grained CUDA Call Interception for GPU Virtualization. | Manos Pavlidakis, Anargyros Argyros, Stelios Mavridis, Giorgos Vasiliadis, Angelos Bilas |
| 2024 | FDPVirt: Investigating the Behavior of FDP SSDs. | Joonyeop Park, Hyeonsang Eom |
| 2024 | Grid-aware Energy Management by Data Center Workload Control Across Multiple Data Centers. | Yoji Ozawa, Satoshi Kaneko, Taku Okamura, Tetsu Moriwake, Yuichi Nabetani, Daiki Shimizu, Soichiro Kumagai |
| 2024 | Cannikin: Optimal Adaptive Distributed DNN Training over Heterogeneous Clusters. | Chengyi Nie, Jessica Maghakian, Zhenhua Liu |
| 2024 | Dexter: A Performance-Cost Efficient Resource Allocation Manager for Serverless Data Analytics. | Anna Maria Nestorov, Diego Marrn, Alberto Gutierrez-Torre, Chen Wang, Claudia Misale, Alaa Youssef, David Carrera, Josep Llus Berral |
| 2024 | Near-Storage Processing in FaaS Environments with Funclets. | Alan Nair, Raven Szewczyk, Donald Jennings, Antonio Barbalace |
| 2024 | UTwinVM: Reliable hints on the effects of hypervisor updates on VMs in the Cloud. | Djob Mvondo, Tong Xing, Antonio Barbalace |
| 2024 | HORSE: Ultra-low latency workloads on FaaS platforms. | Djob Mvondo, Franois Taani, Yrom-David Bromberg |
| 2024 | Optimal Resource Efficiency with Fairness in Heterogeneous GPU Clusters. | Zizhao Mo, Huanle Xu, Wing Cheong Lau |
| 2024 | L3: Latency-aware Load Balancing in Multi-Cluster Service Mesh. | Olivier Michaelis, Stefan Schmid, Habib Mostafaei |
| 2024 | Deep Optimizer States: Towards Scalable Training of Transformer Models using Interleaved Offloading. | Avinash Maurya, Jie Ye, M. Mustafa Rafique, Franck Cappello, Bogdan Nicolae |
| 2024 | ColdPurge: Effecient Metadata Cache Cleaning via Accurate Online Data Hotness Tracking. | Yuhang Li, Rong Gu, Simian Li, Baohan Wang, Wei Yu |
| 2024 | Enhancing Effective Bidirectional Isolation for Function Fusion in Serverless Architectures. | Tianyu Li, Yingpeng Chen, Donghui Yu, Yuanyuan Zhang, Bert Lagaisse |
| 2024 | Towards Interoperability of APIs - an LLM-based approach. | Ren Lehmann |
| 2024 | An LLM-driven Framework for Dynamic Infrastructure as Code Generation. | Junhee Lee, Sungjoo Kang, In-Young Ko |
| 2024 | Targeting Tail Latency in Replicated Systems with Proactive Rejection. | Laura Lawniczak, Tobias Distler |
| 2024 | Creek: A Mixed-Consistency Replication Scheme. | Maciej Kokocinski, Tadeusz Kobus, Pawel T. Wojciechowski |
| 2024 | Towards Embracing Object Granularity in Tiered Memory Management for Big Data. | Maciej Kokocinski, Tadeusz Kobus, Krystian Chmielewski, Rafal Pyzik, Maciej Maciejewski |
| 2024 | KDB: A Persistent Key-Value Data Store with Batch Updates and Snapshots. | Tadeusz Kobus, Maciej Kokocinski, Krzysztof Kortas, Pawel T. Wojciechowski |
| 2024 | AsyncFilter: Detecting Poisoning Attacks in Asynchronous Federated Learning. | Yufei Kang, Baochun Li |
| 2024 | Ripple: Large-Scale Service and Configuration Management in the Cloud. | Shuping Ji, Zhen Tang, Wei Wang, Hui Li, Jianguo Yao, Hans-Arno Jacobsen |
| 2024 | Ripple: Large-Scale Service and Configuration Management in the Cloud. | Shuping Ji, Zhen Tang, Wei Wang, Hui Li, Jianguo Yao, Hans-Arno Jacobsen |
| 2024 | Menos: Split Fine-Tuning Large Language Models with Efficient GPU Memory Sharing. | Chenghao Hu, Baochun Li |
| 2024 | Consensus-Agnostic State-Machine Replication. | Alexander He, Franz J. Hauck, Echo Meiner |
| 2024 | On the Semantic Overlap of Operators in Stream Processing Engines. | Vincenzo Gulisano, Marina Papatriantafilou, Alessandro Margara |