ACM International Conference on Supercomputing
ICS
A
CORE rank
CORE rank (raw)
A
Fields of research
Computer Systems Engineering · Distributed Computing and Systems Software
Papers indexed
1,987
1987–2026
Papers per year
1987135 peak2026
Most published authors
ICS papers
1,987 records sourced from DBLP. Search titles, filter by year, sort by recency.
| Year | Title | Authors |
|---|---|---|
| 2023 | Parallel Software for Million-scale Exact Kernel Regression. | Yu Chen, Lucca Skon, James R. McCombs, Zhenming Liu, Andreas Stathopoulos |
| 2023 | HEAT: A Highly Efficient and Affordable Training System for Collaborative Filtering Based Recommendation on CPUs. | Chengming Zhang, Shaden Smith, Baixi Sun, Jiannan Tian, Jonathan Soifer, Xiaodong Yu, Shuaiwen Leon Song, Yuxiong He, Dingwen Tao |
| 2023 | SPARTA: Spatial Acceleration for Efficient and Scalable Horizontal Diffusion Weather Stencil Computation. | Gagandeep Singh, Alireza Khodamoradi, Kristof Denolf, Jack Lo, Juan Gmez-Luna, Joseph Melber, Andra Bisca, Henk Corporaal, Onur Mutlu |
| 2022 | Low overhead and context sensitive profiling of CPU-accelerated applications. | Keren Zhou, Jonathon M. Anderson, Xiaozhu Meng, John M. Mellor-Crummey |
| 2022 | PAME: precision-aware multi-exit DNN serving for reducing latencies of batched inferences. | Shulai Zhang, Weihao Cui, Quan Chen, Zhengnian Zhang, Yue Guan, Jingwen Leng, Chao Li, Minyi Guo |
| 2022 | LITE: a low-cost practical inter-operable GPU TEE. | Ardhi Wiratama Baskara Yudha, Jake Meyer, Shougang Yuan, Huiyang Zhou, Yan Solihin |
| 2022 | Dense dynamic blocks: optimizing SpMM for processors with vector and matrix units using machine learning techniques. | Serif Yesil, Jos E. Moreira, Josep Torrellas |
| 2022 | Software-defined floating-point number formats and their application to graph processing. | Hans Vandierendonck |
| 2022 | Fast-track cache: a huge racetrack memory L1 data cache. | Hugo Trrega, Alejandro Valero, Vicente Lorente, Salvador Petit, Julio Sahuquillo |
| 2022 | ASAP: automatic synthesis of area-efficient and precision-aware CGRAs. | Cheng Tan, Thierry Tambe, Jeff Jun Zhang, Bo Fang, Tong Geng, Gu-Yeon Wei, David Brooks, Antonino Tumeo, Ganesh Gopalakrishnan, Ang Li |
| 2022 | Rethinking graph data placement for graph neural network training on multiple GPUs. | Shihui Song, Peng Jiang |
| 2022 | Beyond time complexity: data movement complexity analysis for matrix multiplication. | Wesley Smith, Aidan Goldfarb, Chen Ding |
| 2022 | KrakenOnMem: a memristor-augmented HW/SW framework for taxonomic profiling. | Taha Shahroodi, Mahdi Zahedi, Abhairaj Singh, Stephan Wong, Said Hamdioui |
| 2022 | Performance-detective: automatic deduction of cheap and accurate performance models. | Larissa Schmid, Marcin Copik, Alexandru Calotoiu, Dominik Werle, Andreas Reiter, Michael Selzer, Anne Koziolek, Torsten Hoefler |
| 2022 | A data-centric optimization framework for machine learning. | Oliver Rausch, Tal Ben-Nun, Nikoli Dryden, Andrei Ivanov, Shigang Li, Torsten Hoefler |
| 2022 | Dynamic memory management in massively parallel systems: a case on GPUs. | Minh Pham, Hao Li, Yongke Yuan, Chengcheng Mou, Kandethody Ramachandran, Zichen Xu, Yicheng Tu |
| 2022 | SnuQS: scaling quantum circuit simulation using storage devices. | Daeyoung Park, Heehoon Kim, Jinpyo Kim, Taehyun Kim, Jaejin Lee |
| 2022 | Efficient, out-of-memory sparse MTTKRP on massively parallel architectures. | Andy Nguyen, Ahmed E. Helal, Fabio Checconi, Jan Laukemann, Jesmin Jahan Tithi, Yongseok Soh, Teresa M. Ranadive, Fabrizio Petrini, Jee W. Choi |
| 2022 | AnySeq/GPU: a novel approach for faster sequence alignment on GPUs. | Andr Mller, Bertil Schmidt, Richard Membarth, Roland Leia, Sebastian Hack |
| 2022 | Efficiently emulating high-bitwidth computation with low-bitwidth hardware. | Zixuan Ma, Haojie Wang, Guanyu Feng, Chen Zhang, Lei Xie, Jiaao He, Shengqi Chen, Jidong Zhai |
| 2022 | Seamless optimization of the GEMM kernel for task-based programming models. | Arthur Francisco Lorenzon, Sandro Matheus V. N. Marques, Antoni C. Navarro, Vicen Beltran |
| 2022 | Toward accelerated stencil computation by adapting tensor core unit on GPU. | Xiaoyan Liu, Yi Liu, Hailong Yang, Jianjin Liao, Mingzhen Li, Zhongzhi Luan, Depei Qian |
| 2022 | Towards low-latency I/O services for mixed workloads using ultra-low latency SSDs. | Mingzhe Liu, Haikun Liu, Chencheng Ye, Xiaofei Liao, Hai Jin, Yu Zhang, Ran Zheng, Liting Hu |
| 2022 | Cloak: tolerating non-volatile cache read latency. | Apostolos Kokolis, Namrata Mantri, Shrikanth Ganapathy, Josep Torrellas, John Kalamatianos |
| 2022 | SnuHPL: high performance LINPACK for heterogeneous GPUs. | Jinpyo Kim, Hyungdal Kwon, Jintaek Kang, Jihwan Park, Seungwook Lee, Jaejin Lee |
301–325 of 1,987← PreviousNext →
Comparable venues
Other A*/A conferences filed under the same field of research.
- ADATEDesign, Automation and Test in Europe Conference
- A*DACDesign Automation Conf
- ASCInternational Conference for High Performance Computing, Networking, Storage and Analysis (was Supercomputing Conference)
- AICCADIEEE/ACM International Conference on Computer-Aided Design
- AITCIEEE International Test Conference
- A*ISCAACM International Symposium on Computer Architecture
- A*MICROInternational Symposium on Microarchitecture
- AISLPEDIEEE/ACM International Symposium on Low Power Electronics and Design