Skip to content

ACM International Conference on Supercomputing

ICS

A

CORE rank

CORE rank (raw)

A

Fields of research

Computer Systems Engineering · Distributed Computing and Systems Software

Papers indexed

1,987

1987–2026

Papers per year

1987135 peak2026

ICS papers

1,987 records sourced from DBLP. Search titles, filter by year, sort by recency.

YearTitleAuthors
2010An experimental approach to performance measurement of heterogeneous parallel applications using CUDA.Allen D. Malony, Scott Biersdorff, Wyatt Spear, Shangkar Mayanglambam
2010Fast and accurate NCBI BLASTP: acceleration with multiphase FPGA-based prefiltering.Atabak Mahram, Martin C. Herbordt
2010A compiler-automated array compression scheme for optimizing memory intensive programs.Lixia Liu, Zhiyuan Li
2010The auction: optimizing banks usage in Non-Uniform Cache Architectures.Javier Lira, Carlos Molina, Antonio Gonzlez
2010High-throughput Bayesian network learning using heterogeneous multicore computers.Michael D. Linderman, Robert V. Bruggner, Vivek Athalye, Teresa H. Meng, Narges Bani Asadi, Garry P. Nolan
2010Optimal bucket algorithms for large MPI collectives on torus interconnects.Nikhil Jain, Yogish Sabharwal
2010The next-generation supercomputer project and a plan for the advanced institute for computational science.Kimihiko Hirao
2010An empirically tuned 2D and 3D FFT library on CUDA GPU.Liang Gu, Xiaoming Li, Jakob Siegel
2010SAMS multi-layout memory: providing multiple views of data to boost SIMD performance.Chunyang Gou, Georgi Kuzmanov, Georgi Gaydadjiev
2010Clustering performance data efficiently at massive scales.Todd Gamblin, Bronis R. de Supinski, Martin Schulz, Robert J. Fowler, Daniel A. Reed
2010FPGA accelerating double/quad-double high precision floating-point applications for ExaScale computing.Yong Dou, Yuanwu Lei, Guiming Wu, Song Guo, Jie Zhou, Li Shen
2010Throughput computing.William J. Dally
2010Evaluation of parallel H.264 decoding strategies for the Cell Broadband Engine.Chi Ching Chi, Ben H. H. Juurlink, Cor Meenderinck
2010Large-scale FFT on GPU clusters.Yifeng Chen, Xiang Cui, Hong Mei
2010Static reuse distances for locality-based optimizations in MATLAB.Arun Chauhan, Chun-Yu Shei
2010Indemics: an interactive data intensive framework for high performance epidemic simulation.Keith R. Bisset, Jiangzhuo Chen, Xizhou Feng, Yifei Ma, Madhav V. Marathe
2010An approach to resource-aware co-scheduling for CMPs.Major Bhadauria, Sally A. McKee
2010Decomposable and responsive power models for multicore processors using performance counters.Ramon Bertran, Marc Gonzlez, Xavier Martorell, Nacho Navarro, Eduard Ayguad
2010Making nested parallel transactions practical using lightweight hardware support.Woongki Baek, Nathan Grasso Bronson, Christos Kozyrakis, Kunle Olukotun
2010Untitled recordNarges Bani Asadi, Christopher W. Fletcher, Greg Gibeling, John Wawrzynek, Wing H. Wong, Garry P. Nolan
2009Divide-and-conquer: a bubble replacement for low level caches.Chuanjun Zhang, Bing Xue
2009Less reused filter: improving l2 cache performance via filtering less reused lines.Lingxiang Xiang, Tianzhou Chen, Qingsong Shi, Wei Hu
2009Combining thread level speculation helper threads and runahead execution.Polychronis Xekalakis, Nikolas Ioannou, Marcelo Cintra
2009Dynamic parallelization of single-threaded binary programs using speculative slicing.Cheng Wang, Youfeng Wu, Edson Borin, Shiliang Hu, Wei Liu, Dave Sager, Tin-Fook Ngai, Jesse Fang
2009Practice of parallelizing network applications on multi-core architectures.Junchang Wang, Haipeng Cheng, Bei Hua, Xinan Tang
901925 of 1,987← PreviousNext →

Comparable venues

Other A*/A conferences filed under the same field of research.