Skip to content

P. Sadayappan

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

179

Venues

30

Active years

1985–2025

Best venue rank

A*

Where they publish

Papers

179 indexed papers, newest first.

YearVenueTitleAuthors
2025SCFaSTCC: Fast Sparse Tensor Contractions on CPUs.Saurabh Raje, Hunter McCoy, Atanas Rountev, Prashant Pandey, P. Sadayappan
2025SCGuiding Application Users via Estimation of Computational Resources for Massively Parallel Chemistry Computations.Tanzila Tabassum, Omer Subasi, Ajay Panyala, Epiya Ebiapia, Gerald Baumgartner, Erdal Mutlu, P. Sadayappan, Karol Kowalski
2025SCDistributed Sparse Tensor Computations in MLIR.Miheer Vaidya, Shreya Singh, Devanshu Mantri, Michael Shannon Eydenberg, Brian Michael Kelley, Sivasankaran Rajamanickam, Atanas Rountev, P. Sadayappan
2023ICSScalable parallelization for the solution of phonon Boltzmann Transport Equation.Han D. Tran, Siddharth Saurav, P. Sadayappan, Sandip Mazumder, Hari Sundar
2023PPoPPTDC: Towards Extremely Efficient CNNs on GPUs via Hardware-Aware Tucker Decomposition.Lizhi Xiang, Miao Yin, Chengming Zhang, Aravind Sukumaran-Rajam, P. Sadayappan, Bo Yuan, Dingwen Tao
2023SCAutomatic Generation of Distributed-Memory Mappings for Tensor Computations.Martin Kong, Raneem Abu Yosef, Atanas Rountev, P. Sadayappan
2022CCTraining of deep learning pipelines on memory-constrained GPUs via segmented fused-tiled execution.Yufan Xu, Saurabh Raje, Atanas Rountev, Gerald Sabin, Aravind Sukumaran-Rajam, P. Sadayappan
2022CGOComprehensive Accelerator-Dataflow Co-design Optimization for Convolutional Neural Networks.Miheer Vaidya, Aravind Sukumaran-Rajam, Atanas Rountev, P. Sadayappan
2021ASPLOSAnalytical characterization and design space exploration for optimization of CNNs.Rui Li, Yufan Xu, Aravind Sukumaran-Rajam, Atanas Rountev, P. Sadayappan
2021PLDIIOOpt: automatic derivation of I/O complexity bounds for affine programs.Auguste Olivry, Guillaume Iooss, Nicolas Tollenaere, Atanas Rountev, P. Sadayappan, Fabrice Rastello
2021SPAAEfficient Distributed Algorithms for Convolutional Neural Networks.Rui Li, Yufan Xu, Aravind Sukumaran-Rajam, Atanas Rountev, P. Sadayappan
2020KDDALO-NMF: Accelerated Locality-Optimized Non-negative Matrix Factorization.Gordon Euhyun Moon, J. Austin Ellis, Aravind Sukumaran-Rajam, Srinivasan Parthasarathy, P. Sadayappan
2020PLDIAutomated derivation of parametric data movement lower bounds for affine programs.Auguste Olivry, Julien Langou, Louis-Nol Pouchet, P. Sadayappan, Fabrice Rastello
2020SCScalable heterogeneous execution of a coupled-cluster model with perturbative triples.Jinsung Kim, Ajay Panyala, Bo Peng, Karol Kowalski, P. Sadayappan, Sriram Krishnamoorthy
2020SCEfficient tiled sparse matrix multiplication through matrix signatures.Sreyya Emre Kurt, Aravind Sukumaran-Rajam, Fabrice Rastello, P. Sadayappan
2019AAAIATP: Directed Graph Embedding with Asymmetric Transitivity Preservation.Jiankai Sun, Bortik Bandyopadhyay, Armin Bashizade, Jiongqian Liang, P. Sadayappan, Srinivasan Parthasarathy
2019CGOA Code Generator for High-Performance Tensor Contractions on GPUs.Jinsung Kim, Aravind Sukumaran-Rajam, Vineeth Thumma, Sriram Krishnamoorthy, Ajay Panyala, Louis-Nol Pouchet, Atanas Rountev, P. Sadayappan
2019PPoPPAdaptive sparse tiling for sparse matrix multiplication.Changwan Hong, Aravind Sukumaran-Rajam, Israt Nisa, Kunal Singh, P. Sadayappan
2019SCAnalytical cache modeling and tilesize optimization for tensor contractions.Rui Li, Aravind Sukumaran-Rajam, Richard Veras, Tze Meng Low, Fabrice Rastello, Atanas Rountev, P. Sadayappan
2019SCParallel Data-Local Training for Optimizing Word2Vec Embeddings for Word and Graph Embeddings.Gordon Euhyun Moon, Denis Newman-Griffis, Jinsung Kim, Aravind Sukumaran-Rajam, Eric Fosler-Lussier, P. Sadayappan
2019SCAn efficient mixed-mode representation of sparse tensors.Israt Nisa, Jiajia Li, Aravind Sukumaran-Rajam, Prashant Singh Rawat, Sriram Krishnamoorthy, P. Sadayappan
2018HiPCSampled Dense Matrix Multiplication for High-Performance Machine Learning.Israt Nisa, Aravind Sukumaran-Rajam, Sreyya Emre Kurt, Changwan Hong, P. Sadayappan
2018HPDCEfficient sparse-matrix multi-vector product on GPUs.Changwan Hong, Aravind Sukumaran-Rajam, Bortik Bandyopadhyay, Jinsung Kim, Sreyya Emre Kurt, Israt Nisa, Shivani Sabhlok, mit V. atalyrek, Srinivasan Parthasarathy, P. Sadayappan
2018ICCSParallel Latent Dirichlet Allocation on GPUs.Gordon Euhyun Moon, Israt Nisa, Aravind Sukumaran-Rajam, Bortik Bandyopadhyay, Srinivasan Parthasarathy, P. Sadayappan
2018ICSOptimizing Tensor Contractions in CCSD(T) for Efficient Execution on GPUs.Jinsung Kim, Aravind Sukumaran-Rajam, Changwan Hong, Ajay Panyala, Rohit Kumar Srivastava, Sriram Krishnamoorthy, P. Sadayappan
2018PLDIGPU code optimization using abstract kernel emulation and sensitivity analysis.Changwan Hong, Aravind Sukumaran-Rajam, Jinsung Kim, Prashant Singh Rawat, Sriram Krishnamoorthy, Louis-Nol Pouchet, Fabrice Rastello, P. Sadayappan
2018PPoPPPerformance modeling for GPUs using abstract kernel emulation.Changwan Hong, Aravind Sukumaran-Rajam, Jinsung Kim, Prashant Singh Rawat, Sriram Krishnamoorthy, Louis-Nol Pouchet, Fabrice Rastello, P. Sadayappan
2018PPoPPRegister optimizations for stencils on GPUs.Prashant Singh Rawat, Fabrice Rastello, Aravind Sukumaran-Rajam, Louis-Nol Pouchet, Atanas Rountev, P. Sadayappan
2018SCAssociative instruction reordering to alleviate register pressure.Prashant Singh Rawat, Aravind Sukumaran-Rajam, Atanas Rountev, Fabrice Rastello, Louis-Nol Pouchet, P. Sadayappan
2017HiPCCharacterization of Data Movement Requirements for Sparse Matrix Computations on GPUs.Sreyya Emre Kurt, Vineeth Thumma, Changwan Hong, Aravind Sukumaran-Rajam, P. Sadayappan
2017HiPCParallel LDA with Over-Decomposition.Gordon Euhyun Moon, Aravind Sukumaran-Rajam, P. Sadayappan
2017ICSOn improving performance of sparse matrix-matrix multiplication on GPUs.Rakshith Kunchum, Ankur Chaudhry, Aravind Sukumaran-Rajam, Qingpeng Niu, Israt Nisa, P. Sadayappan
2017PPoPPParallel CCD++ on GPU for Matrix Factorization.Israt Nisa, Aravind Sukumaran-Rajam, Rakshith Kunchum, P. Sadayappan
2017PPoPPOptimizing the Four-Index Integral Transform Using Data Movement Lower Bounds Analysis.Samyam Rajbhandari, Fabrice Rastello, Karol Kowalski, Sriram Krishnamoorthy, P. Sadayappan
2016CCRegister allocation and promotion through combined instruction scheduling and loop unrolling.Lukasz Domagala, Duco van Amstel, Fabrice Rastello, P. Sadayappan
2016CCOn fusing recursive traversals of K-d trees.Samyam Rajbhandari, Jinsung Kim, Sriram Krishnamoorthy, Louis-Nol Pouchet, Fabrice Rastello, Robert J. Harrison, P. Sadayappan
2016HiPCCompiler Support for Software Cache Coherence.Sanket Tavarageri, Wooil Kim, Josep Torrellas, P. Sadayappan
2016PLDIEffective padding of multidimensional arrays to avoid cache conflict misses.Changwan Hong, Wenlei Bao, Albert Cohen, Sriram Krishnamoorthy, Louis-Nol Pouchet, Fabrice Rastello, J. Ramanujam, P. Sadayappan
2016POPLPolyCheck: dynamic verification of iteration space transformations on affine programs.Wenlei Bao, Sriram Krishnamoorthy, Louis-Nol Pouchet, Fabrice Rastello, P. Sadayappan
2016PPoPPEffective resource management for enhancing performance of 2D and 3D stencils on GPUs.Prashant Singh Rawat, Changwan Hong, Mahesh Ravishankar, Vinod Grover, Louis-Nol Pouchet, P. Sadayappan
2016SCPIPES: a language and compiler for task-based programming on distributed-memory clusters.Martin Kong, Louis-Nol Pouchet, P. Sadayappan, Vivek Sarkar
2016SCA domain-specific compiler for a parallel multiresolution adaptive numerical simulation environment.Samyam Rajbhandari, Jinsung Kim, Sriram Krishnamoorthy, Louis-Nol Pouchet, Fabrice Rastello, Robert J. Harrison, P. Sadayappan
2016SPAABrief Announcement: Approximating the I/O Complexity of One-Shot Red-Blue Pebbling.Timothy Carpenter, Fabrice Rastello, P. Sadayappan, Anastasios Sidiropoulos
2015CGOCharacterizing and enhancing global memory data coalescing on GPUs.Naznin Fauzia, Louis-Nol Pouchet, P. Sadayappan
2015ICSOptimistic Delinearization of Parametrically Sized Arrays.Tobias Grosser, Jagannathan Ramanujam, Louis-Nol Pouchet, P. Sadayappan, Sebastian Pop
2015ICSAutomatic Selection of Sparse Matrix Representation on GPUs.Naser Sedaghati, Te Mu, Louis-Nol Pouchet, Srinivasan Parthasarathy, P. Sadayappan
2015POPLOn Characterizing the Data Access Complexity of Programs.Venmugil Elango, Fabrice Rastello, Louis-Nol Pouchet, J. Ramanujam, P. Sadayappan
2015PPoPPOn optimizing machine learning workloads via kernel fusion.Arash Ashari, Shirish Tatikonda, Matthias Boehm, Berthold Reinwald, Keith Campbell, John Keenleyside, P. Sadayappan
2015PPoPPDistributed memory code generation for mixed Irregular/Regular computations.Mahesh Ravishankar, Roshan Dathathri, Venmugil Elango, Louis-Nol Pouchet, J. Ramanujam, Atanas Rountev, P. Sadayappan
2015SCAn elegant sufficiency: load-aware differentiated scheduling of data transfers.Rajkumar Kettimuthu, Gayane Vardoyan, Gagan Agrawal, P. Sadayappan, Ian T. Foster
2015SCSDSLc: a multi-target domain-specific compiler for stencil computations.Prashant Singh Rawat, Martin Kong, Thomas Henretty, Justin Holewinski, Kevin Stock, Louis-Nol Pouchet, J. Ramanujam, Atanas Rountev, P. Sadayappan
2014CCGRIDModeling and Optimizing Large-Scale Wide-Area Data Transfers.Rajkumar Kettimuthu, Gayane Vardoyan, Gagan Agrawal, P. Sadayappan
2014CGOHybrid Hexagonal/Classical Tiling for GPUs.Tobias Grosser, Albert Cohen, Justin Holewinski, P. Sadayappan, Sven Verdoolaege
2014HiPCA fast implementation of MLR-MCL algorithm on multi-core processors.Qingpeng Niu, Pai-Wei Lai, S. M. Faisal, Srinivasan Parthasarathy, P. Sadayappan
2014ICPPCAST: Contraction Algorithm for Symmetric Tensors.Samyam Rajbhandari, Akshay Nikam, Pai-Wei Lai, Kevin Stock, Sriram Krishnamoorthy, P. Sadayappan
2014ICSAn efficient two-dimensional blocking strategy for sparse matrix-vector multiplication on GPUs.Arash Ashari, Naser Sedaghati, John Eisenlohr, P. Sadayappan
2014OOPSLAWOSC 2014: second workshop on optimizing stencil computations.Shoaib Kamil, Saman P. Amarasinghe, P. Sadayappan
2014PLDIA framework for enhancing data reuse via associative reordering.Kevin Stock, Martin Kong, Tobias Grosser, Louis-Nol Pouchet, Fabrice Rastello, J. Ramanujam, P. Sadayappan
2014PLDICompiler-assisted detection of transient memory errors.Sanket Tavarageri, Sriram Krishnamoorthy, P. Sadayappan
2014SCFast Sparse Matrix-Vector Multiplication on GPUs for Graph Applications.Arash Ashari, Naser Sedaghati, John Eisenlohr, Srinivasan Parthasarathy, P. Sadayappan
2014SCA Communication-Optimal Framework for Contracting Distributed Tensors.Samyam Rajbhandari, Akshay Nikam, Pai-Wei Lai, Kevin Stock, Sriram Krishnamoorthy, P. Sadayappan
2014SPAAOn characterizing the data movement complexity of computational DAGs for parallel execution.Venmugil Elango, Fabrice Rastello, Louis-Nol Pouchet, J. Ramanujam, P. Sadayappan
2013ASPLOSSplit tiling for GPUs: automatic parallelization using trapezoidal tiles.Tobias Grosser, Albert Cohen, Paul H. J. Kelly, J. Ramanujam, P. Sadayappan, Sven Verdoolaege
2013FPGAPolyhedral-based data reuse optimization for configurable computing.Louis-Nol Pouchet, Peng Zhang, P. Sadayappan, Jason Cong
2013ICDEStratification driven placement of complex data: A framework for distributed data analytics.Ye Wang, Srinivasan Parthasarathy, P. Sadayappan
2013ICSA stencil compiler for short-vector SIMD architectures.Thomas Henretty, Richard Veras, Franz Franchetti, Louis-Nol Pouchet, J. Ramanujam, P. Sadayappan
2013PLDIWhen polyhedral transformations meet SIMD code generation.Martin Kong, Richard Veras, Kevin Stock, Franz Franchetti, Louis-Nol Pouchet, P. Sadayappan
2013SCA framework for load balancing of tensor contraction expressions via dynamic task partitioning.Pai-Wei Lai, Kevin Stock, Samyam Rajbhandari, Sriram Krishnamoorthy, P. Sadayappan
2012ASPLOSHigh-performance sparse matrix-vector multiplication on GPUs for structured grid computations.Jeswin Godwin, Justin Holewinski, P. Sadayappan
2012CCAnalytical Bounds for Optimal Tile Size Selection.Jun Shirako, Kamal Sharma, Naznin Fauzia, Louis-Nol Pouchet, J. Ramanujam, P. Sadayappan, Vivek Sarkar
2012HiPCA global address space approach to automated data management for parallel Quantum Monte Carlo applications.Qingpeng Niu, James Dinan, Sravya Tirukkovalur, Lubos Mitas, Lucas K. Wagner, P. Sadayappan
2012ICSHigh-performance code generation for stencil computations on GPU architectures.Justin Holewinski, Louis-Nol Pouchet, P. Sadayappan
2012PLDIDynamic trace-based analysis of vectorization potential of applications.Justin Holewinski, Ragavendar Ramamurthi, Mahesh Ravishankar, Naznin Fauzia, Louis-Nol Pouchet, Atanas Rountev, P. Sadayappan
2012SCGADBMS: A Framework for Scalable Array Analytics.Tyler Clemons, Srinivasan Parthasarathy, P. Sadayappan
2012SCCode generation for parallel execution of a class of irregular loops on distributed memory systems.Mahesh Ravishankar, John Eisenlohr, Louis-Nol Pouchet, J. Ramanujam, Atanas Rountev, P. Sadayappan
2011CCData Layout Transformation for Stencil Computations on Short-Vector SIMD Architectures.Thomas Henretty, Kevin Stock, Louis-Nol Pouchet, Franz Franchetti, J. Ramanujam, P. Sadayappan
2011CGOPredictive modeling in a polyhedral optimization space.Eunjung Park, Louis-Nol Pouchet, John Cavazos, Albert Cohen, P. Sadayappan
2011HiPCDynamic selection of tile sizes.Sanket Tavarageri, Louis-Nol Pouchet, J. Ramanujam, Atanas Rountev, P. Sadayappan
2011POPLLoop transformations: convexity, pruning and optimization.Louis-Nol Pouchet, Uday Bondhugula, Cdric Bastoul, Albert Cohen, J. Ramanujam, P. Sadayappan, Nicolas Vasilache
2010CCAutomatic C-to-CUDA Code Generation for Affine Programs.Muthu Manikandan Baskaran, J. Ramanujam, P. Sadayappan
2010CCGRIDSelective Recovery from Failures in a Task Parallel Programming Model.James Dinan, Arjun Singri, P. Sadayappan, Sriram Krishnamoorthy
2010CGOParameterized tiling revisited.Muthu Manikandan Baskaran, Albert Hartono, Sanket Tavarageri, Thomas Henretty, J. Ramanujam, P. Sadayappan
2010SCCombined Iterative and Model-driven Optimization in an Automatic Parallelization Framework.Louis-Nol Pouchet, Uday Bondhugula, Cdric Bastoul, Albert Cohen, J. Ramanujam, P. Sadayappan
2009CLUSTERScalable I/O forwarding framework for high-performance computing systems.Nawab Ali, Philip H. Carns, Kamil Iskra, Dries Kimpe, Samuel Lang, Robert Latham, Robert B. Ross, Lee Ward, P. Sadayappan
2009HPDCAn integrated framework for performance-based optimization of scientific workflows.Vijay S. Kumar, P. Sadayappan, Gaurang Mehta, Karan Vahi, Ewa Deelman, Varun Ratnakar, Jihie Kim, Yolanda Gil, Mary W. Hall, Tahsin M. Kur, Joel H. Saltz
2009ICSParametric multi-level tiling of imperfectly nested loops.Albert Hartono, Muthu Manikandan Baskaran, Cdric Bastoul, Albert Cohen, Sriram Krishnamoorthy, Boyana Norris, J. Ramanujam, P. Sadayappan
2009PPoPPCompiler-assisted dynamic scheduling for effective parallelization of loop nests on multicore processors.Muthu Manikandan Baskaran, Nagavijayalakshmi Vydyanathan, Uday Bondhugula, J. Ramanujam, Atanas Rountev, P. Sadayappan
2009SCScalable work stealing.James Dinan, D. Brian Larkins, P. Sadayappan, Sriram Krishnamoorthy, Jarek Nieplocha
2009SCEnabling software management for multicore caches with a lightweight hardware support.Jiang Lin, Qingda Lu, Xiaoning Ding, Zhao Zhang, Xiaodong Zhang, P. Sadayappan
2008CCAutomatic Transformations for Communication-Minimized Parallelization and Locality Optimization in the Polyhedral Model.Uday Bondhugula, Muthu Manikandan Baskaran, Sriram Krishnamoorthy, J. Ramanujam, Atanas Rountev, P. Sadayappan
2008CLUSTERAn OSD-based approach to managing directory operations in parallel file systems.Nawab Ali, Ananth Devulapalli, Dennis Dalessandro, Pete Wyckoff, P. Sadayappan
2008CLUSTERAre nonblocking networks really needed for high-end-computing workloads?Narayan Desai, Pavan Balaji, P. Sadayappan, Mohammad Islam
2008HPCAGaining insights into multicore cache partitioning: Bridging the gap between simulation and real systems.Jiang Lin, Qingda Lu, Xiaoning Ding, Zhao Zhang, Xiaodong Zhang, P. Sadayappan
2008HPDCMulti-hop path splitting and multi-pathing optimizations for data transfers over shared wide-area networks using gridFTP.Gaurav Khanna, mit V. atalyrek, Tahsin M. Kur, P. Sadayappan, Joel H. Saltz, Rajkumar Kettimuthu, Ian T. Foster
2008ICCSIntegrated Data and Task Management for Scientific Applications.Jarek Nieplocha, Sriram Krishnamoorthy, Marat Valiev, Manojkumar Krishnan, Bruce J. Palmer, P. Sadayappan
2008ICPPScioto: A Framework for Global-View Task Parallelism.James Dinan, Sriram Krishnamoorthy, D. Brian Larkins, Jarek Nieplocha, P. Sadayappan
2008ICPPA Duplication Based Algorithm for Optimizing Latency Under Throughput Constraints for Streaming Workflows.Nagavijayalakshmi Vydyanathan, mit V. atalyrek, Tahsin M. Kur, P. Sadayappan, Joel H. Saltz
2008ICSA compiler framework for optimization of affine loop nests for gpgpus.Muthu Manikandan Baskaran, Uday Bondhugula, Sriram Krishnamoorthy, J. Ramanujam, Atanas Rountev, P. Sadayappan
2008PLDIA practical automatic polyhedral parallelizer and locality optimizer.Uday Bondhugula, Albert Hartono, J. Ramanujam, P. Sadayappan
2008PPoPPAutomatic data movement and computation mapping for multi-level parallel architectures with explicitly managed memories.Muthu Manikandan Baskaran, Uday Bondhugula, Sriram Krishnamoorthy, J. Ramanujam, Atanas Rountev, P. Sadayappan
2008SCUsing overlays for efficient data transfer over shared wide-area networks.Gaurav Khanna, mit V. atalyrek, Tahsin M. Kur, Rajkumar Kettimuthu, P. Sadayappan, Ian T. Foster, Joel H. Saltz
2008SCGlobal trees: a framework for linked data structures on distributed memory parallel systems.D. Brian Larkins, James Dinan, Sriram Krishnamoorthy, Srinivasan Parthasarathy, Atanas Rountev, P. Sadayappan
2007CLUSTERNon-collective parallel I/O for global address space programming models.Sriram Krishnamoorthy, Juan Piernas, Vinod Tipparaju, Jarek Nieplocha, P. Sadayappan
2007EuroParScheduling File Transfers for Data-Intensive Jobs on Heterogeneous Clusters.Gaurav Khanna, mit V. atalyrek, Tahsin M. Kur, P. Sadayappan, Joel H. Saltz
2007EuroParToward Optimizing Latency Under Throughput Constraints for Application Workflows on Clusters.Nagavijayalakshmi Vydyanathan, mit V. atalyrek, Tahsin M. Kur, P. Sadayappan, Joel H. Saltz
2007ICPPAnalyzing and Minimizing the Impact of Opportunity Cost in QoS-aware Job Scheduling.Mohammad Islam, Pavan Balaji, Gerald Sabin, P. Sadayappan
2007PLDIEffective automatic parallelization of stencil computations.Sriram Krishnamoorthy, Muthu Manikandan Baskaran, Uday Bondhugula, J. Ramanujam, Atanas Rountev, P. Sadayappan
2007PPoPPAutomatic mapping of nested loops to FPGAS.Uday Bondhugula, J. Ramanujam, P. Sadayappan
2007SCIntegrating parallel file systems with object-based storage devices.Ananth Devulapalli, Dennis Dalessandro, Pete Wyckoff, Nawab Ali, P. Sadayappan
2006CLUSTERA Performance Instrumentation Framework to Characterize Computation-Communication Overlap in Message-Passing Systems.Aniruddha G. Shet, P. Sadayappan, David E. Bernholdt, Jarek Nieplocha, Vinod Tipparaju
2006CLUSTERLocality Conscious Processor Allocation and Scheduling for Mixed Parallel Applications.Nagavijayalakshmi Vydyanathan, Sriram Krishnamoorthy, Gerald Sabin, mit V. atalyrek, Tahsin M. Kur, P. Sadayappan, Joel H. Saltz
2006FCCMHardware/Software Integration for FPGA-based All-Pairs Shortest-Paths.Uday Bondhugula, Ananth Devulapalli, James Dinan, Joseph Fernando, Pete Wyckoff, Eric Stahlberg, P. Sadayappan
2006HPDCTask Scheduling and File Replication for Data-Intensive Jobs with Batch-shared I/O.Gaurav Khanna, Nagavijayalakshmi Vydyanathan, mit V. atalyrek, Tahsin M. Kur, Sriram Krishnamoorthy, P. Sadayappan, Joel H. Saltz
2006ICCSIdentifying Cost-Effective Common Subexpressions to Reduce Operation Count in Tensor Contraction Evaluations.Albert Hartono, Qingda Lu, Xiaoyang Gao, Sriram Krishnamoorthy, Marcel Nooijen, Gerald Baumgartner, David E. Bernholdt, Venkatesh Choppella, Russell M. Pitzer, J. Ramanujam, Atanas Rountev, P. Sadayappan
2006ICPPAn Integrated Approach for Processor Allocation and Scheduling of Mixed-Parallel Applications.Nagavijayalakshmi Vydyanathan, Sriram Krishnamoorthy, Gerald Sabin, mit V. atalyrek, Tahsin M. Kur, P. Sadayappan, Joel H. Saltz
2006JSSPPA Data Locality Aware Online Scheduling Approach for I/O-Intensive Jobs with File Sharing.Gaurav Khanna, mit V. atalyrek, Tahsin M. Kur, P. Sadayappan, Joel H. Saltz
2006JSSPPMoldable Parallel Job Scheduling Using Job Efficiency: An Iterative Approach.Gerald Sabin, Matthew Lang, P. Sadayappan
2006SCData management and query - Hypergraph partitioning for automatic memory hierarchy management.Sriram Krishnamoorthy, mit V. atalyrek, Jarek Nieplocha, Atanas Rountev, P. Sadayappan
2006SCM12 - Overview of the global arrays parallel software development toolkit.Jarek Nieplocha, Bruce J. Palmer, Manojkumar Krishnan, P. Sadayappan
2005CCGRIDA hypergraph partitioning based approach for scheduling of tasks with batch-shared I/O.Gaurav Khanna, Nagavijayalakshmi Vydyanathan, Tahsin M. Kur, mit V. atalyrek, Pete Wyckoff, Joel H. Saltz, P. Sadayappan
2005HiPCData and Computation Abstractions for Dynamic and Irregular Computations.Sriram Krishnamoorthy, Jarek Nieplocha, P. Sadayappan
2005HPDCAssessment and enhancement of meta-schedulers for multi-site job sharing.Gerald Sabin, Vishvesh Sahasrabudhe, P. Sadayappan
2005ICCSAutomated Operation Minimization of Tensor Contraction Expressions in Electronic Structure Calculations.Albert Hartono, Alexander Sibiryakov, Marcel Nooijen, Gerald Baumgartner, David E. Bernholdt, So Hirata, Chi-Chung Lam, Russell M. Pitzer, J. Ramanujam, P. Sadayappan
2005JSSPPUnfairness Metrics for Space-Sharing Parallel Job Schedulers.Gerald Sabin, P. Sadayappan
2005PPoPPPerformance modeling and optimization of parallel out-of-core tensor contractions.Xiaoyang Gao, Swarup Kumar Sahoo, Chi-Chung Lam, J. Ramanujam, Qingda Lu, Gerald Baumgartner, P. Sadayappan
2005SCIntegrated Loop Optimizations for Data Locality Enhancement of Tensor Contraction Expressions.Swarup Kumar Sahoo, Sriram Krishnamoorthy, Rajkiran Panuganti, P. Sadayappan
2004CLUSTERTowards provision of quality of service guarantees in job scheduling.Mohammad Islam, Pavan Balaji, P. Sadayappan, Dhabaleswar K. Panda
2004CLUSTEROn fairness in distributed job scheduling across multiple sites.Gerald Sabin, Vishvesh Sahasrabudhe, P. Sadayappan
2004HiPCEfficient Layout Transformation for Disk-Based Multidimensional Arrays.Sriram Krishnamoorthy, Gerald Baumgartner, Chi-Chung Lam, Jarek Nieplocha, P. Sadayappan
2004ICPPJob Fairness in Non-Preemptive Job Scheduling.Gerald Sabin, Garima Kochhar, P. Sadayappan
2003CLUSTEREfficient Parallel Out-of-Core Matrix Transposition.Sriram Krishnamoorthy, Gerald Baumgartner, Daniel Cociorva, Chi-Chung Lam, P. Sadayappan
2003CLUSTERA Robust Scheduling Strategy for Moldable Scheduling of Parallel Jobs.Sudha Srinivasan, Sriram Krishnamoorthy, P. Sadayappan
2003HiPCData Locality Optimization for Synthesis of Efficient Out-of-Core Algorithms.Sandhya Krishnan, Sriram Krishnamoorthy, Gerald Baumgartner, Daniel Cociorva, Chi-Chung Lam, P. Sadayappan, J. Ramanujam, David E. Bernholdt, Venkatesh Choppella
2003JSSPPQoPS: A QoS Based Scheme for Parallel Job Scheduling.Mohammad Islam, Pavan Balaji, P. Sadayappan, Dhabaleswar K. Panda
2003JSSPPScheduling of Parallel Jobs in a Heterogeneous Multi-site Environement.Gerald Sabin, Rajkumar Kettimuthu, Arun Rajan, P. Sadayappan
2002CLUSTERSelective Buddy Allocation for Scheduling Parallel Jobs on Clusters.Vijay Subramani, Rajkumar Kettimuthu, Srividya Srinivasan, Jeanette Johnston, P. Sadayappan
2002HiPCEffective Selection of Partition Sizes for Moldable Scheduling of Parallel Jobs.Srividya Srinivasan, Vijay Subramani, Rajkumar Kettimuthu, Praveen Holenarsipur, P. Sadayappan
2002HPDCDistributed Job Scheduling on Computational Grids Using Multiple Simultaneous Requests.Vijay Subramani, Rajkumar Kettimuthu, Srividya Srinivasan, P. Sadayappan
2002ICDCSA Reliable Multicast Algorithm for Mobile Ad Hoc Networks.Thiagaraja Gopalsamy, Mukesh Singhal, Dhabaleswar K. Panda, P. Sadayappan
2002JSSPPSelective Reservation Strategies for Backfill Job Scheduling.Srividya Srinivasan, Rajkumar Kettimuthu, Vijay Subramani, P. Sadayappan
2002PLDISpace-Time Trade-Off Optimization for a Class of Electronic Structure Calculations.Daniel Cociorva, Gerald Baumgartner, Chi-Chung Lam, P. Sadayappan, J. Ramanujam, Marcel Nooijen, David E. Bernholdt, Robert J. Harrison
2002SCA high-level approach to synthesis of high-performance codes for quantum chemistry.Gerald Baumgartner, David E. Bernholdt, Daniel Cociorva, Robert J. Harrison, So Hirata, Chi-Chung Lam, Marcel Nooijen, Russell M. Pitzer, J. Ramanujam, P. Sadayappan
2001HiPCTowards Automatic Synthesis of High-Performance Codes for Electronic Structure Calculations: Data Locality Optimization.Daniel Cociorva, J. W. Wilkins, Gerald Baumgartner, P. Sadayappan, J. Ramanujam, Marcel Nooijen, David E. Bernholdt, Robert J. Harrison
2001ICPPImplementing TreadMarksover VIA on Myrinet and Gigabit Ethernet: Challenges, Design Experience, and Performance Evaluation.Mohammad Banikazemi, Jiuxing Liu, Dhabaleswar K. Panda, P. Sadayappan
2001ICPPNIC-Based Rate Control for Proportional Bandwidth Allocation in Myrinet Clusters.Abhishek Gulati, Dhabaleswar K. Panda, P. Sadayappan, Pete Wyckoff
2001ICSLoop optimization for a class of memory-constrained computations.Daniel Cociorva, J. W. Wilkins, Chi-Chung Lam, Gerald Baumgartner, J. Ramanujam, P. Sadayappan
2000HiPCCharacterization and enhancement of Static Mapping Heuristics for Heterogeneous Systems.Praveen Holenarsipur, Vladimir Yarmolenko, Jos Duato, Dhabaleswar K. Panda, P. Sadayappan
1999HCWCommunication Modeling of Heterogeneous Networks of Workstations for Performance Characterization of Collective Operations.Mohammad Banikazemi, Jayanthi Sampathkumar, Sandeep Prabhu, Dhabaleswar K. Panda, P. Sadayappan
1999HiPCMemory-Optimal Evaluation of Expression Trees Involving Large Objects.Chi-Chung Lam, Daniel Cociorva, Gerald Baumgartner, P. Sadayappan
1999ICPPAn Incremental Methodology for Parallelizing Legacy Stencil Codes on Message-Passing Computers.N. S. Sundar, S. Jayanthi, P. Sadayappan, Miguel Visbal
1996ICSHybrid Algorithms for Complete Exchange in 2D Meshes.N. S. Sundar, Doddaballapur Narasimha-Murthy Jayasimha, Dhabaleswar K. Panda, P. Sadayappan
1994ICPADSCommunication-Efficient Implementation of Block Recursive Algorithms on Distributed-Memory Machines.Sandeep K. S. Gupta, Chua-Huang Huang, Rodney W. Johnson, P. Sadayappan
1994ICSAn approach to communication-efficient data redistribution.S. D. Kaushik, Chua-Huang Huang, Rodney W. Johnson, P. Sadayappan
1994ICSOn sparse matrix reordering for parallel factorization.Bharat Kumar, P. Sadayappan, Chua-Huang Huang
1994SCEXTENT: a portable programming environment for designing and implementing high-performance block recursive algorithms.Donglai Dai, Sandeep K. S. Gupta, S. D. Kaushik, J. H. Lu, Raj Verdhan Singh, Chua-Huang Huang, P. Sadayappan, Rodney W. Johnson
1994SPAACommunication Efficient Matrix Multiplication on Hypercubes.Himanshu Gupta, P. Sadayappan
1993DACArchitectural Synthesis of Performance-Driven Multipliers with Accumulator Interleaving.Debabrata Ghosh, S. K. Nandy, P. Sadayappan, K. Parthasarathy
1993ICPPCompile-Time Characterization of Recurrent Patterns in Irregular Computations.Kalluri Eswar, P. Sadayappan, Chua-Huang Huang
1993ICPPSupernodal Sparse Cholesky Facotrization on Distributed-Memory Multiprocessors.Kalluri Eswar, P. Sadayappan, Chua-Huang Huang, V. Visvanathan
1993ICPPOn Compiling Array Expressions for Efficient Execution on Distributed-Memory Machines.Sandeep K. S. Gupta, S. D. Kaushik, S. Mufti, Sanjay Sharma, Chua-Huang Huang, P. Sadayappan
1993ICPPA Parallel Progressive Refinement Image Rendering Algorithm on a Scalable Multithreaded VLSI Processor Array.S. K. Nandy, Ranjani Narayan, V. Visvanathan, P. Sadayappan, Prashant S. Chauhan
1993SCEfficient transposition algorithms for large matrices.S. D. Kaushik, Chua-Huang Huang, John R. Johnson, Rodney W. Johnson, P. Sadayappan
1992SCAn Algebraic Theory for Modeling Direct Interconnection Networks.S. D. Kaushik, Sanjay Sharma, Chua-Huang Huang, Jeremy R. Johnson, Rodney W. Johnson, P. Sadayappan
1991ICPPMultifrontal Factorization of Sparse Matrices on Shared-Memory Multiprocessors.Kalluri Eswar, P. Sadayappan, V. Visvanathan
1991ICPPComputer Graphics Rendering on a Shared Memory Multiprocessor.Scott Whitman, P. Sadayappan
1991PPoPPRemoval of Redundant Dependences in DOACROSS Lops with Constant Dependences.V. Prasad Krothapalli, P. Sadayappan
1991SCTiling multidimensional iteration spaces for nonshared memory machines.J. Ramanujam, P. Sadayappan
1990ICPPTiling of Iteration Spaces for Multicomputers.J. Ramanujam, P. Sadayappan
1989DACEfficient Sparse Matrix Factorization for Circuit Simulation on Vector Supercomputers.P. Sadayappan, V. Visvanathan
1989ICPPOptimal Static Scheduling of Sequential Loops on Multiprocessors.Amr Zaky, P. Sadayappan
1989ICSOne-to-one mapping of process graphs onto a hypercube.Fikret Eral, P. Sadayappan
1989SCA methodology for parallelizing programs for multicomputers and complex memory multiprocessors.J. Ramanujam, P. Sadayappan
1988ICCDComparative analysis of approaches to hardware acceleration for sparse-matrix factorization.P. Sadayappan, V. Visvanathan
1988ICRAA VLSI robotics vector processor for real-time control.Yong-Long Calvin Ling, P. Sadayappan, Karl W. Olson, David E. Orin
1988ICSAn approach to synchronization for parallel computing.V. Prasad Krothapalli, P. Sadayappan
1988ICSParallelization and performance evaluation of circuit simulation on a shared-memory multiprocessor.P. Sadayappan, V. Visvanathan
1987ICPPMapping Finite Element Graphs onto Processor Meshes.P. Sadayappan, Fikret Eral, Steven Martin
1987ICSCluster-Partitioning Approaches to Mapping Parallel Programs onto a Hypercube.P. Sadayappan, Fikret Eral
1985DACModeling switch-level simulation using data flow.V. Ashok, Roger L. Costello, P. Sadayappan