Skip to content

Suvrit Sra

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

66

Venues

16

Active years

2003–2025

Best venue rank

A*

Where they publish

Papers

66 indexed papers, newest first.

YearVenueTitleAuthors
2025ICLRGraph Transformers Dream of Electric Flow.Xiang Cheng, Lawrence Carin, Suvrit Sra
2024ICLRLinear attention is (maybe) all you need (to understand Transformer optimization).Kwangjun Ahn, Xiang Cheng, Minhak Song, Chulhee Yun, Ali Jadbabaie, Suvrit Sra
2024ICMLHow to Escape Sharp Minima with Random Perturbations.Kwangjun Ahn, Ali Jadbabaie, Suvrit Sra
2024ICMLTransformers Implement Functional Gradient Descent to Learn Non-Linear Functions In Context.Xiang Cheng, Yuxin Chen, Suvrit Sra
2023ICLRSign and Basis Invariant Networks for Spectral Graph Representation Learning.Derek Lim, Joshua David Robinson, Lingxiao Zhao, Tess E. Smidt, Suvrit Sra, Haggai Maron, Stefanie Jegelka
2023ICMLGlobal optimality for Euclidean CCCP under Riemannian convexity.Melanie Weber, Suvrit Sra
2023ICMLOn the Training Instability of Shuffling SGD with Batch Normalization.David Xing Wu, Chulhee Yun, Suvrit Sra
2022AAAIMax-Margin Contrastive Learning.Anshul Shah, Suvrit Sra, Rama Chellappa, Anoop Cherian
2022COLTUnderstanding Riemannian Acceleration via a Proximal Extragradient Framework.Jikai Jin, Suvrit Sra
2022ICLRMinibatch vs Local SGD with Shuffling: Tight Convergence Bounds and Beyond.Chulhee Yun, Shashank Rajput, Suvrit Sra
2022ICMLUnderstanding the unstable convergence of gradient descent.Kwangjun Ahn, Jingzhao Zhang, Suvrit Sra
2022ICMLBeyond Worst-Case Analysis in Stochastic Approximation: Moment Estimation Improves Instance Complexity.Jingzhao Zhang, Hongzhou Lin, Subhro Das, Suvrit Sra, Ali Jadbabaie
2022ICMLNeural Network Weights Do Not Converge to Stationary Points: An Invariant Measure Perspective.Jingzhao Zhang, Haochuan Li, Suvrit Sra, Ali Jadbabaie
2021COLTOpen Problem: Can Single-Shuffle SGD be Better than Reshuffling SGD and GD?Chulhee Yun, Suvrit Sra, Ali Jadbabaie
2021ICLRContrastive Learning with Hard Negative Samples.Joshua David Robinson, Ching-Yao Chuang, Suvrit Sra, Stefanie Jegelka
2021ICLRCoping with Label Shift via Distributionally Robust Optimisation.Jingzhao Zhang, Aditya Krishna Menon, Andreas Veit, Srinadh Bhojanapalli, Sanjiv Kumar, Suvrit Sra
2021ICMLOnline Learning in Unknown Markov Games.Yi Tian, Yuanhao Wang, Tiancheng Yu, Suvrit Sra
2021ICMLThree Operator Splitting with a Nonconvex Loss Function.Alp Yurtsever, Varun Mangalick, Suvrit Sra
2021ICMLProvably Efficient Algorithms for Multi-Objective Competitive RL.Tiancheng Yu, Yi Tian, Jingzhao Zhang, Suvrit Sra
2020ACMLGeodesically-convex optimization for averaging partially observed covariance matrices.Florian Yger, Sylvain Chevallier, Quentin Barthlemy, Suvrit Sra
2020COLTFrom Nesterov's Estimate Sequence to Riemannian Acceleration.Kwangjun Ahn, Suvrit Sra
2020ICLRWhy Gradient Clipping Accelerates Training: A Theoretical Justification for Adaptivity.Jingzhao Zhang, Tianxing He, Suvrit Sra, Ali Jadbabaie
2020ICMLLearning Adversarial Markov Decision Processes with Bandit Feedback and Unknown Transition.Chi Jin, Tiancheng Jin, Haipeng Luo, Suvrit Sra, Tiancheng Yu
2020ICMLStrength from Weakness: Fast Learning Using Weak Supervision.Joshua Robinson, Stefanie Jegelka, Suvrit Sra
2020ICMLComplexity of Finding Stationary Points of Nonconvex Nonsmooth Functions.Jingzhao Zhang, Hongzhou Lin, Stefanie Jegelka, Suvrit Sra, Ali Jadbabaie
2019AISTATSLearning Determinantal Point Processes by Corrective Negative Sampling.Zelda Mariet, Mike Gartrell, Suvrit Sra
2019ICLREfficiently testing local optimality and escaping saddles for ReLU networks.Chulhee Yun, Suvrit Sra, Ali Jadbabaie
2019ICLRSmall nonlinearities in activation functions create bad local minima in neural networks.Chulhee Yun, Suvrit Sra, Ali Jadbabaie
2019ICMLRandom Shuffling Beats SGD after Finite Epochs.Jeff Z. HaoChen, Suvrit Sra
2019ICMLEscaping Saddle Points with Adaptive Gradient Methods.Matthew Staib, Sashank J. Reddi, Satyen Kale, Sanjiv Kumar, Suvrit Sra
2019ICMLConditional Gradient Methods via Stochastic Path-Integrated Differential Estimator.Alp Yurtsever, Suvrit Sra, Volkan Cevher
2018AISTATSA Generic Approach for Escaping Saddle points.Sashank J. Reddi, Manzil Zaheer, Suvrit Sra, Barnabs Pczos, Francis R. Bach, Ruslan Salakhutdinov, Alexander J. Smola
2018COLTAn Estimate Sequence for Geodesically Convex Optimization.Hongyi Zhang, Suvrit Sra
2018CVPRNon-Linear Temporal Subspace Representations for Activity Recognition.Anoop Cherian, Suvrit Sra, Stephen Gould, Richard Hartley
2018ICLRDistributional Adversarial Networks.Chengtao Li, David Alvarez-Melis, Keyulu Xu, Stefanie Jegelka, Suvrit Sra
2018ICLRGlobal Optimality Conditions for Deep Neural Networks.Chulhee Yun, Suvrit Sra, Ali Jadbabaie
2017AISTATSCombinatorial Topic Models using Small-Variance Asymptotics.Ke Jiang, Suvrit Sra, Brian Kulis
2016AISTATSEfficient Sampling for k-Determinantal Point Processes.Chengtao Li, Stefanie Jegelka, Suvrit Sra
2016AISTATSAdaDelay: Delay Adaptive Distributed Stochastic Optimization.Suvrit Sra, Adams Wei Yu, Mu Li, Alexander J. Smola
2016COLTFirst-order Methods for Geodesically Convex Optimization.Hongyi Zhang, Suvrit Sra
2016ICMLFast DPP Sampling for Nystrom with Application to Kernel Methods.Chengtao Li, Stefanie Jegelka, Suvrit Sra
2016ICMLGaussian quadrature for matrix inverse forms with applications.Chengtao Li, Suvrit Sra, Stefanie Jegelka
2016ICMLStochastic Variance Reduction for Nonconvex Optimization.Sashank J. Reddi, Ahmed Hefny, Suvrit Sra, Barnabs Pczos, Alexander J. Smola
2016ICMLParallel and Distributed Block-Coordinate Frank-Wolfe Algorithms.Yu-Xiang Wang, Veeranjaneyulu Sadhanala, Wei Dai, Willie Neiswanger, Suvrit Sra, Eric P. Xing
2016ICMLGeometric Mean Metric Learning.Pourya Zadeh, Reshad Hosseini, Suvrit Sra
2015AISTATSData modeling with the elliptical gamma distribution.Suvrit Sra, Reshad Hosseini, Lucas Theis, Matthias Bethge
2015ICMLFixed-point algorithms for learning determinantal point processes.Zelda Mariet, Suvrit Sra
2015UAILarge-scale randomized-coordinate descent methods with non-separable linear constraints.Sashank J. Reddi, Ahmed Hefny, Carlton Downey, Avinava Dubey, Suvrit Sra
2014ECCVRiemannian Sparse Coding for Positive Definite Matrices.Anoop Cherian, Suvrit Sra
2014ICMLTowards an optimal stochastic alternating direction method of multipliers.Samaneh Azadi, Suvrit Sra
2014ICMLRandomized Nonlinear Component Analysis.David Lopez-Paz, Suvrit Sra, Alexander J. Smola, Zoubin Ghahramani, Bernhard Schlkopf
2014UAIFast Newton methods for the group fused lasso.Matt Wytock, Suvrit Sra, Jeremy Z. Kolter
2011ICASSPDenoising sparse noise via online dictionary learning.Anoop Cherian, Suvrit Sra, Nikolaos Papanikolopoulos
2011ICCVEfficient similarity search for covariance matrices via the Jensen-Bregman LogDet Divergence.Anoop Cherian, Suvrit Sra, Arindam Banerjee, Nikolaos Papanikolopoulos
2011ICMLFast Newton-type Methods for Total Variation Regularization.lvaro Barbero Jimnez, Suvrit Sra
2010CVPREfficient filter flow for space-variant multiframe blind deconvolution.Michael Hirsch, Suvrit Sra, Bernhard Schlkopf, Stefan Harmeling
2010ICIPMultiframe blind deconvolution, super-resolution, and saturation correction via incremental EM.Stefan Harmeling, Suvrit Sra, Michael Hirsch, Bernhard Schlkopf
2010ICMLA scalable trust-region algorithm with application to mixed-norm regression.Dongmin Kim, Suvrit Sra, Inderjit S. Dhillon
2009ALTApproximation Algorithms for Tensor Clustering.Stefanie Jegelka, Suvrit Sra, Arindam Banerjee
2009ICMLWorkshop summary: Numerical mathematics in machine learning.Matthias W. Seeger, Suvrit Sra, John P. Cunningham
2008ICDMBlock-Iterative Algorithms for Non-negative Matrix Approximation.Suvrit Sra
2007ICMLInformation-theoretic metric learning.Jason V. Davis, Brian Kulis, Prateek Jain, Suvrit Sra, Inderjit S. Dhillon
2007SDMFast Newton-type Methods for the Least Squares Nonnegative Matrix Approximation Problem.Dongmin Kim, Suvrit Sra, Inderjit S. Dhillon
2006ICASSPRow-Action Methods for Compressed Sensing.Suvrit Sra, Joel A. Tropp
2004SDMMinimum Sum-Squared Residue Co-Clustering of Gene Expression Data.Hyuk Cho, Inderjit S. Dhillon, Yuqiang Guan, Suvrit Sra
2003KDDGenerative model-based clustering of directional data.Arindam Banerjee, Inderjit S. Dhillon, Joydeep Ghosh, Suvrit Sra