Skip to content

Prashanth L. A.

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

20

Venues

9

Active years

2011–2026

Best venue rank

A*

Where they publish

Papers

20 indexed papers, newest first.

YearVenueTitleAuthors
2026AAAIPolicy Newton Methods for Distortion Riskmetrics.Soumen Pachal, Mizhaan Prajit Maniyar, Prashanth L. A.
2025AISTATSRisk-sensitive Bandits: Arm Mixture Optimality and Regret-efficient Algorithms.Meltem Tatli, Arpan Mukherjee, Prashanth L. A., Karthikeyan Shanmugam, Ali Tajer
2024AISTATSA Cubic-regularized Policy Newton Algorithm for Reinforcement Learning.Mizhaan Prajit Maniyar, Prashanth L. A., Akash Mondal, Shalabh Bhatnagar
2024ICMLPolicy Evaluation for Variance in Average Reward Reinforcement Learning.Shubhada Agrawal, Prashanth L. A., Siva Theja Maguluri
2024ICMLRisk Estimation in a Markov Cost Process: Lower and Upper Bounds.Gugan Thoppe, Prashanth L. A., Sanjay P. Bhat
2023AISTATSFinite time analysis of temporal difference learning with linear function approximation: Tail averaging and regularisation.Gandharv Patil, Prashanth L. A., Dheeraj Nagaraj, Doina Precup
2023CISSGeneralized Simultaneous Perturbation Stochastic Approximation with Reduced Estimator Bias.Shalabh Bhatnagar, Prashanth L. A.
2023UAIA policy gradient approach for optimization of smooth risk measures.Nithia Vijayan, Prashanth L. A.
2022IJCAIA Survey of Risk-Aware Multi-Armed Bandits.Vincent Y. F. Tan, Prashanth L. A., Krishna P. Jagannathan
2021AAAIEstimation of Spectral Risk Measures.Ajay Kumar Pandey, Prashanth L. A., Sanjay P. Bhat
2020ICMLConcentration bounds for CVaR estimation: The cases of light-tailed and heavy-tailed distributions.Prashanth L. A., Krishna P. Jagannathan, Ravi Kumar Kolla
2019ICMLCorrelated bandits or: How to minimize mean-squared error online.Vinay Praneeth Boda, Prashanth L. A.
2017AAAIWeighted Bandits or: How Bandits Learn Distorted Values That Are Not Expected.Aditya Gopalan, Prashanth L. A., Michael C. Fu, Steven I. Marcus
2016AISTATS(Bandit) Convex Optimization with Biased Noisy Gradient Oracles.Xiaowei Hu, Prashanth L. A., Andrs Gyrgy, Csaba Szepesvri
2016ICMLCumulative Prospect Theory Meets Reinforcement Learning: Prediction and Control.Prashanth L. A., Cheng Jie, Michael C. Fu, Steven I. Marcus, Csaba Szepesvri
2015AAAIFast Gradient Descent for Drifting Least Squares Regression, with Application to Bandits.Nathaniel Korda, Prashanth L. A., Rmi Munos
2015ICMLOn TD(0) with function approximation: Concentration bounds and a centered variant with exponential convergence.Nathaniel Korda, Prashanth L. A.
2014ALTPolicy Gradients for CVaR-Constrained MDPs.Prashanth L. A.
2014COMSNETSAdaptive sleep-wake control using reinforcement learning in sensor networks.Prashanth L. A., Abhranil Chatterjee, Shalabh Bhatnagar
2011ICSOCStochastic Optimization for Adaptive Labor Staffing in Service Systems.Prashanth L. A., H. L. Prasad, Nirmit Desai, Shalabh Bhatnagar, Gargi Banerjee Dasgupta