Skip to content

Sridhar Mahadevan

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

63

Venues

15

Active years

1985–2026

Best venue rank

A*

Where they publish

Papers

63 indexed papers, newest first.

YearVenueTitleAuthors
2026AAAIRethinking AI: From Functions to Functors.Sridhar Mahadevan
2023AAAISmoothed Online Combinatorial Optimization Using Imperfect Predictions.Kai Wang, Zhao Song, Georgios Theocharous, Sridhar Mahadevan
2023WSDMPrivacy Aware Experiments without Cookies.Shiv Shankar, Ritwik Sinha, Saayan Mitra, Viswanathan (Vishy) Swaminathan, Sridhar Mahadevan, Moumita Sinha
2022WACVGenerating and Controlling Diversity in Image Search.Md. Mehrab Tanjim, Ritwik Sinha, Krishna Kumar Singh, Sridhar Mahadevan, David Arbour, Moumita Sinha, Garrison W. Cottrell
2020ICMLOptimizing for the Future in Non-Stationary MDPs.Yash Chandak, Georgios Theocharous, Shiv Shankar, Martha White, Sridhar Mahadevan, Philip S. Thomas
2018AAAIImagination Machines: A New Challenge for Artificial Intelligence.Sridhar Mahadevan
2017ICLRGenerative Multi-Adversarial Networks.Ishan P. Durugkar, Ian Gemp, Sridhar Mahadevan
2016IJCAIProximal Gradient Temporal Difference Learning Algorithms.Bo Liu, Ji Liu, Mohammad Ghavamzadeh, Sridhar Mahadevan, Marek Petrik
2015AAAIAligning Mixed Manifolds.Thomas Boucher, CJ Carey, Sridhar Mahadevan, Melinda Darby Dyar
2015AAAISolving Large Sustainable Supply Chain Networks Using Variational Inequalities.Ian Gemp, Sridhar Mahadevan, Bo Liu
2015EMNLPEfficient Hyper-parameter Optimization for NLP Applications.Lidan Wang, Minwei Feng, Bowen Zhou, Bing Xiang, Sridhar Mahadevan
2015UAIFinite-Sample Analysis of Proximal Gradient TD Algorithms.Bo Liu, Ji Liu, Mohammad Ghavamzadeh, Sridhar Mahadevan, Marek Petrik
2014AAAIManifold Spanning Graphs.CJ Carey, Sridhar Mahadevan
2013AAAIMultiscale Manifold Learning.Chang Wang, Sridhar Mahadevan
2013AAAIBasis Adaptation for Sparse Nonlinear Reinforcement Learning.Sridhar Mahadevan, Stephen Giguere, Nicholas Jacek
2013IJCAIManifold Alignment Preserving Global Geometry.Chang Wang, Sridhar Mahadevan
2012AAAIManifold Warping: Manifold Alignment over Time.Hoa Trong Vu, Clifton Carey, Sridhar Mahadevan
2012UAISparse Q-learning with Mirror Descent.Sridhar Mahadevan, Bo Liu
2011IJCAIHeterogeneous Domain Adaptation Using Manifold Alignment.Chang Wang, Sridhar Mahadevan
2011IJCAIJointly Learning Data-Dependent Label and Locality-Preserving Projections.Chang Wang, Sridhar Mahadevan
2011PPAMA GPU-Based Approximate SVD Algorithm.Blake Foster, Sridhar Mahadevan, Rui Wang
2010AAAIRepresentation Discovery in Sequential Decision Making.Sridhar Mahadevan
2010AAAICompressing POMDPs Using Locality Preserving Non-Negative Matrix Factorization.Georgios Theocharous, Sridhar Mahadevan
2009AIEDTransfer Learning and Representation Discovery in Intelligent Tutoring Systems.Kimberly Ferguson, Beverly Park Woolf, Sridhar Mahadevan
2009IJCAIManifold Alignment without Correspondence.Chang Wang, Sridhar Mahadevan
2009IJCAIMultiscale Analysis of Document Corpora Based on Diffusion Models.Chang Wang, Sridhar Mahadevan
2008AAAIFast Spectral Learning using Lanczos Eigenspace Projections.Sridhar Mahadevan
2008ICMLManifold alignment using Procrustes analysis.Chang Wang, Sridhar Mahadevan
2007AAAICompact Spectral Bases for Value Function Approximation Using Kronecker Factorization.Jeffrey Johns, Sridhar Mahadevan, Chang Wang
2007AIEDRepairing Disengagement With Non-Invasive Interventions.Ivon Arroyo, Kimberly Ferguson, Jeffrey Johns, Toby Dragon, Hasmik Meheranian, Don Fisher, Andrew G. Barto, Sridhar Mahadevan, Beverly Park Woolf
2007ICMLConstructing basis functions from directed graphs for value function approximation.Jeffrey Johns, Sridhar Mahadevan
2007ICMLAdaptive mesh compression in 3D computer graphics using multiscale manifold learning.Sridhar Mahadevan
2007ICMLLearning state-action basis functions for hierarchical MDPs.Sarah Osentoski, Sridhar Mahadevan
2006AAAILearning Representation and Control in Continuous Markov Decision Processes.Sridhar Mahadevan, Mauro Maggioni, Kimberly Ferguson, Sarah Osentoski
2006ICMLFast direct policy evaluation using multiscale analysis of Markov diffusion processes.Mauro Maggioni, Sridhar Mahadevan
2006ITSImproving Intelligent Tutoring Systems: Using Expectation Maximization to Learn Student Skill Levels.Kimberly Ferguson, Ivon Arroyo, Sridhar Mahadevan, Beverly Park Woolf, Andrew G. Barto
2006ITSEstimating Student Proficiency Using an Item Response Theory Model.Jeffrey Johns, Sridhar Mahadevan, Beverly Park Woolf
2005AAAIA Variational Learning Algorithm for the Abstract Hidden Markov Model.Jeffrey Johns, Sridhar Mahadevan
2005AAAISamuel Meets Amarel: Automating Value Function Approximation Using Global State Space Analysis.Sridhar Mahadevan
2005ICMLProto-value functions: developmental reinforcement learning.Sridhar Mahadevan
2005ICMLCoarticulation: an approach for generating concurrent plans in Markov decision processes.Khashayar Rohanimanesh, Sridhar Mahadevan
2005UAIRepresentation Policy Iteration.Sridhar Mahadevan
2005SECONSwitching kalman filters for prediction and tracking in an adaptive meteorological sensing network.Victoria Manfredi, Sridhar Mahadevan, James F. Kurose
2004IROSLearning hierarchical models of activity.Sarah Osentoski, Victoria Manfredi, Sridhar Mahadevan
2003ICMLHierarchical Policy Gradient Algorithms.Mohammad Ghavamzadeh, Sridhar Mahadevan
2002ICMLHierarchically Optimal Average Reward Reinforcement Learning.Mohammad Ghavamzadeh, Sridhar Mahadevan
2002IROSLearning the hierarchical structure of spatial environments using multiresolution statistical models.Georgios Theocharous, Sridhar Mahadevan
2002ICRAApproximate Planning with Hierarchical Partially Observable Markov Decision Process Models for Robot Navigation.Georgios Theocharous, Sridhar Mahadevan
2001ICMLContinuous-Time Hierarchical Reinforcement Learning.Mohammad Ghavamzadeh, Sridhar Mahadevan
2001ICRALearning Hierarchical Partially Observable Markov Decision Process Models for Robot Navigation.Georgios Theocharous, Khashayar Rohanimanesh, Sridhar Mahadevan
2001UAIDecision-Theoretic Planning with Concurrent Temporally Extended Actions.Khashayar Rohanimanesh, Sridhar Mahadevan
1999ICMLHierarchical Optimization of Policy-Coupled Semi-Markov Decision Processes.Gang Wang, Sridhar Mahadevan
1998FlAIRSOptimizing Production Manufacturing Using Reinforcement Learning.Sridhar Mahadevan, Georgios Theocharous
1996AAAIAn Average-Reward Reinforcement Learning Algorithm for Computing Bias-Optimal Policies.Sridhar Mahadevan
1996ICMLSensitive Discount Optimality: Unifying Discounted and Average Reward Reinforcement Learning.Sridhar Mahadevan
1994ICMLTo Discount or Not to Discount in Reinforcement Learning: A Case Study Comparing R Learning and Q Learning.Sridhar Mahadevan
1992ICMLEnhancing Transfer in Reinforcement Learning by Building Stochastic Models of Robot Actions.Sridhar Mahadevan
1991AAAIAutomatic Programming of Behavior-Based Robots Using Reinforcement Learning.Sridhar Mahadevan, Jonathan Connell
1991ICMLScaling Reinforcement Learning to Robotics by Exploiting the Subsumption Architecture.Sridhar Mahadevan, Jonathan Connell
1989ICMLUsing Determinations in EBL: A Solution to the incomplete Theory Problem.Sridhar Mahadevan
1988ICMLOn the Tractability of Learning from Incomplete Theories.Sridhar Mahadevan, Prasad Tadepalli
1985IJCAIVerification-based Learning: A Generalized Strategy for Inferring Problem-Reduction Methods.Sridhar Mahadevan
1985IJCAILEAP: A Learning Apprentice for VLSI Design.Tom M. Mitchell, Sridhar Mahadevan, Louis I. Steinberg