Skip to content

Ambuj Tewari

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

65

Venues

11

Active years

2002–2026

Best venue rank

A*

Where they publish

Papers

65 indexed papers, newest first.

YearVenueTitleAuthors
2026COLTA Characterization of List Language Identification in the Limit.Moses Charikar, Chirag Pabbaraju, Ambuj Tewari
2025AISTATSLearning Infinite-Horizon Average-Reward Linear Mixture MDPs of Bounded Span.Woojin Chae, Kihyuk Hong, Yufan Zhang, Ambuj Tewari, Dabeen Lee
2025AISTATSReinforcement Learning for Infinite-Horizon Average-Reward Linear MDPs via Approximation by Discounted-Reward MDPs.Kihyuk Hong, Woojin Chae, Yufan Zhang, Dabeen Lee, Ambuj Tewari
2025ALTA Unified Theory of Supervised Online Learnability.Vinod Raman, Unique Subedi, Ambuj Tewari
2025COLTGeneration through the lens of learning theory.Vinod Raman, Jiaxun Li, Ambuj Tewari
2025ICLRA Theoretical Framework for Partially-Observed Reward States in RLHF.Chinmaya Kausik, Mirco Mutti, Aldo Pacchiano, Ambuj Tewari
2025ICMLA Computationally Efficient Algorithm for Infinite-Horizon Average-Reward Linear MDPs.Kihyuk Hong, Ambuj Tewari
2025ICMLLeveraging Offline Data in Linear Latent Contextual Bandits.Chinmaya Kausik, Kevin Tan, Ambuj Tewari
2025ICMLOn the Benefits of Active Data Collection in Operator Learning.Unique Subedi, Ambuj Tewari
2024AISTATSA Primal-Dual-Critic Algorithm for Offline Constrained Reinforcement Learning.Kihyuk Hong, Yuhang Li, Ambuj Tewari
2024AISTATSOffline Policy Evaluation and Optimization Under Confounding.Chinmaya Kausik, Yangyi Lu, Kevin Tan, Maggie Makar, Yixin Wang, Ambuj Tewari
2024AISTATSConformal Contextual Robust Optimization.Yash P. Patel, Sahana Rayan, Ambuj Tewari
2024AISTATSSequence Length Independent Norm-Based Generalization Bounds for Transformers.Jacob Trauger, Ambuj Tewari
2024ALTMulticlass Online Learnability under Bandit Feedback.Ananth Raman, Vinod Raman, Unique Subedi, Idan Mehalel, Ambuj Tewari
2024ALTOnline Infinite-Dimensional Regression: Learning Linear Operators.Unique Subedi, Vinod Raman, Ambuj Tewari
2024COLTApple Tasting: Combinatorial Dimensions and Minimax Rates.Vinod Raman, Unique Subedi, Ananth Raman, Ambuj Tewari
2024COLTOnline Learning with Set-valued Feedback.Vinod Raman, Unique Subedi, Ambuj Tewari
2024ICLRSample Efficient Myopic Exploration Through Multitask Reinforcement Learning with Diverse Tasks.Ziping Xu, Zifan Xu, Runxuan Jiang, Peter Stone, Ambuj Tewari
2024ICMLA Primal-Dual Algorithm for Offline Constrained Reinforcement Learning with Linear MDPs.Kihyuk Hong, Ambuj Tewari
2024ICMLVariational Inference with Coverage Guarantees in Simulation-Based Inference.Yash P. Patel, Declan McNamara, Jackson Loper, Jeffrey Regier, Ambuj Tewari
2023AISTATSAn Optimization-based Algorithm for Non-stationary Kernel Bandits without Prior Knowledge.Kihyuk Hong, Yuhang Li, Ambuj Tewari
2023COLTMulticlass Online Learning and Uniform Convergence.Steve Hanneke, Shay Moran, Vinod Raman, Unique Subedi, Ambuj Tewari
2023ICMLThompson Sampling for High-Dimensional Sparse Linear Contextual Bandits.Sunrit Chakraborty, Saptarshi Roy, Ambuj Tewari
2023ICMLLearning Mixtures of Markov Chains and MDPs.Chinmaya Kausik, Kevin Tan, Ambuj Tewari
2023UAILearning in online MDPs: is there a price for handling the communicating case?Gautam Chandrasekaran, Ambuj Tewari
2022AISTATSWeighted Gaussian Process Bandits for Non-stationary Environments.Yuntian Deng, Xingyu Zhou, Baekjin Kim, Ambuj Tewari, Abhishek Gupta, Ness B. Shroff
2022ICMLOn the Statistical Benefits of Curriculum Learning.Ziping Xu, Ambuj Tewari
2022UAIBalancing adaptability and non-exploitability in repeated games.Anthony DiGiovanni, Ambuj Tewari
2021AISTATSLow-Rank Generalized Linear Bandit Problems.Yangyi Lu, Amirhossein Meisami, Ambuj Tewari
2021AISTATSDecision Making Problems with Funnel Structure: A Multi-Task Learning Approach with Application to Email Marketing Campaigns.Ziping Xu, Amirhossein Meisami, Ambuj Tewari
2021UAIThompson sampling for Markov games with piecewise stationary opponent policies.Anthony DiGiovanni, Ambuj Tewari
2020AISTATSSample Complexity of Reinforcement Learning using Linearly Combined Model Ensembles.Aditya Modi, Nan Jiang, Ambuj Tewari, Satinder Singh
2020UAINo-regret Exploration in Contextual Reinforcement Learning.Aditya Modi, Ambuj Tewari
2020UAIRandomized Exploration for Non-Stationary Stochastic Linear Bandits.Baekjin Kim, Ambuj Tewari
2020UAIRegret Analysis of Bandit Problems with Causal Background Knowledge.Yangyi Lu, Amirhossein Meisami, Ambuj Tewari, William Yan
2020UAIWhat You See May Not Be What You Get: UCB Bandit Algorithms Robust to ε-Contamination.Laura Niss, Ambuj Tewari
2019AISTATSOnline Multiclass Boosting with Bandit Feedback.Daniel T. Zhang, Young Hun Jung, Ambuj Tewari
2018AISTATSOnline Boosting Algorithms for Multi-label Ranking.Young Hun Jung, Ambuj Tewari
2018ALTMarkov Decision Processes with Continuous Side Information.Aditya Modi, Nan Jiang, Satinder Singh, Ambuj Tewari
2016AAAIHandling Class Imbalance in Link Prediction Using Learning to Rank Techniques.Bopeng Li, Sougata Chaudhuri, Ambuj Tewari
2016AISTATSOnline Learning to Rank with Feedback at the Top.Sougata Chaudhuri, Ambuj Tewari
2016ICMLMixture Proportion Estimation via Kernel Embeddings of Distributions.Harish G. Ramaswamy, Clayton Scott, Ambuj Tewari
2016IJCAIOn Structural Properties of MDPs that Bound Loss Due to Shallow Planning.Nan Jiang, Satinder Singh, Ambuj Tewari
2015AISTATSOnline Ranking with Top-1 Feedback.Sougata Chaudhuri, Ambuj Tewari
2015ICMLConvex Calibrated Surrogates for Hierarchical Classification.Harish G. Ramaswamy, Ambuj Tewari, Shivani Agarwal
2015ICMLGeneralization error bounds for learning to rank: Does the length of document lists matter?Ambuj Tewari, Sougata Chaudhuri
2014COLTOnline Linear Optimization via Smoothing.Jacob D. Abernethy, Chansoo Lee, Abhinav Sinha, Ambuj Tewari
2013IJCAIOn Robust Estimation of High Dimensional Generalized Linear Models.Eunho Yang, Ambuj Tewari, Pradeep Ravikumar
2012ICMLOnline Bandit Learning against an Adaptive Adversary: from Regret to Policy Regret.Ofer Dekel, Ambuj Tewari, Raman Arora
2012ICMLPAC Subset Selection in Stochastic Multi-armed Bandits.Shivaram Kalyanakrishnan, Ambuj Tewari, Peter Auer, Peter Stone
2012ICMLScaling Up Coordinate Descent Algorithms for LargeChad Scherrer, Mahantesh Halappanavar, Ambuj Tewari, David Haglin
2012SIGIRParallelizing ListNet training using spark.Shilpa Shukla, Matthew Lease, Ambuj Tewari
2012UAIDeterministic MDPs with Adversarial Rewards and Bandit Feedback.Raman Arora, Ofer Dekel, Ambuj Tewari
2011CIKMExploiting longer cycles for link prediction in signed networks.Kai-Yang Chiang, Nagarajan Natarajan, Ambuj Tewari, Inderjit S. Dhillon
2010COLTComposite Objective Mirror Descent.John C. Duchi, Shai Shalev-Shwartz, Yoram Singer, Ambuj Tewari
2010COLTConvex Games in Banach Spaces.Karthik Sridharan, Ambuj Tewari
2009ICMLStochastic methods forShai Shalev-Shwartz, Ambuj Tewari
2009UAIREGAL: A Regularization based Algorithm for Reinforcement Learning in Weakly Communicating MDPs.Peter L. Bartlett, Ambuj Tewari
2008COLTOptimal Stragies and Minimax Lower Bounds for Online Convex Games.Jacob D. Abernethy, Peter L. Bartlett, Alexander Rakhlin, Ambuj Tewari
2008COLTHigh-Probability Regret Bounds for Bandit Online Linear Optimization.Peter L. Bartlett, Varsha Dani, Thomas P. Hayes, Sham M. Kakade, Alexander Rakhlin, Ambuj Tewari
2008ICMLEfficient bandit algorithms for online multiclass prediction.Sham M. Kakade, Shai Shalev-Shwartz, Ambuj Tewari
2007COLTBounded Parameter Markov Decision Processes with Average Reward Criterion.Ambuj Tewari, Peter L. Bartlett
2005COLTOn the Consistency of Multiclass Classification Methods.Ambuj Tewari, Peter L. Bartlett
2004COLTSparseness Versus Estimating Conditional Probabilities: Some Asymptotic Results.Peter L. Bartlett, Ambuj Tewari
2002HiPCA Parallel DFA Minimization Algorithm.Ambuj Tewari, Utkarsh Srivastava, P. Gupta