Skip to content

Gerald Tesauro

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

34

Venues

12

Active years

1988–2022

Best venue rank

A*

Where they publish

Papers

34 indexed papers, newest first.

YearVenueTitleAuthors
2022AAAIContext-Specific Representation Abstraction for Deep Option Learning.Marwa Abdulhai, Dong-Ki Kim, Matthew Riemer, Miao Liu, Gerald Tesauro, Jonathan P. How
2021AAAIRL Generalization in a Theory of Mind Game Through a Sleep Metaphor (Student Abstract).Tailia Malloy, Tim Klinger, Miao Liu, Gerald Tesauro, Matthew Riemer, Chris R. Sims
2021AAAIText-based RL Agents with Commonsense Knowledge: New Challenges, Environments and Baselines.Keerthiram Murugesan, Mattia Atzeni, Pavan Kapanipathi, Pushkar Shukla, Sadhana Kumaravel, Gerald Tesauro, Kartik Talamadupula, Mrinmaya Sachan, Murray Campbell
2021CogSciModeling Capacity-Limited Decision Making Using a Variational Autoencoder.Tailia Malloy, Tim Klinger, Miao Liu, Gerald Tesauro, Matthew Riemer, Chris R. Sims
2021ICMLA Policy Gradient Algorithm for Learning to Learn in Multiagent Reinforcement Learning.Dong-Ki Kim, Miao Liu, Matthew Riemer, Chuangchuang Sun, Marwa Abdulhai, Golnaz Habibi, Sebastian Lopez-Cot, Gerald Tesauro, Jonathan P. How
2021IJCAIEfficient Black-Box Planning Using Macro-Actions with Focused Effects.Cameron Allen, Michael Katz, Tim Klinger, George Konidaris, Matthew Riemer, Gerald Tesauro
2020AAAIOn the Role of Weight Sharing During Deep Option Learning.Matthew Riemer, Ignacio Cases, Clemens Rosenbaum, Miao Liu, Gerald Tesauro
2019AAAIHybrid Reinforcement Learning with Expert State Sequences.Xiaoxiao Guo, Shiyu Chang, Mo Yu, Gerald Tesauro, Murray Campbell
2019AAAILearning to Teach in Cooperative Multiagent Reinforcement Learning.Shayegan Omidshafiei, Dong-Ki Kim, Miao Liu, Gerald Tesauro, Matthew Riemer, Christopher Amato, Murray Campbell, Jonathan P. How
2019ICLRLearning to Learn without Forgetting by Maximizing Transfer and Minimizing Interference.Matthew Riemer, Ignacio Cases, Robert Ajemian, Miao Liu, Irina Rish, Yuhai Tu, Gerald Tesauro
2018ICLREigenoption Discovery through the Deep Successor Representation.Marlos C. Machado, Clemens Rosenbaum, Xiaoxiao Guo, Miao Liu, Gerald Tesauro, Murray Campbell
2018ICLREvidence Aggregation for Answer Re-Ranking in Open-Domain Question Answering.Shuohang Wang, Mo Yu, Jing Jiang, Wei Zhang, Xiaoxiao Guo, Shiyu Chang, Zhiguo Wang, Tim Klinger, Gerald Tesauro, Murray Campbell
2018NAACLDiverse Few-Shot Text Classification with Multiple Metrics.Mo Yu, Xiaoxiao Guo, Jinfeng Yi, Shiyu Chang, Saloni Potdar, Yu Cheng, Gerald Tesauro, Haoyu Wang, Bowen Zhou
2017AAAIMultiresolution Recurrent Neural Networks: An Application to Dialogue Response Generation.Iulian Vlad Serban, Tim Klinger, Gerald Tesauro, Kartik Talamadupula, Bowen Zhou, Yoshua Bengio, Aaron C. Courville
2017AAAIOptimal Sequential Drilling for Hydrocarbon Field Development Planning.Ruben Rodriguez Torrado, Jesus Rios, Gerald Tesauro
2016AAAISelecting Near-Optimal Learners via Incremental Data Allocation.Ashish Sabharwal, Horst Samulowitz, Gerald Tesauro
2015AAAIBudgeted Prediction with Expert Advice.Kareem Amin, Satyen Kale, Gerald Tesauro, Deepak S. Turaga
2015AAAITowards Cognitive Automation of Data Science.Alain Biem, Maria Butrico, Mark Feblowitz, Tim Klinger, Yuri Malitsky, Kenney Ng, Adam Perer, Chandra Reddy, Anton Riabov, Horst Samulowitz, Daby M. Sow, Gerald Tesauro, Deepak S. Turaga
2012AAMASPlaying repeated Stackelberg games with unknown opponents.Janusz Marecki, Gerald Tesauro, Richard B. Segal
2012WSCApplying a framework for healthcare incentives simulation.Joseph P. Bigus, Ching-Hua Chen-Ritzo, Keith Hermiz, Gerald Tesauro, Robert Sorrentino
2010UAIBayesian Inference in Monte-Carlo Tree Search.Gerald Tesauro, V. T. Rajan, Richard B. Segal
2009ICMLMonte-Carlo simulation balancing.David Silver, Gerald Tesauro
2008ISAIMActive Collaborative Prediction with Maximum Margin Matrix Factorization.Irina Rish, Gerald Tesauro
2007IMEstimating End-to-End Performance by Collaborative Prediction with Active Sampling.Irina Rish, Gerald Tesauro
2005AAAINew Approaches to Optimization and Utility Elicitation in Autonomic Computing.Relu Patrascu, Craig Boutilier, Rajarshi Das, Jeffrey O. Kephart, Gerald Tesauro, William E. Walsh
2005AAAIOnline Resource Allocation Using Decompositional Reinforcement Learning.Gerald Tesauro
2003UAICooperative Negotiation in Autonomic Systems using Incremental Utility Elicitation.Craig Boutilier, Rajarshi Das, Jeffrey O. Kephart, Gerald Tesauro, William E. Walsh
2001IJCAIAgent-Human Interactions in the Continuous Double Auction.Rajarshi Das, James E. Hanson, Jeffrey O. Kephart, Gerald Tesauro
2000ICMLPseudo-convergent Q-Learning by Competitive Pricebots.Jeffrey O. Kephart, Gerald Tesauro
2000ICMLMulti-agent Q-learning and Regression Trees for Automated Pricing Decisions.Manu Sridharan, Gerald Tesauro
1995IJCAIBiologically Inspired Defenses Against Computer Viruses.Jeffrey O. Kephart, Gregory B. Sorkin, William C. Arnold, David M. Chess, Gerald Tesauro, Steve R. White
1992ICMLTemporal Difference Learning of Backgammon Strategy.Gerald Tesauro
1990IJCNNNeurogammon: a neural-network backgammon program.Gerald Tesauro
1988ICMLConnectionist Learning of Expert Backgammon Evaluations.Gerald Tesauro