Gerald Tesauro
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
34
Venues
12
Active years
1988–2022
Best venue rank
A*
Where they publish
Papers
34 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2022 | AAAI | Context-Specific Representation Abstraction for Deep Option Learning. | Marwa Abdulhai, Dong-Ki Kim, Matthew Riemer, Miao Liu, Gerald Tesauro, Jonathan P. How |
| 2021 | AAAI | RL Generalization in a Theory of Mind Game Through a Sleep Metaphor (Student Abstract). | Tailia Malloy, Tim Klinger, Miao Liu, Gerald Tesauro, Matthew Riemer, Chris R. Sims |
| 2021 | AAAI | Text-based RL Agents with Commonsense Knowledge: New Challenges, Environments and Baselines. | Keerthiram Murugesan, Mattia Atzeni, Pavan Kapanipathi, Pushkar Shukla, Sadhana Kumaravel, Gerald Tesauro, Kartik Talamadupula, Mrinmaya Sachan, Murray Campbell |
| 2021 | CogSci | Modeling Capacity-Limited Decision Making Using a Variational Autoencoder. | Tailia Malloy, Tim Klinger, Miao Liu, Gerald Tesauro, Matthew Riemer, Chris R. Sims |
| 2021 | ICML | A Policy Gradient Algorithm for Learning to Learn in Multiagent Reinforcement Learning. | Dong-Ki Kim, Miao Liu, Matthew Riemer, Chuangchuang Sun, Marwa Abdulhai, Golnaz Habibi, Sebastian Lopez-Cot, Gerald Tesauro, Jonathan P. How |
| 2021 | IJCAI | Efficient Black-Box Planning Using Macro-Actions with Focused Effects. | Cameron Allen, Michael Katz, Tim Klinger, George Konidaris, Matthew Riemer, Gerald Tesauro |
| 2020 | AAAI | On the Role of Weight Sharing During Deep Option Learning. | Matthew Riemer, Ignacio Cases, Clemens Rosenbaum, Miao Liu, Gerald Tesauro |
| 2019 | AAAI | Hybrid Reinforcement Learning with Expert State Sequences. | Xiaoxiao Guo, Shiyu Chang, Mo Yu, Gerald Tesauro, Murray Campbell |
| 2019 | AAAI | Learning to Teach in Cooperative Multiagent Reinforcement Learning. | Shayegan Omidshafiei, Dong-Ki Kim, Miao Liu, Gerald Tesauro, Matthew Riemer, Christopher Amato, Murray Campbell, Jonathan P. How |
| 2019 | ICLR | Learning to Learn without Forgetting by Maximizing Transfer and Minimizing Interference. | Matthew Riemer, Ignacio Cases, Robert Ajemian, Miao Liu, Irina Rish, Yuhai Tu, Gerald Tesauro |
| 2018 | ICLR | Eigenoption Discovery through the Deep Successor Representation. | Marlos C. Machado, Clemens Rosenbaum, Xiaoxiao Guo, Miao Liu, Gerald Tesauro, Murray Campbell |
| 2018 | ICLR | Evidence Aggregation for Answer Re-Ranking in Open-Domain Question Answering. | Shuohang Wang, Mo Yu, Jing Jiang, Wei Zhang, Xiaoxiao Guo, Shiyu Chang, Zhiguo Wang, Tim Klinger, Gerald Tesauro, Murray Campbell |
| 2018 | NAACL | Diverse Few-Shot Text Classification with Multiple Metrics. | Mo Yu, Xiaoxiao Guo, Jinfeng Yi, Shiyu Chang, Saloni Potdar, Yu Cheng, Gerald Tesauro, Haoyu Wang, Bowen Zhou |
| 2017 | AAAI | Multiresolution Recurrent Neural Networks: An Application to Dialogue Response Generation. | Iulian Vlad Serban, Tim Klinger, Gerald Tesauro, Kartik Talamadupula, Bowen Zhou, Yoshua Bengio, Aaron C. Courville |
| 2017 | AAAI | Optimal Sequential Drilling for Hydrocarbon Field Development Planning. | Ruben Rodriguez Torrado, Jesus Rios, Gerald Tesauro |
| 2016 | AAAI | Selecting Near-Optimal Learners via Incremental Data Allocation. | Ashish Sabharwal, Horst Samulowitz, Gerald Tesauro |
| 2015 | AAAI | Budgeted Prediction with Expert Advice. | Kareem Amin, Satyen Kale, Gerald Tesauro, Deepak S. Turaga |
| 2015 | AAAI | Towards Cognitive Automation of Data Science. | Alain Biem, Maria Butrico, Mark Feblowitz, Tim Klinger, Yuri Malitsky, Kenney Ng, Adam Perer, Chandra Reddy, Anton Riabov, Horst Samulowitz, Daby M. Sow, Gerald Tesauro, Deepak S. Turaga |
| 2012 | AAMAS | Playing repeated Stackelberg games with unknown opponents. | Janusz Marecki, Gerald Tesauro, Richard B. Segal |
| 2012 | WSC | Applying a framework for healthcare incentives simulation. | Joseph P. Bigus, Ching-Hua Chen-Ritzo, Keith Hermiz, Gerald Tesauro, Robert Sorrentino |
| 2010 | UAI | Bayesian Inference in Monte-Carlo Tree Search. | Gerald Tesauro, V. T. Rajan, Richard B. Segal |
| 2009 | ICML | Monte-Carlo simulation balancing. | David Silver, Gerald Tesauro |
| 2008 | ISAIM | Active Collaborative Prediction with Maximum Margin Matrix Factorization. | Irina Rish, Gerald Tesauro |
| 2007 | IM | Estimating End-to-End Performance by Collaborative Prediction with Active Sampling. | Irina Rish, Gerald Tesauro |
| 2005 | AAAI | New Approaches to Optimization and Utility Elicitation in Autonomic Computing. | Relu Patrascu, Craig Boutilier, Rajarshi Das, Jeffrey O. Kephart, Gerald Tesauro, William E. Walsh |
| 2005 | AAAI | Online Resource Allocation Using Decompositional Reinforcement Learning. | Gerald Tesauro |
| 2003 | UAI | Cooperative Negotiation in Autonomic Systems using Incremental Utility Elicitation. | Craig Boutilier, Rajarshi Das, Jeffrey O. Kephart, Gerald Tesauro, William E. Walsh |
| 2001 | IJCAI | Agent-Human Interactions in the Continuous Double Auction. | Rajarshi Das, James E. Hanson, Jeffrey O. Kephart, Gerald Tesauro |
| 2000 | ICML | Pseudo-convergent Q-Learning by Competitive Pricebots. | Jeffrey O. Kephart, Gerald Tesauro |
| 2000 | ICML | Multi-agent Q-learning and Regression Trees for Automated Pricing Decisions. | Manu Sridharan, Gerald Tesauro |
| 1995 | IJCAI | Biologically Inspired Defenses Against Computer Viruses. | Jeffrey O. Kephart, Gregory B. Sorkin, William C. Arnold, David M. Chess, Gerald Tesauro, Steve R. White |
| 1992 | ICML | Temporal Difference Learning of Backgammon Strategy. | Gerald Tesauro |
| 1990 | IJCNN | Neurogammon: a neural-network backgammon program. | Gerald Tesauro |
| 1988 | ICML | Connectionist Learning of Expert Backgammon Evaluations. | Gerald Tesauro |