Skip to content

Rishabh Joshi

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

15

Venues

7

Active years

2018–2025

Best venue rank

A*

Where they publish

Papers

15 indexed papers, newest first.

YearVenueTitleAuthors
2025ICLRRRM: Robust Reward Model Training Mitigates Reward Hacking.Tianqi Liu, Wei Xiong, Jie Ren, Lichang Chen, Junru Wu, Rishabh Joshi, Yang Gao, Jiaming Shen, Zhen Qin, Tianhe Yu, Daniel Sohn, Anastasia Makarova, Jeremiah Zhe Liu, Yuan Liu, Bilal Piot, Abe Ittycheriah, Aviral Kumar, Mohammad Saleh
2025ICLRBuilding Math Agents with Multi-Turn Iterative Preference Learning.Wei Xiong, Chengshuai Shi, Jiaming Shen, Aviv Rosenberg, Zhen Qin, Daniele Calandriello, Misha Khalman, Rishabh Joshi, Bilal Piot, Mohammad Saleh, Chi Jin, Tong Zhang, Tianqi Liu
2025ICLRLearning from negative feedback, or positive feedback or both.Abbas Abdolmaleki, Bilal Piot, Bobak Shahriari, Jost Tobias Springenberg, Tim Hertweck, Michael Bloesch, Rishabh Joshi, Thomas Lampe, Junhyuk Oh, Nicolas Heess, Jonas Buchli, Martin A. Riedmiller
2025ICMLReward-Guided Prompt Evolving in Reinforcement Learning for LLMs.Ziyu Ye, Rishabh Agarwal, Tianqi Liu, Rishabh Joshi, Sarmishta Velury, Quoc V. Le, Qijun Tan, Yuan Liu
2025NAACLLiPO: Listwise Preference Optimization through Learning-to-Rank.Tianqi Liu, Zhen Qin, Junru Wu, Jiaming Shen, Misha Khalman, Rishabh Joshi, Yao Zhao, Mohammad Saleh, Simon Baumgartner, Jialu Liu, Peter J. Liu, Xuanhui Wang
2024ICLRStatistical Rejection Sampling Improves Preference Optimization.Tianqi Liu, Yao Zhao, Rishabh Joshi, Misha Khalman, Mohammad Saleh, Peter J. Liu, Jialu Liu
2024ICMLHuman Alignment of Large Language Models through Online Preference Optimisation.Daniele Calandriello, Zhaohan Daniel Guo, Rmi Munos, Mark Rowland, Yunhao Tang, Bernardo vila Pires, Pierre Harvey Richemond, Charline Le Lan, Michal Valko, Tianqi Liu, Rishabh Joshi, Zeyu Zheng, Bilal Piot
2023EACLUnsupervised Keyphrase Extraction via Interpretable Neural Networks.Rishabh Joshi, Vidhisha Balachandran, Emily Saldanha, Maria Glenski, Svitlana Volkova, Yulia Tsvetkov
2023ICLRCalibrating Sequence likelihood Improves Conditional Language Generation.Yao Zhao, Misha Khalman, Rishabh Joshi, Shashi Narayan, Mohammad Saleh, Peter J. Liu
2021EACLResPer: Computationally Modelling Resisting Strategies in Persuasive Conversations.Ritam Dutt, Sayan Sinha, Rishabh Joshi, Surya Shekhar Chakraborty, Meredith Riggs, Xinru Yan, Haogang Bao, Carolyn P. Ros
2021ICLRDialoGraph: Incorporating Interpretable Strategy-Graph Networks into Negotiation Dialogues.Rishabh Joshi, Vidhisha Balachandran, Shikhar Vashishth, Alan W. Black, Yulia Tsvetkov
2020EMNLPKeeping Up Appearances: Computational Modeling of Face Acts in Persuasion Oriented Discussions.Ritam Dutt, Rishabh Joshi, Carolyn P. Ros
2020ICWSMAnalysing the Extent of Misinformation in Cancer Related Tweets.Rakesh Bal, Sayan Sinha, Swastika Dutta, Rishabh Joshi, Sayan Ghosh, Ritam Dutt
2020LRECAMUSED: A Multi-Stream Vector Representation Method for Use in Natural Dialogue.Gaurav Kumar, Rishabh Joshi, Jaspreet Singh, Promod Yenigalla
2018EMNLPRESIDE: Improving Distantly-Supervised Neural Relation Extraction using Side Information.Shikhar Vashishth, Rishabh Joshi, Sai Suman Prayaga, Chiranjib Bhattacharyya, Partha P. Talukdar