Rishabh Joshi
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
15
Venues
7
Active years
2018–2025
Best venue rank
A*
Where they publish
Papers
15 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | ICLR | RRM: Robust Reward Model Training Mitigates Reward Hacking. | Tianqi Liu, Wei Xiong, Jie Ren, Lichang Chen, Junru Wu, Rishabh Joshi, Yang Gao, Jiaming Shen, Zhen Qin, Tianhe Yu, Daniel Sohn, Anastasia Makarova, Jeremiah Zhe Liu, Yuan Liu, Bilal Piot, Abe Ittycheriah, Aviral Kumar, Mohammad Saleh |
| 2025 | ICLR | Building Math Agents with Multi-Turn Iterative Preference Learning. | Wei Xiong, Chengshuai Shi, Jiaming Shen, Aviv Rosenberg, Zhen Qin, Daniele Calandriello, Misha Khalman, Rishabh Joshi, Bilal Piot, Mohammad Saleh, Chi Jin, Tong Zhang, Tianqi Liu |
| 2025 | ICLR | Learning from negative feedback, or positive feedback or both. | Abbas Abdolmaleki, Bilal Piot, Bobak Shahriari, Jost Tobias Springenberg, Tim Hertweck, Michael Bloesch, Rishabh Joshi, Thomas Lampe, Junhyuk Oh, Nicolas Heess, Jonas Buchli, Martin A. Riedmiller |
| 2025 | ICML | Reward-Guided Prompt Evolving in Reinforcement Learning for LLMs. | Ziyu Ye, Rishabh Agarwal, Tianqi Liu, Rishabh Joshi, Sarmishta Velury, Quoc V. Le, Qijun Tan, Yuan Liu |
| 2025 | NAACL | LiPO: Listwise Preference Optimization through Learning-to-Rank. | Tianqi Liu, Zhen Qin, Junru Wu, Jiaming Shen, Misha Khalman, Rishabh Joshi, Yao Zhao, Mohammad Saleh, Simon Baumgartner, Jialu Liu, Peter J. Liu, Xuanhui Wang |
| 2024 | ICLR | Statistical Rejection Sampling Improves Preference Optimization. | Tianqi Liu, Yao Zhao, Rishabh Joshi, Misha Khalman, Mohammad Saleh, Peter J. Liu, Jialu Liu |
| 2024 | ICML | Human Alignment of Large Language Models through Online Preference Optimisation. | Daniele Calandriello, Zhaohan Daniel Guo, Rmi Munos, Mark Rowland, Yunhao Tang, Bernardo vila Pires, Pierre Harvey Richemond, Charline Le Lan, Michal Valko, Tianqi Liu, Rishabh Joshi, Zeyu Zheng, Bilal Piot |
| 2023 | EACL | Unsupervised Keyphrase Extraction via Interpretable Neural Networks. | Rishabh Joshi, Vidhisha Balachandran, Emily Saldanha, Maria Glenski, Svitlana Volkova, Yulia Tsvetkov |
| 2023 | ICLR | Calibrating Sequence likelihood Improves Conditional Language Generation. | Yao Zhao, Misha Khalman, Rishabh Joshi, Shashi Narayan, Mohammad Saleh, Peter J. Liu |
| 2021 | EACL | ResPer: Computationally Modelling Resisting Strategies in Persuasive Conversations. | Ritam Dutt, Sayan Sinha, Rishabh Joshi, Surya Shekhar Chakraborty, Meredith Riggs, Xinru Yan, Haogang Bao, Carolyn P. Ros |
| 2021 | ICLR | DialoGraph: Incorporating Interpretable Strategy-Graph Networks into Negotiation Dialogues. | Rishabh Joshi, Vidhisha Balachandran, Shikhar Vashishth, Alan W. Black, Yulia Tsvetkov |
| 2020 | EMNLP | Keeping Up Appearances: Computational Modeling of Face Acts in Persuasion Oriented Discussions. | Ritam Dutt, Rishabh Joshi, Carolyn P. Ros |
| 2020 | ICWSM | Analysing the Extent of Misinformation in Cancer Related Tweets. | Rakesh Bal, Sayan Sinha, Swastika Dutta, Rishabh Joshi, Sayan Ghosh, Ritam Dutt |
| 2020 | LREC | AMUSED: A Multi-Stream Vector Representation Method for Use in Natural Dialogue. | Gaurav Kumar, Rishabh Joshi, Jaspreet Singh, Promod Yenigalla |
| 2018 | EMNLP | RESIDE: Improving Distantly-Supervised Neural Relation Extraction using Side Information. | Shikhar Vashishth, Rishabh Joshi, Sai Suman Prayaga, Chiranjib Bhattacharyya, Partha P. Talukdar |