Subhankar Ghosh
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
23
Venues
14
Active years
2018–2025
Best venue rank
A*
Where they publish
Papers
23 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | ASRU | Open Full-duplex Voice Agent with Speech-to-Speech Language Model. | Edresson Casanova, Chen Chen, Kevin Hu, Ankita Pasad, Elena Rastorgueva, Seelan Lakshmi Narasimhan, Slyne Deng, Ehsan Hosseini-Asl, Piotr Zelasko, Valentin Mendelev, Subhankar Ghosh, Yifan Peng, Zhehuai Chen, Jason Li, Jagadeesh Balam, Vitaly Lavrukhin, Boris Ginsburg |
| 2025 | EMNLP | Koel-TTS: Enhancing LLM based Speech Generation with Preference Alignment and Classifier Free Guidance. | Shehzeen Samarah Hussain, Paarth Neekhara, Xuesong Yang, Edresson Casanova, Subhankar Ghosh, Roy Fejgin, Mikyas T. Desta, Rafael Valle, Jason Li |
| 2025 | ICASSP | TTS-Transducer: End-to-End Speech Synthesis with Neural Transducer. | Vladimir Bataev, Subhankar Ghosh, Vitaly Lavrukhin, Jason Li |
| 2025 | ICASSP | Low Frame-rate Speech Codec: a Codec Designed for Fast High-quality Speech LLM Training and Inference. | Edresson Casanova, Ryan Langman, Paarth Neekhara, Shehzeen Hussain, Jason Li, Subhankar Ghosh, Ante Jukic, Sang-gil Lee |
| 2025 | IJCNLP | The Visual Counter Turing Test (VCT²): A Benchmark for Evaluating AI-Generated Image Detection and the Visual AI Index (V_AI). | Nasrin Imanpour, Abhilekh Borah, Shashwat Bajpai, Subhankar Ghosh, Sainath Reddy Sankepally, Hasnat Md Abdullah, Nishoak Kosaraju, Shreyas Dixit, Ashhar Aziz, Shwetangshu Biswas, Vinija Jain, Aman Chadha, Song Wang, Amit P. Sheth, Amitava Das |
| 2025 | Interspeech | NanoCodec: Towards High-Quality Ultra Fast Speech LLM Inference. | Edresson Casanova, Paarth Neekhara, Ryan Langman, Shehzeen Hussain, Subhankar Ghosh, Xuesong Yang, Ante Jukic, Jason Li, Boris Ginsburg |
| 2025 | Interspeech | Efficient and Direct Duplex Modeling for Speech-to-Speech Language Model. | Ke Hu, Ehsan Hosseini-Asl, Chen Chen, Edresson Casanova, Subhankar Ghosh, Piotr Zelasko, Zhehuai Chen, Jason Li, Jagadeesh Balam, Boris Ginsburg |
| 2025 | WACV | FASTER: A Font-Agnostic Scene Text Editing and Rendering Framework. | Alloy Das, Sanket Biswas, Prasun Roy, Subhankar Ghosh, Umapada Pal, Michael Blumenstein, Josep Llads, Saumik Bhattacharya |
| 2024 | ICASSP | SALM: Speech-Augmented Language Model with in-Context Learning for Speech Recognition and Translation. | Zhehuai Chen, He Huang, Andrei Andrusenko, Oleksii Hrinchuk, Krishna C. Puvvada, Jason Li, Subhankar Ghosh, Jagadeesh Balam, Boris Ginsburg |
| 2024 | ICPR | λ-Color: Amplifying Long-Range Dependencies for Image Colorization. | Subhankar Ghosh, Saumik Bhattacharya, Prasun Roy, Umapada Pal, Michael Blumenstein |
| 2024 | ICPR | d-Sketch: Improving Visual Fidelity of Sketch-to-Image Translation with Pretrained Latent Diffusion Models without Retraining. | Prasun Roy, Saumik Bhattacharya, Subhankar Ghosh, Umapada Pal, Michael Blumenstein |
| 2024 | ICPR | Semantically Consistent Person Image Generation. | Prasun Roy, Saumik Bhattacharya, Subhankar Ghosh, Umapada Pal, Michael Blumenstein |
| 2024 | Interspeech | Improving Robustness of LLM-based Speech Synthesis by Learning Monotonic Alignment. | Paarth Neekhara, Shehzeen Hussain, Subhankar Ghosh, Jason Li, Boris Ginsburg |
| 2023 | AAAI | Improving Uncertainty Quantification of Deep Classifiers via Neighborhood Conformal Prediction: Novel Algorithm and Theoretical Analysis. | Subhankar Ghosh, Taha Belkhouja, Yan Yan, Janardhan Rao Doppa |
| 2023 | ICASSP | Vani: Very-Lightweight Accent-Controllable TTS for Native And Non-Native Speakers With Identity Preservation. | Rohan Badlani, Akshit Arora, Subhankar Ghosh, Rafael Valle, Kevin J. Shih, Joo Felipe Santos, Boris Ginsburg, Bryan Catanzaro |
| 2023 | Interspeech | Adapter-Based Extension of Multi-Speaker Text-To-Speech Model for New Speakers. | Cheng-Ping Hsieh, Subhankar Ghosh, Boris Ginsburg |
| 2023 | UAI | Probabilistically robust conformal prediction. | Subhankar Ghosh, Yuanjie Shi, Taha Belkhouja, Yan Yan, Jana Doppa, Brian Jones |
| 2022 | ECCV | TIPS: Text-Induced Pose Synthesis. | Prasun Roy, Subhankar Ghosh, Saumik Bhattacharya, Umapada Pal, Michael Blumenstein |
| 2022 | ICPR | Scene Aware Person Image Generation through Global Contextual Conditioning. | Prasun Roy, Subhankar Ghosh, Saumik Bhattacharya, Umapada Pal, Michael Blumenstein |
| 2021 | IJCNN | Adversarial Training of Variational Auto-encoders for Continual Zero-shot Learning(A-CZSL). | Subhankar Ghosh |
| 2021 | VRST | Validating Social Distancing through Deep Learning and VR-Based Digital Twins. | Abhishek Mukhopadhyay, G. S. Rajshekar Reddy, Subhankar Ghosh, L. R. D. Murthy, Pradipta Biswas |
| 2020 | CVPR | STEFANN: Scene Text Editor Using Font Adaptive Neural Network. | Prasun Roy, Saumik Bhattacharya, Subhankar Ghosh, Umapada Pal |
| 2018 | ICFHR | A CNN Based Framework for Unistroke Numeral Recognition in Air-Writing. | Prasun Roy, Subhankar Ghosh, Umapada Pal |