R. Manmatha
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
58
Venues
17
Active years
1989–2025
Best venue rank
A*
Where they publish
Papers
58 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | ACL | R-VLM: Region-Aware Vision Language Model for Precise GUI Grounding. | Joonhyung Park, Peng Tang, Sagnik Das, Srikar Appalaraju, Kunwar Yashraj Singh, R. Manmatha, Shabnam Ghadar |
| 2025 | CVPR | Scaling up Image Segmentation across Data and Tasks. | Pei Wang, Zhaowei Cai, Hao Yang, Ashwin Swaminathan, R. Manmatha, Stefano Soatto |
| 2024 | AAAI | DocFormerv2: Local Features for Document Understanding. | Srikar Appalaraju, Peng Tang, Qi Dong, Nishant Sankaran, Yichu Zhou, R. Manmatha |
| 2024 | AAAI | No Head Left Behind - Multi-Head Alignment Distillation for Transformers. | Tianyang Zhao, Kunwar Yashraj Singh, Srikar Appalaraju, Peng Tang, Vijay Mahadevan, R. Manmatha, Ying Nian Wu |
| 2024 | CVPR | On the Scalability of Diffusion-based Text-to-Image Generation. | Hao Li, Yang Zou, Ying Wang, Orchid Majumder, Yusheng Xie, R. Manmatha, Ashwin Swaminathan, Zhuowen Tu, Stefano Ermon, Stefano Soatto |
| 2024 | ECCV | VisFocus: Prompt-Guided Vision Encoders for OCR-Free Dense Document Understanding. | Ofir Abramovich, Niv Nayman, Sharon Fogel, Inbal Lavi, Ron Litman, Shahar Tsiper, Royee Tichauer, Srikar Appalaraju, Shai Mazor, R. Manmatha |
| 2024 | EMNLP | DocKD: Knowledge Distillation from LLMs for Open-World Document Understanding Models. | Sungnyun Kim, Haofu Liao, Srikar Appalaraju, Peng Tang, Zhuowen Tu, Ravi Kumar Satzoda, R. Manmatha, Vijay Mahadevan, Stefano Soatto |
| 2024 | ICDAR | ICDAR 2024 Competition on Recognition and VQA on Handwritten Documents. | Ajoy Mondal, Vijay Mahadevan, R. Manmatha, C. V. Jawahar |
| 2024 | NAACL | Multiple-Question Multiple-Answer Text-VQA. | Peng Tang, Srikar Appalaraju, R. Manmatha, Yusheng Xie, Vijay Mahadevan |
| 2024 | NAACL | DEED: Dynamic Early Exit on Decoder for Accelerating Encoder-Decoder Transformer Models. | Peng Tang, Pengkai Zhu, Tian Li, Srikar Appalaraju, Vijay Mahadevan, R. Manmatha |
| 2023 | CVPR | PolyFormer: Referring Image Segmentation as Sequential Polygon Generation. | Jiang Liu, Hui Ding, Zhaowei Cai, Yuting Zhang, Ravi Kumar Satzoda, Vijay Mahadevan, R. Manmatha |
| 2023 | ICCV | DocTr: Document Transformer for Structured Information Extraction in Documents. | Haofu Liao, Aruni RoyChowdhury, Weijian Li, Ankan Bansal, Yuting Zhang, Zhuowen Tu, Ravi Kumar Satzoda, R. Manmatha, Vijay Mahadevan |
| 2022 | CVPR | ResNeSt: Split-Attention Networks. | Hang Zhang, Chongruo Wu, Zhongyue Zhang, Yi Zhu, Haibin Lin, Zhi Zhang, Yue Sun, Tong He, Jonas Mueller, R. Manmatha, Mu Li, Alexander J. Smola |
| 2022 | CVPR | LaTr: Layout-Aware Transformer for Scene-Text VQA. | Ali Furkan Biten, Ron Litman, Yusheng Xie, Srikar Appalaraju, R. Manmatha |
| 2022 | CVPR | Towards Weakly-Supervised Text Spotting using a Multi-Task Transformer. | Yair Kittenplon, Inbal Lavi, Sharon Fogel, Yarin Bar, R. Manmatha, Pietro Perona |
| 2022 | ECCV | YORO - Lightweight End to End Visual Grounding. | Chih-Hui Ho, Srikar Appalaraju, Bhavan Jasani, R. Manmatha, Nuno Vasconcelos |
| 2022 | ECCV | GLASS: Global to Local Attention for Scene-Text Spotting. | Roi Ronen, Shahar Tsiper, Oron Anschel, Inbal Lavi, Amir Markovitz, R. Manmatha |
| 2022 | ECCV | On Calibration of Scene-Text Recognition Models. | Ron Slossberg, Oron Anschel, Amir Markovitz, Ron Litman, Aviad Aberdam, Shahar Tsiper, Shai Mazor, Jon Wu, R. Manmatha |
| 2021 | CVPR | Sequence-to-Sequence Contrastive Learning for Text Recognition. | Aviad Aberdam, Ron Litman, Shahar Tsiper, Oron Anschel, Ron Slossberg, Shai Mazor, R. Manmatha, Pietro Perona |
| 2021 | ICCV | DocFormer: End-to-End Transformer for Document Understanding. | Srikar Appalaraju, Bhavan Jasani, Bhargava Urala Kota, Yusheng Xie, R. Manmatha |
| 2021 | WACV | Saliency Driven Perceptual Image Compression. | Yash Patel, Srikar Appalaraju, R. Manmatha |
| 2020 | CVPR | SCATTER: Selective Context Attentional Scene Text Recognizer. | Ron Litman, Oron Anschel, Shahar Tsiper, Roee Litman, Shai Mazor, R. Manmatha |
| 2018 | CVPR | Compressed Video Action Recognition. | Chao-Yuan Wu, Manzil Zaheer, Hexiang Hu, R. Manmatha, Alexander J. Smola, Philipp Krhenbhl |
| 2017 | ICCV | Sampling Matters in Deep Embedding Learning. | R. Manmatha, Chao-Yuan Wu, Alexander J. Smola, Philipp Krhenbhl |
| 2016 | CVPR | Deep Decision Network for Multi-class Image Classification. | Venkatesh N. Murthy, Vivek K. Singh, Terrence Chen, R. Manmatha, Dorin Comaniciu |
| 2016 | ECCV | Efficient Exploration of Text Regions in Natural Scene Images Using Adaptive Image Sampling. | Ismet Zeki Yalniz, Douglas Gray, R. Manmatha |
| 2014 | DAS | Sequential Word Spotting in Historical Handwritten Documents. | David Fernndez Mota, R. Manmatha, Alicia Forns, Josep Llads |
| 2014 | SIGIR | Incorporating query-specific feedback into learning-to-rank models. | Ethem F. Can, W. Bruce Croft, R. Manmatha |
| 2013 | CIKM | Predicting retweet count using visual cues. | Ethem F. Can, Hseyin Oktay, R. Manmatha |
| 2013 | CVPR | Formulating Action Recognition as a Ranking Problem. | Ethem F. Can, R. Manmatha |
| 2013 | ICDAR | Creating an Improved Version Using Noisy OCR from Multiple Editions. | David Wemhoener, Ismet Zeki Yalniz, R. Manmatha |
| 2012 | DAS | An Efficient Framework for Searching Text in Noisy Document Images. | Ismet Zeki Yalniz, R. Manmatha |
| 2012 | ICFHR | On Influence of Line Segmentation in Efficient Word Segmentation in Old Manuscripts. | David Fernndez, Josep Llads, Alicia Forns, R. Manmatha |
| 2012 | SIGIR | A framework for manipulating and searching multiple retrieval types. | Marc-Allen Cartright, Ethem F. Can, William Dabney, Jeff Dalton, Logan Giorda, Kriste Krstovski, Xiaoye Wu, Ismet Zeki Yalniz, James Allan, R. Manmatha, David A. Smith |
| 2012 | SIGIR | Finding translations in scanned book collections. | Ismet Zeki Yalniz, R. Manmatha |
| 2011 | CIKM | Mining relational structure from millions of books: position paper. | David A. Smith, R. Manmatha, James Allan |
| 2011 | CIKM | Partial duplicate detection for large book collections. | Ismet Zeki Yalniz, Ethem F. Can, R. Manmatha |
| 2010 | ICFHR | Adapting BLSTM Neural Network Based Keyword Spotting Trained on Modern Data to Historical Documents. | Volkmar Frinken, Andreas Fischer, Horst Bunke, R. Manmatha |
| 2008 | SENSYS | Distributed image search in camera sensor networks. | Tingxin Yan, Deepak Ganesan, R. Manmatha |
| 2007 | ACCV | Efficient Search in Document Image Collections. | Anand Kumar, C. V. Jawahar, R. Manmatha |
| 2006 | DAS | Aligning Transcripts to Automatically Segmented Handwritten Manuscripts. | Jamie L. Rothfeder, R. Manmatha, Toni M. Rath |
| 2005 | ICASSP | Combining text and audio-visual features in video indexing. | Shih-Fu Chang, R. Manmatha, Tat-Seng Chua |
| 2005 | SIGIR | Boosted decision trees for word recognition in handwritten document retrieval. | Nicholas R. Howe, Toni M. Rath, R. Manmatha |
| 2004 | SIGIR | A search engine for historical manuscript images. | Toni M. Rath, R. Manmatha, Victor Lavrenko |
| 2003 | CVPR | Word Image Matching Using Dynamic Time Warping. | Toni M. Rath, R. Manmatha |
| 2003 | ICDAR | Features for Word Spotting in Historical Manuscripts. | Toni M. Rath, R. Manmatha |
| 2003 | ICNP | Mobile Distributed Information Retrieval for Highly-Partitioned Networks. | Katrina M. Hanna, Brian Neil Levine, R. Manmatha |
| 2003 | SIGIR | Automatic image annotation and retrieval using cross-media relevance models. | Jiwoon Jeon, Victor Lavrenko, R. Manmatha |
| 2002 | SIGIR | A critical examination of TDT's cost function. | R. Manmatha, Ao Feng, James Allan |
| 2001 | ICCV | Automatic Segmentation and Indexing in a Database of Bird Images. | Madirakshi Das, R. Manmatha |
| 2001 | SIGIR | Modeling Score Distributions for Combining the Outputs of Search Engines. | R. Manmatha, Toni M. Rath, Fangfang Feng |
| 1998 | ICCV | Retrieving Images by Appearance. | Srinivas Ravela, R. Manmatha |
| 1997 | SIGIR | Image Retrieval by Appearance. | Srinivas Ravela, R. Manmatha |
| 1996 | CVPR | Word Spotting: A New Approach to Indexing Handwriting. | R. Manmatha, Chengfeng Han, Edward M. Riseman |
| 1996 | ECCV | Image Retrieval Using Scale-Space Matching. | Srinivas Ravela, R. Manmatha, Edward M. Riseman |
| 1994 | CVPR | A framework for recovering affine transforms using points, lines or image brightnesses. | R. Manmatha |
| 1994 | ECCV | Measuring the Affine Transform Using Gaussian Filters. | R. Manmatha |
| 1989 | CVPR | A data set for quantitative motion analysis. | Rabindranath Dutta, R. Manmatha, Lance R. Williams, Edward M. Riseman |