Skip to content

R. Manmatha

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

58

Venues

17

Active years

1989–2025

Best venue rank

A*

Where they publish

Papers

58 indexed papers, newest first.

YearVenueTitleAuthors
2025ACLR-VLM: Region-Aware Vision Language Model for Precise GUI Grounding.Joonhyung Park, Peng Tang, Sagnik Das, Srikar Appalaraju, Kunwar Yashraj Singh, R. Manmatha, Shabnam Ghadar
2025CVPRScaling up Image Segmentation across Data and Tasks.Pei Wang, Zhaowei Cai, Hao Yang, Ashwin Swaminathan, R. Manmatha, Stefano Soatto
2024AAAIDocFormerv2: Local Features for Document Understanding.Srikar Appalaraju, Peng Tang, Qi Dong, Nishant Sankaran, Yichu Zhou, R. Manmatha
2024AAAINo Head Left Behind - Multi-Head Alignment Distillation for Transformers.Tianyang Zhao, Kunwar Yashraj Singh, Srikar Appalaraju, Peng Tang, Vijay Mahadevan, R. Manmatha, Ying Nian Wu
2024CVPROn the Scalability of Diffusion-based Text-to-Image Generation.Hao Li, Yang Zou, Ying Wang, Orchid Majumder, Yusheng Xie, R. Manmatha, Ashwin Swaminathan, Zhuowen Tu, Stefano Ermon, Stefano Soatto
2024ECCVVisFocus: Prompt-Guided Vision Encoders for OCR-Free Dense Document Understanding.Ofir Abramovich, Niv Nayman, Sharon Fogel, Inbal Lavi, Ron Litman, Shahar Tsiper, Royee Tichauer, Srikar Appalaraju, Shai Mazor, R. Manmatha
2024EMNLPDocKD: Knowledge Distillation from LLMs for Open-World Document Understanding Models.Sungnyun Kim, Haofu Liao, Srikar Appalaraju, Peng Tang, Zhuowen Tu, Ravi Kumar Satzoda, R. Manmatha, Vijay Mahadevan, Stefano Soatto
2024ICDARICDAR 2024 Competition on Recognition and VQA on Handwritten Documents.Ajoy Mondal, Vijay Mahadevan, R. Manmatha, C. V. Jawahar
2024NAACLMultiple-Question Multiple-Answer Text-VQA.Peng Tang, Srikar Appalaraju, R. Manmatha, Yusheng Xie, Vijay Mahadevan
2024NAACLDEED: Dynamic Early Exit on Decoder for Accelerating Encoder-Decoder Transformer Models.Peng Tang, Pengkai Zhu, Tian Li, Srikar Appalaraju, Vijay Mahadevan, R. Manmatha
2023CVPRPolyFormer: Referring Image Segmentation as Sequential Polygon Generation.Jiang Liu, Hui Ding, Zhaowei Cai, Yuting Zhang, Ravi Kumar Satzoda, Vijay Mahadevan, R. Manmatha
2023ICCVDocTr: Document Transformer for Structured Information Extraction in Documents.Haofu Liao, Aruni RoyChowdhury, Weijian Li, Ankan Bansal, Yuting Zhang, Zhuowen Tu, Ravi Kumar Satzoda, R. Manmatha, Vijay Mahadevan
2022CVPRResNeSt: Split-Attention Networks.Hang Zhang, Chongruo Wu, Zhongyue Zhang, Yi Zhu, Haibin Lin, Zhi Zhang, Yue Sun, Tong He, Jonas Mueller, R. Manmatha, Mu Li, Alexander J. Smola
2022CVPRLaTr: Layout-Aware Transformer for Scene-Text VQA.Ali Furkan Biten, Ron Litman, Yusheng Xie, Srikar Appalaraju, R. Manmatha
2022CVPRTowards Weakly-Supervised Text Spotting using a Multi-Task Transformer.Yair Kittenplon, Inbal Lavi, Sharon Fogel, Yarin Bar, R. Manmatha, Pietro Perona
2022ECCVYORO - Lightweight End to End Visual Grounding.Chih-Hui Ho, Srikar Appalaraju, Bhavan Jasani, R. Manmatha, Nuno Vasconcelos
2022ECCVGLASS: Global to Local Attention for Scene-Text Spotting.Roi Ronen, Shahar Tsiper, Oron Anschel, Inbal Lavi, Amir Markovitz, R. Manmatha
2022ECCVOn Calibration of Scene-Text Recognition Models.Ron Slossberg, Oron Anschel, Amir Markovitz, Ron Litman, Aviad Aberdam, Shahar Tsiper, Shai Mazor, Jon Wu, R. Manmatha
2021CVPRSequence-to-Sequence Contrastive Learning for Text Recognition.Aviad Aberdam, Ron Litman, Shahar Tsiper, Oron Anschel, Ron Slossberg, Shai Mazor, R. Manmatha, Pietro Perona
2021ICCVDocFormer: End-to-End Transformer for Document Understanding.Srikar Appalaraju, Bhavan Jasani, Bhargava Urala Kota, Yusheng Xie, R. Manmatha
2021WACVSaliency Driven Perceptual Image Compression.Yash Patel, Srikar Appalaraju, R. Manmatha
2020CVPRSCATTER: Selective Context Attentional Scene Text Recognizer.Ron Litman, Oron Anschel, Shahar Tsiper, Roee Litman, Shai Mazor, R. Manmatha
2018CVPRCompressed Video Action Recognition.Chao-Yuan Wu, Manzil Zaheer, Hexiang Hu, R. Manmatha, Alexander J. Smola, Philipp Krhenbhl
2017ICCVSampling Matters in Deep Embedding Learning.R. Manmatha, Chao-Yuan Wu, Alexander J. Smola, Philipp Krhenbhl
2016CVPRDeep Decision Network for Multi-class Image Classification.Venkatesh N. Murthy, Vivek K. Singh, Terrence Chen, R. Manmatha, Dorin Comaniciu
2016ECCVEfficient Exploration of Text Regions in Natural Scene Images Using Adaptive Image Sampling.Ismet Zeki Yalniz, Douglas Gray, R. Manmatha
2014DASSequential Word Spotting in Historical Handwritten Documents.David Fernndez Mota, R. Manmatha, Alicia Forns, Josep Llads
2014SIGIRIncorporating query-specific feedback into learning-to-rank models.Ethem F. Can, W. Bruce Croft, R. Manmatha
2013CIKMPredicting retweet count using visual cues.Ethem F. Can, Hseyin Oktay, R. Manmatha
2013CVPRFormulating Action Recognition as a Ranking Problem.Ethem F. Can, R. Manmatha
2013ICDARCreating an Improved Version Using Noisy OCR from Multiple Editions.David Wemhoener, Ismet Zeki Yalniz, R. Manmatha
2012DASAn Efficient Framework for Searching Text in Noisy Document Images.Ismet Zeki Yalniz, R. Manmatha
2012ICFHROn Influence of Line Segmentation in Efficient Word Segmentation in Old Manuscripts.David Fernndez, Josep Llads, Alicia Forns, R. Manmatha
2012SIGIRA framework for manipulating and searching multiple retrieval types.Marc-Allen Cartright, Ethem F. Can, William Dabney, Jeff Dalton, Logan Giorda, Kriste Krstovski, Xiaoye Wu, Ismet Zeki Yalniz, James Allan, R. Manmatha, David A. Smith
2012SIGIRFinding translations in scanned book collections.Ismet Zeki Yalniz, R. Manmatha
2011CIKMMining relational structure from millions of books: position paper.David A. Smith, R. Manmatha, James Allan
2011CIKMPartial duplicate detection for large book collections.Ismet Zeki Yalniz, Ethem F. Can, R. Manmatha
2010ICFHRAdapting BLSTM Neural Network Based Keyword Spotting Trained on Modern Data to Historical Documents.Volkmar Frinken, Andreas Fischer, Horst Bunke, R. Manmatha
2008SENSYSDistributed image search in camera sensor networks.Tingxin Yan, Deepak Ganesan, R. Manmatha
2007ACCVEfficient Search in Document Image Collections.Anand Kumar, C. V. Jawahar, R. Manmatha
2006DASAligning Transcripts to Automatically Segmented Handwritten Manuscripts.Jamie L. Rothfeder, R. Manmatha, Toni M. Rath
2005ICASSPCombining text and audio-visual features in video indexing.Shih-Fu Chang, R. Manmatha, Tat-Seng Chua
2005SIGIRBoosted decision trees for word recognition in handwritten document retrieval.Nicholas R. Howe, Toni M. Rath, R. Manmatha
2004SIGIRA search engine for historical manuscript images.Toni M. Rath, R. Manmatha, Victor Lavrenko
2003CVPRWord Image Matching Using Dynamic Time Warping.Toni M. Rath, R. Manmatha
2003ICDARFeatures for Word Spotting in Historical Manuscripts.Toni M. Rath, R. Manmatha
2003ICNPMobile Distributed Information Retrieval for Highly-Partitioned Networks.Katrina M. Hanna, Brian Neil Levine, R. Manmatha
2003SIGIRAutomatic image annotation and retrieval using cross-media relevance models.Jiwoon Jeon, Victor Lavrenko, R. Manmatha
2002SIGIRA critical examination of TDT's cost function.R. Manmatha, Ao Feng, James Allan
2001ICCVAutomatic Segmentation and Indexing in a Database of Bird Images.Madirakshi Das, R. Manmatha
2001SIGIRModeling Score Distributions for Combining the Outputs of Search Engines.R. Manmatha, Toni M. Rath, Fangfang Feng
1998ICCVRetrieving Images by Appearance.Srinivas Ravela, R. Manmatha
1997SIGIRImage Retrieval by Appearance.Srinivas Ravela, R. Manmatha
1996CVPRWord Spotting: A New Approach to Indexing Handwriting.R. Manmatha, Chengfeng Han, Edward M. Riseman
1996ECCVImage Retrieval Using Scale-Space Matching.Srinivas Ravela, R. Manmatha, Edward M. Riseman
1994CVPRA framework for recovering affine transforms using points, lines or image brightnesses.R. Manmatha
1994ECCVMeasuring the Affine Transform Using Gaussian Filters.R. Manmatha
1989CVPRA data set for quantitative motion analysis.Rabindranath Dutta, R. Manmatha, Lance R. Williams, Edward M. Riseman