Mohamed Elhoseiny
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
96
Venues
18
Active years
2012–2026
Best venue rank
A*
Where they publish
Papers
96 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2026 | AAAI | Step-by-step Layered Design Generation. | Faizan Farooq Khan, K. J. Joseph, Koustava Goswami, Mohamed Elhoseiny, Balaji Vasan Srinivasan |
| 2026 | ECIR | XProvence: Zero-Cost Multilingual Context Pruning for Retrieval-Augmented Generation. | Youssef Mohamed, Mohamed Elhoseiny, Thibault Formal, Nadezhda Chirkova |
| 2026 | WACV | iMotion-LLM: Instruction-Conditioned Trajectory Generation. | Abdulwahab Felemban, Nussair Hroub, Jian Ding, Eslam Abdelrahman, Xiaoqian Shen, Abduallah A. Mohamed, Mohamed Elhoseiny |
| 2026 | WACV | Sketch2Stitch: GANs for Abstract Sketch-Based Dress Synthesis. | Faizan Farooq Khan, Eslam Abdelrahman, Davide Morelli, Marcella Cornia, Rita Cucchiara, Mohamed Elhoseiny |
| 2025 | CVPR | Document Haystacks: Vision-Language Reasoning Over Piles of 1000+ Documents. | Jun Chen, Dannong Xu, Junjie Fei, Chun-Mei Feng, Mohamed Elhoseiny |
| 2025 | CVPR | StoryGPT-V: Large Language Models as Consistent Story Visualizers. | Xiaoqian Shen, Mohamed Elhoseiny |
| 2025 | EMNLP | InfiniBench: A Benchmark for Large Multi-Modal Models in Long-Form Movies and TV Shows. | Kirolos Ataallah, Eslam Mohamed Bakr, Mahmoud Ahmed, Chenhui Gou, Khushbu Pahwa, Jian Ding, Mohamed Elhoseiny |
| 2025 | EMNLP | Towards AI-Assisted Psychotherapy: Emotion-Guided Generative Interventions. | Kilichbek Haydarov, Youssef Mohamed, Emilio Goldenhersch, Paul OCallaghan, Li-jia Li, Mohamed Elhoseiny |
| 2025 | ICCV | Kestrel: 3D Multimodal LLM for Part-Aware Grounded Description. | Mahmoud Ahmed, Junjie Fei, Jian Ding, Eslam Mohamed Bakr, Mohamed Elhoseiny |
| 2025 | ICCV | Aurelia: Test-Time Reasoning Distillation in Audio-Visual LLMs. | Sanjoy Chowdhury, Hanan Gani, Nishit Anand, Sayan Nag, Ruohan Gao, Mohamed Elhoseiny, Salman Khan, Dinesh Manocha |
| 2025 | ICCV | AVTrustBench: Assessing and Enhancing Reliability and Robustness in Audio-Visual LLMs. | Sanjoy Chowdhury, Sayan Nag, Subhrajyoti Dasgupta, Yaoting Wang, Mohamed Elhoseiny, Ruohan Gao, Dinesh Manocha |
| 2025 | ICCV | A Survey on Long-Video Storytelling Generation: Architectures, Consistency, and Cinematic Quality. | Mohamed Elmoghany, Ryan A. Rossi, Seunghyun Yoon, Subhojyoti Mukherjee, Eslam Mohamed Bakr, Puneet Mathur, Gang Wu, Viet Dac Lai, Nedim Lipka, Ruiyi Zhang, Varun Manjunatha, Chien Nguyen, Daksh Dangi, Abel Salinas, Hongjie Chen, Xiaolei Huang, Joe Barrow, Nesreen K. Ahmed, Hoda Eldardiry, Namyong Park, Yu Wang, Zhengzhong Tu, Thien Huu Nguyen, Dinesh Manocha, Mohamed Elhoseiny, Franck Dernoncourt |
| 2025 | ICCV | Diffusion-Based Imaginative Coordination for Bimanual Manipulation. | Huilin Xu, Jian Ding, Jiakun Xu, Ruixiang Wang, Jun Chen, Jinjie Mai, Yanwei Fu, Bernard Ghanem, Feng Xu, Mohamed Elhoseiny |
| 2025 | ICCV | WikiAutoGen: Towards Multi-Modal Wikipedia-Style Article Generation. | Zhongyu Yang, Jun Chen, Dannong Xu, Junjie Fei, Xiaoqian Shen, Liangbing Zhao, Chun-Mei Feng, Mohamed Elhoseiny |
| 2025 | ICCV | 4D-Bench: Benchmarking Multi-Modal Large Language Models for 4D Object Understanding. | Wenxuan Zhu, Bing Li, Cheng Zheng, Jinjie Mai, Jun Chen, Letian Jiang, Abdullah Hamdi, Sara Rojas Martinez, Chia-Wen Lin, Mohamed Elhoseiny, Bernard Ghanem |
| 2025 | ICCV | From Reflection to Perfection: Scaling Inference-Time Optimization for Text-to-Image Diffusion Models via Reflection Tuning. | Le Zhuo, Liangbing Zhao, Sayak Paul, Yue Liao, Renrui Zhang, Yi Xin, Peng Gao, Mohamed Elhoseiny, Hongsheng Li |
| 2025 | ICLR | Bi-Factorial Preference Optimization: Balancing Safety-Helpfulness in Language Models. | Wenxuan Zhang, Philip Torr, Mohamed Elhoseiny, Adel Bibi |
| 2025 | ICLR | Query-based Knowledge Transfer for Heterogeneous Learning Environments. | Norah Alballa, Wenxuan Zhang, Ziquan Liu, Ahmed M. Abdelmoniem, Mohamed Elhoseiny, Marco Canini |
| 2025 | ICLR | ToddlerDiffusion: Interactive Structured Image Generation with Cascaded Schrdinger Bridge. | Eslam Mohamed Bakr, Liangbing Zhao, Vincent Tao Hu, Matthieu Cord, Patrick Prez, Mohamed Elhoseiny |
| 2025 | ICML | LongVU: Spatiotemporal Adaptive Compression for Long Video-Language Understanding. | Xiaoqian Shen, Yunyang Xiong, Changsheng Zhao, Lemeng Wu, Jun Chen, Chenchen Zhu, Zechun Liu, Fanyi Xiao, Balakrishnan Varadarajan, Florian Bordes, Zhuang Liu, Hu Xu, Hyunwoo J. Kim, Bilge Soran, Raghuraman Krishnamoorthi, Mohamed Elhoseiny, Vikas Chandra |
| 2025 | MICCAI | Temporal Model-Based Federated Active Medical Image Classification. | Yunlu Yan, Chun-Mei Feng, Yuexiang Li, Jinheng Xie, Jun Chen, Mohamed Elhoseiny, Ming Hu, Kaishun Wu, Lei Zhu |
| 2025 | WACV | Local Masked Reconstruction for Efficient Self-Supervised Learning on High-Resolution Images. | Jun Chen, Faizan Farooq Khan, Ming Hu, Ammar Sherif, Zongyuan Ge, Boyang Li, Mohamed Elhoseiny |
| 2024 | AAAI | ImageCaptioner2: Image Captioner for Image Captioning Bias Amplification Assessment. | Eslam Abdelrahman, Pengzhan Sun, Li Erran Li, Mohamed Elhoseiny |
| 2024 | ACCV | EmoTalker: Audio Driven Emotion Aware Talking Head Generation. | Xiaoqian Shen, Faizan Farooq Khan, Mohamed Elhoseiny |
| 2024 | CVPR | Adversarial Text to Continuous Image Generation. | Kilichbek Haydarov, Aashiq Muhamed, Xiaoqian Shen, Jovana Lazarevic, Ivan Skorokhodov, Chamuditha Jayanga Galappaththige, Mohamed Elhoseiny |
| 2024 | CVPR | AI Art Neural Constellation: Revealing the Collective and Contrastive State of AI-Generated and Human Art. | Faizan Farooq Khan, Diana Kim, Divyansh Jha, Youssef Mohamed, Hanna H. Chang, Ahmed Elgammal, Luba Elliott, Mohamed Elhoseiny |
| 2024 | CVPR | ShapeWalk: Compositional Shape Editing Through Language-Guided Chains. | Habib Slim, Mohamed Elhoseiny |
| 2024 | CVPR | Overcoming Generic Knowledge Loss with Selective Parameter Update. | Wenxuan Zhang, Paul Janson, Rahaf Aljundi, Mohamed Elhoseiny |
| 2024 | ECCV | Goldfish: Vision-Language Understanding of Arbitrarily Long Videos. | Kirolos Ataallah, Xiaoqian Shen, Eslam Abdelrahman, Essam Sleiman, Mingchen Zhuge, Jian Ding, Deyao Zhu, Jrgen Schmidhuber, Mohamed Elhoseiny |
| 2024 | ECCV | MEERKAT: Audio-Visual Large Language Model for Grounding in Space and Time. | Sanjoy Chowdhury, Sayan Nag, Subhrajyoti Dasgupta, Jun Chen, Mohamed Elhoseiny, Ruohan Gao, Dinesh Manocha |
| 2024 | ECCV | Affective Visual Dialog: A Large-Scale Benchmark for Emotional Reasoning Based on Visually Grounded Conversations. | Kilichbek Haydarov, Xiaoqian Shen, Avinash Madasu, Mahmoud Salem, Li-Jia Li, Gamaleldin Elsayed, Mohamed Elhoseiny |
| 2024 | ECCV | Uni3DL: A Unified Model for 3D Vision-Language Understanding. | Xiang Li, Jian Ding, Zhaoyang Chen, Mohamed Elhoseiny |
| 2024 | EMNLP | No Culture Left Behind: ArtELingo-28, a Benchmark of WikiArt with Captions in 28 Languages. | Youssef Mohamed, Runjia Li, Ibrahim Said Ahmad, Kilichbek Haydarov, Philip Torr, Kenneth Church, Mohamed Elhoseiny |
| 2024 | ICLR | CoT3DRef: Chain-of-Thoughts Data-Efficient 3D Visual Grounding. | Eslam Mohamed Bakr, Mohamed Ayman, Mahmoud Ahmed, Habib Slim, Mohamed Elhoseiny |
| 2024 | ICLR | Continual Learning on a Diet: Learning from Sparsely Labeled Streams Under Constrained Computation. | Wenxuan Zhang, Youssef Mohamed, Bernard Ghanem, Philip Torr, Adel Bibi, Mohamed Elhoseiny |
| 2024 | ICLR | MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models. | Deyao Zhu, Jun Chen, Xiaoqian Shen, Xiang Li, Mohamed Elhoseiny |
| 2024 | SIGIR | Multimodal Representation and Retrieval [MRR 2024]. | Xinliang Zhu, Arnab Dhua, Douglas Gray, I. Zeki Yalniz, Tan Yu, Mohamed Elhoseiny, Bryan A. Plummer |
| 2024 | WACV | A Hybrid Graph Network for Complex Activity Detection in Video. | Salman Khan, Izzeddin Teeti, Andrew Bradley, Mohamed Elhoseiny, Fabio Cuzzolin |
| 2023 | CVPR | MammalNet: A Large-Scale Video Benchmark for Mammal Recognition and Behavior Understanding. | Jun Chen, Ming Hu, Darren J. Coker, Michael L. Berumen, Blair R. Costelloe, Sara Beery, Anna Rohrbach, Mohamed Elhoseiny |
| 2023 | CVPR | MoStGAN-V: Video Generation with Temporal Motion Styles. | Xiaoqian Shen, Xiang Li, Mohamed Elhoseiny |
| 2023 | ICCV | HRS-Bench: Holistic, Reliable and Scalable Benchmark for Text-to-Image Models. | Eslam Mohamed Bakr, Pengzhan Sun, Xiaoqian Shen, Faizan Farooq Khan, Li Erran Li, Mohamed Elhoseiny |
| 2023 | ICCV | Exploring Open-Vocabulary Semantic Segmentation from CLIP Vision Encoder Distillation Only. | Jun Chen, Deyao Zhu, Guocheng Qian, Bernard Ghanem, Zhicheng Yan, Chenchen Zhu, Fanyi Xiao, Sean Chang Culatana, Mohamed Elhoseiny |
| 2023 | ICCV | FishNet: A Large-scale Dataset and Benchmark for Fish Recognition, Detection, and Functional Trait Prediction. | Faizan Farooq Khan, Xiang Li, Andrew J. Temple, Mohamed Elhoseiny |
| 2023 | ICCV | OxfordTVG-HIC: Can Machine Make Humorous Captions from Images? | Runjia Li, Shuyang Sun, Mohamed Elhoseiny, Philip H. S. Torr |
| 2023 | ICCV | Continual Zero-Shot Learning through Semantically Guided Generative Random Walks. | Wenxuan Zhang, Paul Janson, Kai Yi, Ivan Skorokhodov, Mohamed Elhoseiny |
| 2023 | ICLR | Value Memory Graph: A Graph-Structured World Model for Offline Reinforcement Learning. | Deyao Zhu, Li Erran Li, Mohamed Elhoseiny |
| 2023 | ICML | SLAMB: Accelerated Large Batch Training with Sparse Communication. | Hang Xu, Wenxuan Zhang, Jiawei Fei, Yuzhe Wu, Tingwen Xie, Jun Huang, Yuchen Xie, Mohamed Elhoseiny, Panos Kalnis |
| 2022 | CVPR | RelTransformer: A Transformer-Based Long-Tail Visual Relationship Recognition. | Jun Chen, Aniket Agarwal, Sherif Abdelkarim, Deyao Zhu, Mohamed Elhoseiny |
| 2022 | CVPR | VisualGPT: Data-efficient Adaptation of Pretrained Language Models for Image Captioning. | Jun Chen, Han Guo, Kai Yi, Boyang Li, Mohamed Elhoseiny |
| 2022 | CVPR | It is Okay to Not Be Okay: Overcoming Emotional Bias in Affective Image Captioning by Contrastive Data Collection. | Youssef Mohamed, Faizan Farooq Khan, Kilichbek Haydarov, Mohamed Elhoseiny |
| 2022 | CVPR | StyleGAN-V: A Continuous Video Generator with the Price, Image Quality and Perks of StyleGAN2. | Ivan Skorokhodov, Sergey Tulyakov, Mohamed Elhoseiny |
| 2022 | ECCV | 3D CoMPaT: Composition of Materials on Parts of 3D Things. | Yuchen Li, Ujjwal Upadhyay, Habib Slim, Ahmed Abdelreheem, Arpit Prajapati, Suhail Pothigara, Peter Wonka, Mohamed Elhoseiny |
| 2022 | ECCV | Social-Implicit: Rethinking Trajectory Prediction Evaluation and The Effectiveness of Implicit Maximum Likelihood Estimation. | Abduallah A. Mohamed, Deyao Zhu, Warren Vu, Mohamed Elhoseiny, Christian G. Claudel |
| 2022 | ECCV | Exploring Hierarchical Graph Representation for Large-Scale Zero-Shot Image Classification. | Kai Yi, Xiaoqian Shen, Yunhao Gou, Mohamed Elhoseiny |
| 2022 | EMNLP | ArtELingo: A Million Emotion Annotations of WikiArt with Emphasis on Diversity over Language and Culture. | Youssef Mohamed, Mohamed Abdelfattah, Shyma Alhuwaider, Feifan Li, Xiangliang Zhang, Kenneth Church, Mohamed Elhoseiny |
| 2022 | WACV | 3DRefTransformer: Fine-Grained Object Identification in Real-World Scenes Using Natural Language. | Ahmed Abdelreheem, Ujjwal Upadhyay, Ivan Skorokhodov, Rawan Al Yahya, Jun Chen, Mohamed Elhoseiny |
| 2021 | AAAI | Semi-Supervised Few-Shot Learning with Prototypical Random Walks. | Ahmed Ayyad, Yuchen Li, Raden Muaz, Shadi Albarqouni, Mohamed Elhoseiny |
| 2021 | CoRL | Motion Forecasting with Unlikelihood Training in Continuous Space. | Deyao Zhu, Mohamed Zahran, Li Erran Li, Mohamed Elhoseiny |
| 2021 | CVPR | ArtEmis: Affective Language for Visual Art. | Panos Achlioptas, Maks Ovsjanikov, Kilichbek Haydarov, Mohamed Elhoseiny, Leonidas J. Guibas |
| 2021 | CVPR | Adversarial Generation of Continuous Images. | Ivan Skorokhodov, Savva Ignatyev, Mohamed Elhoseiny |
| 2021 | ICCV | Exploring Long Tail Visual Relationship Recognition with Large Vocabulary. | Sherif Abdelkarim, Aniket Agarwal, Panos Achlioptas, Jun Chen, Jiaji Huang, Boyang Li, Kenneth Church, Mohamed Elhoseiny |
| 2021 | ICCV | Aligning Latent and Image Spaces to Connect the Unconnectable. | Ivan Skorokhodov, Grigorii Sotnikov, Mohamed Elhoseiny |
| 2021 | ICLR | Class Normalization for (Continual)? Generalized Zero-Shot Learning. | Ivan Skorokhodov, Mohamed Elhoseiny |
| 2021 | ICLR | HalentNet: Multimodal Trajectory Forecasting with Hallucinative Intents. | Deyao Zhu, Mohamed Zahran, Li Erran Li, Mohamed Elhoseiny |
| 2020 | CVPR | Social-STGCNN: A Social Spatio-Temporal Graph Convolutional Neural Network for Human Trajectory Prediction. | Abduallah A. Mohamed, Kun Qian, Mohamed Elhoseiny, Christian G. Claudel |
| 2020 | ECCV | ReferIt3D: Neural Listeners for Fine-Grained 3D Object Identification in Real-World Scenes. | Panos Achlioptas, Ahmed Abdelreheem, Fei Xia, Mohamed Elhoseiny, Leonidas J. Guibas |
| 2020 | ICLR | Uncertainty-guided Continual Learning with Bayesian Neural Networks. | Sayna Ebrahimi, Mohamed Elhoseiny, Trevor Darrell, Marcus Rohrbach |
| 2020 | ICLR | Compositional Language Continual Learning. | Yuanpeng Li, Liang Zhao, Kenneth Church, Mohamed Elhoseiny |
| 2019 | AAAI | Large-Scale Visual Relationship Understanding. | Ji Zhang, Yannis Kalantidis, Marcus Rohrbach, Manohar Paluri, Ahmed Elgammal, Mohamed Elhoseiny |
| 2019 | CVPR | Uncertainty-Guided Continual Learning in Bayesian Neural Networks - Extended Abstract. | Sayna Ebrahimi, Mohamed Elhoseiny, Trevor Darrell, Marcus Rohrbach |
| 2019 | ICCV | Creativity Inspired Zero-Shot Learning. | Mohamed Elhoseiny, Mohamed Elfeki |
| 2019 | ICLR | Efficient Lifelong Learning with A-GEM. | Arslan Chaudhry, Marc'Aurelio Ranzato, Marcus Rohrbach, Mohamed Elhoseiny |
| 2019 | ICML | GDPP: Learning Diverse Generations using Determinantal Point Processes. | Mohamed Elfeki, Camille Couprie, Morgane Rivire, Mohamed Elhoseiny |
| 2019 | ICRA | Video Object Segmentation using Teacher-Student Adaptation in a Human Robot Interaction (HRI) Setting. | Mennatullah Siam, Chen Jiang, Steven Weikai Lu, Laura Petrich, Mahmoud Gamal, Mohamed Elhoseiny, Martin Jgersand |
| 2018 | AAAI | The Shape of Art History in the Eyes of the Machine. | Ahmed Elgammal, Bingchen Liu, Diana Kim, Mohamed Elhoseiny, Marian Mazzone |
| 2018 | ACCV | Exploring the Challenges Towards Lifelong Fact Learning. | Mohamed Elhoseiny, Francesca Babiloni, Rahaf Aljundi, Marcus Rohrbach, Manohar Paluri, Tinne Tuytelaars |
| 2018 | CVPR | A Generative Adversarial Approach for Zero-Shot Learning From Noisy Texts. | Yizhe Zhu, Mohamed Elhoseiny, Bingchen Liu, Xi Peng, Ahmed Elgammal |
| 2018 | ECCV | Memory Aware Synapses: Learning What (not) to Forget. | Rahaf Aljundi, Francesca Babiloni, Mohamed Elhoseiny, Marcus Rohrbach, Tinne Tuytelaars |
| 2018 | ECCV | DesIGN: Design Inspiration from Generative Networks. | Othman Sbai, Mohamed Elhoseiny, Antoine Bordes, Yann LeCun, Camille Couprie |
| 2018 | ECCV | Choose Your Neuron: Incorporating Domain Knowledge Through Neuron-Importance. | Ramprasaath R. Selvaraju, Prithvijit Chattopadhyay, Mohamed Elhoseiny, Tilak Sharma, Dhruv Batra, Devi Parikh, Stefan Lee |
| 2017 | AAAI | Sherlock: Scalable Fact Learning in Images. | Mohamed Elhoseiny, Scott Cohen, Walter Chang, Brian L. Price, Ahmed M. Elgammal |
| 2017 | CVPR | Link the Head to the "Beak": Zero Shot Learning from Noisy Text Description at Part Precision. | Mohamed Elhoseiny, Yizhe Zhu, Han Zhang, Ahmed M. Elgammal |
| 2017 | CVPR | Relationship Proposal Networks. | Ji Zhang, Mohamed Elhoseiny, Scott Cohen, Walter Chang, Ahmed M. Elgammal |
| 2016 | AAAI | Zero-Shot Event Detection by Multimodal Distributional Semantic Embedding of Videos. | Mohamed Elhoseiny, Jingen Liu, Hui Cheng, Harpreet S. Sawhney, Ahmed M. Elgammal |
| 2016 | ACL | Automatic Annotation of Structured Facts in Images. | Mohamed Elhoseiny, Scott Cohen, Walter Chang, Brian L. Price, Ahmed M. Elgammal |
| 2016 | CVPR | SPDA-CNN: Unifying Semantic Part Detection and Abstraction for Fine-Grained Recognition. | Han Zhang, Tao Xu, Mohamed Elhoseiny, Xiaolei Huang, Shaoting Zhang, Ahmed M. Elgammal, Dimitris N. Metaxas |
| 2016 | ICML | A Comparative Analysis and Study of Multiview CNN Models for Joint Object Categorization and Pose Estimation. | Mohamed Elhoseiny, Tarek El-Gaaly, Amr Bakry, Ahmed M. Elgammal |
| 2016 | WACV | Joint object recognition and pose estimation using a nonlinear view-invariant latent generative model. | Amr Bakry, Tarek El-Gaaly, Mohamed Elhoseiny, Ahmed M. Elgammal |
| 2015 | BMVC | Overlapping Domain Cover for Scalable and Accurate Regression Kernel Machines. | Mohamed Elhoseiny, Ahmed M. Elgammal |
| 2015 | CVPR | Learning Hypergraph-regularized Attribute Predictors. | Sheng Huang, Mohamed Elhoseiny, Ahmed M. Elgammal, Dan Yang |
| 2015 | ICIP | Weather classification with deep convolutional neural networks. | Mohamed Elhoseiny, Sheng Huang, Ahmed M. Elgammal |
| 2014 | ICIP | Improving non-negative matrix factorization via ranking its bases. | Sheng Huang, Mohamed Elhoseiny, Ahmed M. Elgammal, Dan Yang |
| 2013 | CVPR | MultiClass Object Classification in Video Surveillance Systems - Experimental Study. | Mohamed Elhoseiny, Amr Bakry, Ahmed M. Elgammal |
| 2013 | ICCV | Write a Classifier: Zero-Shot Learning Using Purely Textual Descriptions. | Mohamed Elhoseiny, Babak Saleh, Ahmed M. Elgammal |
| 2013 | ICIP | Low-bitrate benefits of JPEG compression on sift recognition. | Mohamed Elhoseiny, Bing Song, Jeremi Sudol, David McKinnon |
| 2012 | ISM | English2MindMap: An Automated System for MindMap Generation from English Text. | Mohamed Elhoseiny, Ahmed M. Elgammal |