| 2026 | ICPR | HOI-R1: Exploring the Potential of Multimodal Large Language Models for Human-Object Interaction Detection. | Junwen Chen, Peilin Xiong, Keiji Yanai |
| 2025 | ICDAR | BengaliDiff: Diffusion Model for Few-Shot Bengali Font Generation. | Md Bilayet Hossain, Honghui Yuan, Shabnur Anonna Akhy, Keiji Yanai |
| 2025 | ICDAR | Japanese Kuzushiji Font Generation Employing Differentiable Renderer. | Honghui Yuan, Junwen Chen, Keiji Yanai |
| 2025 | MMM | WaveFontStyler: Font Style Transfer Based on Sound. | Kota Izumi, Keiji Yanai |
| 2025 | MMM | CalorieVoL: Integrating Volumetric Context Into Multimodal Large Language Models for Image-Based Calorie Estimation. | Hikaru Tanabe, Keiji Yanai |
| 2025 | MMM | KuzushijiDiffuser: Japanese Kuzushiji Font Generation with FontDiffuser. | Honghui Yuan, Keiji Yanai |
| 2025 | MMM | KuzushijiFontDiff: Diffusion Model for Japanese Kuzushiji Font Generation. | Honghui Yuan, Keiji Yanai |
| 2025 | MMM | SceneTextStyler: Editing Text with Style Transformation. | Honghui Yuan, Keiji Yanai |
| 2025 | WACV | Focusing on what to Decode and what to Train: SOV Decoding with Specific Target Guided DeNoising and Vision Language Advisor. | Junwen Chen, Yingcheng Wang, Keiji Yanai |
| 2024 | ACCV | Exploring Cross-Attention Maps in Multi-modal Diffusion Transformers for Training-Free Semantic Segmentation. | Rento Yamaguchi, Keiji Yanai |
| 2024 | ACCV | Vector Logo Image Synthesis Using Differentiable Renderer. | Ryuta Yamakura, Keiji Yanai |
| 2024 | ICPR | Act-ChatGPT: Introducing Action Features into Multi-modal Large Language Models for Video Understanding. | Yuto Nakamizo, Keiji Yanai |
| 2024 | ICPR | CalorieLLaVA: Image-Based Calorie Estimation with Multimodal Large Language Models. | Hikaru Tanabe, Keiji Yanai |
| 2024 | ICPR | Calorie-Aware Food Image Editing with Image Generation Models. | Kohei Yamamoto, Honghui Yuan, Keiji Yanai |
| 2024 | ICPR | Font Style Translation in Scene Text Images with CLIPstyler. | Honghui Yuan, Keiji Yanai |
| 2024 | MMM | Training-Free Region Prediction with Stable Diffusion. | Yuma Honbu, Keiji Yanai |
| 2023 | MMM | Virtual Try-On Considering Temporal Consistency for Videoconferencing. | Daiki Shimizu, Keiji Yanai |
| 2023 | MMM | Transformer-Based Cross-Modal Recipe Embeddings with Large Batch Training. | Jing Yang, Junwen Chen, Keiji Yanai |
| 2023 | MVA | QAHOI: Query-Based Anchors for Human-Object Interaction Detection. | Junwen Chen, Keiji Yanai |
| 2022 | CBMI | StyleGAN-based CLIP-guided Image Shape Manipulation. | Yuchen Qian, Kohei Yamamoto, Keiji Yanai |
| 2022 | ICIP | Continual Learning in Vision Transformer. | Mana Takeda, Keiji Yanai |
| 2020 | ICPR | IPN Hand: A Video Dataset and Benchmark for Real-Time Continuous Hand Gesture Recognition. | Gibran Benitez-Garcia, Jesus Olivares-Mercado, Gabriel Sanchez-Perez, Keiji Yanai |
| 2020 | ICPR | Mask-based Style-Controlled Image Synthesis Using a Mask Style Encoder. | Jaehyeong Cho, Wataru Shimoda, Keiji Yanai |
| 2020 | ICPR | Rescue Dog Action Recognition by Integrating Ego-Centric Video, Sound and Sensor Information. | Yuta Ide, Tsuyohito Araki, Ryunosuke Hamada, Kazunori Ohno, Keiji Yanai |
| 2020 | ICPR | UEC-FoodPix Complete: A Large-Scale Food Image Segmentation Dataset. | Kaimu Okamoto, Keiji Yanai |
| 2020 | ICPR | Fast and Accurate Real-Time Semantic Segmentation with Dilated Asymmetric Convolutions. | Leonel Rosas-Arias, Gibran Benitez-Garcia, Jos Portillo-Portillo, Gabriel Snchez-Prez, Keiji Yanai |
| 2020 | ICPR | Training of Multiple and Mixed Tasks with a Single Network Using Feature Modulation. | Mana Takeda, Gibran Benitez-Garcia, Keiji Yanai |
| 2020 | VR | CalorieCaptorGlass: Food Calorie Estimation based on Actual Size using HoloLens and Deep Learning. | Shu Naritomi, Keiji Yanai |
| 2019 | CBMI | Analyzing Regional Food Trends with Geo-tagged Twitter Food Photos. | Kaimu Okamoto, Keiji Yanai |
| 2019 | CVPR | Self-supervised Difference Detection for Refinement CRF and Seed Interpolation. | Wataru Shimoda, Keiji Yanai |
| 2019 | IBPRIA | Mosquito Larvae Image Classification Based on DenseNet and Guided Grad-CAM. | Zaira Garca-Nonoal, Keiji Yanai, Mariko Nakano, Antonio Arista-Jalife, Laura Cleofas-Snchez, Hctor Prez |
| 2019 | ICCV | Self-Supervised Difference Detection for Weakly-Supervised Semantic Segmentation. | Wataru Shimoda, Keiji Yanai |
| 2019 | ISMAR | DeepTaste: Augmented Reality Gustatory Manipulation with GAN-Based Real-Time Food-to-Food Translation. | Kizashi Nakano, Daichi Horita, Nobuchika Sakata, Kiyoshi Kiyokawa, Keiji Yanai, Takuji Narumi |
| 2019 | VR | Enchanting Your Noodles: GAN-based Real-time Food-to-Food Translation and Its Impact on Vision-induced Gustatory Manipulation. | Kizashi Nakano, Kiyoshi Kiyokawa, Daichi Horita, Keiji Yanai, Nobuchika Sakata, Takuji Narumi |
| 2019 | VR | Enchanting Your Noodles: A Gustatory Manipulation Interface by Using GAN-based Real-time Food-to-Food Translation. | Kizashi Nakano, Kiyoshi Kiyokawa, Daichi Horita, Keiji Yanai, Nobuchika Sakata, Takuji Narurni |
| 2018 | ACCV | Font Style Transfer Using Neural Style Transfer and Unsupervised Cross-domain Transfer. | Atsushi Narusawa, Wataru Shimoda, Keiji Yanai |
| 2018 | ACCV | Word-Conditioned Image Style Transfer. | Yu Sugiyama, Keiji Yanai |
| 2018 | IJCAI | Multi-task learning of dish detection and calorie estimation. | Takumi Ege, Keiji Yanai |
| 2018 | IJCAI | Food category transfer with conditional cycleGAN and a large-scale food image dataset. | Daichi Horita, Ryosuke Tanno, Wataru Shimoda, Keiji Yanai |
| 2018 | IJCAI | Food image generation using a large amount of food images with conditional GAN: ramenGAN and recipeGAN. | Yoshifumi Ito, Wataru Shimoda, Keiji Yanai |
| 2018 | MMM | AR DeepCalorieCam: An iOS App for Food Calorie Estimation with Augmented Reality. | Ryosuke Tanno, Takumi Ege, Keiji Yanai |
| 2018 | VRST | Ramen spoon eraser: CNN-based photo transformation for improving attractiveness of ramen photos. | Daichi Horita, Jaehyeong Cho, Takumi Ege, Keiji Yanai |
| 2018 | VRST | AR DeepCalorieCam V2: food calorie estimation with CNN and AR-based actual size estimation. | Ryosuke Tanno, Takumi Ege, Keiji Yanai |
| 2017 | ICDAR | Neural Font Style Transfer. | Gantugs Atarsaikhan, Brian Kenji Iwana, Atsushi Narusawa, Keiji Yanai, Seiichi Uchida |
| 2017 | ICDAR | Scene Text Eraser. | Toshiki Nakamura, Anna Zhu, Keiji Yanai, Seiichi Uchida |
| 2017 | ICIAP | Comparison of Two Approaches for Direct Food Calorie Estimation. | Takumi Ege, Keiji Yanai |
| 2017 | ICLR | Unseen Style Transfer Based on a Conditional Fast Style Transfer Network. | Keiji Yanai |
| 2017 | MMM | DeepStyleCam: A Real-Time Style Transfer App on iOS. | Ryosuke Tanno, Shin Matsuo, Wataru Shimoda, Keiji Yanai |
| 2017 | MVA | Simultaneous estimation of food categories and calories with multi-task CNN. | Takumi Ege, Keiji Yanai |
| 2016 | ECCV | Distinct Class-Specific Saliency Maps for Weakly Supervised Semantic Segmentation. | Wataru Shimoda, Keiji Yanai |
| 2016 | ICPR | Weakly-supervised segmentation by combining CNN feature maps and object saliency maps. | Wataru Shimoda, Keiji Yanai |
| 2016 | MMM | GrillCam: A Real-Time Eating Action Recognition System. | Koichi Okamoto, Keiji Yanai |
| 2016 | Mobiquitous | Caffe2C: A Framework for Easy Implementation of CNN-based Mobile Applications. | Ryosuke Tanno, Keiji Yanai |
| 2016 | WWW | Visual Event Mining from the Twitter Stream. | Takamu Kaneko, Keiji Yanai |
| 2015 | ICIAP | CNN-Based Food Image Segmentation Without Pixel-Wise Annotation. | Wataru Shimoda, Keiji Yanai |
| 2015 | PSIVT | Automatic Construction of Action Datasets Using Web Videos with Density-Based Cluster Analysis and Outlier Detection. | Do Hang Nga, Keiji Yanai |
| 2015 | SIGGRAPH | A system to support the amateurs to take a delicious-looking picture of foods. | Takao Kakimori, Makoto Okabe, Keiji Yanai, Rikio Onai |
| 2014 | ACCV | Hand Detection and Tracking in Videos for Fine-Grained Action Recognition. | Do Hang Nga, Keiji Yanai |
| 2014 | CVPR | Offline 1000-Class Classification on a Smartphone. | Yoshiyuki Kawano, Keiji Yanai |
| 2014 | ECCV | Automatic Expansion of a Food Image Dataset Leveraging Existing Categories with Domain Adaptation. | Yoshiyuki Kawano, Keiji Yanai |
| 2014 | ISM | Real-Time Photo Mining from the Twitter Stream: Event Photo Discovery and Food Photo Detection. | Keiji Yanai, Takamu Kaneko, Yoshiyuki Kawano |
| 2014 | MMM | FoodCam: A Real-Time Mobile Food Recognition System Employing Fisher Vector. | Yoshiyuki Kawano, Keiji Yanai |
| 2014 | MMM | A Dense SURF and Triangulation Based Spatio-temporal Feature for Action Recognition. | Do Hang Nga, Keiji Yanai |
| 2013 | CVPR | Real-Time Mobile Food Recognition System. | Yoshiyuki Kawano, Keiji Yanai |
| 2013 | MMM | Visual Analysis of Tag Co-occurrence on Nouns and Adjectives. | Yuya Kohara, Keiji Yanai |
| 2013 | PSIVT | Summarization of Egocentric Moving Videos for Generating Walking Route Guidance. | Masaya Okamoto, Keiji Yanai |
| 2012 | CVPR | Automatic collection of Web video shots corresponding to specific actions using Web images. | Do Hang Nga, Keiji Yanai |
| 2012 | ICPR | Multiple-food recognition considering co-occurrence employing manifold ranking. | Yuji Matsuda, Keiji Yanai |
| 2011 | ICCV | Automatic construction of an action video shot database using web videos. | Do Hang Nga, Keiji Yanai |
| 2011 | WWW | GeoVisualRank: a ranking method of geotagged imagesconsidering visual similarity and geo-location proximity. | Hidetoshi Kawakubo, Keiji Yanai |
| 2010 | ACCV | Geotagged Image Recognition by Combining Three Different Kinds of Geolocation Features. | Keita Yaegashi, Keiji Yanai |
| 2010 | ECCV | A SURF-Based Spatio-Temporal Feature for Feature-Fusion-Based Action Recognition. | Akitsugu Noguchi, Keiji Yanai |
| 2010 | ICPR | Geotagged Photo Recognition Using Corresponding Aerial Photos with Multiple Kernel Learning. | Keita Yaegashi, Keiji Yanai |
| 2010 | ISM | Image Recognition of 85 Food Categories by Feature Fusion. | Hajime Hoashi, Taichi Joutou, Keiji Yanai |
| 2010 | ISM | Automatic Construction of a Folksonomy-Based Visual Ontology. | Hidetoshi Kawakubo, Yuuta Akima, Keiji Yanai |
| 2009 | ACCV | Extracting Spatio-temporal Local Features Considering Consecutiveness of Motions. | Akitsugu Noguchi, Keiji Yanai |
| 2009 | ICIP | A food image recognition system with Multiple Kernel Learning. | Taichi Joutou, Keiji Yanai |
| 2009 | PSIVT | Can Geotags Help Image Recognition?. | Keita Yaegashi, Keiji Yanai |
| 2009 | WWW | Mining cultural differences from a large number of geotagged photos. | Keiji Yanai, Bingyu Qiu |
| 2008 | AINA | Associating Faces and Names in Japanese Photo News Articles on the Web. | Akio Kitahara, Taichi Joutou, Keiji Yanai |
| 2008 | ICIP | Web video retrieval based on the Earth Mover's Distance by integrating color, motion and sound. | Keisuke Takada, Keiji Yanai |
| 2008 | MMM | Web Image Gathering with a Part-Based Object Recognition Method. | Keiji Yanai |
| 2008 | WWW | Automatic web image selection with a probabilistic latent topic model. | Keiji Yanai |
| 2007 | WWW | Image collector III: a web image-gathering system with bag-of-keypoints. | Keiji Yanai |
| 2006 | MMM | Automatic "Go" record generation from a TV program. | Keiji Yanai, Takehisa Hayashiyama |
| 2006 | WWW | Finding visual concepts by web image mining. | Keiji Yanai, Kobus Barnard |
| 2003 | WWW | Image Collector II : An Over-One-Thousand-Image-Gathering System. | Keiji Yanai |
| 2003 | WWW | Web Image Mining toward Generic Image Recognition. | Keiji Yanai |
| 2002 | PRICAI | Image Classification by Web Images. | Keiji Yanai |
| 2000 | ICPR | Recognition of Indoor Images Employing Qualitative Model Fitting and Supporting Relation between Objects. | Keiji Yanai, Koichiro Deguchi |
| 2000 | MVA | A Multi-Resolution Image Understanding System Based on Multi-agent Architecture for High-Resolution Images. | Keiji Yanai, Koichiro Deguchi |
| 1998 | ICPR | An architecture of object recognition system for various images based on multi-agents. | Keiji Yanai, Koichiro Deguchi |