| 2024 | ICMLA | WindVibraTransformer: A Foundational Model for Precise and Robust Wind Turbine Condition Monitoring via Vibration Signals. | Takuya Wakayama, Taiki Inoue, Jun Ogata, Makoto Iida, Tetsuji Ogawa |
| 2023 | ICMLA | Learning Discriminative Feature Representations via Metric Learning for Early Operation of Wind Turbine Anomaly Detection Systems. | Taiki Inoue, Jun Ogata, Makoto Iida, Tetsuji Ogawa |
| 2022 | HCI | Applying Generative Adversarial Networks and Vision Transformers in Speech Emotion Recognition. | Panikos Heracleous, Satoru Fukayama, Jun Ogata, Yasser Mohammad |
| 2022 | Interspeech | Exploiting Fine-tuning of Self-supervised Learning Models for Improving Bi-modal Sentiment Analysis and Emotion Recognition. | Wei Yang, Satoru Fukayama, Panikos Heracleous, Jun Ogata |
| 2021 | PACLIC | Stronger Baseline for Robust Results in Multimodal Sentiment Analysis. | Wei Yang, Jun Ogata |
| 2019 | Interspeech | Knowledge Distillation for Throat Microphone Speech Recognition. | Takahito Suzuki, Jun Ogata, Takashi Tsunakawa, Masafumi Nishida, Masafumi Nishimura |
| 2018 | ISCAS | Fast Intra Mode Decision Method Based on Outliers of DCT Coefficients and Neighboring Block Information for H.265/HEVC. | Jun Ogata, Koichi Ichige |
| 2012 | Interspeech | PodCastle: Collaborative Training of Language Models on the Basis of Wisdom of Crowds. | Jun Ogata, Masataka Goto |
| 2012 | ITA | PodCastle and songle: Crowdsourcing-based web services for spoken document retrieval and active music listening. | Masataka Goto, Jun Ogata, Kazuyoshi Yoshii, Hiromasa Fujihara, Matthias Mauch, Tomoyasu Nakano |
| 2012 | WWW | PodCastle and Songle: Crowdsourcing-Based Web Services for Retrieval and Browsing of Speech and Music Content. | Masataka Goto, Jun Ogata, Kazuyoshi Yoshii, Hiromasa Fujihara, Matthias Mauch, Tomoyasu Nakano |
| 2011 | Interspeech | PodCastle: Recent Advances of a Spoken Document Retrieval Service Improved by Anonymous User Contributions. | Masataka Goto, Jun Ogata |
| 2010 | PACLIC | PodCastle: A Spoken Document Retrieval Service Improved by Anonymous User Contributions. | Masataka Goto, Jun Ogata |
| 2009 | ICASSP | The use of acoustically detected filled and silent pauses in spontaneous speech recognition. | Jun Ogata, Masataka Goto, Katunobu Itou |
| 2009 | Interspeech | Podcastle: collaborative training of acoustic models on the basis of wisdom of crowds for podcast transcription. | Jun Ogata, Masataka Goto |
| 2009 | Interspeech | Acoustic event detection for spotting "hot spots" in podcasts. | Kouhei Sumi, Tatsuya Kawahara, Jun Ogata, Masataka Goto |
| 2007 | ICMI | Presentation sensei: a presentation training system using speech and image processing. | Kazutaka Kurihara, Masataka Goto, Jun Ogata, Yosuke Matsusaka, Takeo Igarashi |
| 2007 | Interspeech | Podcastle: a web 2.0 approach to speech recognition research. | Masataka Goto, Jun Ogata, Kouichirou Eto |
| 2007 | Interspeech | Automatic transcription for a web 2.0 service to search podcasts. | Jun Ogata, Masataka Goto, Kouichirou Eto |
| 2006 | CHI | Speech pen: predictive handwriting based on ambient multimodal recognition. | Kazutaka Kurihara, Masataka Goto, Jun Ogata, Takeo Igarashi |
| 2006 | Interspeech | Detection and separation of speech events in meeting recordings. | Futoshi Asano, Jun Ogata |
| 2006 | ISM | Automatic Synchronization between Lyrics and Music CD Recordings Based on Viterbi Alignment of Segregated Vocal Signals. | Hiromasa Fujihara, Masataka Goto, Jun Ogata, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2005 | Interspeech | State estimation of meetings by information fusion using Bayesian network. | Michiaki Katoh, Kiyoshi Yamamoto, Jun Ogata, Takashi Yoshimura, Futoshi Asano, Hideki Asoh, Nobuhiko Kitawaki |
| 2005 | Interspeech | Speech repair: quick error correction just by using selection operation for speech input interfaces. | Jun Ogata, Masataka Goto |
| 2004 | IROS | Robust speech interface based on audio and video information fusion for humanoid HRP-2. | Isao Hara, Futoshi Asano, Hideki Asoh, Jun Ogata, Naoyuki Ichimura, Yoshihiro Kawai, Fumio Kanehiro, Hirohisa Hirukawa, Kiyoshi Yamamoto |
| 2003 | Interspeech | Live speech recognition in sports games by adaptation of acoustic model and language model. | Yasuo Ariki, Takeru Shigemori, Tsuyoshi Kaneko, Jun Ogata, Masakiyo Fujimoto |
| 2003 | Interspeech | Syllable-based acoustic modeling for Japanese spontaneous speech recognition. | Jun Ogata, Yasuo Ariki |
| 2003 | Interspeech | Topic segmentation and retrieval system for lecture videos based on spontaneous speech recognition. | Natsuo Yamamoto, Jun Ogata, Yasuo Ariki |
| 2002 | Interspeech | English call system with functions of speech segmentation and pronunciation evaluation using speech recognition technology. | Yasuo Ariki, Jun Ogata |
| 2002 | Interspeech | Unsupervised acoustic model adaptation based on phoneme error minimization. | Jun Ogata, Yasuo Ariki |
| 2001 | Interspeech | Improved speech recognition using iterative decoding based on confidence measures. | Jun Ogata, Yasuo Ariki |
| 2000 | Interspeech | Large vocabulary continuous speech recognition under real environments using adaptive sub-band spectral subtraction. | Masahiro Fujimoto, Jun Ogata, Yasuo Ariki |
| 2000 | Interspeech | An efficient lexical tree search for large vocabulary continuous speech recognition. | Jun Ogata, Yasuo Ariki |
| 2000 | Interspeech | Expanded vector space model based on word space in cross media retrieval of news speech data. | Seiichi Takao, Jun Ogata, Yasuo Ariki |
| 1998 | Interspeech | Indexing and classification of TV news articles based on speech dictation using word bigram. | Jun Ogata, Yasuo Ariki |
| 1992 | MVA | Neural Network Approaches for Attractive Area Extraction from Video Images. | Jun Ogata, Mikiya Sase, Yukio Kosugi |