| 2021 | IROS | Alternating Drive-and-Glide Flight Navigation of a Kiteplane for Sound Source Position Estimation. | Makoto Kumon, Hiroshi G. Okuno, Shuichi Tajima |
| 2020 | IROS | Computational Design of Balanced Open Link Planar Mechanisms with Counterweights from User Sketches. | Takuto Takahashi, Hiroshi G. Okuno, Shigeki Sugano, Stelian Coros, Bernhard Thomaszewski |
| 2018 | IROS | Extracting the Relationship between the Spatial Distribution and Types of Bird Vocalizations Using Robot Audition System HARK. | Shinji Sumitani, Reiji Suzuki, Shiho Matsubayashi, Takaya Arita, Kazuhiro Nakadai, Hiroshi G. Okuno |
| 2018 | IROS | Design and Implementation of Programmable Drawing Automata based on Cam Mechanisms for Representing Spatial Trajectory. | Takuto Takahashi, Hiroshi G. Okuno |
| 2017 | IROS | Development of microphone-array-embedded UAV for search and rescue task. | Kazuhiro Nakadai, Makoto Kumon, Hiroshi G. Okuno, Kotaro Hoshiba, Mizuho Wakabayashi, Kai Washizaki, Takahiro Ishiki, Daniel Gabriel, Yoshiaki Bando, Takayuki Morito, Ryosuke Kojima, Osamu Sugiyama |
| 2016 | Interspeech | Call Alternation Between Specific Pairs of Male Frogs Revealed by a Sound-Imaging Method in Their Natural Habitat. | Ikkyu Aihara, Takeshi Mizumoto, Hiromitsu Awano, Hiroshi G. Okuno |
| 2016 | Interspeech | Localizing Bird Songs Using an Open Source Robot Audition System with a Microphone Array. | Reiji Suzuki, Shiho Matsubayashi, Kazuhiro Nakadai, Hiroshi G. Okuno |
| 2016 | LREC | Parallel Speech Corpora of Japanese Dialects. | Koichiro Yoshino, Naoki Hirayama, Shinsuke Mori, Fumihiko Takahashi, Katsutoshi Itoyama, Hiroshi G. Okuno |
| 2015 | ICASSP | Challenges in deploying a microphone array to localize and separate sound sources in real auditory scenes. | Yoshiaki Bando, Takuma Otsuka, Katsutoshi Itoyama, Kazuyoshi Yoshii, Yoko Sasaki, Satoshi Kagami, Hiroshi G. Okuno |
| 2015 | ICASSP | Robot audition: Its rise and perspectives. | Hiroshi G. Okuno, Kazuhiro Nakadai |
| 2015 | IROS | Microphone-accelerometer based 3D posture estimation for a hose-shaped rescue robot. | Yoshiaki Bando, Katsutoshi Itoyama, Masashi Konyo, Satoshi Tadokoro, Kazuhiro Nakadai, Kazuyoshi Yoshii, Hiroshi G. Okuno |
| 2014 | ICASSP | Transcribing vocal expression from polyphonic music. | Yukara Ikemiya, Katsutoshi Itoyama, Hiroshi G. Okuno |
| 2014 | ICASSP | Audio part mixture alignment based on hierarchical nonparametric Bayesian model of musical audio sequence collection. | Akira Maezawa, Hiroshi G. Okuno |
| 2014 | ICASSP | Automatic transcription of guitar tablature from audio signals in accordance with player's proficiency. | Kazuki Yazawa, Katsutoshi Itoyama, Hiroshi G. Okuno |
| 2014 | ICRA | Insertion of pause in drawing from babbling for robot's developmental imitation learning. | Shun Nishide, Keita Mochizuki, Hiroshi G. Okuno, Tetsuya Ogata |
| 2014 | Interspeech | Lipreading using convolutional neural network. | Kuniaki Noda, Yuki Yamaguchi, Kazuhiro Nakadai, Hiroshi G. Okuno, Tetsuya Ogata |
| 2014 | IROS | Visualization of auditory awareness based on sound source positions estimated by depth sensor and microphone array. | Takahiro Iyama, Osamu Sugiyama, Takuma Otsuka, Katsutoshi Itoyama, Hiroshi G. Okuno |
| 2014 | IROS | Making a robot dance to diverse musical genre in noisy environments. | Joo Lobato Oliveira, Keisuke Nakamura, Thibault Langlois, Fabien Gouyon, Kazuhiro Nakadai, Angelica Lim, Lus Paulo Reis, Hiroshi G. Okuno |
| 2014 | SMC | Sound annotation tool for multidirectional sounds based on spatial information extracted by HARK robot audition software. | Osamu Sugiyama, Katsutoshi Itoyama, Kazuhiro Nakadai, Hiroshi G. Okuno |
| 2013 | ICASSP | Multiple index combination for Japanese spoken term detection with optimum index selection based on OOV-region classifier. | Naoyuki Kanda, Katsutoshi Itoyama, Hiroshi G. Okuno |
| 2013 | ICASSP | Initialization-robust Bayesian multipitch analyzer based on psychoacoustical and musical criteria. | Daichi Sakaue, Takuma Otsuka, Katsutoshi Itoyama, Hiroshi G. Okuno |
| 2013 | ICASSP | Audio-based guitar tablature transcription using multipitch analysis and playability constraints. | Kazuki Yazawa, Daichi Sakaue, Kohei Nagira, Katsutoshi Itoyama, Hiroshi G. Okuno |
| 2013 | ICRA | Hands-free human-robot communication robust to speaker's radial position. | Randy Gomez, Keisuke Nakamura, Kazuhiro Nakadai, Ui-Hyun Kim, Hiroshi G. Okuno, Tatsuya Kawahara |
| 2013 | Interspeech | Automatic estimation of dialect mixing ratio for dialect speech recognition. | Naoki Hirayama, Koichiro Yoshino, Katsutoshi Itoyama, Shinsuke Mori, Hiroshi G. Okuno |
| 2013 | IROS | Posture estimation of hose-shaped robot using microphone array localization. | Yoshiaki Bando, Takeshi Mizumoto, Katsutoshi Itoyama, Kazuhiro Nakadai, Hiroshi G. Okuno |
| 2013 | IROS | Noise correlation matrix estimation for improving sound source localization by multirotor UAV. | Koutarou Furukawa, Keita Okutani, Kohei Nagira, Takuma Otsuka, Katsutoshi Itoyama, Kazuhiro Nakadai, Hiroshi G. Okuno |
| 2013 | IWSEC | Solving Google's Continuous Audio CAPTCHA with HMM-Based Automatic Speech Recognition. | Shotaro Sano, Takuma Otsuka, Hiroshi G. Okuno |
| 2013 | SMC | Developmental Human-Robot Imitation Learning of Drawing with a Neuro Dynamical System. | Keita Mochizuki, Shun Nishide, Hiroshi G. Okuno, Tetsuya Ogata |
| 2012 | AAAI | Bayesian Unification of Sound Source Localization and Separation with Permutation Resolution. | Takuma Otsuka, Katsuhiko Ishiguro, Hiroshi Sawada, Hiroshi G. Okuno |
| 2012 | COLING | Statistical Method of Building Dialect Language Models for ASR Systems. | Naoki Hirayama, Shinsuke Mori, Hiroshi G. Okuno |
| 2012 | ICASSP | Initialization-robust multipitch estimation based on latent harmonic allocation using overtone corpus. | Daichi Sakaue, Katsutoshi Itoyama, Tetsuya Ogata, Hiroshi G. Okuno |
| 2012 | ICRA | Incremental probabilistic geometry estimation for robot scene understanding. | Louis-Kenzo Cahier, Tetsuya Ogata, Hiroshi G. Okuno |
| 2012 | IJCNN | Self-organization of object features representing motion using Multiple Timescales Recurrent Neural Network. | Shun Nishide, Jun Tani, Hiroshi G. Okuno, Tetsuya Ogata |
| 2012 | IJCNN | Body area segmentation from visual scene based on predictability of neuro-dynamical system. | Harumitsu Nobuta, Kenta Kawamoto, Kuniaki Noda, Kohtaro Sabe, Shun Nishide, Hiroshi G. Okuno, Tetsuya Ogata |
| 2012 | IROS | Who is the leader in a multiperson ensemble? - Multiperson human-robot ensemble model with leaderness -. | Takeshi Mizumoto, Tetsuya Ogata, Hiroshi G. Okuno |
| 2012 | IROS | Live assessment of beat tracking for robot audition. | Joo Lobato Oliveira, Gkhan Ince, Keisuke Nakamura, Kazuhiro Nakadai, Hiroshi G. Okuno, Lus Paulo Reis, Fabien Gouyon |
| 2012 | IROS | Unified auditory functions based on Bayesian topic model. | Takuma Otsuka, Katsuhiko Ishiguro, Hiroshi Sawada, Hiroshi G. Okuno |
| 2012 | IROS | Sound sources selection system by using onomatopoeic querries from multiple sound sources. | Yusuke Yamamura, Toru Takahashi, Tetsuya Ogata, Hiroshi G. Okuno |
| 2012 | RO-MAN | An active audition framework for auditory-driven HRI: Application to interactive robot dancing. | Joo Lobato Oliveira, Gkhan Ince, Keisuke Nakamura, Kazuhiro Nakadai, Hiroshi G. Okuno, Lus Paulo Reis, Fabien Gouyon |
| 2012 | SSPR | Infinite Sparse Factor Analysis for Blind Source Separation in Reverberant Environments. | Kohei Nagira, Takuma Otsuka, Hiroshi G. Okuno |
| 2011 | ICANN | Cluster Self-organization of Known and Unknown Environmental Sounds Using Recurrent Neural Network. | Yang Zhang, Shun Nishide, Toru Takahashi, Hiroshi G. Okuno, Tetsuya Ogata |
| 2011 | ICASSP | Simultaneous processing of sound source separation and musical instrument identification using Bayesian spectral modeling. | Katsutoshi Itoyama, Masataka Goto, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2011 | ICASSP | Polyphonic audio-to-score alignment based on Bayesian Latent Harmonic Allocation Hidden Markov Model. | Akira Maezawa, Hiroshi G. Okuno, Tetsuya Ogata, Masataka Goto |
| 2011 | ICASSP | I-Divergence-based dereverberation method with auxiliary function approach. | Naoki Yasuraoka, Hirokazu Kameoka, Takuya Yoshioka, Hiroshi G. Okuno |
| 2011 | ICONIP | Use of a Sparse Structure to Improve Learning Performance of Recurrent Neural Networks. | Hiromitsu Awano, Shun Nishide, Hiroaki Arie, Jun Tani, Toru Takahashi, Hiroshi G. Okuno, Tetsuya Ogata |
| 2011 | ICRA | Design and implementation of selectable sound separation on the Texai telepresence system using HARK. | Takeshi Mizumoto, Kazuhiro Nakadai, Takami Yoshida, Ryu Takeda, Takuma Otsuka, Toru Takahashi, Hiroshi G. Okuno |
| 2011 | Interspeech | Fast and Simple Iterative Algorithm of Lp-Norm Minimization for Under-Determined Speech Separation. | Yasuharu Hirasawa, Naoki Yasuraoka, Toru Takahashi, Tetsuya Ogata, Hiroshi G. Okuno |
| 2011 | Interspeech | Bayesian Extension of MUSIC for Sound Source Localization and Tracking. | Takuma Otsuka, Kazuhiro Nakadai, Tetsuya Ogata, Hiroshi G. Okuno |
| 2011 | IROS | Particle-filter based audio-visual beat-tracking for music robot ensemble with human guitarist. | Tatsuhiko Itohara, Takuma Otsuka, Takeshi Mizumoto, Tetsuya Ogata, Hiroshi G. Okuno |
| 2011 | IROS | Improvement of speaker localization by considering multipath interference of sound wave for binaural robot audition. | Ui-Hyun Kim, Takeshi Mizumoto, Tetsuya Ogata, Hiroshi G. Okuno |
| 2011 | SMC | Handwriting prediction based character recognition using recurrent neural network. | Shun Nishide, Hiroshi G. Okuno, Tetsuya Ogata, Jun Tani |
| 2011 | SIGdial | A Two-Stage Domain Selection Framework for Extensible Multi-Domain Spoken Dialogue Systems. | Mikio Nakano, Shun Sato, Kazunori Komatani, Kyoko Matsuyama, Kotaro Funakoshi, Hiroshi G. Okuno |
| 2010 | AAAI | Design and Implementation of Two-level Synchronization for Interactive Music Robot. | Takuma Otsuka, Kazuhiro Nakadai, Toru Takahashi, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2010 | COLING | Automatic Allocation of Training Data for Rapid Prototyping of Speech Understanding based on Multiple Model Combination. | Kazunori Komatani, Masaki Katsumaru, Mikio Nakano, Kotaro Funakoshi, Tetsuya Ogata, Hiroshi G. Okuno |
| 2010 | ICASSP | Music dereverberation using harmonic structure source model and Wiener filter. | Naoki Yasuraoka, Takuya Yoshioka, Tomohiro Nakatani, Atsushi Nakamura, Hiroshi G. Okuno |
| 2010 | ICASSP | Noisy speech enhancement based on prior knowledge about spectral envelope and harmonic structure. | Takuya Yoshioka, Tomohiro Nakatani, Hiroshi G. Okuno |
| 2010 | Interspeech | Analyzing user utterances in barge-in-able spoken dialogue system for improving identification accuracy. | Kyoko Matsuyama, Kazunori Komatani, Ryu Takeda, Toru Takahashi, Tetsuya Ogata, Hiroshi G. Okuno |
| 2010 | Interspeech | Effects of modelling within- and between-frame temporal variations in power spectra on non-verbal sound recognition. | Nobuhide Yamakawa, Tetsuro Kitahara, Toru Takahashi, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2010 | IROS | Exploiting harmonic structures to improve separating simultaneous speech in under-determined conditions. | Yasuharu Hirasawa, Toru Takahashi, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2010 | IROS | Robot musical accompaniment: integrating audio and visual cues for real-time synchronization with a human flutist. | Angelica Lim, Takeshi Mizumoto, Louis-Kenzo Cahier, Takuma Otsuka, Toru Takahashi, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2010 | IROS | Human-robot ensemble between robot thereminist and human percussionist using coupled oscillator model. | Takeshi Mizumoto, Takuma Otsuka, Kazuhiro Nakadai, Toru Takahashi, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2010 | IROS | Motion generation based on reliable predictability using self-organized object features. | Shun Nishide, Tetsuya Ogata, Jun Tani, Toru Takahashi, Kazunori Komatani, Hiroshi G. Okuno |
| 2010 | IROS | An improvement in automatic speech recognition using soft missing feature masks for robot audition. | Toru Takahashi, Kazuhiro Nakadai, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2010 | IROS | Speedup and performance improvement of ICA-based robot audition by parallel and resampling-based block-wise processing. | Ryu Takeda, Kazuhiro Nakadai, Toru Takahashi, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2010 | IROS | Two-layered audio-visual speech recognition for robots in noisy environments. | Takami Yoshida, Kazuhiro Nakadai, Hiroshi G. Okuno |
| 2010 | ICRA | Improvement in listening capability for humanoid robot HRP-2. | Toru Takahashi, Kazuhiro Nakadai, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2010 | ICRA | Upper-limit evaluation of robot audition based on ICA-BSS in multi-source, barge-in and highly reverberant conditions. | Ryu Takeda, Kazuhiro Nakadai, Toru Takahashi, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2010 | SMC | Human-robot cooperation in arrangement of objects using confidence measure of neuro-dynamical system. | Hiromitsu Awano, Tetsuya Ogata, Shun Nishide, Toru Takahashi, Kazunori Komatani, Hiroshi G. Okuno |
| 2010 | SIGdial | Online Error Detection of Barge-In Utterances by Using Individual Users' Utterance Histories in Spoken Dialogue System. | Kazunori Komatani, Hiroshi G. Okuno |
| 2009 | ICASSP | ICA-based efficient blind dereverberation and echo cancellation method for barge-in-able robot audition. | Ryu Takeda, Kazuhiro Nakadai, Toru Takahashi, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2009 | ICRA | Continuous vocal imitation with self-organized vowel spaces in Recurrent Neural Network. | Hisashi Kanda, Tetsuya Ogata, Toru Takahashi, Kazunori Komatani, Hiroshi G. Okuno |
| 2009 | ICRA | Prediction and imitation of other's motions by reusing own forward-inverse model in robots. | Tetsuya Ogata, Ryunosuke Yokoya, Jun Tani, Kazunori Komatani, Hiroshi G. Okuno |
| 2009 | Interspeech | Improving speech understanding accuracy with limited training data using multiple language models and multiple understanding models. | Masaki Katsumaru, Mikio Nakano, Kazunori Komatani, Kotaro Funakoshi, Tetsuya Ogata, Hiroshi G. Okuno |
| 2009 | Interspeech | Enabling a user to specify an item at any time during system enumeration - item identification for barge-in-able conversational dialogue systems. | Kyoko Matsuyama, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2009 | IROS | Emergence of evolutionary interaction with voice and motion between two robots using RNN. | Wataru Hinoshita, Tetsuya Ogata, Hideki Kozima, Hisashi Kanda, Toru Takahashi, Hiroshi G. Okuno |
| 2009 | IROS | Phoneme acquisition model based on vowel imitation using Recurrent Neural Network. | Hisashi Kanda, Tetsuya Ogata, Toru Takahashi, Kazunori Komatani, Hiroshi G. Okuno |
| 2009 | IROS | Thereminist robot: Development of a robot theremin player with feedforward and feedback arm control based on a Theremin's pitch model. | Takeshi Mizumoto, Hiroshi Tsujino, Toru Takahashi, Tetsuya Ogata, Hiroshi G. Okuno |
| 2009 | IROS | Modeling tool-body assimilation using second-order Recurrent Neural Network. | Shun Nishide, Tatsuhiro Nakagawa, Tetsuya Ogata, Jun Tani, Toru Takahashi, Hiroshi G. Okuno |
| 2009 | IROS | Incremental polyphonic audio to score alignment using beat tracking for singer robots. | Takuma Otsuka, Toru Takahashi, Hiroshi G. Okuno, Kazunori Komatani, Tetsuya Ogata, Kazumasa Murata, Kazuhiro Nakadai |
| 2009 | IROS | Missing-feature-theory-based robust simultaneous speech recognition system with non-clean speech acoustic model. | Toru Takahashi, Kazuhiro Nakadai, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2009 | IROS | Step-size parameter adaptation of multi-channel semi-blind ICA with piecewise linear model for barge-in-able robot audition. | Ryu Takeda, Kazuhiro Nakadai, Toru Takahashi, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2009 | ISM | Bowed String Sequence Estimation of a Violin Based on Adaptive Audio Signal Classification and Context-Dependent Error Correction. | Akira Maezawa, Katsutoshi Itoyama, Toru Takahashi, Tetsuya Ogata, Hiroshi G. Okuno |
| 2009 | ISRR | Robot Audition: Missing Feature Theory Approach and Active Audition. | Hiroshi G. Okuno, Kazuhiro Nakadai, Hyun-Don Kim |
| 2009 | NAACL | A Speech Understanding Framework that Uses Multiple Language Models and Multiple Understanding Models. | Masaki Katsumaru, Mikio Nakano, Kazunori Komatani, Kotaro Funakoshi, Tetsuya Ogata, Hiroshi G. Okuno |
| 2009 | SIGdial | Ranking Help Message Candidates Based on Robust Grammar Verification Results and Utterance History in Spoken Dialogue Systems. | Kazunori Komatani, Satoshi Ikeda, Yuichiro Fukubayashi, Tetsuya Ogata, Hiroshi G. Okuno |
| 2008 | IJCNLP | Rapid Prototyping of Robust Language Understanding Modules for Spoken Dialogue Systems. | Yuichiro Fukubayashi, Kazunori Komatani, Mikio Nakano, Kotaro Funakoshi, Hiroshi Tsujino, Tetsuya Ogata, Hiroshi G. Okuno |
| 2008 | ICRA | Two-channel-based voice activity detection for humanoid robots in noisy home environments. | Hyun-Don Kim, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2008 | ICRA | A robot referee for rock-paper-scissors sound games. | Kazuhiro Nakadai, Shun'ichi Yamamoto, Hiroshi G. Okuno, Hirofumi Nakajima, Yuji Hasegawa, Hiroshi Tsujino |
| 2008 | ICRA | Object dynamics prediction and motion generation based on reliable predictability. | Shun Nishide, Tetsuya Ogata, Ryunosuke Yokoya, Jun Tani, Kazunori Komatani, Hiroshi G. Okuno |
| 2008 | Interspeech | Extensibility verification of robust domain selection against out-of-grammar utterances in multi-domain spoken dialogue system. | Satoshi Ikeda, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2008 | Interspeech | Expanding vocabulary for recognizing user's abbreviations of proper nouns without increasing ASR error rates in spoken dialogue systems. | Masaki Katsumaru, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2008 | Interspeech | Predicting ASR errors by exploiting barge-in rate of individual users for spoken dialogue systems. | Kazunori Komatani, Tatsuya Kawahara, Hiroshi G. Okuno |
| 2008 | Interspeech | Soft missing-feature mask generation for simultaneous speech recognition system in robots. | Toru Takahashi, Shun'ichi Yamamoto, Kazuhiro Nakadai, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2008 | IROS | Segmenting acoustic signal with articulatory movement using Recurrent Neural Network for phoneme acquisition. | Hisashi Kanda, Tetsuya Ogata, Kazunori Komatani, Hiroshi G. Okuno |
| 2008 | IROS | Target speech detection and separation for humanoid robots in sparse dialogue with noisy home environments. | Hyun-Don Kim, Jinsung Kim, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2008 | IROS | Design and evaluation of two-channel-based sound source localization over entire azimuth range for moving talkers. | Hyun-Don Kim, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2008 | IROS | A robot listens to music and counts its beats aloud by separating music from counting voice. | Takeshi Mizumoto, Ryu Takeda, Kazuyoshi Yoshii, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2008 | IROS | A robot uses its own microphone to synchronize its steps to musical beats while scatting and singing. | Kazumasa Murata, Kazuhiro Nakadai, Kazuyoshi Yoshii, Ryu Takeda, Toyotaka Torii, Hiroshi G. Okuno, Yuji Hasegawa, Hiroshi Tsujino |
| 2008 | IROS | Active sensing based dynamical object feature extraction. | Shun Nishide, Tetsuya Ogata, Ryunosuke Yokoya, Jun Tani, Kazunori Komatani, Hiroshi G. Okuno |
| 2008 | IROS | Barge-in-able robot audition based on ICA and missing feature theory under semi-blind situation. | Ryu Takeda, Kazuhiro Nakadai, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2008 | ISM | Design and Implementation of 3D Auditory Scene Visualizer towards Auditory Awareness with Face Tracking. | Yuji Kubota, Masatoshi Yoshida, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2008 | PRICAI | SalienceGraph: Visualizing Salience Dynamics of Written Discourse by Using Reference Probability and PLSA. | Shun Shiramatsu, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2007 | ASRU | Design and implementation of a robot audition system for automatic speech recognition of simultaneous speech. | Shun'ichi Yamamoto, Kazuhiro Nakadai, Mikio Nakano, Hiroshi Tsujino, Jean-Marc Valin, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2007 | ICASSP | Integration and Adaptation of Harmonic and Inharmonic Models for Separating Polyphonic Musical Signals. | Katsutoshi Itoyama, Masataka Goto, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2007 | ICONIP | Vowel Imitation Using Vocal Tract Model and Recurrent Neural Network. | Hisashi Kanda, Tetsuya Ogata, Kazunori Komatani, Hiroshi G. Okuno |
| 2007 | ICRA | Predicting Object Dynamics from Visual Images through Active Sensing Experiences. | Shun Nishide, Tetsuya Ogata, Jun Tani, Kazunori Komatani, Hiroshi G. Okuno |
| 2007 | ICRA | Distance Estimation of Hidden Objects Based on Acoustical Holography by applying Acoustic Diffraction of Audible Sound. | Haruhiko Niwa, Tetsuya Ogata, Kazunori Komatani, Hiroshi G. Okuno |
| 2007 | ICRA | Human-Robot Cooperation using Quasi-symbols Generated by RNNPB Model. | Tetsuya Ogata, Shohei Matsumoto, Jun Tani, Kazunori Komatani, Hiroshi G. Okuno |
| 2007 | Interspeech | Topic estimation with domain extensibility for guiding user's out-of-grammar utterances in multi-domain spoken dialogue systems. | Satoshi Ikeda, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2007 | Interspeech | Analyzing temporal transition of real user's behaviors in a spoken dialogue system. | Kazunori Komatani, Tatsuya Kawahara, Hiroshi G. Okuno |
| 2007 | IROS | Vocal imitation using physical vocal tract model. | Hisashi Kanda, Tetsuya Ogata, Kazunori Komatani, Hiroshi G. Okuno |
| 2007 | IROS | Auditory and visual integration based localization and tracking of humans in daily-life environments. | Hyun-Don Kim, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2007 | IROS | Two-way translation of compound sentences and arm motions by recurrent neural networks. | Tetsuya Ogata, Masamitsu Murase, Jun Tani, Kazunori Komatani, Hiroshi G. Okuno |
| 2007 | IROS | Exploiting known sound source signals to improve ICA-based robot audition in speech separation and recognition. | Ryu Takeda, Kazuhiro Nakadai, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2007 | IROS | Discovery of other individuals by projecting a self-model through imitation. | Ryunosuke Yokoya, Tetsuya Ogata, Jun Tani, Kazunori Komatani, Hiroshi G. Okuno |
| 2007 | IROS | A biped robot that keeps steps in time with musical beats while listening to music with its own ears. | Kazuyoshi Yoshii, Kazuhiro Nakadai, Toyotaka Torii, Yuji Hasegawa, Hiroshi Tsujino, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2007 | RO-MAN | Auditory and Visual Integration based Localization and Tracking of Multiple Moving Sounds in Daily-life Environments. | Hyun-Don Kim, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2007 | SIGdial | Introducing Utterance Verification in Spoken Dialogue System to Improve Dynamic Help Generation for Novice Users. | Kazunori Komatani, Yuichiro Fukubayashi, Tetsuya Ogata, Hiroshi G. Okuno |
| 2006 | ICASSP | F0 Estimation Method for Singing Voice in Polyphonic Audio Signal Based on Statistical Vocal Model and Viterbi Search. | Hiromasa Fujihara, Tetsuro Kitahara, Masataka Goto, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2006 | ICASSP | Instrogram: A New Musical Instrument Recognition Technique Without Using Onset Detection NOR F0 Estimation. | Tetsuro Kitahara, Masataka Goto, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2006 | ICASSP | Robust Tracking of Multiple Sound Sources by Spatial Integration of Room And Robot Microphone Arrays. | Kazuhiro Nakadai, Hirofumi Nakajima, Masamitsu Murase, Satoshi Kaijiri, Kentaro Yamada, Takahiro Nakamura, Yuji Hasegawa, Hiroshi G. Okuno, Hiroshi Tsujino |
| 2006 | ICASSP | An Error Correction Framework Based on Drum Pattern Periodicity for Improving Drum Sound Detection. | Kazuyoshi Yoshii, Masataka Goto, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2006 | Interspeech | Speaker identification under noisy environments by using harmonic structure extraction and reliable frame weighting. | Hiromasa Fujihara, Tetsuro Kitahara, Masataka Goto, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2006 | Interspeech | Dynamic help generation by estimating user²s mental model in spoken dialogue systems. | Yuichiro Fukubayashi, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2006 | Interspeech | Improving speech recognition of two simultaneous speech signals by integrating ICA BSS and automatic missing feature mask generation. | Ryu Takeda, Shun'ichi Yamamoto, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2006 | Interspeech | Leak energy based missing feature mask generation for ICA and GSS and its evaluation with simultaneous speech recognition. | Shun'ichi Yamamoto, Ryu Takeda, Kazuhiro Nakadai, Mikio Nakano, Hiroshi Tsujino, Jean-Marc Valin, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2006 | IROS | Real-Time Tracking of Multiple Sound Sources by Integration of In-Room and Robot-Embedded Microphone Arrays. | Kazuhiro Nakadai, Hirofumi Nakajima, Masamitsu Murase, Hiroshi G. Okuno, Yuji Hasegawa, Hiroshi Tsujino |
| 2006 | IROS | Multiple Acoustical Holography Method for Localization of Objects in Broad Range using Audible Sound. | Haruhiko Niwa, Tetsuya Ogata, Kazunori Komatani, Hiroshi G. Okuno |
| 2006 | IROS | Missing-Feature based Speech Recognition for Two Simultaneous Speech Signals Separated by ICA with a pair of Humanoid Ears. | Ryu Takeda, Shun'ichi Yamamoto, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2006 | IROS | Real-Time Robot Audition System That Recognizes Simultaneous Speech in The Real World. | Shun'ichi Yamamoto, Kazuhiro Nakadai, Mikio Nakano, Hiroshi Tsujino, Jean-Marc Valin, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2006 | IROS | Experience Based Imitation Using RNNPB. | Ryunosuke Yokoya, Tetsuya Ogata, Jun Tani, Kazunori Komatani, Hiroshi G. Okuno |
| 2006 | ISM | Automatic Synchronization between Lyrics and Music CD Recordings Based on Viterbi Alignment of Segregated Vocal Signals. | Hiromasa Fujihara, Masataka Goto, Jun Ogata, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2006 | ISM | Musical Instrument Recognizer "Instrogram" and Its Application to Music Retrieval Based on Instrumentation Similarity. | Tetsuro Kitahara, Masataka Goto, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2006 | PRICAI | Recognition of Simultaneous Speech by Estimating Reliability of Separated Signals for Robot Audition. | Shun'ichi Yamamoto, Ryu Takeda, Kazuhiro Nakadai, Mikio Nakano, Hiroshi Tsujino, Jean-Marc Valin, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2006 | SIGdial | Multi-Domain Spoken Dialogue System with Extensibility and Robustness against Speech Recognition Errors. | Kazunori Komatani, Naoyuki Kanda, Mikio Nakano, Kazuhiro Nakadai, Hiroshi Tsujino, Tetsuya Ogata, Hiroshi G. Okuno |
| 2005 | Interspeech | Contextual constraints based on dialogue models in database search task for spoken dialogue systems. | Kazunori Komatani, Naoyuki Kanda, Tetsuya Ogata, Hiroshi G. Okuno |
| 2005 | Interspeech | Multiple moving speaker tracking by microphone array on mobile robot. | Masamitsu Murase, Shun'ichi Yamamoto, Jean-Marc Valin, Kazuhiro Nakadai, Kentaro Yamada, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2005 | IROS | Implementation of active direction-pass filter on dynamically reconfigurable processor. | Shunsuke Kurotaki, Noriaki Suzuki, Kazuhiro Nakadai, Hiroshi G. Okuno, Hideharu Amano |
| 2005 | IROS | A two-layer model for behavior and dialogue planning in conversational service robots. | Mikio Nakano, Yuji Hasegawa, Kazuhiro Nakadai, Takahiro Nakamura, Johane Takeuchi, Toyotaka Torii, Hiroshi Tsujino, Naoyuki Kanda, Hiroshi G. Okuno |
| 2005 | IROS | Extracting multi-modal dynamics of objects using RNNPB. | Tetsuya Ogata, Hayato Ohba, Jun Tani, Kazunori Komatani, Hiroshi G. Okuno |
| 2005 | IROS | Spatially mapping of friendliness for human-robot interaction. | Tsuyoshi Tasaki, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2005 | IROS | Making a robot recognize three simultaneous sentences in real-time. | Shun'ichi Yamamoto, Kazuhiro Nakadai, Jean-Marc Valin, Jean Rouat, Franois Michaud, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2005 | ICRA | Enhanced Robot Speech Recognition Based on Microphone Array Source Separation and Missing Feature Theory. | Shun'ichi Yamamoto, Jean-Marc Valin, Kazuhiro Nakadai, Jean Rouat, Franois Michaud, Tetsuya Ogata, Hiroshi G. Okuno |
| 2005 | PACLIC | Empirical Verification of Meaning-Game-based Generalization of Centering Theory with Large Japanese Corpus. | Shun Shiramatsu, Kazunori Komatani, Takashi Miyata, Koichi Hashida, Hiroshi G. Okuno |
| 2005 | SMC | Walking with body-sense in virtual space using the nonlinear oscillator. | Kenri Kodaka, Tetsuya Ogata, Hiroshi G. Okuno |
| 2004 | COLING | Using a Mixture of N-Best Lists from Multiple MT Systems in Rank-Sum-Based Confidence Measure for MT Outputs. | Yasuhiro Akiba, Eiichiro Sumita, Hiromi Nakaiwa, Seiichi Yamamoto, Hiroshi G. Okuno |
| 2004 | COLING | Efficient Confirmation Strategy for Large-scale Text Retrieval Systems with Spoken Dialogue Interface. | Kazunori Komatani, Teruhisa Misu, Tatsuya Kawahara, Hiroshi G. Okuno |
| 2004 | ICASSP | Category-level identification of non-registered musical instrument sounds. | Tetsuro Kitahara, Masataka Goto, Hiroshi G. Okuno |
| 2004 | ICASSP | Comparing features for forming music streams in automatic music transcription. | Yohei Sakuraba, Tetsuro Kitahara, Hiroshi G. Okuno |
| 2004 | Interspeech | Disambiguation in determining phonemes of sound-imitation words for environmental sound recognition. | Kazushi Ishihara, Yuya Hattori, Tomohiro Nakatani, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2004 | Interspeech | Robot motion control using listener's back-channels and head gesture information. | Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno, Tsuyoshi Tasaki, Takeshi Yamaguchi |
| 2004 | Interspeech | Drum sound identification for polyphonic music using template adaptation and matching methods. | Kazuyoshi Yoshii, Masataka Goto, Hiroshi G. Okuno |
| 2004 | IROS | Assessment of general applicability of robot audition system by recognizing three simultaneous speeches. | Shun'ichi Yamamoto, Kazuhiro Nakadai, Hiroshi Tsujino, Hiroshi G. Okuno |
| 2004 | ICRA | Improvement of Robot Audition by Interfacing Sound Source Separation and Automatic Speech Recognition with Missing Feature Theory. | Shun'ichi Yamamoto, Kazuhiro Nakadai, Hiroshi Tsujino, Toshio Yokoyama, Hiroshi G. Okuno |
| 2004 | LREC | Incremental Methods to Select Test Sentences for Evaluating Translation Ability. | Yasuhiro Akiba, Eiichiro Sumita, Hiromi Nakaiwa, Seiichi Yamamoto, Hiroshi G. Okuno |
| 2004 | PRICAI | Automatic Sound-Imitation Word Recognition from Environmental Sounds Focusing on Ambiguity Problem in Determining Phonemes. | Kazushi Ishihara, Tomohiro Nakatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2003 | ACL | Flexible Guidance Generation Using User Model in Spoken Dialogue Systems. | Kazunori Komatani, Shinichi Ueno, Tatsuya Kawahara, Hiroshi G. Okuno |
| 2003 | ACL | Chunk-Based Statistical Translation. | Taro Watanabe, Eiichiro Sumita, Hiroshi G. Okuno |
| 2003 | AINA | Privacy-Enhanced SPKI Access Control on PKIX and Its Application to Web Server . | Takamichi Saito, Kentaro Umesawa, Toshiyuki Kito, Hiroshi G. Okuno |
| 2003 | ICASSP | Musical instrument identification based on F0-dependent multivariate normal distribution. | Tetsuro Kitahara, Masataka Goto, Hiroshi G. Okuno |
| 2003 | ICRA | Robot recognizes three simultaneous speech by active audition. | Kazuhiro Nakadai, Hiroshi G. Okuno, Hiroaki Kitano |
| 2003 | ICRA | Realizing personality in audio-visually triggered non-verbal behaviors. | Hiroshi G. Okuno, Kazuhiro Nakadai, Hiroaki Kitano |
| 2003 | Interspeech | Automatic transformation of environmental sounds into sound-imitation words based on Japanese syllable structure. | Kazushi Ishihara, Yasushi Tsubota, Hiroshi G. Okuno |
| 2003 | Interspeech | User modeling in spoken dialogue systems for flexible guidance generation. | Kazunori Komatani, Shinichi Ueno, Tatsuya Kawahara, Hiroshi G. Okuno |
| 2003 | Interspeech | Three simultaneous speech recognition by integration of active audition and face recognition for humanoid. | Kazuhiro Nakadai, Daisuke Matsuura, Hiroshi G. Okuno, Hiroshi Tsujino |
| 2003 | IROS | Applying scattering theory to robot audition system: robust sound source localization and extraction. | Kazuhiro Nakadai, Daisuke Matsuura, Hiroshi G. Okuno, Hiroaki Kitano |
| 2003 | IWANN | Real-Time Sound Source Localization and Separation Based on Active Audio-Visual Integration. | Hiroshi G. Okuno, Kazuhiro Nakadai |
| 2003 | SIGdial | Flexible Spoken Dialogue System based on User Models and Dynamic Generation of VoiceXML Scripts. | Kazunori Komatani, Fumihiro Adachi, Shinichi Ueno, Tatsuya Kawahara, Hiroshi G. Okuno |
| 2002 | AAAI | Exploiting Auditory Fovea in Humanoid-Human Interaction. | Kazuhiro Nakadai, Hiroshi G. Okuno, Hiroaki Kitano |
| 2002 | COLING | Efficient Dialogue Strategy to Find Users' Intended Items from Information Query Results. | Kazunori Komatani, Tatsuya Kawahara, Ryosuke Ito, Hiroshi G. Okuno |
| 2002 | ICRA | Real-Time Speaker Localization and Speech Separation by Audio-Visual Integration. | Kazuhiro Nakadai, Ken-ichi Hidai, Hiroshi G. Okuno, Hiroaki Kitano |
| 2002 | Interspeech | Real-time sound source localization and separation for robot audition. | Kazuhiro Nakadai, Hiroshi G. Okuno, Hiroaki Kitano |
| 2002 | Interspeech | Auditory fovea based speech enhancement and its application to human-robot dialog system. | Kazuhiro Nakadai, Hiroshi G. Okuno, Hiroaki Kitano |
| 2002 | Interspeech | Belief network based disambiguation of object reference in spoken dialogue system for robot. | Yoko Yamakata, Tatsuya Kawahara, Hiroshi G. Okuno |
| 2002 | IROS | Auditory fovea based speech separation and its application to dialog system. | Kazuhiro Nakadai, Hiroshi G. Okuno, Hiroaki Kitano |
| 2002 | PRICAI | Realizing Audio-Visually Triggered ELIZA-Like Non-verbal Behaviors. | Hiroshi G. Okuno, Kazuhiro Nakadai, Hiroaki Kitano |
| 2001 | ESANN | A computational model of monkey grating cells for oriented repetitive alternating patterns. | Tino Lourens, Kazuhiro Nakadai, Hiroshi G. Okuno, Hiroaki Kitano |
| 2001 | ESANN | Graph extraction from color images. | Tino Lourens, Kazuhiro Nakadai, Hiroshi G. Okuno, Hiroaki Kitano |
| 2001 | ICIAP | Automatic Graph Extraction from Color Images. | Tino Lourens, Hiroshi G. Okuno, Hiroaki Kitano |
| 2001 | IJCAI | Real-Time Auditory and Visual Multiple-Object Tracking for Humanoids. | Kazuhiro Nakadai, Ken-ichi Hidai, Hiroshi Mizoguchi, Hiroshi G. Okuno, Hiroaki Kitano |
| 2001 | Interspeech | Real-time multiple speaker tracking by multi-modal integration for mobile robots. | Kazuhiro Nakadai, Ken-ichi Hidai, Hiroshi G. Okuno, Hiroaki Kitano |
| 2001 | Interspeech | Separating three simultaneous speeches with two microphones by integrating auditory and visual processing. | Hiroshi G. Okuno, Kazuhiro Nakadai, Tino Lourens, Hiroaki Kitano |
| 2001 | IROS | Epipolar geometry based sound localization and extraction for humanoid audition. | Kazuhiro Nakadai, Hiroshi G. Okuno, Hiroaki Kitano |
| 2001 | IROS | Human-robot interaction through real-time auditory and visual multiple-talker tracking. | Hiroshi G. Okuno, Kazuhiro Nakadai, Ken-ichi Hidai, Hiroshi Mizoguchi, Hiroaki Kitano |
| 2001 | IWANN | Detection of Oriented Repetitive Alternating Patterns in Color Images (A Computational Model of Monkey Grating Cells). | Tino Lourens, Hiroshi G. Okuno, Hiroaki Kitano |
| 2000 | AAAI | Active Audition for Humanoid. | Kazuhiro Nakadai, Tino Lourens, Hiroshi G. Okuno, Hiroaki Kitano |
| 2000 | ICPADS | Privacy enhanced access control by SPKI. | Takamichi Saito, Kentaro Umesawa, Hiroshi G. Okuno |
| 2000 | IROS | A framework for integrating sensory information in a humanoid robot. | Iris Fermin, Hiroshi G. Okuno, Hiroshi Ishiguro, Hiroaki Kitano |
| 2000 | IROS | Design and architecture of SIG the humanoid: an experimental platform for integrated perception in RoboCup humanoid challenge. | Hiroaki Kitano, Hiroshi G. Okuno, Kazuhiro Nakadai, Theo Sabisch, Tatsuya Matsui |
| 2000 | IROS | Active audition system and humanoid exterior design. | Kazuhiro Nakadai, Tatsuya Matsui, Hiroshi G. Okuno, Hiroaki Kitano |
| 2000 | PRICAI | Humanoid Active Audition System Improved by the Cover Acoustics. | Kazuhiro Nakadai, Hiroshi G. Okuno, Hiroaki Kitano |
| 2000 | WETICE | Privacy-Enhanced Access Control by SPKI and Its Application to Web Server. | Takamichi Saito, Kentaro Umesawa, Hiroshi G. Okuno |
| 2000 | RoboCup | And the Fans Are Going Wild! SIG plus MIKE. | Ian Frank, Kumiko Tanaka-Ishii, Hiroshi G. Okuno, Junichi Akita, Yukiko Nakagawa, Kazuaki Maeda, Kazuhiro Nakadai, Hiroaki Kitano |
| 2000 | RoboCup | Bridging Gap between the Simulation and Robotics with a Global Vision System. | Yukiko Nakagawa, Hiroshi G. Okuno, Hiroaki Kitano |
| 1999 | AAAI | Using Vision to Improve Sound Source Separation. | Yukiko Nakagawa, Hiroshi G. Okuno, Hiroaki Kitano |
| 1998 | AAAI | Sound Ontology for Computational Auditory Scence Analysis. | Tomohiro Nakatani, Hiroshi G. Okuno |
| 1997 | IJCAI | Understanding Three Simultaneous Speeches. | Hiroshi G. Okuno, Tomohiro Nakatani, Takeshi Kawabata |
| 1996 | AAAI | Interfacing Sound Stream Segregation to Automatic Speech Recognition - Preliminary Results on Listening to Several Sounds Simultaneously. | Hiroshi G. Okuno, Tomohiro Nakatani, Takeshi Kawabata |
| 1996 | ICASSP | Localization by harmonic structure and its application to harmonic sound stream segregation. | Tomohiro Nakatani, Masataka Goto, Hiroshi G. Okuno |
| 1996 | Interspeech | A new speech enhancement: speech stream segregation. | Hiroshi G. Okuno, Tomohiro Nakatani, Takeshi Kawabata |
| 1995 | ICASSP | A computational model of sound stream segregation with multi-agent paradigm. | Tomohiro Nakatani, Takeshi Kawabata, Hiroshi G. Okuno |
| 1995 | IJCAI | Residue-Driven Architecture for Computational Auditory Scene Analysis. | Tomohiro Nakatani, Hiroshi G. Okuno, Takeshi Kawabata |
| 1994 | AAAI | Auditory Stream Segregation in Auditory Scene Analysis with a Multi-Agent System. | Tomohiro Nakatani, Hiroshi G. Okuno, Takeshi Kawabata |
| 1994 | Interspeech | Unified architecture for auditory scene analysis and spoken language processing. | Tomohiro Nakatani, Takeshi Kawabata, Hiroshi G. Okuno |
| 1987 | MICRO | Firmware approach to fast Lisp interpreter. | Hiroshi G. Okuno, Nobuyasu Osato, Ikuo Takeuchi |