| 2004 | A design of audio-visual talker tracking system based on CSP analysis and frame difference in real noisy environments. | Nishiura Denda, Takanobu Nishiura, Hideki Kawahara, Toshio Irino |
| 2004 | 3D model search and retrieval based on the spherical trace transform. | Petros Daras, Dimitrios Zarpalas, Dimitrios Tzovaras, Michael G. Strintzis |
| 2004 | Inlier modeling for multimedia data analysis. | Rozenn Dahyot, Niall Rea, Anil C. Kokaram, Nick G. Kingsbury |
| 2004 | A novel scheme for merging digital audio watermarking and authentication. | Nedeljko Cvejic, Tapio Seppnen |
| 2004 | Invariant action classification with volumetric data. | Fabio Cuzzolin, Augusto Sarti, Stefano Tubaro |
| 2004 | Optimal energy distribution in embedded packet video transmission over wireless channels. | Cristina Emilia Costa, Fabrizio Granelli, Francesco G. B. De Natale |
| 2004 | Video summarization using a neurodynamical model of visual attention. | Silvia Corchs, Gianluigi Ciocca, Raimondo Schettini |
| 2004 | Tennis video abstraction from audio and visual cues. | Franois Coldefy, Patrick Bouthemy, Michael Betser, Guillaume Gravier |
| 2004 | A significant motion vector protection-based error-resilient scheme in H.264. | Jan-Ru Chen, Chun-Shien Lu, Kuo-Chin Fan |
| 2004 | Sender-based rate-distortion optimized streaming of 3-D wavelet video with low latency. | Chuo-Ling Chang, Sangeun Han, Bernd Girod |
| 2004 | On optimal selection of lip-motion features for speaker identification. | Hasan Ertan etingl, Engin Erzin, Ycel Yemez, A. Murat Tekalp |
| 2004 | Fully automatic, real-time detection of facial gestures from generic video. | Marco La Cascia, Lorenzo Valenti, Stan Sclaroff |
| 2004 | Adaptive image data fusion for consumer devices application. | Alessandro Capra, Alfio Castorina, Paolo Vivirito, Sebastiano Battiato |
| 2004 | A smoothly scalable and fully JPEG2000-compatible video coder. | Marco Cagnazzo, Thomas Andr, Marc Antonini, Michel Barlaud |
| 2004 | Gait recognition using dynamic time warping. | Nikolaos V. Boulgouris, Konstantinos N. Plataniotis, Dimitrios Hatzinakos |
| 2004 | A system for real-time synthesis of subtle expressivity for life-like MPEG-4 based virtual characters. | Carlo Bonamico, Carlo Braccini, Fabio Lavagetto, Maurizio Costa |
| 2004 | Bandwidth estimation in prerecorded VBR-video distribution systems exploiting stream correlation. | Gennaro Boggia, Pietro Camarda, Domenico Striccoli |
| 2004 | Automatic annotation of video streams. | Marco Bertini, Alberto Del Bimbo, Walter Nunziati |
| 2004 | Soft-input decoding of variable-length codes applied to the H.264 standard. | Cyril Bergeron, Catherine Lamy-Bergot |
| 2004 | Joint optimization of scale factors and Huffman code books for MPEG-4 AAC. | Claus Bauer, Mark Vinton |
| 2004 | Evaluation of distance measures for MPEG-7 melody contours. | Jan-Mark Batke, Gunnar Eisenberg, Gunnar Weishaupt, Thomas Sikora |
| 2004 | A low complexity concealment algorithm for the whole-frame loss in H.264/AVC. | Pierpaolo Baccichet, Antonio Chimienti |
| 2004 | A novel iterative approach for JPEG2000 error concealment. | Luigi Atzori, Giaime Ginesu, Alessio Raccis, Daniele D. Giusto |
| 2004 | Accurate and fast audio-realistic rendering of sounds in virtual environments. | Fabio Antonacci, Marco Foco, Augusto Sarti, Stefano Tubaro |
| 2004 | Multiple camera image acquisition models for multi-view 3D display interaction. | Zahir Y. Alpaslan, Alexander A. Sawchuk |