| 2007 | Joint Analysis of the Emotional Fingerprint in the Face and Speech: A single subject study. | Carlos Busso, Shrikanth S. Narayanan |
| 2007 | New Directions in Image and Video Quality Assessment Plenary Talk. | Alan C. Bovik |
| 2007 | Segmentation of Document Images Using Higher Order Statistics. | Paulo Vinicius Koerich Borges, Joceli Mayer, Ebroul Izquierdo |
| 2007 | Key frame extraction in video sequences: a vantage points approach. | Dimitrios Besiris, Nikolaos A. Laskaris, Fotini Fotopoulou, George Economou |
| 2007 | Dual-Mode Wideband Speech Compression. | Visar Berisha, Andreas Spanias |
| 2007 | Toward a 3D watermarking benchmark. | Jihane Bennour, Jean-Luc Dugelay |
| 2007 | Multimodal Sensor Analysis of Sitar Performance: Where is the Beat? | Manjinder Singh Benning, Ajay Kapur, Bernie C. Till, George Tzanetakis |
| 2007 | Impact of Additional Noise on Subjective and Objective Quality Assessement in VoIP. | Zdenek Becvar, Lukas Novak, Jan Zelenka, Miloslav Brada, Pavel Slepicka |
| 2007 | Smart Transcoding between CELP Speech Codecs through Voiced Oriented Pitch Mapping. | Christophe Beaugeant |
| 2007 | Enhancing Social Communication in High-Functioning Children with Autism through a Co-Located Interface. | Nirit Bauminger, Dina Goren-Bar, Eynat Gal, Patrice L. Weiss, Judi Kupersmitt, Fabio Pianesi, Oliviero Stock, Massimo Zancanaro |
| 2007 | Temporal registration using 3D phase correlation and a maximum likelihood approach in the perceptual evaluation of video quality. | Marcus Barkowsky, Jens Bialkowski, Roland Bitto, Andr Kaup |
| 2007 | Adaptive Frame Interpolation for Wyner-Ziv Video Coding. | Savvas Argyropoulos, Nikolaos Thomos, Nikolaos V. Boulgouris, Michael G. Strintzis |
| 2007 | Cross-Layer Adaptive ARQ for Uplink Video Streaming in Tandem Wireless/Wireline Networks. | Antonios Argyriou |
| 2007 | Wyner-Ziv Stereo Video Coding using a Side Information Fusion Approach. | Jos Diogo Areia, Joo Ascenso, Catarina Brites, Fernando Pereira |
| 2007 | A System for Technology Based Assessment of Language and Literacy in Young Children: the Role of Multiple Information Sources. | Abeer Alwan, Yijian Bai, Matthew Black, Larry Casey, Matteo Gerosa, Margaret Heritage, Markus Iseli, Barbara Jones, Abe Kazemzadeh, Sungbok Lee, Shrikanth S. Narayanan, Patti Price, Joseph Tepperman, Shizhen Wang |
| 2007 | Unsupervised Discovery of Action Hierarchies in Large Collections of Activity Videos. | Parvez Ahammad, Chuohao Yeo, Kannan Ramchandran, S. Shankar Sastry |
| 2007 | R-Flow: An Extensible XML Based Multimodal Dialog System Architecture. | Li Li, Quanzhi Li, Wu Chou, Feng Liu |
| 2006 | Enhanced SLCCA Image Compression. | Xinhua Zhuang, Shu-Mei Guo, Wei-Nong Lee |
| 2006 | Adaptive Multi-path Prediction for Error Resilient H.264 Coding. | Xiaosong Zhou, C.-C. Jay Kuo |
| 2006 | Boosting-Based Multimodal Speaker Detection for Distributed Meetings. | Cha Zhang, Pei Yin, Yong Rui, Ross Cutler, Paul A. Viola |
| 2006 | Joint Data Partition and Rate-Distortion Optimized Mode Selection for H.264 Error-Resilient Coding. | Yuan Zhang, Wen Gao, Debin Zhao |
| 2006 | Server Policies For Interactive Transmission Of 3D Scenes. | Pietro Zanuttigh, Nicola Brusco, David Taubman, Guido Maria Cortelazzo |
| 2006 | Compressed Domain Real-time Action Recognition. | Chuohao Yeo, Parvez Ahammad, Kannan Ramchandran, S. Shankar Sastry |
| 2006 | Fast Image/Video Contrast Enhancement Based on WTHE. | Qing Wang, Rabab K. Ward |
| 2006 | Region-Based Stereo Panorama Disparity Adjusting. | Chiao Wang, Alexander A. Sawchuk |