| 2021 | HAT-Net: A Hierarchical Transformer Graph Neural Network for Grading of Colorectal Cancer Histology Images. | Yihan Su, Yu Bai, Bo Zhang, Zheng Zhang, Wendong Wang |
| 2021 | Spatial-Temporal Residual Aggregation for High Resolution Video Inpainting. | Vishnu Sanjay Ramiya Srinivasan, Rui Ma, Qiang Tang, Zili Yi, Zhan Xu |
| 2021 | DISCO: accurate Discrete Scale Convolutions. | Ivan Sosnovik, Artem Moskalev, Arnold W. M. Smeulders |
| 2021 | IB-MVS: An Iterative Algorithm for Deep Multi-View Stereo based on Binary Decisions. | Christian Sormann, Mattia Rossi, Andreas Kuhn, Friedrich Fraundorfer |
| 2021 | Exploiting Scene Depth for Object Detection with Multimodal Transformers. | Hwanjun Song, Eunyoung Kim, Varun Jampani, Deqing Sun, Jae-Gil Lee, Ming-Hsuan Yang |
| 2021 | Image-Text Alignment using Adaptive Cross-attention with Transformer Encoder for Scene Graphs. | Juyong Song, Sunghyun Choi |
| 2021 | Progressive Growing of Points with Tree-structured Generators. | Hyeontae Son, Young Min Kim |
| 2021 | Localizing Objects with Self-supervised Transformers and no Labels. | Oriane Simoni, Gilles Puy, Huy V. Vo, Simon Roburin, Spyros Gidaris, Andrei Bursuc, Patrick Prez, Renaud Marlet, Jean Ponce |
| 2021 | Looking at the whole picture: constrained unsupervised anomaly segmentation. | Julio Silva-Rodrguez, Valery Naranjo, Jose Dolz |
| 2021 | Unsupervised Spatio-temporal Latent Feature Clustering for Multiple-object Tracking and Segmentation. | Abubakar Siddique, Reza Jalil Mozhdehi, Henry Medeiros |
| 2021 | Lightweight HDR Camera ISP for Robust Perception in Dynamic Illumination Conditions via Fourier Adversarial Networks. | Pranjay Shyam, Sandeep Singh Sengar, Kuk-Jin Yoon, Kyung-Soo Kim |
| 2021 | 3D-RETR: End-to-End Single and Multi-View 3D Reconstruction with Transformers. | Zai Shi, Zhao Meng, Yiran Xing, Yunpu Ma, Roger Wattenhofer |
| 2021 | Efficient Cross-Modal Retrieval via Deep Binary Hashing and Quantization. | Yang Shi, Young-joo Chung |
| 2021 | DKMA-ULD: Domain Knowledge augmented Multi-head Attention based Robust Universal Lesion Detection. | Manu Sheoran, Meghal Dani, Monika Sharma, Lovekesh Vig |
| 2021 | Lane Line Detection based on Parallel Spatial Separation Convolution. | Xile Shen, Zongqing Lu, Youcheng Zhang, Jing-Hao Xue |
| 2021 | TridentAdapt: Learning Domain-invariance via Source-Target Confrontation and Self-induced Cross-domain Augmentation. | Fengyi Shen, Akhil Gurram, Ahmet Faruk Tuna, Onay Urfalioglu, Alois C. Knoll |
| 2021 | SAGAN: Adversarial Spatial-asymmetric Attention for Noisy Nona-Bayer Reconstruction. | S. M. A. Sharif, Rizwan Ali Naqvi, Mithun Biswas |
| 2021 | WP2-GAN: Wavelet-based Multi-level GAN for Progressive Facial Expression Translation with Parallel Generators. | Jun Shao, Tien Bui |
| 2021 | Learning Neural Transmittance for Efficient Rendering of Reflectance Fields. | Mohammad Shafiei, Sai Bi, Zhengqin Li, Aidas Liaudanskas, Rodrigo Ortiz Cayon, Ravi Ramamoorthi |
| 2021 | Pose-Transformation and Radial Distance Clustering for Unsupervised Person Re-identification. | Siddharth Seth, Akash Sonth, Anirban Chakraborty |
| 2021 | Probabilistic Estimation of 3D Human Shape and Pose with a Semantic Local Parametric Model. | Akash Sengupta, Ignas Budvytis, Roberto Cipolla |
| 2021 | Personalized One-Shot Lipreading for an ALS Patient. | Bipasha Sen, Aditya Agarwal, Rudrabha Mukhopadhyay, Vinay P. Namboodiri, C. V. Jawahar |
| 2021 | Learning to Sparsify Differences of Synaptic Signal for Efficient Event Processing. | Yusuke Sekikawa, Keisuke Uto |
| 2021 | Multi-Source Domain Adaptation via supervised contrastive learning and confident consistency regularization. | Marin Scalbert, Florent Couzini-Devy, Maria Vakalopoulou |
| 2021 | Deep Knowledge Distillation using Trainable Dense Attention. | Bharat Bhusan Sau, Soumya Roy, Vinay P. Namboodiri, Raghu Sesha Iyengar |