Skip to content

Lorenzo Baraldi

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

89

Venues

18

Active years

2014–2026

Best venue rank

A*

Where they publish

Papers

89 indexed papers, newest first.

YearVenueTitleAuthors
2026ICPRImproving LLM First-Token Predictions in Multiple-Choice Question Answering via Output Prefilling.Silvia Cappelletti, Tobia Poppi, Samuele Poppi, Zheng Xin Yong, Diego Garcia-Olano, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
2026ICPRGramSR: Visual Feature Conditioning for Diffusion-Based Super-Resolution.Fabio D'Oronzio, Federico Putamorsi, Leonardo Zini, Marcella Cornia, Lorenzo Baraldi
2026ICPRRaTA-Tool: Retrieval-Based Tool Selection with Multimodal Large Language Models.Gabriele Mattioli, Evelyn Turri, Sara Sarto, Lorenzo Baraldi, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
2026WACVFG-Tracer: Tracing Information Flow in Multimodal Large Language Models in Free-Form Generation.Alessia Saporita, Vittorio Pipoli, Federico Bolelli, Lorenzo Baraldi, Andrea Acquaviva, Elisa Ficarra
2025CAIPTracing Information Flow in LLaMA Vision: A Step Toward Multimodal Understanding.Alessia Saporita, Vittorio Pipoli, Federico Bolelli, Lorenzo Baraldi, Andrea Acquaviva, Elisa Ficarra
2025CVPRRecurrence-Enhanced Vision-and-Language Transformers for Robust Multimodal Document Retrieval.Davide Caffagni, Sara Sarto, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
2025CVPRAugmenting Multimodal LLMs with Self-Reflective Tokens for Knowledge-based Visual Question Answering.Federico Cocchi, Nicholas Moratelli, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
2025CVPRHyperbolic Safety-Aware Vision-Language Models.Tobia Poppi, Tejaswi Kasarla, Pascal Mettes, Lorenzo Baraldi, Rita Cucchiara
2025ICASSPMultimodal Emotion Recognition in Conversation via Possible Speaker's Audio and Visual Sequence Selection.Rahul Singh Maharjan, Niyati Rawal, Marta Romeo, Lorenzo Baraldi, Rita Cucchiara, Angelo Cangelosi
2025ICCVWhat Changed? Detecting and Evaluating Instruction-Guided Image Edits with Multimodal Large Language Models.Lorenzo Baraldi, Davide Bucciarelli, Federico Betti, Marcella Cornia, Lorenzo Baraldi, Nicu Sebe, Rita Cucchiara
2025ICCVTalking to DINO: Bridging Self-Supervised Vision Backbones with Language for Open-Vocabulary Segmentation.Luca Barsellotti, Lorenzo Bianchi, Nicola Messina, Fabio Carrara, Marcella Cornia, Lorenzo Baraldi, Fabrizio Falchi, Rita Cucchiara
2025ICCVLLaVA-MORE: A Comparative Study of LLMs and Visual Backbones for Enhanced Visual Instruction Tuning.Federico Cocchi, Nicholas Moratelli, Davide Caffagni, Sara Sarto, Lorenzo Baraldi, Marcella Cornia, Rita Cucchiara
2025ICCVMISSRAG: Addressing the Missing Modality Challenge in Multimodal Large Language Models.Vittorio Pipoli, Alessia Saporita, Federico Bolelli, Marcella Cornia, Lorenzo Baraldi, Costantino Grana, Rita Cucchiara, Elisa Ficarra
2025ICIAPMATE: Multimodal Agent that Talks and Empathizes.Niyati Rawal, Matteo Xia, David Tessaro, Lorenzo Baraldi, Rita Cucchiara
2025ICIAPSVGauge: Towards Human-Aligned Evaluation for SVG Generation.Leonardo Zini, Elia Frigieri, Sebastiano Aloscari, Marcello Generali, Lorenzo Dodi, Robert Dosen, Lorenzo Baraldi
2025ICLRCausal Graphical Models for Vision-Language Compositional Understanding.Fiorenzo Parascandolo, Nicholas Moratelli, Enver Sangineto, Lorenzo Baraldi, Rita Cucchiara
2025WACVPerceive. Query & Reason: Enhancing Video QA with Question-Guided Temporal Queries.Roberto Amoroso, Gengyuan Zhang, Rajat Koner, Lorenzo Baraldi, Rita Cucchiara, Volker Tresp
2025WACVSemantically Conditioned Prompts for Visual Recognition Under Missing Modality Scenarios.Vittorio Pipoli, Federico Bolelli, Sara Sarto, Marcella Cornia, Lorenzo Baraldi, Costantino Grana, Rita Cucchiara, Elisa Ficarra
2024ACLThe Revolution of Multimodal Large Language Models: A Survey.Davide Caffagni, Federico Cocchi, Luca Barsellotti, Nicholas Moratelli, Sara Sarto, Lorenzo Baraldi, Marcella Cornia, Rita Cucchiara
2024BMVCRevisiting Image Captioning Training Paradigm via Direct CLIP-based Optimization.Nicholas Moratelli, Davide Caffagni, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
2024CVPRTraining-Free Open-Vocabulary Segmentation with Offline Diffusion-Augmented Prototype Generation.Luca Barsellotti, Roberto Amoroso, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
2024CVPRWiki-LLaVA: Hierarchical Retrieval-Augmented Generation for Multimodal LLMs.Davide Caffagni, Federico Cocchi, Nicholas Moratelli, Sara Sarto, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
2024CVPRAIGeN: An Adversarial Approach for Instruction Generation in VLN.Niyati Rawal, Roberto Bigazzi, Lorenzo Baraldi, Rita Cucchiara
2024ECCVContrasting Deepfakes Diffusion via Contrastive Learning and Global-Local Similarities.Federico Cocchi, Marcella Cornia, Lorenzo Baraldi, Alessandro Nicolosi, Rita Cucchiara
2024ECCVOptimizing Resource Consumption in Diffusion Models Through Hallucination Early Detection.Federico Betti, Lorenzo Baraldi, Lorenzo Baraldi, Rita Cucchiara, Nicu Sebe
2024ECCVPersonalizing Multimodal Large Language Models for Image Captioning: An Experimental Analysis.Davide Bucciarelli, Nicholas Moratelli, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
2024ECCVSafe-CLIP: Removing NSFW Concepts from Vision-and-Language Models.Samuele Poppi, Tobia Poppi, Federico Cocchi, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
2024ECCVBRIDGE: Bridging Gaps in Image Captioning Evaluation with Stronger Visual Cues.Sara Sarto, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
2024ICPRAdapt to Scarcity: Few-Shot Deepfake Detection via Low-Rank Adaptation.Silvia Cappelletti, Lorenzo Baraldi, Federico Cocchi, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
2024ICPRFluent and Accurate Image Captioning with a Self-trained Reward Model.Nicholas Moratelli, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
2024ICPRUnlearning Vision Transformers Without Retaining Data via Low-Rank Decompositions.Samuele Poppi, Sara Sarto, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
2024ICRAMapping High-level Semantic Regions in Indoor Environments without Object Recognition.Roberto Bigazzi, Lorenzo Baraldi, Shreyas Kousik, Rita Cucchiara, Marco Pavone
2024WACVFOSSIL: Free Open-Vocabulary Semantic Segmentation through Synthetic References Retrieval.Luca Barsellotti, Roberto Amoroso, Lorenzo Baraldi, Rita Cucchiara
2024WACVWhat's Outside the Intersection? Fine-grained Error Analysis for Semantic Segmentation Beyond IoU.Maximilian Bernhard, Roberto Amoroso, Yannic Kindermann, Lorenzo Baraldi, Rita Cucchiara, Volker Tresp, Matthias Schubert
2023BMVCSuperpixel Positional Encoding to Improve ViT-based Semantic Segmentation Models.Roberto Amoroso, Matteo Tomei, Lorenzo Baraldi, Rita Cucchiara
2023CVPRPositive-Augmented Contrastive Learning for Image and Video Captioning Evaluation.Sara Sarto, Manuele Barraco, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
2023ICCVWith a Little Help from your own Past: Prototypical Memory Networks for Image Captioning.Manuele Barraco, Sara Sarto, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
2023ICIAPEnhancing Open-Vocabulary Semantic Segmentation with Prototype Retrieval.Luca Barsellotti, Roberto Amoroso, Lorenzo Baraldi, Rita Cucchiara
2023ICIAPSynthCap: Augmenting Transformers with Synthetic Data for Image Captioning.Davide Caffagni, Manuele Barraco, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
2023ICIAPUnveiling the Impact of Image Transformations on Deepfake Detection: An Experimental Analysis.Federico Cocchi, Lorenzo Baraldi, Samuele Poppi, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
2023ICIAPTowards Explainable Navigation and Recounting.Samuele Poppi, Roberto Bigazzi, Niyati Rawal, Marcella Cornia, Silvia Cascianelli, Lorenzo Baraldi, Rita Cucchiara
2023ICRAEmbodied Agents for Efficient Exploration and Smart Scene Description.Roberto Bigazzi, Marcella Cornia, Silvia Cascianelli, Lorenzo Baraldi, Rita Cucchiara
2022CBMIALADIN: Distilling Fine-grained Alignment Scores for Efficient Image-Text Matching and Retrieval.Nicola Messina, Matteo Stefanini, Marcella Cornia, Lorenzo Baraldi, Fabrizio Falchi, Giuseppe Amato, Rita Cucchiara
2022CBMIRetrieval-Augmented Transformer for Image Captioning.Sara Sarto, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
2022CVPRThe Unreasonable Effectiveness of CLIP Features for Image Captioning: An Experimental Analysis.Manuele Barraco, Marcella Cornia, Silvia Cascianelli, Lorenzo Baraldi, Rita Cucchiara
2022CVPRDual-Branch Collaborative Transformer for Virtual Try-On.Emanuele Fenocchi, Davide Morelli, Marcella Cornia, Lorenzo Baraldi, Fabio Cesari, Rita Cucchiara
2022ICIAPEmbodied Navigation at the Art Gallery.Roberto Bigazzi, Federico Landi, Silvia Cascianelli, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
2022ICIAPInvestigating Bidimensional Downsampling in Vision Transformer Models.Paolo Bruno, Roberto Amoroso, Marcella Cornia, Silvia Cascianelli, Lorenzo Baraldi, Rita Cucchiara
2022ICPRCaMEL: Mean Teacher Learning for Image Captioning.Manuele Barraco, Matteo Stefanini, Marcella Cornia, Silvia Cascianelli, Lorenzo Baraldi, Rita Cucchiara
2022ICPRThe LAM Dataset: A Novel Benchmark for Line-Level Handwritten Text Recognition.Silvia Cascianelli, Vittorio Pippi, Martin Maarand, Marcella Cornia, Lorenzo Baraldi, Christopher Kermorvant, Rita Cucchiara
2022ICPRSpot the Difference: A Novel Task for Embodied Agents in Changing Environments.Federico Landi, Roberto Bigazzi, Marcella Cornia, Silvia Cascianelli, Lorenzo Baraldi, Rita Cucchiara
2021CAIPAssessing the Role of Boundary-Level Objectives in Indoor Semantic Segmentation.Roberto Amoroso, Lorenzo Baraldi, Rita Cucchiara
2021CAIPOut of the Box: Embodied Navigation in the Real World.Roberto Bigazzi, Federico Landi, Marcella Cornia, Silvia Cascianelli, Lorenzo Baraldi, Rita Cucchiara
2021CAIPLearning to Read L'Infinito: Handwritten Text Recognition with Synthetic Training Data.Silvia Cascianelli, Marcella Cornia, Lorenzo Baraldi, Maria Ludovica Piazzi, Rosiana Schiuma, Rita Cucchiara
2021CVPRRevisiting the Evaluation of Class Activation Mapping for Explainability: A Novel Metric and Experimental Analysis.Samuele Poppi, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
2021CVPREstimating (and Fixing) the Effect of Face Obfuscation in Video Recognition.Matteo Tomei, Lorenzo Baraldi, Simone Bronzin, Rita Cucchiara
2021IWANNImproving Indoor Semantic Segmentation with Boundary-Level Objectives.Roberto Amoroso, Lorenzo Baraldi, Rita Cucchiara
2020CVPRMeshed-Memory Transformer for Image Captioning.Marcella Cornia, Matteo Stefanini, Lorenzo Baraldi, Rita Cucchiara
2020ICPRExplore and Explain: Self-supervised Navigation and Recounting.Roberto Bigazzi, Federico Landi, Marcella Cornia, Silvia Cascianelli, Lorenzo Baraldi, Rita Cucchiara
2020ICPRWatch Your Strokes: Improving Handwritten Text Recognition with Deformable Convolutions.Iulian Cojocaru, Silvia Cascianelli, Lorenzo Baraldi, Massimiliano Corsini, Rita Cucchiara
2020ICPRA Novel Attention-based Aggregation Function to Combine Vision and Language.Matteo Stefanini, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
2020ICPRRMS-Net: Regression and Masking for Soccer Event Spotting.Matteo Tomei, Lorenzo Baraldi, Simone Calderara, Simone Bronzin, Rita Cucchiara
2020ICRASMArT: Training Shallow Memory-aware Transformers for Robotic Explainability.Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
2019BMVCEmbodied Vision-and-Language Navigation with Dynamic Convolutional Filters.Federico Landi, Lorenzo Baraldi, Massimiliano Corsini, Rita Cucchiara
2019CLOSERA Deep-learning-based approach to VM behavior Identification in Cloud Systems.Matteo Stefanini, Riccardo Lancellotti, Lorenzo Baraldi, Simone Calderara
2019CVPRShow, Control and Tell: A Framework for Generating Controllable and Grounded Captions.Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
2019CVPRArt2Real: Unfolding the Reality of Artworks via Semantically-Aware Image-To-Image Translation.Matteo Tomei, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
2019ICIAPArtpedia: A New Visual-Semantic Dataset with Visual and Contextual Sentences in the Artistic Domain.Matteo Stefanini, Marcella Cornia, Lorenzo Baraldi, Massimiliano Corsini, Rita Cucchiara
2019ICIAPImage-to-Image Translation to Unfold the Reality of Artworks: An Empirical Analysis.Matteo Tomei, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
2018CVPRLAMV: Learning to Align and Match Videos With Kernelized Temporal Layers.Lorenzo Baraldi, Matthijs Douze, Rita Cucchiara, Herv Jgou
2018CVPRSAM: Pushing the Limits of Saliency Prediction Models.Marcella Cornia, Lorenzo Baraldi, Giuseppe Serra, Rita Cucchiara
2018ECCVVisual-Semantic Alignment Across Domains Using a Semi-Supervised Approach.Angelo Carraggi, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
2018ECCVTowards Cycle-Consistent Models for Text and Image Retrieval.Marcella Cornia, Lorenzo Baraldi, Hamed R. Tavakoli, Rita Cucchiara
2018ECCVWhat Was Monet Seeing While Painting? Translating Artworks to Photo-Realistic Images.Matteo Tomei, Lorenzo Baraldi, Marcella Cornia, Rita Cucchiara
2018ICPRAligning Text and Document Illustrations: Towards Visually Explainable Digital Humanities.Lorenzo Baraldi, Marcella Cornia, Costantino Grana, Rita Cucchiara
2018ICPRConnected Components Labeling on DRAGs.Federico Bolelli, Lorenzo Baraldi, Michele Cancilla, Costantino Grana
2017CBMINeuralStory: an Interactive Multimedia System for Video Indexing and Re-use.Lorenzo Baraldi, Costantino Grana, Rita Cucchiara
2017CVPRHierarchical Boundary-Aware Neural Encoder for Video Captioning.Lorenzo Baraldi, Costantino Grana, Rita Cucchiara
2017ICIAPTowards Video Captioning with Naming: A Novel Dataset and a Multi-modal Approach.Stefano Pini, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
2017ICMIModeling multimodal cues in a deep learning-based framework for emotion recognition in the wild.Stefano Pini, Olfa Ben Ahmed, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara, Benoit Huet
2016ACIVSOptimized Connected Components Labeling with Pixel Prediction.Costantino Grana, Lorenzo Baraldi, Federico Bolelli
2016ECCVMulti-level Net: A Visual Saliency Prediction Model.Marcella Cornia, Lorenzo Baraldi, Giuseppe Serra, Rita Cucchiara
2016ECCVContext Change Detection for an Ultra-Low Power Low-Resolution Ego-Vision Imager.Francesco Paci, Lorenzo Baraldi, Giuseppe Serra, Rita Cucchiara, Luca Benini
2016ICPRHistorical document digitization through layout analysis and deep content classification.Andrea Corbelli, Lorenzo Baraldi, Costantino Grana, Rita Cucchiara
2016ICPRA deep multi-level network for saliency prediction.Marcella Cornia, Lorenzo Baraldi, Giuseppe Serra, Rita Cucchiara
2016ICPRYACCLAB - Yet Another Connected Components Labeling Benchmark.Costantino Grana, Federico Bolelli, Lorenzo Baraldi, Roberto Vezzani
2015CAIPShot and Scene Detection via Hierarchical Clustering for Re-using Broadcast Video.Lorenzo Baraldi, Costantino Grana, Rita Cucchiara
2015IBPRIAMeasuring Scene Detection Performance.Lorenzo Baraldi, Costantino Grana, Rita Cucchiara
2014CVPRGesture Recognition in Ego-centric Videos Using Dense Trajectories and Hand Segmentation.Lorenzo Baraldi, Francesco Paci, Giuseppe Serra, Luca Benini, Rita Cucchiara