Search Results for author: Matteo Stefanini

Found 8 papers, 4 papers with code

CaMEL: Mean Teacher Learning for Image Captioning

1 code implementation21 Feb 2022 Manuele Barraco, Matteo Stefanini, Marcella Cornia, Silvia Cascianelli, Lorenzo Baraldi, Rita Cucchiara

Describing images in natural language is a fundamental step towards the automatic modeling of connections between the visual and textual modalities.

Image Captioning Knowledge Distillation

From Show to Tell: A Survey on Deep Learning-based Image Captioning

no code implementations14 Jul 2021 Matteo Stefanini, Marcella Cornia, Lorenzo Baraldi, Silvia Cascianelli, Giuseppe Fiameni, Rita Cucchiara

Starting from 2015 the task has generally been addressed with pipelines composed of a visual encoder and a language model for text generation.

Image Captioning Language Modelling +1

Learning to Select: A Fully Attentive Approach for Novel Object Captioning

no code implementations2 Jun 2021 Marco Cagrandi, Marcella Cornia, Matteo Stefanini, Lorenzo Baraldi, Rita Cucchiara

In this paper, we present a novel approach for NOC that learns to select the most relevant objects of an image, regardless of their adherence to the training set, and to constrain the generative process of a language model accordingly.

Image Captioning Language Modelling

A Novel Attention-based Aggregation Function to Combine Vision and Language

no code implementations27 Apr 2020 Matteo Stefanini, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara

The joint understanding of vision and language has been recently gaining a lot of attention in both the Computer Vision and Natural Language Processing communities, with the emergence of tasks such as image captioning, image-text matching, and visual question answering.

General Classification Image Captioning +4

Meshed-Memory Transformer for Image Captioning

2 code implementations CVPR 2020 Marcella Cornia, Matteo Stefanini, Lorenzo Baraldi, Rita Cucchiara

Transformer-based architectures represent the state of the art in sequence modeling tasks like machine translation and language understanding.

Image Captioning Machine Translation +2

Artpedia

no code implementations International Conference on Image Analysis and Processing 2019 Matteo Stefanini, Marcella Cornia, Lorenzo Baraldi, Massimiliano Corsini, and Rita Cucchiara

As vision and language techniques are widely applied to realistic images , there is a growing interest in designing visual-semantic models suitable for more complex and challenging scenarios.

Cross-Modal Retrieval Retrieval

A Deep Learning based approach to VM behavior identification in cloud systems

1 code implementation5 Mar 2019 Matteo Stefanini, Riccardo Lancellotti, Lorenzo Baraldi, Simone Calderara

The experiments compare our proposal with state-of-the-art solutions available in literature, demonstrating that our proposal achieve better performance.

Cloud Computing Clustering +1

Cannot find the paper you are looking for? You can Submit a new open access paper.