Evaluation of automatic video captioning using direct assessment.

Yvette Graham; George Awad; Alan Smeaton

Repository landing page

oai:doaj.org/article:dfb055f536f947ea96873e4f936889a9

Evaluation of automatic video captioning using direct assessment.

Authors: Yvette Graham
George Awad
Alan Smeaton
Publication date: 1 January 2018
Publisher: 'Public Library of Science (PLoS)'
Doi

Abstract

We present Direct Assessment, a method for manually assessing the quality of automatically-generated captions for video. Evaluating the accuracy of video captions is particularly difficult because for any given video clip there is no definitive ground truth or correct answer against which to measure. Metrics for comparing automatic video captions against a manual caption such as BLEU and METEOR, drawn from techniques used in evaluating machine translation, were used in the TRECVid video captioning task in 2016 but these are shown to have weaknesses. The work presented here brings human assessment into the evaluation by crowd sourcing how well a caption describes a video. We automatically degrade the quality of some sample captions which are assessed manually and from this we are able to rate the quality of the human assessors, a factor we take into account in the evaluation. Using data from the TRECVid video-to-text task in 2016, we show how our direct assessment method is replicable and robust and scales to where there are many caption-generation techniques to be evaluated including the TRECVid video-to-text task in 2017

Similar works

Full text

Directory of Open Access Journals

oai:doaj.org/article:dfb055f53...

Last time updated on 03/06/2019

This paper was published in Directory of Open Access Journals.

Having an issue?

Is data on this page outdated, violates copyrights or anything else? Report the problem now and we will take corresponding actions after reviewing your request.