Visual Storytelling

NAACL 2016 Ting-HaoHuangFrancis FerraroNasrin MostafazadehIshan MisraAishwarya AgrawalJacob DevlinRoss GirshickXiaodong HePushmeet KohliDhruv BatraC. Lawrence ZitnickDevi ParikhLucy VanderwendeMichel GalleyMargaret Mitchell

We introduce the first dataset for sequential vision-to-language, and explore how this data may be used for the task of visual storytelling. The first release of this dataset, SIND v.1, includes 81,743 unique photos in 20,211 sequences, aligned to both descriptive (caption) and story language... (read more)

PDF Abstract

Results from the Paper


  Submit results from this paper to get state-of-the-art GitHub badges and help the community compare results to other papers.