Automatic Comic Generation with Stylistic Multi-page Layouts and Emotion-driven Text Balloon Generation

Research output: Journal Publications and Reviews (RGC: 21, 22, 62)21_Publication in refereed journalpeer-review

View graph of relations

Author(s)

  • Xin YANG
  • Zongliang MA
  • Letian YU
  • Ying CAO
  • Baocai YIN
  • Xiaopeng WEI
  • Qiang ZHANG

Related Research Unit(s)

Detail(s)

Original languageEnglish
Article number55
Journal / PublicationACM Transactions on Multimedia Computing, Communications and Applications
Volume17
Issue number2
Online published28 May 2021
Publication statusPublished - Jun 2021

Abstract

In this article, we propose a fully automatic system for generating comic books from videos without any human intervention. Given an input video along with its subtitles, our approach first extracts informative keyframes by analyzing the subtitles and stylizes keyframes into comic-style images. Then, we propose a novel automatic multi-page layout framework that can allocate the images across multiple pages and synthesize visually interesting layouts based on the rich semantics of the images (e.g., importance and inter-image relation). Finally, as opposed to using the same type of balloon as in previous works, we propose an emotion-aware balloon generation method to create different types of word balloons by analyzing the emotion of subtitles and audio. Our method is able to vary balloon shapes and word sizes in balloons in response to different emotions, leading to more enriched reading experience. Once the balloons are generated, they are placed adjacent to their corresponding speakers via speaker detection. Our results show that our method, without requiring any user inputs, can generate high-quality comic pages with visually rich layouts and balloons. Our user studies also demonstrate that users prefer our generated results over those by state-of-the-art comic generation systems.

Research Area(s)

  • Automatic, comic books, keyframes, layout, multi-page, stylizing

Bibliographic Note

Full text of this publication does not contain sufficient affiliation information. With consent from the author(s) concerned, the Research Unit(s) information for this record is based on the existing academic department affiliation of the author(s).

Citation Format(s)

Automatic Comic Generation with Stylistic Multi-page Layouts and Emotion-driven Text Balloon Generation. / YANG, Xin; MA, Zongliang; YU, Letian; CAO, Ying; YIN, Baocai; WEI, Xiaopeng; ZHANG, Qiang; LAU, Rynson W. H.

In: ACM Transactions on Multimedia Computing, Communications and Applications, Vol. 17, No. 2, 55, 06.2021.

Research output: Journal Publications and Reviews (RGC: 21, 22, 62)21_Publication in refereed journalpeer-review