Automatic Comic Generation with Stylistic Multi-page Layouts and Emotion-driven Text Balloon Generation

01/26/2021
by   Xin Yang, et al.
0

In this paper, we propose a fully automatic system for generating comic books from videos without any human intervention. Given an input video along with its subtitles, our approach first extracts informative keyframes by analyzing the subtitles, and stylizes keyframes into comic-style images. Then, we propose a novel automatic multi-page layout framework, which can allocate the images across multiple pages and synthesize visually interesting layouts based on the rich semantics of the images (e.g., importance and inter-image relation). Finally, as opposed to using the same type of balloon as in previous works, we propose an emotion-aware balloon generation method to create different types of word balloons by analyzing the emotion of subtitles and audios. Our method is able to vary balloon shapes and word sizes in balloons in response to different emotions, leading to more enriched reading experience. Once the balloons are generated, they are placed adjacent to their corresponding speakers via speaker detection. Our results show that our method, without requiring any user inputs, can generate high-quality comic pages with visually rich layouts and balloons. Our user studies also demonstrate that users prefer our generated results over those by state-of-the-art comic generation systems.

READ FULL TEXT

page 5

page 6

page 7

page 8

page 9

page 11

page 13

page 15

research
05/07/2022

Empathetic Response Generation with State Management

The goal of empathetic response generation is to enhance the ability of ...
research
11/07/2021

Emotional Prosody Control for Speech Generation

Machine-generated speech is characterized by its limited or unnatural em...
research
10/04/2020

MIME: MIMicking Emotions for Empathetic Response Generation

Current approaches to empathetic response generation view the set of emo...
research
04/25/2021

3D-TalkEmo: Learning to Synthesize 3D Emotional Talking Head

Impressive progress has been made in audio-driven 3D facial animation re...
research
08/02/2023

AutoPoster: A Highly Automatic and Content-aware Design System for Advertising Poster Generation

Advertising posters, a form of information presentation, combine visual ...
research
09/02/2020

Automatic cinematography for 360 video

We describe our method for automatic generation of a visually interestin...
research
02/15/2022

ViNTER: Image Narrative Generation with Emotion-Arc-Aware Transformer

Image narrative generation describes the creation of stories regarding t...

Please sign up or login with your details

Forgot password? Click here to reset