Horizontal-to-Vertical Video Conversion

Zhu, Tun; Zhang, Daoxin; Hu, Yao; Wang, Tianran; Jiang, Xiaolong; Zhu, Jianke; Li, Jiawei

Computer Science > Computer Vision and Pattern Recognition

arXiv:2101.04051 (cs)

[Submitted on 11 Jan 2021 (v1), last revised 23 Jun 2021 (this version, v2)]

Title:Horizontal-to-Vertical Video Conversion

Authors:Tun Zhu, Daoxin Zhang, Yao Hu, Tianran Wang, Xiaolong Jiang, Jianke Zhu, Jiawei Li

View PDF

Abstract:Alongside the prevalence of mobile videos, the general public leans towards consuming vertical videos on hand-held devices. To revitalize the exposure of horizontal contents, we hereby set forth the exploration of automated horizontal-to-vertical (abbreviated as H2V) video conversion with our proposed H2V framework, accompanied by an accurately annotated H2V-142K dataset. Concretely, H2V framework integrates video shot boundary detection, subject selection and multi-object tracking to facilitate the subject-preserving conversion, wherein the key is subject selection. To achieve so, we propose a Rank-SS module that detects human objects, then selects the subject-to-preserve via exploiting location, appearance, and salient cues. Afterward, the framework automatically crops the video around the subject to produce vertical contents from horizontal sources. To build and evaluate our H2V framework, H2V-142K dataset is densely annotated with subject bounding boxes for 125 videos with 132K frames and 9,500 video covers, upon which we demonstrate superior subject selection performance comparing to traditional salient approaches, and exhibit promising horizontal-to-vertical conversion performance overall. By publicizing this dataset as well as our approach, we wish to pave the way for more valuable endeavors on the horizontal-to-vertical video conversion task.

Comments:	Accept by IEEE Transactions on Multimedia
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2101.04051 [cs.CV]
	(or arXiv:2101.04051v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2101.04051

Submission history

From: Tun Zhu [view email]
[v1] Mon, 11 Jan 2021 17:37:31 UTC (12,221 KB)
[v2] Wed, 23 Jun 2021 15:37:45 UTC (12,215 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Horizontal-to-Vertical Video Conversion

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Horizontal-to-Vertical Video Conversion

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators