Semantic Embedding Space for Zero-Shot Action Recognition

Xu, Xun; Hospedales, Timothy; Gong, Shaogang

Computer Science > Computer Vision and Pattern Recognition

arXiv:1502.01540 (cs)

[Submitted on 5 Feb 2015]

Title:Semantic Embedding Space for Zero-Shot Action Recognition

Authors:Xun Xu, Timothy Hospedales, Shaogang Gong

View PDF

Abstract:The number of categories for action recognition is growing rapidly. It is thus becoming increasingly hard to collect sufficient training data to learn conventional models for each category. This issue may be ameliorated by the increasingly popular 'zero-shot learning' (ZSL) paradigm. In this framework a mapping is constructed between visual features and a human interpretable semantic description of each category, allowing categories to be recognised in the absence of any training data. Existing ZSL studies focus primarily on image data, and attribute-based semantic representations. In this paper, we address zero-shot recognition in contemporary video action recognition tasks, using semantic word vector space as the common space to embed videos and category labels. This is more challenging because the mapping between the semantic space and space-time features of videos containing complex actions is more complex and harder to learn. We demonstrate that a simple self-training and data augmentation strategy can significantly improve the efficacy of this mapping. Experiments on human action datasets including HMDB51 and UCF101 demonstrate that our approach achieves the state-of-the-art zero-shot action recognition performance.

Comments:	5 pages
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:1502.01540 [cs.CV]
	(or arXiv:1502.01540v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.1502.01540

Submission history

From: Xun Xu [view email]
[v1] Thu, 5 Feb 2015 13:34:48 UTC (492 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.CV

< prev | next >

new | recent | 2015-02

Change to browse by:

References & Citations

DBLP - CS Bibliography

listing | bibtex

Xun Xu
Timothy M. Hospedales
Shaogang Gong

export BibTeX citation

Computer Science > Computer Vision and Pattern Recognition

Title:Semantic Embedding Space for Zero-Shot Action Recognition

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Semantic Embedding Space for Zero-Shot Action Recognition

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators