Abstract
We address the task of unsupervised domain adaptation (UDA) for videos with self-supervised learning. While UDA for images is a widely studied problem, UDA for videos is relatively unexplored. In this paper, we propose a novel self-supervised loss for the task of video UDA. The method is motivated by inverted reasoning. Many works on video classification have shown success with representations based on events in videos, e.g., 'reaching', 'picking', and 'drinking' events for 'drinking coffee'. We argue that if we have event-based representations, we should be able to predict the relative distances between clips in videos. Inverting that, we propose a self-supervised task to predict the difference of the distance between two clips from the source video and the distance between two clips from the target video. We hope that such a task would encourage learning event-based representations of the videos, which is known to be beneficial for classification. Since we predict the difference of clip distances between clips from source videos and target videos, we 'tie' the two domains and expect to achieve well-adapted representations. We combine this purely self-supervised loss and the source classification loss to learn the model parameters. We give extensive empirical results on challenging video UDA benchmarks, i.e., UCF-HMDB and EPIC-Kitchens. The presented qualitative and quantitative results support our motivations and method.
| Original language | English |
|---|---|
| Title of host publication | 2022 26th International Conference on Pattern Recognition, ICPR 2022 |
| Publisher | Institute of Electrical and Electronics Engineers Inc. |
| Pages | 3464-3470 |
| Number of pages | 7 |
| ISBN (Electronic) | 9781665490627 |
| DOIs | |
| Publication status | Published - 2022 |
| Event | 26th International Conference on Pattern Recognition, ICPR 2022 - Montreal, Canada Duration: 21 Aug 2022 → 25 Aug 2022 |
Publication series
| Name | Proceedings - International Conference on Pattern Recognition |
|---|---|
| Volume | 2022-August |
| ISSN (Print) | 1051-4651 |
Conference
| Conference | 26th International Conference on Pattern Recognition, ICPR 2022 |
|---|---|
| Country/Territory | Canada |
| City | Montreal |
| Period | 21/08/22 → 25/08/22 |
Bibliographical note
Publisher Copyright:© 2022 IEEE.
Fingerprint
Dive into the research topics of 'Self-Supervised Cross-Video Temporal Learning for Unsupervised Video Domain Adaptation'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver