Concept Drift Detection for Multivariate Data Streams and Temporal Segmentation of Daylong Egocentric Videos
Pravin Nagar, Mansi Khemka, Chetan Arora
Abstract
The long and unconstrained nature of egocentric videos makes it imperative to use temporal segmentation as an important pre-processing step for many higher-level inference tasks. Activities of the wearer in an egocentric video typically span over hours and are often separated by slow, gradual changes. Furthermore, the change of camera viewpoint due to the wearer's head motion causes frequent and extreme, but, spurious scene changes. The continuous nature of boundaries makes it difficult to apply traditional Markov Random Field (MRF) pipelines relying on temporal discontinuity, whereas deep Long Short Term Memory (LSTM) networks gather context only upto a few hundred frames, rendering them ineffective for egocentric videos. In this paper, we present a novel unsupervised temporal segmentation technique especially suited for day-long egocentric videos. We formulate the problem as detecting concept drift in a time-varying, non i.i.d. sequence of frames. Statistically bounded thresholds are calculated to detect concept drift between two temporally adjacent multivariate data segments with different underlying distributions while establishing guarantees on false positives. Since the derived threshold indicates confidence in the prediction, it can also be used to control the granularity of the output segmentation. Using our technique, we report significantly improved state of the art f-measure for daylong egocentric video datasets, as well as photostream datasets derived from them: HUJI (73.01%, 59.44%), UTEgo (58.41%, 60.61%) and Disney (67.63%, 68.83%).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6cfe5e30-b6ac-4fd0-a0df-9ab2e6443775Cited by top-tier papers2
- Anonymizing Egocentric VideosDaksh Thapar, Aditya Nigam, Chetan AroraICCV 2021 · 9 citations
- Merry Go Round: Rotate a Frame and Fool a DNNDaksh Thapar, Aditya Nigam, Chetan AroraCVPR 2022 · 1 citation
Builds on1
Related papers
- Online Anomaly Detection over Live Social Video StreamingChengkun He, Xiangmin Zhou, Chen Wang, Iqbal Gondal et al.ICDE 2024 · 2 citations
- From ViT Features to Training-free Video Object Segmentation via Streaming-data Mixture ModelsRoy Uziel, Or Dinari, Oren FreifeldNeurIPS 2023 · 6 citations
- A Large-Scale Study on Unsupervised Spatiotemporal Representation LearningChristoph Feichtenhofer, Haoqi Fan, Bo Xiong, Ross B. Girshick et al.CVPR 2021
- Ego-Grounding for Personalized Question-Answering in Egocentric VideosJunbin Xiao, Shenglang Zhang, Pengxiang Zhu, Angela YaoCVPR 2026 · 7 citations
- Anchor Diffusion for Unsupervised Video Object SegmentationZhao Yang, Qiang Wang, Luca Bertinetto, Song Bai et al.ICCV 2019 · 127 citations
