Identifying relevant frames in weakly labeled videos for training concept detectors
Generate an AI Snapshot to get a quick, structured summary of this paper.
A concise AI-generated summary of the paper will appear here once you click Generate AI Snapshot.
TL;DR
A probabilistic framework for learning from weakly annotated training videos in the presence of irrelevant content is presented, and the relevance of keyframes is modeled as a latent random variable that is estimated during training.
Abstract
A key problem with the automatic detection of semantic concepts (like 'interview' or 'soccer') in video streams is the manual acquisition of adequate training sets. Recently, we have proposed to use online videos downloaded from portals like youtube.com for this purpose, whereas tags provided by users during video upload serve as ground truth annotations.
