Sentence-Guided Activity Recognition In Video.pdf

cvpr2014.pdf
Preview of Sentence-Guided Activity Recognition in Video
🔗 Source: cbmm.mit.edu
📊 Size: 453 KB
📄 Pages: 8 pages
⬇️ Downloads: 479

Summary

Sentence-Guided Activity Recognition In Video" presents a system that integrates language understanding with video processing. The system uses the compositional structure of language to guide activity recognition in multi-activity videos. Here's a concise summary:

1. Problem: Current object detection in videos is unreliable, with false positives and negatives. This hinders event recognition, which requires detecting and tracking objects and applying semantic predicates to their tracks.

2. Solution: The authors propose a 'Sentence Tracker' that takes over-generated detections, a sentence, and a lexicon as inputs. It outputs a score and a set of tracks that satisfy the sentence while optimizing a linear combination of detection scores and temporal coherence.

3. Key Features:
- Represents sentence semantics using predicates over tracks, formulated as finite-state recognizers over features extracted from detections.
- Handles nouns, adjectives, verbs, adverbs, and prepositions uniformly using this mechanism.
- Supports three tasks: focus of attention (identifying tracks for different sentences in a video), generation (finding the best sentence to describe a video), and retrieval (searching videos based on a sentential query).

4. Contributions: The paper introduces the Sentence Tracker and demonstrates its utility through these three tasks. It also extends previous work on detection-based tracking by selecting detections that collectively satisfy a complex, sentence-level predicate.

Description

We present a system that integrates video action recognition with language, using sentence-guided activity recognition to interplay with focusing mechanisms in video. The system uses whole-sentence descriptions to guide the activity-recognition process, incorporating roles, characteristics, actions, and spatial relations. This framework enables top-down and bottom-up integration as well as multi-modal integration between vision and language.

Technical Information

  • File Format: PDF
  • File Size: 453 KB
  • Pages: 8
  • Language: EN
  • Total Downloads: 479
  • Last Updated: 2 weeks ago

Document Overview

This PDF document about Sentence-Guided Activity Recognition in Video provides comprehensive information and guidance. Whether you're a beginner or advanced user, this resource offers valuable insights into Sentence-Guided Activity Recognition in Video.

Related Topics

If you're interested in Sentence-Guided Activity Recognition in Video, you might also want to explore:

Download Sentence-Guided Activity Recognition in Video eBooks for free and learn more about Sentence-Guided Activity Recognition in Video. These books contain exercises and tutorials to improve your practical skills, at all levels!

Not satisfied with this document? We have related documents to Sentence-Guided Activity Recognition in Video, try searching with similar keywords: Sentence-Guided Activity Recognition in Video, Guided Activity Chapter 8 Summary Activity The Pre, Guided Activity Reteaching Activity The Federal System, Application For Recognition Renewal Of Recognition, Board Recognition - HS Seniors - June 17 Recognition (pdf), Sentence Recognition, Sentence Recognition Practice, Sentence Recognition Worksheets

You can download PDF versions of the user's guide, manuals and ebooks about Sentence-Guided Activity Recognition in Video, you can also find and download for free A free online manual (notices) with beginner and intermediate, Downloads Documentation, You can download PDF files (or DOC and PPT) about Sentence-Guided Activity Recognition in Video for free, but please respect copyrighted ebooks.