Papers
Topics
Authors
Recent
Gemini 2.5 Flash
Gemini 2.5 Flash
119 tokens/sec
GPT-4o
56 tokens/sec
Gemini 2.5 Pro Pro
43 tokens/sec
o3 Pro
6 tokens/sec
GPT-4.1 Pro
47 tokens/sec
DeepSeek R1 via Azure Pro
28 tokens/sec
2000 character limit reached

Inferring Temporal Compositions of Actions Using Probabilistic Automata (2004.13217v1)

Published 28 Apr 2020 in cs.CV

Abstract: This paper presents a framework to recognize temporal compositions of atomic actions in videos. Specifically, we propose to express temporal compositions of actions as semantic regular expressions and derive an inference framework using probabilistic automata to recognize complex actions as satisfying these expressions on the input video features. Our approach is different from existing works that either predict long-range complex activities as unordered sets of atomic actions, or retrieve videos using natural language sentences. Instead, the proposed approach allows recognizing complex fine-grained activities using only pretrained action classifiers, without requiring any additional data, annotations or neural network training. To evaluate the potential of our approach, we provide experiments on synthetic datasets and challenging real action recognition datasets, such as MultiTHUMOS and Charades. We conclude that the proposed approach can extend state-of-the-art primitive action classifiers to vastly more complex activities without large performance degradation.

User Edit Pencil Streamline Icon: https://streamlinehq.com
Authors (5)
  1. Rodrigo Santa Cruz (14 papers)
  2. Anoop Cherian (65 papers)
  3. Basura Fernando (60 papers)
  4. Dylan Campbell (44 papers)
  5. Stephen Gould (104 papers)
Citations (2)

Summary

We haven't generated a summary for this paper yet.