TasksSotADatasetsPapersMethodsSubmitAbout
Papers With Code 2

A community resource for machine learning research: papers, code, benchmarks, and state-of-the-art results.

Explore

Notable BenchmarksAll SotADatasetsPapersMethods

Community

Submit ResultsAbout

Data sourced from the PWC Archive (CC-BY-SA 4.0). Built by the community, for the community.

Papers/All About Knowledge Graphs for Actions

All About Knowledge Graphs for Actions

Pallabi Ghosh, Nirat Saini, Larry S. Davis, Abhinav Shrivastava

2020-08-28Knowledge GraphsFew-Shot LearningObject RecognitionTransfer LearningZero-Shot Action RecognitionFew Shot Action RecognitionAllAction Recognition
PaperPDF

Abstract

Current action recognition systems require large amounts of training data for recognizing an action. Recent works have explored the paradigm of zero-shot and few-shot learning to learn classifiers for unseen categories or categories with few labels. Following similar paradigms in object recognition, these approaches utilize external sources of knowledge (eg. knowledge graphs from language domains). However, unlike objects, it is unclear what is the best knowledge representation for actions. In this paper, we intend to gain a better understanding of knowledge graphs (KGs) that can be utilized for zero-shot and few-shot action recognition. In particular, we study three different construction mechanisms for KGs: action embeddings, action-object embeddings, visual embeddings. We present extensive analysis of the impact of different KGs in different experimental setups. Finally, to enable a systematic study of zero-shot and few-shot approaches, we propose an improved evaluation paradigm based on UCF101, HMDB51, and Charades datasets for knowledge transfer from models trained on Kinetics.

Results

TaskDatasetMetricValueModel
Zero-Shot Action RecognitionKineticsTop-1 Accuracy22.3GCN
Zero-Shot Action RecognitionKineticsTop-5 Accuracy49.7GCN

Related Papers

RaMen: Multi-Strategy Multi-Modal Learning for Bundle Construction2025-07-18SMART: Relation-Aware Learning of Geometric Representations for Knowledge Graphs2025-07-17GLAD: Generalizable Tuning for Vision-Language Models2025-07-17Disentangling coincident cell events using deep transfer learning and compressive sensing2025-07-17A Real-Time System for Egocentric Hand-Object Interaction Detection in Industrial Domains2025-07-17Best Practices for Large-Scale, Pixel-Wise Crop Mapping and Transfer Learning Workflows2025-07-16Robust-Multi-Task Gradient Boosting2025-07-15Modeling Code: Is Text All You Need?2025-07-15