Abductive Action Inference

Tan, Clement; Yeo, Chai Kiat; Tan, Cheston; Fernando, Basura

Computer Science > Computer Vision and Pattern Recognition

arXiv:2210.13984v4 (cs)

[Submitted on 24 Oct 2022 (v1), last revised 7 Aug 2023 (this version, v4)]

Title:Abductive Action Inference

Authors:Clement Tan, Chai Kiat Yeo, Cheston Tan, Basura Fernando

View PDF

Abstract:Abductive reasoning aims to make the most likely inference for a given set of incomplete observations. In this paper, we introduce a novel research task known as "abductive action inference" which addresses the question of which actions were executed by a human to reach a specific state shown in a single snapshot. The research explores three key abductive inference problems: action set prediction, action sequence prediction, and abductive action verification. To tackle these challenging tasks, we investigate various models, including established ones such as Transformers, Graph Neural Networks, CLIP, BLIP, GPT3, end-to-end trained Slow-Fast, Resnet50-3D, and ViT models. Furthermore, the paper introduces several innovative models tailored for abductive action inference, including a relational graph neural network, a relational bilinear pooling model, a relational rule-based inference model, a relational GPT-3 prompt method, and a relational Transformer model. Notably, the newly proposed object-relational bilinear graph encoder-decoder (BiGED) model emerges as the most effective among all methods evaluated, demonstrating good proficiency in handling the intricacies of the Action Genome dataset. The contributions of this research offer significant progress toward comprehending the implications of human actions and making highly plausible inferences concerning the outcomes of these actions.

Comments:	16 pages, 9 figures
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2210.13984 [cs.CV]
	(or arXiv:2210.13984v4 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2210.13984

Submission history

From: Clement Tan [view email]
[v1] Mon, 24 Oct 2022 07:43:59 UTC (2,019 KB)
[v2] Mon, 3 Apr 2023 10:28:38 UTC (13,194 KB)
[v3] Mon, 17 Apr 2023 07:08:12 UTC (13,195 KB)
[v4] Mon, 7 Aug 2023 11:29:26 UTC (14,553 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Abductive Action Inference

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Abductive Action Inference

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators