• Home
  • Search
  • Intentonomy: a Dataset and Study towards Human Intent Understanding
  • Open Access IconOpen Access
  • Cite Icon33
  • https://doi.org/10.1109/cvpr46437.2021.01279Copy DOI Icon

Intentonomy: a Dataset and Study towards Human Intent Understanding

  • Jun 1, 2021
  • Menglin Jia +5 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

An image is worth a thousand words, conveying information that goes beyond the mere visual content therein. In this paper, we study the intent behind social media images with an aim to analyze how visual information can facilitate recognition of human intent. Towards this goal, we introduce an intent dataset, Intentonomy, comprising 14K images covering a wide range of everyday scenes. These images are manually annotated with 28 intent categories derived from a social psychology taxonomy. We then systematically study whether, and to what extent, commonly used visual information, i.e., object and context, contribute to human motive understanding. Based on our findings, we conduct further study to quantify the effect of attending to object and context classes as well as textual information in the form of hashtags when training an intent classifier. Our results quantitatively and qualitatively shed light on how visual and textual information can produce observable effects when predicting intent. <sup xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">1</sup>

Similar Papers
  • Research Article
  • Citations9

CE-DCVSI: Multimodal relational extraction based on collaborative enhancement of dual-channel visual semantic information

  • Nov 04, 2024
  • Expert Systems With Applications
  • Yunchao Gong +7
  • Conference Article

Effective multiple feature fusion using topic model for social image visualization

  • Oct 01, 2014
  • Kouhei Tateno +2
  • Research Article

Optimized Hypergraph Based Social Image Search Using VisualTextual Joint Relevance Learning

  • Jan 01, 2014
  • IOSR Journal of Computer Engineering
  • Arya S
  • Research Article
  • Citations23

A comprehensive review of the video-to-text problem

  • Jan 16, 2022
  • Artificial Intelligence Review
  • Jesus Perez-Martin +5
  • Research Article

Order Matters: The Marginal Value of Visual and Textual Information in Sequential Product Information Presentation

  • Jan 30, 2026
  • International Journal of Human–Computer Interaction
  • Xiaohan Hu +1
  • Book Chapter
  • Citations59

Multimodal Fake News Detection with Textual, Visual and Semantic Information

  • Jan 01, 2020
  • Anastasia Giachanou +2
  • Research Article
  • Citations15

Evaluating multimodal relevance feedback techniques for medical image retrieval

  • Aug 01, 2015
  • Information Retrieval Journal
  • Dimitrios Markonis +2
  • Research Article
  • Citations3

An investigation of the visual advantage effect on objective decision quality in choice tasks

  • Jun 02, 2025
  • Journal of Global Scholars of Marketing Science
  • Sereikhuoch Eng
  • Research Article
  • Citations70

Sentiment-aware multimodal pre-training for multimodal sentiment analysis

  • Oct 19, 2022
  • Knowledge-Based Systems
  • Junjie Ye +7
  • Research Article
  • Citations5

Edge data based trailer inception probabilistic matrix factorization for context-aware movie recommendation

  • Dec 08, 2021
  • World Wide Web
  • Honglong Chen +7
  • Book Chapter

Sentiment Analysis of Images with Tensor Factorization

  • Jan 01, 2019
  • Ayumu Sakaguchi +1
  • Conference Article
  • Citations19

Medical image retrieval based on visual contents and text information

  • Oct 10, 2004
  • Hong Shao +2
  • Conference Article

Research on multimodal intent recognition model based on cross-modal attention fusion

  • Dec 18, 2025
  • Hao Qu +2
  • Research Article
  • Citations19

Socially meaningful visual context either enhances or inhibits vocalisation processing in the macaque brain

  • Aug 19, 2022
  • Nature Communications
  • Mathilda Froesel +5
  • Book Chapter
  • Citations2

Adaptive Metadata Generation for Integration of Visual and Semantic Information

  • Jul 10, 2015
  • Hideyasu Sasaki +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.