Analyzing Image Beyond Visual Aspect: Image Emotion Classification via Multiple-Affective Captioning
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhou, Zibo, Zhai, Zhengjun, Chen, Huimin, Dai, Wei, Yang, Hansen |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Evaluating Remote Sensing Image Captions Beyond Metric Biases
por: Chen, Ziyun, et al.
Publicado: (2026)
por: Chen, Ziyun, et al.
Publicado: (2026)
Emotion-Director: Bridging Affective Shortcut in Emotion-Oriented Image Generation
por: Jia, Guoli, et al.
Publicado: (2025)
por: Jia, Guoli, et al.
Publicado: (2025)
GRACE: Estimating Geometry-level 3D Human-Scene Contact from 2D Images
por: Wang, Chengfeng, et al.
Publicado: (2025)
por: Wang, Chengfeng, et al.
Publicado: (2025)
CaptionQA: Is Your Caption as Useful as the Image Itself?
por: Yang, Shijia, et al.
Publicado: (2025)
por: Yang, Shijia, et al.
Publicado: (2025)
Affective Image Editing: Shaping Emotional Factors via Text Descriptions
por: Zhang, Peixuan, et al.
Publicado: (2025)
por: Zhang, Peixuan, et al.
Publicado: (2025)
CapArena: Benchmarking and Analyzing Detailed Image Captioning in the LLM Era
por: Cheng, Kanzhi, et al.
Publicado: (2025)
por: Cheng, Kanzhi, et al.
Publicado: (2025)
ITIScore: An Image-to-Text-to-Image Rating Framework for the Image Captioning Ability of MLLMs
por: Xu, Zitong, et al.
Publicado: (2026)
por: Xu, Zitong, et al.
Publicado: (2026)
Degradation-Aware Image Enhancement via Vision-Language Classification
por: Cai, Jie, et al.
Publicado: (2025)
por: Cai, Jie, et al.
Publicado: (2025)
Benchmarking Large Vision-Language Models via Directed Scene Graph for Comprehensive Image Captioning
por: Lu, Fan, et al.
Publicado: (2024)
por: Lu, Fan, et al.
Publicado: (2024)
From Image Captioning to Visual Storytelling
por: Passadakis, Admitos, et al.
Publicado: (2025)
por: Passadakis, Admitos, et al.
Publicado: (2025)
Image Captioning via Dynamic Path Customization
por: Ma, Yiwei, et al.
Publicado: (2024)
por: Ma, Yiwei, et al.
Publicado: (2024)
Bridging Visual Affective Gap: Borrowing Textual Knowledge by Learning from Noisy Image-Text Pairs
por: Wu, Daiqing, et al.
Publicado: (2025)
por: Wu, Daiqing, et al.
Publicado: (2025)
Where Do Images Come From? Analyzing Captions to Geographically Profile Datasets
por: Basu, Abhipsa, et al.
Publicado: (2026)
por: Basu, Abhipsa, et al.
Publicado: (2026)
Analyzing Transformer Models and Knowledge Distillation Approaches for Image Captioning on Edge AI
por: Kwok, Wing Man Casca, et al.
Publicado: (2025)
por: Kwok, Wing Man Casca, et al.
Publicado: (2025)
Visually-Aware Context Modeling for News Image Captioning
por: Qu, Tingyu, et al.
Publicado: (2023)
por: Qu, Tingyu, et al.
Publicado: (2023)
Towards Deeper Emotional Reflection: Crafting Affective Image Filters with Generative Priors
por: Zhang, Peixuan, et al.
Publicado: (2025)
por: Zhang, Peixuan, et al.
Publicado: (2025)
Exploring Diverse In-Context Configurations for Image Captioning
por: Yang, Xu, et al.
Publicado: (2023)
por: Yang, Xu, et al.
Publicado: (2023)
EMoTive: Event-guided Trajectory Modeling for 3D Motion Estimation
por: Wan, Zengyu, et al.
Publicado: (2025)
por: Wan, Zengyu, et al.
Publicado: (2025)
Evaluating Image Caption via Cycle-consistent Text-to-Image Generation
por: Cui, Tianyu, et al.
Publicado: (2025)
por: Cui, Tianyu, et al.
Publicado: (2025)
Image Captioning via Compact Bidirectional Architecture
por: Song, Zijie, et al.
Publicado: (2022)
por: Song, Zijie, et al.
Publicado: (2022)
Towards LLM-centric Affective Visual Customization via Efficient and Precise Emotion Manipulating
por: Luo, Jiamin, et al.
Publicado: (2026)
por: Luo, Jiamin, et al.
Publicado: (2026)
DMS-Net:Dual-Modal Multi-Scale Siamese Network for Binocular Fundus Image Classification
por: Huo, Guohao, et al.
Publicado: (2025)
por: Huo, Guohao, et al.
Publicado: (2025)
Recognizing Multiple Ingredients in Food Images Using a Single-Ingredient Classification Model
por: Fu, Kun, et al.
Publicado: (2024)
por: Fu, Kun, et al.
Publicado: (2024)
TopoImages: Incorporating Local Topology Encoding into Deep Learning Models for Medical Image Classification
por: Gu, Pengfei, et al.
Publicado: (2025)
por: Gu, Pengfei, et al.
Publicado: (2025)
SC-Captioner: Improving Image Captioning with Self-Correction by Reinforcement Learning
por: Zhang, Lin, et al.
Publicado: (2025)
por: Zhang, Lin, et al.
Publicado: (2025)
Dual-Stream Collaborative Transformer for Image Captioning
por: Wan, Jun, et al.
Publicado: (2026)
por: Wan, Jun, et al.
Publicado: (2026)
New Encoder Learning for Captioning Heavy Rain Images via Semantic Visual Feature Matching
por: Son, Chang-Hwan, et al.
Publicado: (2021)
por: Son, Chang-Hwan, et al.
Publicado: (2021)
Exploiting Multiple Sequence Lengths in Fast End to End Training for Image Captioning
por: Hu, Jia Cheng, et al.
Publicado: (2022)
por: Hu, Jia Cheng, et al.
Publicado: (2022)
DreamLIP: Language-Image Pre-training with Long Captions
por: Zheng, Kecheng, et al.
Publicado: (2024)
por: Zheng, Kecheng, et al.
Publicado: (2024)
What Makes for Good Image Captions?
por: Chen, Delong, et al.
Publicado: (2024)
por: Chen, Delong, et al.
Publicado: (2024)
VIXEN: Visual Text Comparison Network for Image Difference Captioning
por: Black, Alexander, et al.
Publicado: (2024)
por: Black, Alexander, et al.
Publicado: (2024)
PathM3: A Multimodal Multi-Task Multiple Instance Learning Framework for Whole Slide Image Classification and Captioning
por: Zhou, Qifeng, et al.
Publicado: (2024)
por: Zhou, Qifeng, et al.
Publicado: (2024)
BACON: Improving Clarity of Image Captions via Bag-of-Concept Graphs
por: Yang, Zhantao, et al.
Publicado: (2024)
por: Yang, Zhantao, et al.
Publicado: (2024)
Fourier Transform Multiple Instance Learning for Whole Slide Image Classification
por: Bilic, Anthony, et al.
Publicado: (2025)
por: Bilic, Anthony, et al.
Publicado: (2025)
CaptionSmiths: Flexibly Controlling Language Pattern in Image Captioning
por: Saito, Kuniaki, et al.
Publicado: (2025)
por: Saito, Kuniaki, et al.
Publicado: (2025)
Few-shot Image Generation via Masked Discrimination
por: Zhu, Jingyuan, et al.
Publicado: (2022)
por: Zhu, Jingyuan, et al.
Publicado: (2022)
Low-Light Image Enhancement via Generative Perceptual Priors
por: Zhou, Han, et al.
Publicado: (2024)
por: Zhou, Han, et al.
Publicado: (2024)
MatE: Material Extraction from Single-Image via Geometric Prior
por: Zhang, Zeyu, et al.
Publicado: (2025)
por: Zhang, Zeyu, et al.
Publicado: (2025)
DEVICE: Depth and Visual Concepts Aware Transformer for OCR-based Image Captioning
por: Xu, Dongsheng, et al.
Publicado: (2023)
por: Xu, Dongsheng, et al.
Publicado: (2023)
Image Generation from Image Captioning -- Invertible Approach
por: Menon, Nandakishore S, et al.
Publicado: (2024)
por: Menon, Nandakishore S, et al.
Publicado: (2024)
Ejemplares similares
-
Evaluating Remote Sensing Image Captions Beyond Metric Biases
por: Chen, Ziyun, et al.
Publicado: (2026) -
Emotion-Director: Bridging Affective Shortcut in Emotion-Oriented Image Generation
por: Jia, Guoli, et al.
Publicado: (2025) -
GRACE: Estimating Geometry-level 3D Human-Scene Contact from 2D Images
por: Wang, Chengfeng, et al.
Publicado: (2025) -
CaptionQA: Is Your Caption as Useful as the Image Itself?
por: Yang, Shijia, et al.
Publicado: (2025) -
Affective Image Editing: Shaping Emotional Factors via Text Descriptions
por: Zhang, Peixuan, et al.
Publicado: (2025)