Seeing the Undefined: Chain-of-Action for Generative Semantic Labels
Fuente:
arXiv
Saved in:
| Main Authors: | Wei, Meng, Li, Zhongnian, Ying, Peng, Xu, Xinzheng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning from True-False Labels via Multi-modal Prompt Retrieving
by: Li, Zhongnian, et al.
Published: (2024)
by: Li, Zhongnian, et al.
Published: (2024)
Learning from Reduced Labels for Long-Tailed Data
by: Wei, Meng, et al.
Published: (2024)
by: Wei, Meng, et al.
Published: (2024)
Beyond Label Semantics: Language-Guided Action Anatomy for Few-shot Action Recognition
by: Qian, Zefeng, et al.
Published: (2025)
by: Qian, Zefeng, et al.
Published: (2025)
Energy Score-based Pseudo-Label Filtering and Adaptive Loss for Imbalanced Semi-supervised SAR target recognition
by: Zhang, Xinzheng, et al.
Published: (2024)
by: Zhang, Xinzheng, et al.
Published: (2024)
Seeing Through Fog: Towards Fog-Invariant Action Recognition
by: Liu, Enqi, et al.
Published: (2026)
by: Liu, Enqi, et al.
Published: (2026)
Seeing Space and Motion: Enhancing Latent Actions with Geometric and Dynamic Awareness for Vision-Language-Action Models
by: Cai, Zhejia, et al.
Published: (2025)
by: Cai, Zhejia, et al.
Published: (2025)
Seeing the Poem: Image-Semantic Detection of AI-Generated Modern Chinese Poetry with MLLMs
by: Wang, Shanshan, et al.
Published: (2026)
by: Wang, Shanshan, et al.
Published: (2026)
ChainHOI: Joint-based Kinematic Chain Modeling for Human-Object Interaction Generation
by: Zeng, Ling-An, et al.
Published: (2025)
by: Zeng, Ling-An, et al.
Published: (2025)
Towards Completeness: A Generalizable Action Proposal Generator for Zero-Shot Temporal Action Localization
by: Du, Jia-Run, et al.
Published: (2024)
by: Du, Jia-Run, et al.
Published: (2024)
Learning Semantic-Aware Threshold for Multi-Label Image Recognition with Partial Labels
by: Ruan, Haoxian, et al.
Published: (2025)
by: Ruan, Haoxian, et al.
Published: (2025)
Capturing Rich Behavior Representations: A Dynamic Action Semantic-Aware Graph Transformer for Video Captioning
by: Liu, Caihua, et al.
Published: (2025)
by: Liu, Caihua, et al.
Published: (2025)
Seeing the Unseen: Visual Common Sense for Semantic Placement
by: Ramrakhya, Ram, et al.
Published: (2024)
by: Ramrakhya, Ram, et al.
Published: (2024)
Towards Micro-Action Recognition with Limited Annotations: An Asynchronous Pseudo Labeling and Training Approach
by: Zhang, Yan, et al.
Published: (2025)
by: Zhang, Yan, et al.
Published: (2025)
See&Say: Vision Language Guided Safe Zone Detection for Autonomous Package Delivery Drones
by: Ghazanfari, Mahyar, et al.
Published: (2026)
by: Ghazanfari, Mahyar, et al.
Published: (2026)
Learning Semantic-Aware Representation in Visual-Language Models for Multi-Label Recognition with Partial Labels
by: Ruan, Haoxian, et al.
Published: (2024)
by: Ruan, Haoxian, et al.
Published: (2024)
Hard to See, Hard to Label: Generative and Symbolic Acquisition for Subtle Visual Phenomena
by: Prasad, Renjith, et al.
Published: (2026)
by: Prasad, Renjith, et al.
Published: (2026)
See No Evil: Semantic Context-Aware Privacy Risk Detection for AR
by: Liu, Jialu, et al.
Published: (2026)
by: Liu, Jialu, et al.
Published: (2026)
SeeSR: Towards Semantics-Aware Real-World Image Super-Resolution
by: Wu, Rongyuan, et al.
Published: (2023)
by: Wu, Rongyuan, et al.
Published: (2023)
UniTriGen: Unified Triplet Generation of Aligned Visible-Infrared-Label for Few-Shot RGB-T Semantic Segmentation
by: Zhou, Ping, et al.
Published: (2026)
by: Zhou, Ping, et al.
Published: (2026)
OccLE: Label-Efficient 3D Semantic Occupancy Prediction
by: Fang, Naiyu, et al.
Published: (2025)
by: Fang, Naiyu, et al.
Published: (2025)
Using Unreliable Pseudo-Labels for Label-Efficient Semantic Segmentation
by: Wang, Haochen, et al.
Published: (2023)
by: Wang, Haochen, et al.
Published: (2023)
Noisy Label Refinement with Semantically Reliable Synthetic Images
by: Li, Yingxuan, et al.
Published: (2025)
by: Li, Yingxuan, et al.
Published: (2025)
Are Multimodal Large Language Models Good Annotators for Image Tagging?
by: Xie, Ming-Kun, et al.
Published: (2026)
by: Xie, Ming-Kun, et al.
Published: (2026)
Semantic-Consistent Bidirectional Contrastive Hashing for Noisy Multi-Label Cross-Modal Retrieval
by: Peng, Likang, et al.
Published: (2025)
by: Peng, Likang, et al.
Published: (2025)
Multi-graph Graph Matching for Coronary Artery Semantic Labeling
by: Zhao, Chen, et al.
Published: (2024)
by: Zhao, Chen, et al.
Published: (2024)
Multi-Level Label Correction by Distilling Proximate Patterns for Semi-supervised Semantic Segmentation
by: Xiao, Hui, et al.
Published: (2024)
by: Xiao, Hui, et al.
Published: (2024)
You Only Speak Once to See
by: Yang, Wenhao, et al.
Published: (2024)
by: Yang, Wenhao, et al.
Published: (2024)
Repetitive Action Counting with Hybrid Temporal Relation Modeling
by: Li, Kun, et al.
Published: (2024)
by: Li, Kun, et al.
Published: (2024)
Tokenization Allows Multimodal Large Language Models to Understand, Generate and Edit Architectural Floor Plans
by: Qin, Sizhong, et al.
Published: (2026)
by: Qin, Sizhong, et al.
Published: (2026)
Seeing Realism from Simulation: Efficient Video Transfer for Vision-Language-Action Data Augmentation
by: Hui, Chenyu, et al.
Published: (2026)
by: Hui, Chenyu, et al.
Published: (2026)
Towards Multimodal Domain Generalization with Few Labels
by: Li, Hongzhao, et al.
Published: (2026)
by: Li, Hongzhao, et al.
Published: (2026)
See It Before You Grab It: Deep Learning-based Action Anticipation in Basketball
by: Roy, Arnau Barrera, et al.
Published: (2025)
by: Roy, Arnau Barrera, et al.
Published: (2025)
Towards Open-Vocabulary Semantic Segmentation Without Semantic Labels
by: Shin, Heeseong, et al.
Published: (2024)
by: Shin, Heeseong, et al.
Published: (2024)
Incomplete Multi-Label Image Recognition by Co-learning Semantic-Aware Features and Label Recovery
by: He, Zhi-Fen, et al.
Published: (2025)
by: He, Zhi-Fen, et al.
Published: (2025)
Bidirectional Learning of Facial Action Units and Expressions via Structured Semantic Mapping across Heterogeneous Datasets
by: Li, Jia, et al.
Published: (2026)
by: Li, Jia, et al.
Published: (2026)
Semantic-Aware Label Placement for Augmented Reality in Street View
by: Jia, Jianqing, et al.
Published: (2019)
by: Jia, Jianqing, et al.
Published: (2019)
SeeClear: Reliable Transparent Object Depth Estimation via Generative Opacification
by: Wang, Xiaoying, et al.
Published: (2026)
by: Wang, Xiaoying, et al.
Published: (2026)
Towards Robust Pseudo-Label Learning in Semantic Segmentation: An Encoding Perspective
by: Li, Wangkai, et al.
Published: (2025)
by: Li, Wangkai, et al.
Published: (2025)
Test-time Ego-Exo-centric Adaptation for Action Anticipation via Multi-Label Prototype Growing and Dual-Clue Consistency
by: Shi, Zhaofeng, et al.
Published: (2026)
by: Shi, Zhaofeng, et al.
Published: (2026)
HPL-ESS: Hybrid Pseudo-Labeling for Unsupervised Event-based Semantic Segmentation
by: Jing, Linglin, et al.
Published: (2024)
by: Jing, Linglin, et al.
Published: (2024)
Similar Items
-
Learning from True-False Labels via Multi-modal Prompt Retrieving
by: Li, Zhongnian, et al.
Published: (2024) -
Learning from Reduced Labels for Long-Tailed Data
by: Wei, Meng, et al.
Published: (2024) -
Beyond Label Semantics: Language-Guided Action Anatomy for Few-shot Action Recognition
by: Qian, Zefeng, et al.
Published: (2025) -
Energy Score-based Pseudo-Label Filtering and Adaptive Loss for Imbalanced Semi-supervised SAR target recognition
by: Zhang, Xinzheng, et al.
Published: (2024) -
Seeing Through Fog: Towards Fog-Invariant Action Recognition
by: Liu, Enqi, et al.
Published: (2026)