Make Graph-based Referring Expression Comprehension Great Again through Expression-guided Dynamic Gating and Regression
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ke, Jingcheng, Wang, Dele, Chen, Jun-Cheng, Jhuo, I-Hong, Lin, Chia-Wen, Lin, Yen-Yu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Interpretable Zero-shot Referring Expression Comprehension with Query-driven Scene Graphs
von: Wu, Yike, et al.
Veröffentlicht: (2026)
von: Wu, Yike, et al.
Veröffentlicht: (2026)
Improving Visual Object Tracking through Visual Prompting
von: Chen, Shih-Fang, et al.
Veröffentlicht: (2024)
von: Chen, Shih-Fang, et al.
Veröffentlicht: (2024)
CK-Transformer: Commonsense Knowledge Enhanced Transformers for Referring Expression Comprehension
von: Zhang, Zhi, et al.
Veröffentlicht: (2023)
von: Zhang, Zhi, et al.
Veröffentlicht: (2023)
Understanding What Is Not Said:Referring Remote Sensing Image Segmentation with Scarce Expressions
von: Ye, Kai, et al.
Veröffentlicht: (2025)
von: Ye, Kai, et al.
Veröffentlicht: (2025)
GOT-JEPA: Generic Object Tracking with Model Adaptation and Occlusion Handling using Joint-Embedding Predictive Architecture
von: Chen, Shih-Fang, et al.
Veröffentlicht: (2026)
von: Chen, Shih-Fang, et al.
Veröffentlicht: (2026)
GOT-Edit: Geometry-Aware Generic Object Tracking via Online Model Editing
von: Chen, Shih-Fang, et al.
Veröffentlicht: (2026)
von: Chen, Shih-Fang, et al.
Veröffentlicht: (2026)
Dynamic Resolution Guidance for Facial Expression Recognition
von: Wang, Songpan, et al.
Veröffentlicht: (2024)
von: Wang, Songpan, et al.
Veröffentlicht: (2024)
Anchoring Trends: Mitigating Social Media Popularity Prediction Drift via Feature Clustering and Expansion
von: Lee, Chia-Ming, et al.
Veröffentlicht: (2025)
von: Lee, Chia-Ming, et al.
Veröffentlicht: (2025)
SAM as the Guide: Mastering Pseudo-Label Refinement in Semi-Supervised Referring Expression Segmentation
von: Yang, Danni, et al.
Veröffentlicht: (2024)
von: Yang, Danni, et al.
Veröffentlicht: (2024)
MEGC2025: Micro-Expression Grand Challenge on Spot Then Recognize and Visual Question Answering
von: Fan, Xinqi, et al.
Veröffentlicht: (2025)
von: Fan, Xinqi, et al.
Veröffentlicht: (2025)
ExpLLM: Towards Chain of Thought for Facial Expression Recognition
von: Lan, Xing, et al.
Veröffentlicht: (2024)
von: Lan, Xing, et al.
Veröffentlicht: (2024)
Generalized Jersey Number Recognition Using Multi-task Learning With Orientation-guided Weight Refinement
von: Lin, Yung-Hui, et al.
Veröffentlicht: (2024)
von: Lin, Yung-Hui, et al.
Veröffentlicht: (2024)
MMED: A Multimodal Micro-Expression Dataset based on Audio-Visual Fusion
von: Wang, Junbo, et al.
Veröffentlicht: (2025)
von: Wang, Junbo, et al.
Veröffentlicht: (2025)
PRM-BAS: Enhancing Multimodal Reasoning through PRM-guided Beam Annealing Search
von: Hu, Pengfei, et al.
Veröffentlicht: (2025)
von: Hu, Pengfei, et al.
Veröffentlicht: (2025)
SFE-Net: Harnessing Biological Principles of Differential Gene Expression for Improved Feature Selection in Deep Learning Networks
von: Li, Yuqi, et al.
Veröffentlicht: (2024)
von: Li, Yuqi, et al.
Veröffentlicht: (2024)
Dynamic Multimodal Expression Generation for LLM-Driven Pedagogical Agents: From User Experience Perspective
von: Wan, Ninghao, et al.
Veröffentlicht: (2026)
von: Wan, Ninghao, et al.
Veröffentlicht: (2026)
MicroEmo: Time-Sensitive Multimodal Emotion Recognition with Micro-Expression Dynamics in Video Dialogues
von: Zhang, Liyun
Veröffentlicht: (2024)
von: Zhang, Liyun
Veröffentlicht: (2024)
AllTheDocks road safety dataset: A cyclist's perspective and experience
von: Chiang, Chia-Yen, et al.
Veröffentlicht: (2024)
von: Chiang, Chia-Yen, et al.
Veröffentlicht: (2024)
M2ORT: Many-To-One Regression Transformer for Spatial Transcriptomics Prediction from Histopathology Images
von: Wang, Hongyi, et al.
Veröffentlicht: (2024)
von: Wang, Hongyi, et al.
Veröffentlicht: (2024)
Optimized Learned Image Compression for Facial Expression Recognition
von: Li, Xiumei, et al.
Veröffentlicht: (2025)
von: Li, Xiumei, et al.
Veröffentlicht: (2025)
AI TrackMate: Finally, Someone Who Will Give Your Music More Than Just "Sounds Great!"
von: Jiang, Yi-Lin, et al.
Veröffentlicht: (2024)
von: Jiang, Yi-Lin, et al.
Veröffentlicht: (2024)
MTCAE-DFER: Multi-Task Cascaded Autoencoder for Dynamic Facial Expression Recognition
von: Xiang, Peihao, et al.
Veröffentlicht: (2024)
von: Xiang, Peihao, et al.
Veröffentlicht: (2024)
Multimodal Interaction Modeling via Self-Supervised Multi-Task Learning for Review Helpfulness Prediction
von: Gong, HongLin, et al.
Veröffentlicht: (2024)
von: Gong, HongLin, et al.
Veröffentlicht: (2024)
Speech2AffectiveGestures: Synthesizing Co-Speech Gestures with Generative Adversarial Affective Expression Learning
von: Bhattacharya, Uttaran, et al.
Veröffentlicht: (2021)
von: Bhattacharya, Uttaran, et al.
Veröffentlicht: (2021)
SafaRi:Adaptive Sequence Transformer for Weakly Supervised Referring Expression Segmentation
von: Nag, Sayan, et al.
Veröffentlicht: (2024)
von: Nag, Sayan, et al.
Veröffentlicht: (2024)
MEGC2026: Micro-Expression Grand Challenge on Visual Question Answering
von: Fan, Xinqi, et al.
Veröffentlicht: (2026)
von: Fan, Xinqi, et al.
Veröffentlicht: (2026)
Multimodal Graph Neural Network for Recommendation with Dynamic De-redundancy and Modality-Guided Feature De-noisy
von: Mo, Feng, et al.
Veröffentlicht: (2024)
von: Mo, Feng, et al.
Veröffentlicht: (2024)
Lost in Overlap: Exploring Logit-based Watermark Collision in LLMs
von: Luo, Yiyang, et al.
Veröffentlicht: (2024)
von: Luo, Yiyang, et al.
Veröffentlicht: (2024)
GANonymization: A GAN-based Face Anonymization Framework for Preserving Emotional Expressions
von: Hellmann, Fabio, et al.
Veröffentlicht: (2023)
von: Hellmann, Fabio, et al.
Veröffentlicht: (2023)
Revisiting Vision-Language Features Adaptation and Inconsistency for Social Media Popularity Prediction
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024)
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024)
Routing Experts: Learning to Route Dynamic Experts in Multi-modal Large Language Models
von: Wu, Qiong, et al.
Veröffentlicht: (2024)
von: Wu, Qiong, et al.
Veröffentlicht: (2024)
CueNet: Robust Audio-Visual Speaker Extraction through Cross-Modal Cue Mining and Interaction
von: Wang, Jiadong, et al.
Veröffentlicht: (2026)
von: Wang, Jiadong, et al.
Veröffentlicht: (2026)
SemanticGarment: Semantic-Controlled Generation and Editing of 3D Gaussian Garments
von: Wang, Ruiyan, et al.
Veröffentlicht: (2025)
von: Wang, Ruiyan, et al.
Veröffentlicht: (2025)
Improving the Efficiency of VVC using Partitioning of Reference Frames
von: Qureshi, Kamran, et al.
Veröffentlicht: (2025)
von: Qureshi, Kamran, et al.
Veröffentlicht: (2025)
EditEmoTalk: Controllable Speech-Driven 3D Facial Animation with Continuous Expression Editing
von: Jiang, Diqiong, et al.
Veröffentlicht: (2026)
von: Jiang, Diqiong, et al.
Veröffentlicht: (2026)
MMPKUBase: A Comprehensive and High-quality Chinese Multi-modal Knowledge Graph
von: Yi, Xuan, et al.
Veröffentlicht: (2024)
von: Yi, Xuan, et al.
Veröffentlicht: (2024)
Multi-Reference Generative Face Video Compression with Contrastive Learning
von: Konuko, Goluck, et al.
Veröffentlicht: (2024)
von: Konuko, Goluck, et al.
Veröffentlicht: (2024)
Identity-Driven Multimedia Forgery Detection via Reference Assistance
von: Xu, Junhao, et al.
Veröffentlicht: (2024)
von: Xu, Junhao, et al.
Veröffentlicht: (2024)
HER2 Expression Prediction with Flexible Multi-Modal Inputs via Dynamic Bidirectional Reconstruction
von: Qin, Jie, et al.
Veröffentlicht: (2025)
von: Qin, Jie, et al.
Veröffentlicht: (2025)
LLM2Manim: Pedagogy-Aware AI Generation of STEM Animations
von: Joshi, Aastha, et al.
Veröffentlicht: (2026)
von: Joshi, Aastha, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Interpretable Zero-shot Referring Expression Comprehension with Query-driven Scene Graphs
von: Wu, Yike, et al.
Veröffentlicht: (2026) -
Improving Visual Object Tracking through Visual Prompting
von: Chen, Shih-Fang, et al.
Veröffentlicht: (2024) -
CK-Transformer: Commonsense Knowledge Enhanced Transformers for Referring Expression Comprehension
von: Zhang, Zhi, et al.
Veröffentlicht: (2023) -
Understanding What Is Not Said:Referring Remote Sensing Image Segmentation with Scarce Expressions
von: Ye, Kai, et al.
Veröffentlicht: (2025) -
GOT-JEPA: Generic Object Tracking with Model Adaptation and Occlusion Handling using Joint-Embedding Predictive Architecture
von: Chen, Shih-Fang, et al.
Veröffentlicht: (2026)