Saved in:
| Main Authors: | Zhang, Wenqian, Wang, Zehao |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2604.27445 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HumanGenesis: Agent-Based Geometric and Generative Modeling for Synthetic Human Dynamics
by: Li, Weiqi, et al.
Published: (2025)
by: Li, Weiqi, et al.
Published: (2025)
Enhancing 3D Lane Detection and Topology Reasoning with 2D Lane Priors
by: Li, Han, et al.
Published: (2024)
by: Li, Han, et al.
Published: (2024)
EmbodiedHead: Real-Time Listening and Speaking Avatar for Conversational Agents
by: Zhang, Yu, et al.
Published: (2026)
by: Zhang, Yu, et al.
Published: (2026)
EgoSpeak: Learning When to Speak for Egocentric Conversational Agents in the Wild
by: Kim, Junhyeok, et al.
Published: (2025)
by: Kim, Junhyeok, et al.
Published: (2025)
VizDefender: Unmasking Visualization Tampering through Proactive Localization and Intent Inference
by: Song, Sicheng, et al.
Published: (2025)
by: Song, Sicheng, et al.
Published: (2025)
Breaking Down and Building Up: Mixture of Skill-Based Vision-and-Language Navigation Agents
by: Ma, Tianyi, et al.
Published: (2025)
by: Ma, Tianyi, et al.
Published: (2025)
DVP-MVS++: Synergize Depth-Normal-Edge and Harmonized Visibility Prior for Multi-View Stereo
by: Yuan, Zhenlong, et al.
Published: (2025)
by: Yuan, Zhenlong, et al.
Published: (2025)
Spuriousness-Aware Meta-Learning for Learning Robust Classifiers
by: Zheng, Guangtao, et al.
Published: (2024)
by: Zheng, Guangtao, et al.
Published: (2024)
Context-aware Multi-task Learning for Pedestrian Intent and Trajectory Prediction
by: Munir, Farzeen, et al.
Published: (2024)
by: Munir, Farzeen, et al.
Published: (2024)
Probabilistic Human Intent Prediction for Mobile Manipulation: An Evaluation with Human-Inspired Constraints
by: Contreras, Cesar Alan, et al.
Published: (2025)
by: Contreras, Cesar Alan, et al.
Published: (2025)
RemoteAgent: Bridging Vague Human Intents and Earth Observation with RL-based Agentic MLLMs
by: Yao, Liang, et al.
Published: (2026)
by: Yao, Liang, et al.
Published: (2026)
GS-CLIP: Zero-shot 3D Anomaly Detection by Geometry-Aware Prompt and Synergistic View Representation Learning
by: Deng, Zehao, et al.
Published: (2026)
by: Deng, Zehao, et al.
Published: (2026)
Image Corruption-Inspired Membership Inference Attacks against Large Vision-Language Models
by: Wu, Zongyu, et al.
Published: (2025)
by: Wu, Zongyu, et al.
Published: (2025)
Embodied Laser Attack:Leveraging Scene Priors to Achieve Agent-based Robust Non-contact Attacks
by: Sun, Yitong, et al.
Published: (2023)
by: Sun, Yitong, et al.
Published: (2023)
TaskGround: Structured Executable Task Inference for Full-Scene Household Reasoning
by: Feng, ZhiYuan, et al.
Published: (2026)
by: Feng, ZhiYuan, et al.
Published: (2026)
CIFAR-10-Warehouse: Broad and More Realistic Testbeds in Model Generalization Analysis
by: Sun, Xiaoxiao, et al.
Published: (2023)
by: Sun, Xiaoxiao, et al.
Published: (2023)
Generalizable Non-Line-of-Sight Imaging with Learnable Physical Priors
by: Sun, Shida, et al.
Published: (2024)
by: Sun, Shida, et al.
Published: (2024)
Variational Bayesian Imaging with an Efficient Surrogate Score-based Prior
by: Feng, Berthy T., et al.
Published: (2023)
by: Feng, Berthy T., et al.
Published: (2023)
ACD: Direct Conditional Control for Video Diffusion Models via Attention Supervision
by: Li, Weiqi, et al.
Published: (2025)
by: Li, Weiqi, et al.
Published: (2025)
Learnable Graph Matching: A Practical Paradigm for Data Association
by: He, Jiawei, et al.
Published: (2023)
by: He, Jiawei, et al.
Published: (2023)
The DeepSpeak Dataset
by: Barrington, Sarah, et al.
Published: (2024)
by: Barrington, Sarah, et al.
Published: (2024)
Integrating Specialized and Generic Agent Motion Prediction with Dynamic Occupancy Grid Maps
by: Asghar, Rabbia, et al.
Published: (2026)
by: Asghar, Rabbia, et al.
Published: (2026)
ContactArt: Learning 3D Interaction Priors for Category-level Articulated Object and Hand Poses Estimation
by: Zhu, Zehao, et al.
Published: (2023)
by: Zhu, Zehao, et al.
Published: (2023)
LongCat-Video Technical Report
by: Meituan LongCat Team, et al.
Published: (2025)
by: Meituan LongCat Team, et al.
Published: (2025)
Training Like a Medical Resident: Context-Prior Learning Toward Universal Medical Image Segmentation
by: Gao, Yunhe, et al.
Published: (2023)
by: Gao, Yunhe, et al.
Published: (2023)
Posture-Driven Action Intent Inference for Playing style and Fatigue Assessment
by: Jaiswal, Abhishek, et al.
Published: (2025)
by: Jaiswal, Abhishek, et al.
Published: (2025)
Pose Prior Learner: Unsupervised Categorical Prior Learning for Pose Estimation
by: Wang, Ziyu, et al.
Published: (2024)
by: Wang, Ziyu, et al.
Published: (2024)
What to Say and When to Say it: Live Fitness Coaching as a Testbed for Situated Interaction
by: Panchal, Sunny, et al.
Published: (2024)
by: Panchal, Sunny, et al.
Published: (2024)
Proxy-GS: Unified Occlusion Priors for Training and Inference in Structured 3D Gaussian Splatting
by: Gao, Yuanyuan, et al.
Published: (2025)
by: Gao, Yuanyuan, et al.
Published: (2025)
NeuroBridge: Bio-Inspired Self-Supervised EEG-to-Image Decoding via Cognitive Priors and Bidirectional Semantic Alignment
by: Zhang, Wenjiang, et al.
Published: (2025)
by: Zhang, Wenjiang, et al.
Published: (2025)
IRIS: Intent Resolution via Inference-time Saccades for Open-Ended VQA in Large Vision-Language Models
by: Madinei, Parsa, et al.
Published: (2026)
by: Madinei, Parsa, et al.
Published: (2026)
UMO: Unified In-Context Learning Unlocks Motion Foundation Model Priors
by: Cong, Xiaoyan, et al.
Published: (2026)
by: Cong, Xiaoyan, et al.
Published: (2026)
Diffusion Prior-Based Amortized Variational Inference for Noisy Inverse Problems
by: Lee, Sojin, et al.
Published: (2024)
by: Lee, Sojin, et al.
Published: (2024)
MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation
by: Rong, Fu, et al.
Published: (2025)
by: Rong, Fu, et al.
Published: (2025)
From Clinical Intent to Clinical Model: Autonomous Coding-Agents for Clinician-driven AI Development
by: Zhao, Zihao, et al.
Published: (2026)
by: Zhao, Zihao, et al.
Published: (2026)
SecAgent: Efficient Mobile GUI Agent with Semantic Context
by: Xie, Yiping, et al.
Published: (2026)
by: Xie, Yiping, et al.
Published: (2026)
Rethinking Efficient Mixture-of-Experts for Remote Sensing Modality-Missing Classification
by: Gao, Qinghao, et al.
Published: (2025)
by: Gao, Qinghao, et al.
Published: (2025)
ChangeChat: An Interactive Model for Remote Sensing Change Analysis via Multimodal Instruction Tuning
by: Deng, Pei, et al.
Published: (2024)
by: Deng, Pei, et al.
Published: (2024)
TDS-CLIP: Temporal Difference Side Network for Efficient VideoAction Recognition
by: Wang, Bin, et al.
Published: (2024)
by: Wang, Bin, et al.
Published: (2024)
LongCat-Image Technical Report
by: Meituan LongCat Team, et al.
Published: (2025)
by: Meituan LongCat Team, et al.
Published: (2025)
Similar Items
-
HumanGenesis: Agent-Based Geometric and Generative Modeling for Synthetic Human Dynamics
by: Li, Weiqi, et al.
Published: (2025) -
Enhancing 3D Lane Detection and Topology Reasoning with 2D Lane Priors
by: Li, Han, et al.
Published: (2024) -
EmbodiedHead: Real-Time Listening and Speaking Avatar for Conversational Agents
by: Zhang, Yu, et al.
Published: (2026) -
EgoSpeak: Learning When to Speak for Egocentric Conversational Agents in the Wild
by: Kim, Junhyeok, et al.
Published: (2025) -
VizDefender: Unmasking Visualization Tampering through Proactive Localization and Intent Inference
by: Song, Sicheng, et al.
Published: (2025)