M3PT: A Transformer for Multimodal, Multi-Party Social Signal Prediction with Person-aware Blockwise Attention
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tang, Yiming, Anwar, Abrar, Thomason, Jesse |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection
von: Anwar, Abrar, et al.
Veröffentlicht: (2025)
von: Anwar, Abrar, et al.
Veröffentlicht: (2025)
Contrast Sets for Evaluating Language-Guided Robot Policies
von: Anwar, Abrar, et al.
Veröffentlicht: (2024)
von: Anwar, Abrar, et al.
Veröffentlicht: (2024)
RobotFleet: An Open-Source Framework for Centralized Multi-Robot Task Planning
von: Gupta, Rohan, et al.
Veröffentlicht: (2025)
von: Gupta, Rohan, et al.
Veröffentlicht: (2025)
Self-Discovered Intention-aware Transformer for Multi-modal Vehicle Trajectory Prediction
von: Liu, Diyi, et al.
Veröffentlicht: (2026)
von: Liu, Diyi, et al.
Veröffentlicht: (2026)
THE COLOSSEUM: A Benchmark for Evaluating Generalization for Robotic Manipulation
von: Pumacay, Wilbert, et al.
Veröffentlicht: (2024)
von: Pumacay, Wilbert, et al.
Veröffentlicht: (2024)
PEEK: Guiding and Minimal Image Representations for Zero-Shot Generalization of Robot Manipulation Policies
von: Zhang, Jesse, et al.
Veröffentlicht: (2025)
von: Zhang, Jesse, et al.
Veröffentlicht: (2025)
Crossfusor: A Cross-Attention Transformer Enhanced Conditional Diffusion Model for Car-Following Trajectory Prediction
von: You, Junwei, et al.
Veröffentlicht: (2024)
von: You, Junwei, et al.
Veröffentlicht: (2024)
Robometer: Scaling General-Purpose Robotic Reward Models via Trajectory Comparisons
von: Liang, Anthony, et al.
Veröffentlicht: (2026)
von: Liang, Anthony, et al.
Veröffentlicht: (2026)
CroSTAta: Cross-State Transition Attention Transformer for Robotic Manipulation
von: Minelli, Giovanni, et al.
Veröffentlicht: (2025)
von: Minelli, Giovanni, et al.
Veröffentlicht: (2025)
Which One? Leveraging Context Between Objects and Multiple Views for Language Grounding
von: Mitra, Chancharik, et al.
Veröffentlicht: (2023)
von: Mitra, Chancharik, et al.
Veröffentlicht: (2023)
Multi-modal Synthetic Data Training and Model Collapse: Insights from VLMs and Diffusion Models
von: Hu, Zizhao, et al.
Veröffentlicht: (2025)
von: Hu, Zizhao, et al.
Veröffentlicht: (2025)
Neural-Network-Driven Reward Prediction as a Heuristic: Advancing Q-Learning for Mobile Robot Path Planning
von: Ji, Yiming, et al.
Veröffentlicht: (2024)
von: Ji, Yiming, et al.
Veröffentlicht: (2024)
Multimodal Gender Fairness in Depression Prediction: Insights on Data from the USA & China
von: Cameron, Joseph, et al.
Veröffentlicht: (2024)
von: Cameron, Joseph, et al.
Veröffentlicht: (2024)
Target Return Optimizer for Multi-Game Decision Transformer
von: Tatematsu, Kensuke, et al.
Veröffentlicht: (2025)
von: Tatematsu, Kensuke, et al.
Veröffentlicht: (2025)
Signaling and Social Learning in Swarms of Robots
von: Cazenille, Leo, et al.
Veröffentlicht: (2024)
von: Cazenille, Leo, et al.
Veröffentlicht: (2024)
KI-GAN: Knowledge-Informed Generative Adversarial Networks for Enhanced Multi-Vehicle Trajectory Forecasting at Signalized Intersections
von: Wei, Chuheng, et al.
Veröffentlicht: (2024)
von: Wei, Chuheng, et al.
Veröffentlicht: (2024)
Learning Long-Context Diffusion Policies via Past-Token Prediction
von: Torne, Marcel, et al.
Veröffentlicht: (2025)
von: Torne, Marcel, et al.
Veröffentlicht: (2025)
DR-MPC: Deep Residual Model Predictive Control for Real-world Social Navigation
von: Han, James R., et al.
Veröffentlicht: (2024)
von: Han, James R., et al.
Veröffentlicht: (2024)
Multi-vessel Interaction-Aware Trajectory Prediction and Collision Risk Assessment
von: Alam, Md Mahbub, et al.
Veröffentlicht: (2025)
von: Alam, Md Mahbub, et al.
Veröffentlicht: (2025)
Skill-aware Mutual Information Optimisation for Generalisation in Reinforcement Learning
von: Yu, Xuehui, et al.
Veröffentlicht: (2024)
von: Yu, Xuehui, et al.
Veröffentlicht: (2024)
Beyond Features: How Dataset Design Influences Multi-Agent Trajectory Prediction Performance
von: Demmler, Tobias, et al.
Veröffentlicht: (2025)
von: Demmler, Tobias, et al.
Veröffentlicht: (2025)
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning
von: Lee, Dongsu, et al.
Veröffentlicht: (2025)
von: Lee, Dongsu, et al.
Veröffentlicht: (2025)
Safety-aware Causal Representation for Trustworthy Offline Reinforcement Learning in Autonomous Driving
von: Lin, Haohong, et al.
Veröffentlicht: (2023)
von: Lin, Haohong, et al.
Veröffentlicht: (2023)
TMRL: Diffusion Timestep-Modulated Pretraining Enables Exploration for Efficient Policy Finetuning
von: Hong, Matthew M., et al.
Veröffentlicht: (2026)
von: Hong, Matthew M., et al.
Veröffentlicht: (2026)
SPRINT: Scalable Policy Pre-Training via Language Instruction Relabeling
von: Zhang, Jesse, et al.
Veröffentlicht: (2023)
von: Zhang, Jesse, et al.
Veröffentlicht: (2023)
State- and context-dependent robotic manipulation and grasping via uncertainty-aware imitation learning
von: Winter, Tim R., et al.
Veröffentlicht: (2024)
von: Winter, Tim R., et al.
Veröffentlicht: (2024)
MetaFollower: Adaptable Personalized Autonomous Car Following
von: Chen, Xianda, et al.
Veröffentlicht: (2024)
von: Chen, Xianda, et al.
Veröffentlicht: (2024)
JAM: Keypoint-Guided Joint Prediction after Classification-Aware Marginal Proposal for Multi-Agent Interaction
von: Lin, Fangze, et al.
Veröffentlicht: (2025)
von: Lin, Fangze, et al.
Veröffentlicht: (2025)
ViSaRL: Visual Reinforcement Learning Guided by Human Saliency
von: Liang, Anthony, et al.
Veröffentlicht: (2024)
von: Liang, Anthony, et al.
Veröffentlicht: (2024)
Symmetry-aware Reinforcement Learning for Robotic Assembly under Partial Observability with a Soft Wrist
von: Nguyen, Hai, et al.
Veröffentlicht: (2024)
von: Nguyen, Hai, et al.
Veröffentlicht: (2024)
Multimodal Diffusion Forcing for Forceful Manipulation
von: Huang, Zixuan, et al.
Veröffentlicht: (2025)
von: Huang, Zixuan, et al.
Veröffentlicht: (2025)
SENSOR: Imitate Third-Person Expert's Behaviors via Active Sensoring
von: Huang, Kaichen, et al.
Veröffentlicht: (2024)
von: Huang, Kaichen, et al.
Veröffentlicht: (2024)
GOPT: Generalizable Online 3D Bin Packing via Transformer-based Deep Reinforcement Learning
von: Xiong, Heng, et al.
Veröffentlicht: (2024)
von: Xiong, Heng, et al.
Veröffentlicht: (2024)
Graph Attention-Guided Search for Dense Multi-Agent Pathfinding
von: Jain, Rishabh, et al.
Veröffentlicht: (2025)
von: Jain, Rishabh, et al.
Veröffentlicht: (2025)
Subequivariant Reinforcement Learning in 3D Multi-Entity Physical Environments
von: Chen, Runfa, et al.
Veröffentlicht: (2024)
von: Chen, Runfa, et al.
Veröffentlicht: (2024)
Understanding Multimodal Failure in Action-Chunking Behavioral Cloning
von: Mazza, Lorenzo, et al.
Veröffentlicht: (2026)
von: Mazza, Lorenzo, et al.
Veröffentlicht: (2026)
Attention-Based Neural-Augmented Kalman Filter for Legged Robot State Estimation
von: Lee, Seokju, et al.
Veröffentlicht: (2026)
von: Lee, Seokju, et al.
Veröffentlicht: (2026)
PP-TIL: Personalized Planning for Autonomous Driving with Instance-based Transfer Imitation Learning
von: Lin, Fangze, et al.
Veröffentlicht: (2024)
von: Lin, Fangze, et al.
Veröffentlicht: (2024)
Model Predictive Robustness of Signal Temporal Logic Predicates
von: Lin, Yuanfei, et al.
Veröffentlicht: (2022)
von: Lin, Yuanfei, et al.
Veröffentlicht: (2022)
RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
von: Wang, Yufei, et al.
Veröffentlicht: (2024)
von: Wang, Yufei, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection
von: Anwar, Abrar, et al.
Veröffentlicht: (2025) -
Contrast Sets for Evaluating Language-Guided Robot Policies
von: Anwar, Abrar, et al.
Veröffentlicht: (2024) -
RobotFleet: An Open-Source Framework for Centralized Multi-Robot Task Planning
von: Gupta, Rohan, et al.
Veröffentlicht: (2025) -
Self-Discovered Intention-aware Transformer for Multi-modal Vehicle Trajectory Prediction
von: Liu, Diyi, et al.
Veröffentlicht: (2026) -
THE COLOSSEUM: A Benchmark for Evaluating Generalization for Robotic Manipulation
von: Pumacay, Wilbert, et al.
Veröffentlicht: (2024)