Transforming Multimodal Models into Action Models for Radiotherapy
Fuente:
arXiv
Saved in:
| Main Authors: | Ferrante, Matteo, Carosi, Alessandra, Angelillo, Rolando Maria D, Toschi, Nicola |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Beam angle optimization for radiotherapy using LLMs via reinforcement‐learning inspired iterative refinement
by: Sara Cammarota, et al.
Published: (2026)
by: Sara Cammarota, et al.
Published: (2026)
Bridging Auditory Perception and Language Comprehension through MEG-Driven Encoding Models
by: Ciferri, Matteo, et al.
Published: (2024)
by: Ciferri, Matteo, et al.
Published: (2024)
R&B -- Rhythm and Brain: Cross-subject Decoding of Music from Human Brain Activity
by: Ferrante, Matteo, et al.
Published: (2024)
by: Ferrante, Matteo, et al.
Published: (2024)
Towards Neural Foundation Models for Vision: Aligning EEG, MEG, and fMRI Representations for Decoding, Encoding, and Modality Conversion
by: Ferrante, Matteo, et al.
Published: (2024)
by: Ferrante, Matteo, et al.
Published: (2024)
MARIA: a Multimodal Transformer Model for Incomplete Healthcare Data
by: Caruso, Camillo Maria, et al.
Published: (2024)
by: Caruso, Camillo Maria, et al.
Published: (2024)
Inference-Time Toxicity Mitigation in Protein Language Models
by: Burda, Manuel Fernández, et al.
Published: (2026)
by: Burda, Manuel Fernández, et al.
Published: (2026)
Context-Selective State Space Models: Feedback is All You Need
by: Zattra, Riccardo, et al.
Published: (2025)
by: Zattra, Riccardo, et al.
Published: (2025)
Back To The Future: A Hybrid Transformer-XGBoost Model for Action-oriented Future-proofing Nowcasting
by: Sun, Ziheng
Published: (2024)
by: Sun, Ziheng
Published: (2024)
Recurrent Action Transformer with Memory
by: Cherepanov, Egor, et al.
Published: (2023)
by: Cherepanov, Egor, et al.
Published: (2023)
Looping Back to Move Forward: Recursive Transformers for Efficient and Flexible Large Multimodal Models
by: Xu, Ruihan, et al.
Published: (2026)
by: Xu, Ruihan, et al.
Published: (2026)
MicroFlow: An Efficient Rust-Based Inference Engine for TinyML
by: Carnelos, Matteo, et al.
Published: (2024)
by: Carnelos, Matteo, et al.
Published: (2024)
villa-X: Enhancing Latent Action Modeling in Vision-Language-Action Models
by: Chen, Xiaoyu, et al.
Published: (2025)
by: Chen, Xiaoyu, et al.
Published: (2025)
Adjusting the Output of Decision Transformer with Action Gradient
by: Lin, Rui, et al.
Published: (2025)
by: Lin, Rui, et al.
Published: (2025)
In-Context Learning for MIMO Equalization Using Transformer-Based Sequence Models
by: Zecchin, Matteo, et al.
Published: (2023)
by: Zecchin, Matteo, et al.
Published: (2023)
Multi-Armed Bandits With Best-Action Queries
by: Bacchiocchi, Francesco, et al.
Published: (2026)
by: Bacchiocchi, Francesco, et al.
Published: (2026)
Benchmarking Multimodal Knowledge Conflict for Large Multimodal Models
by: Jia, Yifan, et al.
Published: (2025)
by: Jia, Yifan, et al.
Published: (2025)
Recurrence-Complete Frame-based Action Models
by: Keiblinger, Michael
Published: (2025)
by: Keiblinger, Michael
Published: (2025)
Learning Safe Numeric Planning Action Models
by: Mordoch, Argaman, et al.
Published: (2023)
by: Mordoch, Argaman, et al.
Published: (2023)
Expanding Expressivity in Transformer Models with MöbiusAttention
by: Halacheva, Anna-Maria, et al.
Published: (2024)
by: Halacheva, Anna-Maria, et al.
Published: (2024)
Linear Model Merging Unlocks Simple and Scalable Multimodal Data Mixture Optimization
by: Berasi, Davide, et al.
Published: (2026)
by: Berasi, Davide, et al.
Published: (2026)
Multimodal Cancer Modeling in the Age of Foundation Model Embeddings
by: Song, Steven, et al.
Published: (2025)
by: Song, Steven, et al.
Published: (2025)
DAWM: Diffusion Action World Models for Offline Reinforcement Learning via Action-Inferred Transitions
by: Li, Zongyue, et al.
Published: (2025)
by: Li, Zongyue, et al.
Published: (2025)
Scalable Offline Model-Based RL with Action Chunks
by: Park, Kwanyoung, et al.
Published: (2025)
by: Park, Kwanyoung, et al.
Published: (2025)
What Do Latent Action Models Actually Learn?
by: Zhang, Chuheng, et al.
Published: (2025)
by: Zhang, Chuheng, et al.
Published: (2025)
Model-based Reinforcement Learning for Parameterized Action Spaces
by: Zhang, Renhao, et al.
Published: (2024)
by: Zhang, Renhao, et al.
Published: (2024)
From Efficient Multimodal Models to World Models: A Survey
by: Mai, Xinji, et al.
Published: (2024)
by: Mai, Xinji, et al.
Published: (2024)
Disentanglement of Variations with Multimodal Generative Modeling
by: Zhang, Yijie, et al.
Published: (2025)
by: Zhang, Yijie, et al.
Published: (2025)
Scaling Law Hypothesis for Multimodal Model
by: Sun, Qingyun, et al.
Published: (2024)
by: Sun, Qingyun, et al.
Published: (2024)
Empty SPACE: Cross-Attention Sparsity for Concept Erasure in Diffusion Models
by: Novello, Nicola, et al.
Published: (2026)
by: Novello, Nicola, et al.
Published: (2026)
Noise Stability of Transformer Models
by: Haris, Themistoklis, et al.
Published: (2026)
by: Haris, Themistoklis, et al.
Published: (2026)
Feature-level Interaction Explanations in Multimodal Transformers
by: Kim, Yeji, et al.
Published: (2026)
by: Kim, Yeji, et al.
Published: (2026)
Understanding Multimodal Failure in Action-Chunking Behavioral Cloning
by: Mazza, Lorenzo, et al.
Published: (2026)
by: Mazza, Lorenzo, et al.
Published: (2026)
The Role of Foundation Models in Neuro-Symbolic Learning and Reasoning
by: Cunnington, Daniel, et al.
Published: (2024)
by: Cunnington, Daniel, et al.
Published: (2024)
Scaling Offline Model-Based RL via Jointly-Optimized World-Action Model Pretraining
by: Cheng, Jie, et al.
Published: (2024)
by: Cheng, Jie, et al.
Published: (2024)
Do Transformer World Models Give Better Policy Gradients?
by: Ma, Michel, et al.
Published: (2024)
by: Ma, Michel, et al.
Published: (2024)
Traj-Transformer: Diffusion Models with Transformer for GPS Trajectory Generation
by: Zhang, Zhiyang, et al.
Published: (2025)
by: Zhang, Zhiyang, et al.
Published: (2025)
Do What You Say: Steering Vision-Language-Action Models via Runtime Reasoning-Action Alignment Verification
by: Wu, Yilin, et al.
Published: (2025)
by: Wu, Yilin, et al.
Published: (2025)
Revealing Multimodal Causality with Large Language Models
by: Li, Jin, et al.
Published: (2025)
by: Li, Jin, et al.
Published: (2025)
Hyperbolic Learning with Multimodal Large Language Models
by: Mandica, Paolo, et al.
Published: (2024)
by: Mandica, Paolo, et al.
Published: (2024)
Multimodal Attack Detection for Action Recognition Models
by: Mumcu, Furkan, et al.
Published: (2024)
by: Mumcu, Furkan, et al.
Published: (2024)
Similar Items
-
Beam angle optimization for radiotherapy using LLMs via reinforcement‐learning inspired iterative refinement
by: Sara Cammarota, et al.
Published: (2026) -
Bridging Auditory Perception and Language Comprehension through MEG-Driven Encoding Models
by: Ciferri, Matteo, et al.
Published: (2024) -
R&B -- Rhythm and Brain: Cross-subject Decoding of Music from Human Brain Activity
by: Ferrante, Matteo, et al.
Published: (2024) -
Towards Neural Foundation Models for Vision: Aligning EEG, MEG, and fMRI Representations for Decoding, Encoding, and Modality Conversion
by: Ferrante, Matteo, et al.
Published: (2024) -
MARIA: a Multimodal Transformer Model for Incomplete Healthcare Data
by: Caruso, Camillo Maria, et al.
Published: (2024)