Resolving Token-Space Gradient Conflicts: Token Space Manipulation for Transformer-Based Multi-Task Learning
Fuente:
arXiv
Salvato in:
| Autori principali: | Jeong, Wooseong, Yoon, Kuk-Jin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Label-Free Cross-Task LoRA Merging with Null-Space Compression
di: Lee, Wonyoung, et al.
Pubblicazione: (2026)
di: Lee, Wonyoung, et al.
Pubblicazione: (2026)
Synchronizing Task Behavior: Aligning Multiple Tasks during Test-Time Training
di: Jeong, Wooseong, et al.
Pubblicazione: (2025)
di: Jeong, Wooseong, et al.
Pubblicazione: (2025)
Quantifying Task Priority for Multi-Task Optimization
di: Jeong, Wooseong, et al.
Pubblicazione: (2024)
di: Jeong, Wooseong, et al.
Pubblicazione: (2024)
Preference-Aligned LoRA Merging: Preserving Subspace Coverage and Addressing Directional Anisotropy
di: Jeong, Wooseong, et al.
Pubblicazione: (2026)
di: Jeong, Wooseong, et al.
Pubblicazione: (2026)
Interaction-Merged Motion Planning: Effectively Leveraging Diverse Motion Datasets for Robust Planning
di: Lee, Giwon, et al.
Pubblicazione: (2025)
di: Lee, Giwon, et al.
Pubblicazione: (2025)
FACL-Attack: Frequency-Aware Contrastive Learning for Transferable Adversarial Attacks
di: Yang, Hunmin, et al.
Pubblicazione: (2024)
di: Yang, Hunmin, et al.
Pubblicazione: (2024)
Prompt-Driven Contrastive Learning for Transferable Adversarial Attacks
di: Yang, Hunmin, et al.
Pubblicazione: (2024)
di: Yang, Hunmin, et al.
Pubblicazione: (2024)
Exploring Token Pruning in Vision State Space Models
di: Zhan, Zheng, et al.
Pubblicazione: (2024)
di: Zhan, Zheng, et al.
Pubblicazione: (2024)
Proactive Gradient Conflict Mitigation in Multi-Task Learning: A Sparse Training Perspective
di: Zhang, Zhi, et al.
Pubblicazione: (2024)
di: Zhang, Zhi, et al.
Pubblicazione: (2024)
SceneTok: A Compressed, Diffusable Token Space for 3D Scenes
di: Asim, Mohammad, et al.
Pubblicazione: (2026)
di: Asim, Mohammad, et al.
Pubblicazione: (2026)
MambaMixer: Efficient Selective State Space Models with Dual Token and Channel Selection
di: Behrouz, Ali, et al.
Pubblicazione: (2024)
di: Behrouz, Ali, et al.
Pubblicazione: (2024)
Taming Outlier Tokens in Diffusion Transformers
di: Wu, Xiaoyu, et al.
Pubblicazione: (2026)
di: Wu, Xiaoyu, et al.
Pubblicazione: (2026)
ULTra: Unveiling Latent Token Interpretability in Transformer-Based Understanding and Segmentation
di: Hosseini, Hesam, et al.
Pubblicazione: (2024)
di: Hosseini, Hesam, et al.
Pubblicazione: (2024)
Robustness Tokens: Towards Adversarial Robustness of Transformers
di: Pulfer, Brian, et al.
Pubblicazione: (2025)
di: Pulfer, Brian, et al.
Pubblicazione: (2025)
Hourglass Tokenizer for Efficient Transformer-Based 3D Human Pose Estimation
di: Li, Wenhao, et al.
Pubblicazione: (2023)
di: Li, Wenhao, et al.
Pubblicazione: (2023)
Accelerating Diffusion Transformers with Token-wise Feature Caching
di: Zou, Chang, et al.
Pubblicazione: (2024)
di: Zou, Chang, et al.
Pubblicazione: (2024)
A More Word-like Image Tokenization for MLLMs
di: Lee, Hyun, et al.
Pubblicazione: (2026)
di: Lee, Hyun, et al.
Pubblicazione: (2026)
Random Token Fusion for Multi-View Medical Diagnosis
di: Guo, Jingyu, et al.
Pubblicazione: (2024)
di: Guo, Jingyu, et al.
Pubblicazione: (2024)
Multi-agent Long-term 3D Human Pose Forecasting via Interaction-aware Trajectory Conditioning
di: Jeong, Jaewoo, et al.
Pubblicazione: (2024)
di: Jeong, Jaewoo, et al.
Pubblicazione: (2024)
GeoToken: Hierarchical Geolocalization of Images via Next Token Prediction
di: Ghasemi, Narges, et al.
Pubblicazione: (2025)
di: Ghasemi, Narges, et al.
Pubblicazione: (2025)
TORE: Token Recycling in Vision Transformers for Efficient Active Visual Exploration
di: Olszewski, Jan, et al.
Pubblicazione: (2023)
di: Olszewski, Jan, et al.
Pubblicazione: (2023)
MCAT: Visual Query-Based Localization of Standard Anatomical Clips in Fetal Ultrasound Videos Using Multi-Tier Class-Aware Token Transformer
di: Mishra, Divyanshu, et al.
Pubblicazione: (2025)
di: Mishra, Divyanshu, et al.
Pubblicazione: (2025)
Multi-Token Prediction Needs Registers
di: Gerontopoulos, Anastasios, et al.
Pubblicazione: (2025)
di: Gerontopoulos, Anastasios, et al.
Pubblicazione: (2025)
Good Token Hunting: A Hitchhiker's Guide to Token Selection for Visual Geometry Transformers
di: Zheng, Shuhong, et al.
Pubblicazione: (2026)
di: Zheng, Shuhong, et al.
Pubblicazione: (2026)
Towards Resolving Optimization Conflicts Between Image- and Text-Based Person Re-Identification
di: Kvanchiani, Karina, et al.
Pubblicazione: (2026)
di: Kvanchiani, Karina, et al.
Pubblicazione: (2026)
Point-RTD: Replaced Token Denoising for Pretraining Transformer Models on Point Clouds
di: Stone, Gunner, et al.
Pubblicazione: (2025)
di: Stone, Gunner, et al.
Pubblicazione: (2025)
H$_{2}$OT: Hierarchical Hourglass Tokenizer for Efficient Video Pose Transformers
di: Li, Wenhao, et al.
Pubblicazione: (2025)
di: Li, Wenhao, et al.
Pubblicazione: (2025)
Layer- and Timestep-Adaptive Differentiable Token Compression Ratios for Efficient Diffusion Transformers
di: You, Haoran, et al.
Pubblicazione: (2024)
di: You, Haoran, et al.
Pubblicazione: (2024)
Robust Multimodal Learning via Cross-Modal Proxy Tokens
di: Reza, Md Kaykobad, et al.
Pubblicazione: (2025)
di: Reza, Md Kaykobad, et al.
Pubblicazione: (2025)
SkipViT: Speeding Up Vision Transformers with a Token-Level Skip Connection
di: Ataiefard, Foozhan, et al.
Pubblicazione: (2024)
di: Ataiefard, Foozhan, et al.
Pubblicazione: (2024)
Rethinking Token-wise Feature Caching: Accelerating Diffusion Transformers with Dual Feature Caching
di: Zou, Chang, et al.
Pubblicazione: (2024)
di: Zou, Chang, et al.
Pubblicazione: (2024)
Moto: Latent Motion Token as the Bridging Language for Learning Robot Manipulation from Videos
di: Chen, Yi, et al.
Pubblicazione: (2024)
di: Chen, Yi, et al.
Pubblicazione: (2024)
A General and Efficient Training for Transformer via Token Expansion
di: Huang, Wenxuan, et al.
Pubblicazione: (2024)
di: Huang, Wenxuan, et al.
Pubblicazione: (2024)
Diffusion Autoencoders are Scalable Image Tokenizers
di: Chen, Yinbo, et al.
Pubblicazione: (2025)
di: Chen, Yinbo, et al.
Pubblicazione: (2025)
Identifiable Token Correspondence for World Models
di: Kim, Youngin, et al.
Pubblicazione: (2026)
di: Kim, Youngin, et al.
Pubblicazione: (2026)
Language-Guided Image Tokenization for Generation
di: Zha, Kaiwen, et al.
Pubblicazione: (2024)
di: Zha, Kaiwen, et al.
Pubblicazione: (2024)
Uncertainty-DTW for Sequences and Visual Tokens
di: Wang, Lei, et al.
Pubblicazione: (2026)
di: Wang, Lei, et al.
Pubblicazione: (2026)
VORTEX: Challenging CNNs at Texture Recognition by using Vision Transformers with Orderless and Randomized Token Encodings
di: Scabini, Leonardo, et al.
Pubblicazione: (2025)
di: Scabini, Leonardo, et al.
Pubblicazione: (2025)
Masked Autoencoders Are Effective Tokenizers for Diffusion Models
di: Chen, Hao, et al.
Pubblicazione: (2025)
di: Chen, Hao, et al.
Pubblicazione: (2025)
Efficient World Models with Context-Aware Tokenization
di: Micheli, Vincent, et al.
Pubblicazione: (2024)
di: Micheli, Vincent, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Label-Free Cross-Task LoRA Merging with Null-Space Compression
di: Lee, Wonyoung, et al.
Pubblicazione: (2026) -
Synchronizing Task Behavior: Aligning Multiple Tasks during Test-Time Training
di: Jeong, Wooseong, et al.
Pubblicazione: (2025) -
Quantifying Task Priority for Multi-Task Optimization
di: Jeong, Wooseong, et al.
Pubblicazione: (2024) -
Preference-Aligned LoRA Merging: Preserving Subspace Coverage and Addressing Directional Anisotropy
di: Jeong, Wooseong, et al.
Pubblicazione: (2026) -
Interaction-Merged Motion Planning: Effectively Leveraging Diverse Motion Datasets for Robust Planning
di: Lee, Giwon, et al.
Pubblicazione: (2025)