Unprejudiced Training Auxiliary Tasks Makes Primary Better: A Multi-Task Learning Perspective
Fuente:
arXiv
Guardado en:
| Autores principales: | Li, Yuanze, Feng, Chun-Mei, Wang, Qilong, Yang, Guanglei, Zuo, Wangmeng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Segmenting Objectiveness and Task-awareness Unknown Region for Autonomous Driving
por: Zheng, Mi, et al.
Publicado: (2025)
por: Zheng, Mi, et al.
Publicado: (2025)
UniM$^2$AE: Multi-modal Masked Autoencoders with Unified 3D Representation for 3D Perception in Autonomous Driving
por: Zou, Jian, et al.
Publicado: (2023)
por: Zou, Jian, et al.
Publicado: (2023)
MetricDepth: Enhancing Monocular Depth Estimation with Deep Metric Learning
por: Liu, Chunpu, et al.
Publicado: (2024)
por: Liu, Chunpu, et al.
Publicado: (2024)
FedSmoothLoRA: Toward Smoother and Faster Convergence in Federated Low-Rank Adaptation
por: Wang, Zehao, et al.
Publicado: (2026)
por: Wang, Zehao, et al.
Publicado: (2026)
ConSept: Continual Semantic Segmentation via Adapter-based Vision Transformer
por: Dong, Bowen, et al.
Publicado: (2024)
por: Dong, Bowen, et al.
Publicado: (2024)
Multi-Modality Driven LoRA for Adverse Condition Depth Estimation
por: Yang, Guanglei, et al.
Publicado: (2024)
por: Yang, Guanglei, et al.
Publicado: (2024)
Class Balance Matters to Active Class-Incremental Learning
por: Huang, Zitong, et al.
Publicado: (2024)
por: Huang, Zitong, et al.
Publicado: (2024)
MR-GDINO: Efficient Open-World Continual Object Detection
por: Dong, Bowen, et al.
Publicado: (2024)
por: Dong, Bowen, et al.
Publicado: (2024)
FILP-3D: Enhancing 3D Few-shot Class-incremental Learning with Pre-trained Vision-Language Models
por: Xu, Wan, et al.
Publicado: (2023)
por: Xu, Wan, et al.
Publicado: (2023)
Bridging Geometry-Coherent Text-to-3D Generation with Multi-View Diffusion Priors and Gaussian Splatting
por: Yang, Feng, et al.
Publicado: (2025)
por: Yang, Feng, et al.
Publicado: (2025)
Generative Inbetweening through Frame-wise Conditions-Driven Video Generation
por: Zhu, Tianyi, et al.
Publicado: (2024)
por: Zhu, Tianyi, et al.
Publicado: (2024)
Improving Vessel Segmentation with Multi-Task Learning and Auxiliary Data Available Only During Model Training
por: Sobotka, Daniel, et al.
Publicado: (2025)
por: Sobotka, Daniel, et al.
Publicado: (2025)
Text to Image for Multi-Label Image Recognition with Joint Prompt-Adapter Learning
por: Feng, Chun-Mei, et al.
Publicado: (2025)
por: Feng, Chun-Mei, et al.
Publicado: (2025)
Enhancing Monocular Depth Estimation with Multi-Source Auxiliary Tasks
por: Quercia, Alessio, et al.
Publicado: (2025)
por: Quercia, Alessio, et al.
Publicado: (2025)
Enhancing Visual Planning with Auxiliary Tasks and Multi-token Prediction
por: Zhang, Ce, et al.
Publicado: (2025)
por: Zhang, Ce, et al.
Publicado: (2025)
IMWA: Iterative Model Weight Averaging Benefits Class-Imbalanced Learning Tasks
por: Huang, Zitong, et al.
Publicado: (2024)
por: Huang, Zitong, et al.
Publicado: (2024)
Triad: Empowering LMM-based Anomaly Detection with Vision Expert-guided Visual Tokenizer and Manufacturing Process
por: Li, Yuanze, et al.
Publicado: (2025)
por: Li, Yuanze, et al.
Publicado: (2025)
DISPLAY: Directable Human-Object Interaction Video Generation via Sparse Motion Guidance and Multi-Task Auxiliary
por: Guan, Jiazhi, et al.
Publicado: (2026)
por: Guan, Jiazhi, et al.
Publicado: (2026)
MVS-TTA: Test-Time Adaptation for Multi-View Stereo via Meta-Auxiliary Learning
por: Zhang, Hannuo, et al.
Publicado: (2025)
por: Zhang, Hannuo, et al.
Publicado: (2025)
RoomEditor++: A Parameter-Sharing Diffusion Architecture for High-Fidelity Furniture Synthesis
por: Wang, Qilong, et al.
Publicado: (2025)
por: Wang, Qilong, et al.
Publicado: (2025)
MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM
por: Dong, Bowen, et al.
Publicado: (2025)
por: Dong, Bowen, et al.
Publicado: (2025)
SelfDRSC++: Self-Supervised Learning for Dual Reversed Rolling Shutter Correction
por: Shang, Wei, et al.
Publicado: (2024)
por: Shang, Wei, et al.
Publicado: (2024)
Image Demoiréing Using Dual Camera Fusion on Mobile Phones
por: Mei, Yanting, et al.
Publicado: (2025)
por: Mei, Yanting, et al.
Publicado: (2025)
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning
por: Li, Wenrui, et al.
Publicado: (2025)
por: Li, Wenrui, et al.
Publicado: (2025)
Olympus: A Universal Task Router for Computer Vision Tasks
por: Lin, Yuanze, et al.
Publicado: (2024)
por: Lin, Yuanze, et al.
Publicado: (2024)
Reblurring-Guided Single Image Defocus Deblurring: A Learning Framework with Misaligned Training Pairs
por: Ren, Dongwei, et al.
Publicado: (2024)
por: Ren, Dongwei, et al.
Publicado: (2024)
LPT++: Efficient Training on Mixture of Long-tailed Experts
por: Dong, Bowen, et al.
Publicado: (2024)
por: Dong, Bowen, et al.
Publicado: (2024)
Auxiliary Tasks Enhanced Dual-affinity Learning for Weakly Supervised Semantic Segmentation
por: Xu, Lian, et al.
Publicado: (2024)
por: Xu, Lian, et al.
Publicado: (2024)
MAN++: Scaling Momentum Auxiliary Network for Supervised Local Learning in Vision Tasks
por: Su, Junhao, et al.
Publicado: (2025)
por: Su, Junhao, et al.
Publicado: (2025)
Myriad: Large Multimodal Model by Applying Vision Experts for Industrial Anomaly Detection
por: Li, Yuanze, et al.
Publicado: (2023)
por: Li, Yuanze, et al.
Publicado: (2023)
Visual-O1: Understanding Ambiguous Instructions via Multi-modal Multi-turn Chain-of-thoughts Reasoning
por: Ni, Minheng, et al.
Publicado: (2024)
por: Ni, Minheng, et al.
Publicado: (2024)
SplatWeaver: Learning to Allocate Gaussian Primitives for Generalizable Novel View Synthesis
por: Wan, Yecong, et al.
Publicado: (2026)
por: Wan, Yecong, et al.
Publicado: (2026)
Lie Flow: Video Dynamic Fields Modeling and Predicting with Lie Algebra as Geometric Physics Principle
por: Qiao, Weidong, et al.
Publicado: (2026)
por: Qiao, Weidong, et al.
Publicado: (2026)
AR-GRPO: Training Autoregressive Image Generation Models via Reinforcement Learning
por: Yuan, Shihao, et al.
Publicado: (2025)
por: Yuan, Shihao, et al.
Publicado: (2025)
GLAD: Towards Better Reconstruction with Global and Local Adaptive Diffusion Models for Unsupervised Anomaly Detection
por: Yao, Hang, et al.
Publicado: (2024)
por: Yao, Hang, et al.
Publicado: (2024)
Self-Supervised Learning for Real-World Super-Resolution from Dual and Multiple Zoomed Observations
por: Zhang, Zhilu, et al.
Publicado: (2024)
por: Zhang, Zhilu, et al.
Publicado: (2024)
CTA: Cross-Task Alignment for Better Test Time Training
por: Barbeau, Samuel, et al.
Publicado: (2025)
por: Barbeau, Samuel, et al.
Publicado: (2025)
Embodied Navigation with Auxiliary Task of Action Description Prediction
por: Kondoh, Haru, et al.
Publicado: (2025)
por: Kondoh, Haru, et al.
Publicado: (2025)
LiSD: An Efficient Multi-Task Learning Framework for LiDAR Segmentation and Detection
por: Xu, Jiahua, et al.
Publicado: (2024)
por: Xu, Jiahua, et al.
Publicado: (2024)
On the Limits of Multi-modal Meta-Learning with Auxiliary Task Modulation Using Conditional Batch Normalization
por: Armengol-Estapé, Jordi, et al.
Publicado: (2024)
por: Armengol-Estapé, Jordi, et al.
Publicado: (2024)
Ejemplares similares
-
Segmenting Objectiveness and Task-awareness Unknown Region for Autonomous Driving
por: Zheng, Mi, et al.
Publicado: (2025) -
UniM$^2$AE: Multi-modal Masked Autoencoders with Unified 3D Representation for 3D Perception in Autonomous Driving
por: Zou, Jian, et al.
Publicado: (2023) -
MetricDepth: Enhancing Monocular Depth Estimation with Deep Metric Learning
por: Liu, Chunpu, et al.
Publicado: (2024) -
FedSmoothLoRA: Toward Smoother and Faster Convergence in Federated Low-Rank Adaptation
por: Wang, Zehao, et al.
Publicado: (2026) -
ConSept: Continual Semantic Segmentation via Adapter-based Vision Transformer
por: Dong, Bowen, et al.
Publicado: (2024)