Align then Adapt: Rethinking Parameter-Efficient Transfer Learning in 4D Perception
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Yiding, Zhu, Jihua, Cheng, Haozhe, Lu, Chaoyi, Yang, Zhichuan, Chen, Lin, Wang, Yaonan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PointDico: Contrastive 3D Representation Learning Guided by Diffusion Models
von: Li, Pengbo, et al.
Veröffentlicht: (2025)
von: Li, Pengbo, et al.
Veröffentlicht: (2025)
Point-SRA: Self-Representation Alignment for 3D Representation Learning
von: Wei, Lintong, et al.
Veröffentlicht: (2026)
von: Wei, Lintong, et al.
Veröffentlicht: (2026)
Hyperbolic Image-and-Pointcloud Contrastive Learning for 3D Classification
von: Hu, Naiwen, et al.
Veröffentlicht: (2024)
von: Hu, Naiwen, et al.
Veröffentlicht: (2024)
3D-JEPA: A Joint Embedding Predictive Architecture for 3D Self-Supervised Representation Learning
von: Hu, Naiwen, et al.
Veröffentlicht: (2024)
von: Hu, Naiwen, et al.
Veröffentlicht: (2024)
Chain-of-Thought Compression Should Not Be Blind: V-Skip for Efficient Multimodal Reasoning via Dual-Path Anchoring
von: Zhang, Dongxu, et al.
Veröffentlicht: (2026)
von: Zhang, Dongxu, et al.
Veröffentlicht: (2026)
PointRFT: Explicit Reinforcement Fine-tuning for Point Cloud Few-shot Learning
von: Wang, Yankai, et al.
Veröffentlicht: (2026)
von: Wang, Yankai, et al.
Veröffentlicht: (2026)
TikArt: Stabilizing Aperture-Guided Fine-Grained Visual Reasoning with Reinforcement Learning
von: Ding, Hao, et al.
Veröffentlicht: (2026)
von: Ding, Hao, et al.
Veröffentlicht: (2026)
Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval
von: Wang, Zhichuan, et al.
Veröffentlicht: (2025)
von: Wang, Zhichuan, et al.
Veröffentlicht: (2025)
Learning Coherent Matrixized Representation in Latent Space for Volumetric 4D Generation
von: Yang, Qitong, et al.
Veröffentlicht: (2024)
von: Yang, Qitong, et al.
Veröffentlicht: (2024)
CMHANet: A Cross-Modal Hybrid Attention Network for Point Cloud Registration
von: Zhang, Dongxu, et al.
Veröffentlicht: (2026)
von: Zhang, Dongxu, et al.
Veröffentlicht: (2026)
Learning to Learn Transferable Generative Attack for Person Re-Identification
von: Bian, Yuan, et al.
Veröffentlicht: (2024)
von: Bian, Yuan, et al.
Veröffentlicht: (2024)
Mantis: Mamba-native Tuning is Efficient for 3D Point Cloud Foundation Models
von: Guo, Zihao, et al.
Veröffentlicht: (2026)
von: Guo, Zihao, et al.
Veröffentlicht: (2026)
DINO Eats CLIP: Adapting Beyond Knowns for Open-set 3D Object Retrieval
von: He, Xinwei, et al.
Veröffentlicht: (2026)
von: He, Xinwei, et al.
Veröffentlicht: (2026)
Efficient Multimodal 3D Object Detector via Instance-Level Contrastive Distillation
von: Su, Zhuoqun, et al.
Veröffentlicht: (2025)
von: Su, Zhuoqun, et al.
Veröffentlicht: (2025)
Towards Real-World Aerial Vision Guidance with Categorical 6D Pose Tracker
von: Sun, Jingtao, et al.
Veröffentlicht: (2024)
von: Sun, Jingtao, et al.
Veröffentlicht: (2024)
Topo4D: Topology-Preserving Gaussian Splatting for High-Fidelity 4D Head Capture
von: Li, Xuanchen, et al.
Veröffentlicht: (2024)
von: Li, Xuanchen, et al.
Veröffentlicht: (2024)
Rethinking Efficient Crack Segmentation with Task-Aligned Structural-Directional Modeling
von: Liu, Shipeng, et al.
Veröffentlicht: (2026)
von: Liu, Shipeng, et al.
Veröffentlicht: (2026)
AgentAlign: Misalignment-Adapted Multi-Agent Perception for Resilient Inter-Agent Sensor Correlations
von: Meng, Zonglin, et al.
Veröffentlicht: (2024)
von: Meng, Zonglin, et al.
Veröffentlicht: (2024)
AdaptPrompt: Parameter-Efficient Adaptation of VLMs for Generalizable Deepfake Detection
von: Jiang, Yichen, et al.
Veröffentlicht: (2025)
von: Jiang, Yichen, et al.
Veröffentlicht: (2025)
UniPT: Universal Parallel Tuning for Transfer Learning with Efficient Parameter and Memory
von: Diao, Haiwen, et al.
Veröffentlicht: (2023)
von: Diao, Haiwen, et al.
Veröffentlicht: (2023)
Rethinking the Spatio-Temporal Alignment of End-to-End 3D Perception
von: Li, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Li, Xiaoyu, et al.
Veröffentlicht: (2025)
DENOISER: Rethinking the Robustness for Open-Vocabulary Action Recognition
von: Cheng, Haozhe, et al.
Veröffentlicht: (2024)
von: Cheng, Haozhe, et al.
Veröffentlicht: (2024)
Cross-Modal Adapter: Parameter-Efficient Transfer Learning Approach for Vision-Language Models
von: Yang, Juncheng, et al.
Veröffentlicht: (2024)
von: Yang, Juncheng, et al.
Veröffentlicht: (2024)
UP-Person: Unified Parameter-Efficient Transfer Learning for Text-based Person Retrieval
von: Liu, Yating, et al.
Veröffentlicht: (2025)
von: Liu, Yating, et al.
Veröffentlicht: (2025)
Can Video Diffusion Model Reconstruct 4D Geometry?
von: Mai, Jinjie, et al.
Veröffentlicht: (2025)
von: Mai, Jinjie, et al.
Veröffentlicht: (2025)
Percept, Chat, and then Adapt: Multimodal Knowledge Transfer of Foundation Models for Open-World Video Recognition
von: Chen, Boyu, et al.
Veröffentlicht: (2024)
von: Chen, Boyu, et al.
Veröffentlicht: (2024)
Dynamic Adapter Meets Prompt Tuning: Parameter-Efficient Transfer Learning for Point Cloud Analysis
von: Zhou, Xin, et al.
Veröffentlicht: (2024)
von: Zhou, Xin, et al.
Veröffentlicht: (2024)
PTQAT: A Hybrid Parameter-Efficient Quantization Algorithm for 3D Perception Tasks
von: Wang, Xinhao, et al.
Veröffentlicht: (2025)
von: Wang, Xinhao, et al.
Veröffentlicht: (2025)
Learning More by Seeing Less: Structure First Learning for Efficient, Transferable, and Human-Aligned Vision
von: Li, Tianqin, et al.
Veröffentlicht: (2025)
von: Li, Tianqin, et al.
Veröffentlicht: (2025)
Rethinking Model Ensemble in Transfer-based Adversarial Attacks
von: Chen, Huanran, et al.
Veröffentlicht: (2023)
von: Chen, Huanran, et al.
Veröffentlicht: (2023)
VMBench: A Benchmark for Perception-Aligned Video Motion Generation
von: Ling, Xinran, et al.
Veröffentlicht: (2025)
von: Ling, Xinran, et al.
Veröffentlicht: (2025)
Rethink Arbitrary Style Transfer with Transformer and Contrastive Learning
von: Zhang, Zhanjie, et al.
Veröffentlicht: (2024)
von: Zhang, Zhanjie, et al.
Veröffentlicht: (2024)
IV-tuning: Parameter-Efficient Transfer Learning for Infrared-Visible Tasks
von: Zhang, Yaming, et al.
Veröffentlicht: (2024)
von: Zhang, Yaming, et al.
Veröffentlicht: (2024)
PETALface: Parameter Efficient Transfer Learning for Low-resolution Face Recognition
von: Narayan, Kartik, et al.
Veröffentlicht: (2024)
von: Narayan, Kartik, et al.
Veröffentlicht: (2024)
AFBS:Buffer Gradient Selection in Semi-asynchronous Federated Learning
von: Lu, Chaoyi, et al.
Veröffentlicht: (2025)
von: Lu, Chaoyi, et al.
Veröffentlicht: (2025)
FedPSA: Modeling Behavioral Staleness in Asynchronous Federated Learning
von: Lu, Chaoyi, et al.
Veröffentlicht: (2026)
von: Lu, Chaoyi, et al.
Veröffentlicht: (2026)
VolSplat: Rethinking Feed-Forward 3D Gaussian Splatting with Voxel-Aligned Prediction
von: Wang, Weijie, et al.
Veröffentlicht: (2025)
von: Wang, Weijie, et al.
Veröffentlicht: (2025)
SurgPETL: Parameter-Efficient Image-to-Surgical-Video Transfer Learning for Surgical Phase Recognition
von: Yang, Shu, et al.
Veröffentlicht: (2024)
von: Yang, Shu, et al.
Veröffentlicht: (2024)
Starve to Perceive: Taming Lazy Perception in VLMs with Constrained Visual Bandwidth
von: Wu, Yuhuan, et al.
Veröffentlicht: (2026)
von: Wu, Yuhuan, et al.
Veröffentlicht: (2026)
Parameters as Experts: Adapting Vision Models with Dynamic Parameter Routing
von: Lou, Meng, et al.
Veröffentlicht: (2026)
von: Lou, Meng, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
PointDico: Contrastive 3D Representation Learning Guided by Diffusion Models
von: Li, Pengbo, et al.
Veröffentlicht: (2025) -
Point-SRA: Self-Representation Alignment for 3D Representation Learning
von: Wei, Lintong, et al.
Veröffentlicht: (2026) -
Hyperbolic Image-and-Pointcloud Contrastive Learning for 3D Classification
von: Hu, Naiwen, et al.
Veröffentlicht: (2024) -
3D-JEPA: A Joint Embedding Predictive Architecture for 3D Self-Supervised Representation Learning
von: Hu, Naiwen, et al.
Veröffentlicht: (2024) -
Chain-of-Thought Compression Should Not Be Blind: V-Skip for Efficient Multimodal Reasoning via Dual-Path Anchoring
von: Zhang, Dongxu, et al.
Veröffentlicht: (2026)