Text-Guided Multi-Scale Frequency Representation Adaptation
Fuente:
arXiv
Guardado en:
| Autores principales: | Yan, Weicai, Ma, Xinhua, Lin, Wang, Jin, Tao |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Efficient Prompting for Continual Adaptation to Missing Modalities
por: Guo, Zirun, et al.
Publicado: (2025)
por: Guo, Zirun, et al.
Publicado: (2025)
Advancing Comprehensive Aesthetic Insight with Multi-Scale Text-Guided Self-Supervised Learning
por: Liu, Yuti, et al.
Publicado: (2024)
por: Liu, Yuti, et al.
Publicado: (2024)
Representation Surgery for Multi-Task Model Merging
por: Yang, Enneng, et al.
Publicado: (2024)
por: Yang, Enneng, et al.
Publicado: (2024)
Visual Explanations of Image-Text Representations via Multi-Modal Information Bottleneck Attribution
por: Wang, Ying, et al.
Publicado: (2023)
por: Wang, Ying, et al.
Publicado: (2023)
Joint Memory Frequency and Computing Frequency Scaling for Energy-efficient DNN Inference
por: Han, Yunchu, et al.
Publicado: (2025)
por: Han, Yunchu, et al.
Publicado: (2025)
Revisiting Multimodal KV Cache Compression: A Frequency-Domain-Guided Outlier-KV-Aware Approach
por: Yang, Yaoxin, et al.
Publicado: (2025)
por: Yang, Yaoxin, et al.
Publicado: (2025)
M4V: Multi-Modal Mamba for Text-to-Video Generation
por: Huang, Jiancheng, et al.
Publicado: (2025)
por: Huang, Jiancheng, et al.
Publicado: (2025)
CycleNet: Rethinking Cycle Consistency in Text-Guided Diffusion for Image Manipulation
por: Xu, Sihan, et al.
Publicado: (2023)
por: Xu, Sihan, et al.
Publicado: (2023)
Advancing Multimodal Large Language Models with Quantization-Aware Scale Learning for Efficient Adaptation
por: Xie, Jingjing, et al.
Publicado: (2024)
por: Xie, Jingjing, et al.
Publicado: (2024)
Text-to-Image GAN with Pretrained Representations
por: You, Xiaozhou, et al.
Publicado: (2024)
por: You, Xiaozhou, et al.
Publicado: (2024)
Uncertainty-Guided Selective Adaptation Enables Cross-Platform Predictive Fluorescence Microscopy
por: Yang, Kai-Wen K., et al.
Publicado: (2025)
por: Yang, Kai-Wen K., et al.
Publicado: (2025)
Integrating Frequency Guidance into Multi-source Domain Generalization for Bearing Fault Diagnosis
por: Tu, Xiaotong, et al.
Publicado: (2025)
por: Tu, Xiaotong, et al.
Publicado: (2025)
Scaling 4D Representations
por: Carreira, João, et al.
Publicado: (2024)
por: Carreira, João, et al.
Publicado: (2024)
DohaScript: A Large-Scale Multi-Writer Dataset for Continuous Handwritten Hindi Text
por: Singh, Kunwar Arpit, et al.
Publicado: (2026)
por: Singh, Kunwar Arpit, et al.
Publicado: (2026)
Semantically Guided Representation Learning For Action Anticipation
por: Diko, Anxhelo, et al.
Publicado: (2024)
por: Diko, Anxhelo, et al.
Publicado: (2024)
MedSegFactory: Text-Guided Generation of Medical Image-Mask Pairs
por: Mao, Jiawei, et al.
Publicado: (2025)
por: Mao, Jiawei, et al.
Publicado: (2025)
Invariant Representation Guided Multimodal Sentiment Decoding with Sequential Variation Regularization
por: Xu, Guoyang, et al.
Publicado: (2024)
por: Xu, Guoyang, et al.
Publicado: (2024)
SurgeryV2: Bridging the Gap Between Model Merging and Multi-Task Learning with Deep Representation Surgery
por: Yang, Enneng, et al.
Publicado: (2024)
por: Yang, Enneng, et al.
Publicado: (2024)
Decoupling Amplitude and Phase Attention in Frequency Domain for RGB-Event based Visual Object Tracking
por: Wang, Shiao, et al.
Publicado: (2026)
por: Wang, Shiao, et al.
Publicado: (2026)
Compositional Text-to-Image Generation with Dense Blob Representations
por: Nie, Weili, et al.
Publicado: (2024)
por: Nie, Weili, et al.
Publicado: (2024)
Source-Free Domain Adaptation with Diffusion-Guided Source Data Generation
por: Chopra, Shivang, et al.
Publicado: (2024)
por: Chopra, Shivang, et al.
Publicado: (2024)
MFAF: An EVA02-Based Multi-scale Frequency Attention Fusion Method for Cross-View Geo-Localization
por: Liu, YiTong, et al.
Publicado: (2025)
por: Liu, YiTong, et al.
Publicado: (2025)
Implicit Contrastive Representation Learning with Guided Stop-gradient
por: Lee, Byeongchan, et al.
Publicado: (2025)
por: Lee, Byeongchan, et al.
Publicado: (2025)
Superclass-Guided Representation Disentanglement for Spurious Correlation Mitigation
por: Liu, Chenruo, et al.
Publicado: (2025)
por: Liu, Chenruo, et al.
Publicado: (2025)
Uncertainty Quantification via Hölder Divergence for Multi-View Representation Learning
por: Zhang, Yan, et al.
Publicado: (2024)
por: Zhang, Yan, et al.
Publicado: (2024)
SuperLoRA: Parameter-Efficient Unified Adaptation of Multi-Layer Attention Modules
por: Chen, Xiangyu, et al.
Publicado: (2024)
por: Chen, Xiangyu, et al.
Publicado: (2024)
Point Cloud Matters: Rethinking the Impact of Different Observation Spaces on Robot Learning
por: Zhu, Haoyi, et al.
Publicado: (2024)
por: Zhu, Haoyi, et al.
Publicado: (2024)
Consistent Flow Distillation for Text-to-3D Generation
por: Yan, Runjie, et al.
Publicado: (2025)
por: Yan, Runjie, et al.
Publicado: (2025)
Contextualized Diffusion Models for Text-Guided Image and Video Generation
por: Yang, Ling, et al.
Publicado: (2024)
por: Yang, Ling, et al.
Publicado: (2024)
DenseTRF: Texture-Aware Unsupervised Representation Adaptation for Surgical Scene Dense Prediction
por: Liao, Guiqiu, et al.
Publicado: (2026)
por: Liao, Guiqiu, et al.
Publicado: (2026)
ShapeWords: Guiding Text-to-Image Synthesis with 3D Shape-Aware Prompts
por: Petrov, Dmitry, et al.
Publicado: (2024)
por: Petrov, Dmitry, et al.
Publicado: (2024)
Alignment-Guided Score Matching for Text-to-Image Alignment in Diffusion Models
por: Lee, Jaa-Yeon, et al.
Publicado: (2026)
por: Lee, Jaa-Yeon, et al.
Publicado: (2026)
EditInspector: A Benchmark for Evaluation of Text-Guided Image Edits
por: Yosef, Ron, et al.
Publicado: (2025)
por: Yosef, Ron, et al.
Publicado: (2025)
RL for Consistency Models: Faster Reward Guided Text-to-Image Generation
por: Oertell, Owen, et al.
Publicado: (2024)
por: Oertell, Owen, et al.
Publicado: (2024)
MMCORE: MultiModal COnnection with Representation Aligned Latent Embeddings
por: Li, Zijie, et al.
Publicado: (2026)
por: Li, Zijie, et al.
Publicado: (2026)
Guiding Diffusion-based Reconstruction with Contrastive Signals for Balanced Visual Representation
por: Han, Boyu, et al.
Publicado: (2026)
por: Han, Boyu, et al.
Publicado: (2026)
A Survey on Cache Methods in Diffusion Models: Toward Efficient Multi-Modal Generation
por: Liu, Jiacheng, et al.
Publicado: (2025)
por: Liu, Jiacheng, et al.
Publicado: (2025)
FACL-Attack: Frequency-Aware Contrastive Learning for Transferable Adversarial Attacks
por: Yang, Hunmin, et al.
Publicado: (2024)
por: Yang, Hunmin, et al.
Publicado: (2024)
Orchestrate Latent Expertise: Advancing Online Continual Learning with Multi-Level Supervision and Reverse Self-Distillation
por: Yan, HongWei, et al.
Publicado: (2024)
por: Yan, HongWei, et al.
Publicado: (2024)
Exploring Text-to-Motion Generation with Human Preference
por: Sheng, Jenny, et al.
Publicado: (2024)
por: Sheng, Jenny, et al.
Publicado: (2024)
Ejemplares similares
-
Efficient Prompting for Continual Adaptation to Missing Modalities
por: Guo, Zirun, et al.
Publicado: (2025) -
Advancing Comprehensive Aesthetic Insight with Multi-Scale Text-Guided Self-Supervised Learning
por: Liu, Yuti, et al.
Publicado: (2024) -
Representation Surgery for Multi-Task Model Merging
por: Yang, Enneng, et al.
Publicado: (2024) -
Visual Explanations of Image-Text Representations via Multi-Modal Information Bottleneck Attribution
por: Wang, Ying, et al.
Publicado: (2023) -
Joint Memory Frequency and Computing Frequency Scaling for Energy-efficient DNN Inference
por: Han, Yunchu, et al.
Publicado: (2025)