Spectral-Adaptive Modulation Networks for Visual Perception
Fuente:
arXiv
Salvato in:
| Autori principali: | Yun, Guhnoo, Yoo, Juhan, Kim, Kijung, Lee, Jeongho, Seo, Paul Hongsuck, Kim, Dong Hwan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Initiation of Interaction Detection Framework using a Nonverbal Cue for Human-Robot Interaction
di: Yun, Guhnoo, et al.
Pubblicazione: (2026)
di: Yun, Guhnoo, et al.
Pubblicazione: (2026)
Robust Image Self-Recovery against Tampering using Watermark Generation with Pixel Shuffling
di: Kim, Minyoung, et al.
Pubblicazione: (2025)
di: Kim, Minyoung, et al.
Pubblicazione: (2025)
Breaking the Visual Shortcuts in Multimodal Knowledge-Based Visual Question Answering
di: Lee, Dosung, et al.
Pubblicazione: (2025)
di: Lee, Dosung, et al.
Pubblicazione: (2025)
Image Diffusion Models Exhibit Emergent Temporal Propagation in Videos
di: Kim, Youngseo, et al.
Pubblicazione: (2025)
di: Kim, Youngseo, et al.
Pubblicazione: (2025)
Direct Diffusion Score Preference Optimization via Stepwise Contrastive Policy-Pair Supervision
di: Kim, Dohyun, et al.
Pubblicazione: (2025)
di: Kim, Dohyun, et al.
Pubblicazione: (2025)
Learning Correlation Structures for Vision Transformers
di: Kim, Manjin, et al.
Pubblicazione: (2024)
di: Kim, Manjin, et al.
Pubblicazione: (2024)
Multi-Granularity Video Object Segmentation
di: Lim, Sangbeom, et al.
Pubblicazione: (2024)
di: Lim, Sangbeom, et al.
Pubblicazione: (2024)
Bridging Audio and Vision: Zero-Shot Audiovisual Segmentation by Connecting Pretrained Models
di: Lee, Seung-jae, et al.
Pubblicazione: (2025)
di: Lee, Seung-jae, et al.
Pubblicazione: (2025)
Bridging the Domain Gap: A Simple Domain Matching Method for Reference-based Image Super-Resolution in Remote Sensing
di: Min, Jeongho, et al.
Pubblicazione: (2024)
di: Min, Jeongho, et al.
Pubblicazione: (2024)
Towards Open-Vocabulary Semantic Segmentation Without Semantic Labels
di: Shin, Heeseong, et al.
Pubblicazione: (2024)
di: Shin, Heeseong, et al.
Pubblicazione: (2024)
TranSplat: Surface Embedding-guided 3D Gaussian Splatting for Transparent Object Manipulation
di: Kim, Jeongyun, et al.
Pubblicazione: (2025)
di: Kim, Jeongyun, et al.
Pubblicazione: (2025)
CAT-Seg: Cost Aggregation for Open-Vocabulary Semantic Segmentation
di: Cho, Seokju, et al.
Pubblicazione: (2023)
di: Cho, Seokju, et al.
Pubblicazione: (2023)
Pseudo-RIS: Distinctive Pseudo-supervision Generation for Referring Image Segmentation
di: Yu, Seonghoon, et al.
Pubblicazione: (2024)
di: Yu, Seonghoon, et al.
Pubblicazione: (2024)
Hybrid-Vector Retrieval for Visually Rich Documents: Combining Single-Vector Efficiency and Multi-Vector Accuracy
di: Kim, Juyeon, et al.
Pubblicazione: (2025)
di: Kim, Juyeon, et al.
Pubblicazione: (2025)
Seg4Diff: Unveiling Open-Vocabulary Segmentation in Text-to-Image Diffusion Transformers
di: Kim, Chaehyun, et al.
Pubblicazione: (2025)
di: Kim, Chaehyun, et al.
Pubblicazione: (2025)
From Street to Orbit: Training-Free Cross-View Retrieval via Location Semantics and LLM Guidance
di: Min, Jeongho, et al.
Pubblicazione: (2025)
di: Min, Jeongho, et al.
Pubblicazione: (2025)
DialNav: Multi-turn Dialog Navigation with a Remote Guide
di: Han, Leekyeung, et al.
Pubblicazione: (2025)
di: Han, Leekyeung, et al.
Pubblicazione: (2025)
TCAN: Animating Human Images with Temporally Consistent Pose Guidance using Diffusion Models
di: Kim, Jeongho, et al.
Pubblicazione: (2024)
di: Kim, Jeongho, et al.
Pubblicazione: (2024)
H2G: Hierarchy-Aware Hyperbolic Grouping for 3D Scenes
di: Ko, ByungHa, et al.
Pubblicazione: (2026)
di: Ko, ByungHa, et al.
Pubblicazione: (2026)
Clutt3R-Seg: Sparse-view 3D Instance Segmentation for Language-grounded Grasping in Cluttered Scenes
di: Noh, Jeongho, et al.
Pubblicazione: (2026)
di: Noh, Jeongho, et al.
Pubblicazione: (2026)
From Wardrobe to Canvas: Wardrobe Polyptych LoRA for Part-level Controllable Human Image Generation
di: Kim, Jeongho, et al.
Pubblicazione: (2025)
di: Kim, Jeongho, et al.
Pubblicazione: (2025)
VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models
di: Ju, Jeongho, et al.
Pubblicazione: (2024)
di: Ju, Jeongho, et al.
Pubblicazione: (2024)
TERDNet: Transformer Encoder-Recurrent Decoder Network for Scene Change Detection
di: Yoon, Jiae, et al.
Pubblicazione: (2026)
di: Yoon, Jiae, et al.
Pubblicazione: (2026)
What to Preserve and What to Transfer: Faithful, Identity-Preserving Diffusion-based Hairstyle Transfer
di: Chung, Chaeyeon, et al.
Pubblicazione: (2024)
di: Chung, Chaeyeon, et al.
Pubblicazione: (2024)
RePaintGS: Reference-Guided Gaussian Splatting for Realistic and View-Consistent 3D Scene Inpainting
di: Seo, Ji Hyun, et al.
Pubblicazione: (2025)
di: Seo, Ji Hyun, et al.
Pubblicazione: (2025)
Robust Multimodal 3D Object Detection via Modality-Agnostic Decoding and Proximity-based Modality Ensemble
di: Cha, Juhan, et al.
Pubblicazione: (2024)
di: Cha, Juhan, et al.
Pubblicazione: (2024)
UniSpector: Towards Universal Open-set Defect Recognition via Spectral-Contrastive Visual Prompting
di: Kim, Geonuk, et al.
Pubblicazione: (2026)
di: Kim, Geonuk, et al.
Pubblicazione: (2026)
BF-STVSR: B-Splines and Fourier-Best Friends for High Fidelity Spatial-Temporal Video Super-Resolution
di: Kim, Eunjin, et al.
Pubblicazione: (2025)
di: Kim, Eunjin, et al.
Pubblicazione: (2025)
Infinite-Homography as Robust Conditioning for Camera-Controlled Video Generation
di: Kim, Min-Jung, et al.
Pubblicazione: (2025)
di: Kim, Min-Jung, et al.
Pubblicazione: (2025)
Fieldscale: Locality-Aware Field-based Adaptive Rescaling for Thermal Infrared Image
di: Gil, Hyeonjae, et al.
Pubblicazione: (2024)
di: Gil, Hyeonjae, et al.
Pubblicazione: (2024)
Memory-Efficient Personalization of Text-to-Image Diffusion Models via Selective Optimization Strategies
di: Choi, Seokeon, et al.
Pubblicazione: (2025)
di: Choi, Seokeon, et al.
Pubblicazione: (2025)
PlugTrack: Multi-Perceptive Motion Analysis for Adaptive Fusion in Multi-Object Tracking
di: Kim, Seungjae, et al.
Pubblicazione: (2025)
di: Kim, Seungjae, et al.
Pubblicazione: (2025)
ERASE: Eliminating Redundant Visual Tokens via Adaptive Two-Stage Token Pruning
di: Lee, Yuna, et al.
Pubblicazione: (2026)
di: Lee, Yuna, et al.
Pubblicazione: (2026)
On Efficient Language and Vision Assistants for Visually-Situated Natural Language Understanding: What Matters in Reading and Reasoning
di: Kim, Geewook, et al.
Pubblicazione: (2024)
di: Kim, Geewook, et al.
Pubblicazione: (2024)
Degradation-Agnostic Statistical Facial Feature Transformation for Blind Face Restoration in Adverse Weather Conditions
di: Son, Chang-Hwan, et al.
Pubblicazione: (2025)
di: Son, Chang-Hwan, et al.
Pubblicazione: (2025)
Superpixel Tokenization for Vision Transformers: Preserving Semantic Integrity in Visual Tokens
di: Lew, Jaihyun, et al.
Pubblicazione: (2024)
di: Lew, Jaihyun, et al.
Pubblicazione: (2024)
JoVALE: Detecting Human Actions in Video Using Audiovisual and Language Contexts
di: Son, Taein, et al.
Pubblicazione: (2024)
di: Son, Taein, et al.
Pubblicazione: (2024)
Masked Autoregressive Model for Weather Forecasting
di: Kim, Doyi, et al.
Pubblicazione: (2024)
di: Kim, Doyi, et al.
Pubblicazione: (2024)
From Adaptation to Generalization: Adaptive Visual Prompting for Medical Image Segmentation
di: Çetinkaya, Evren, et al.
Pubblicazione: (2026)
di: Çetinkaya, Evren, et al.
Pubblicazione: (2026)
A Decoding Scheme with Successive Aggregation of Multi-Level Features for Light-Weight Semantic Segmentation
di: Yoo, Jiwon, et al.
Pubblicazione: (2024)
di: Yoo, Jiwon, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Initiation of Interaction Detection Framework using a Nonverbal Cue for Human-Robot Interaction
di: Yun, Guhnoo, et al.
Pubblicazione: (2026) -
Robust Image Self-Recovery against Tampering using Watermark Generation with Pixel Shuffling
di: Kim, Minyoung, et al.
Pubblicazione: (2025) -
Breaking the Visual Shortcuts in Multimodal Knowledge-Based Visual Question Answering
di: Lee, Dosung, et al.
Pubblicazione: (2025) -
Image Diffusion Models Exhibit Emergent Temporal Propagation in Videos
di: Kim, Youngseo, et al.
Pubblicazione: (2025) -
Direct Diffusion Score Preference Optimization via Stepwise Contrastive Policy-Pair Supervision
di: Kim, Dohyun, et al.
Pubblicazione: (2025)