Gespeichert in:
| Hauptverfasser: | Jang, Jiho, Kim, Jinyoung, Baek, Kyungjune, Kwak, Nojun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2512.10237 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MomentMix Augmentation with Length-Aware DETR for Temporally Robust Moment Retrieval
von: Park, Seojeong, et al.
Veröffentlicht: (2024)
von: Park, Seojeong, et al.
Veröffentlicht: (2024)
PedaCo-Gen: Scaffolding Pedagogical Agency in Human-AI Collaborative Video Authoring
von: Baek, Injun, et al.
Veröffentlicht: (2026)
von: Baek, Injun, et al.
Veröffentlicht: (2026)
Rethinking Direct Preference Optimization in Diffusion Models
von: Kang, Junyong, et al.
Veröffentlicht: (2025)
von: Kang, Junyong, et al.
Veröffentlicht: (2025)
ConcatPlexer: Additional Dim1 Batching for Faster ViTs
von: Han, Donghoon, et al.
Veröffentlicht: (2023)
von: Han, Donghoon, et al.
Veröffentlicht: (2023)
From Local to Global to Mechanistic: An iERF-Centered Unified Framework for Interpreting Vision Models
von: Kim, Yearim, et al.
Veröffentlicht: (2026)
von: Kim, Yearim, et al.
Veröffentlicht: (2026)
Conservative Generator, Progressive Discriminator: Coordination of Adversaries in Few-shot Incremental Image Synthesis
von: Kong, Chaerin, et al.
Veröffentlicht: (2022)
von: Kong, Chaerin, et al.
Veröffentlicht: (2022)
CAMEO: Correspondence-Attention Alignment for Multi-View Diffusion Models
von: Kwon, Minkyung, et al.
Veröffentlicht: (2025)
von: Kwon, Minkyung, et al.
Veröffentlicht: (2025)
DivCon-NeRF: Diverse and Consistent Ray Augmentation for Few-Shot NeRF
von: Lee, Ingyun, et al.
Veröffentlicht: (2025)
von: Lee, Ingyun, et al.
Veröffentlicht: (2025)
Causal Interpretation of Sparse Autoencoder Features in Vision
von: Han, Sangyu, et al.
Veröffentlicht: (2025)
von: Han, Sangyu, et al.
Veröffentlicht: (2025)
Respect the model: Fine-grained and Robust Explanation with Sharing Ratio Decomposition
von: Han, Sangyu, et al.
Veröffentlicht: (2024)
von: Han, Sangyu, et al.
Veröffentlicht: (2024)
4DGS360: 360° Gaussian Reconstruction of Dynamic Objects from a Single Video
von: Jang, Jae Won, et al.
Veröffentlicht: (2026)
von: Jang, Jae Won, et al.
Veröffentlicht: (2026)
Projected Representation Conditioning for High-fidelity Novel View Synthesis
von: Kwak, Min-Seop, et al.
Veröffentlicht: (2026)
von: Kwak, Min-Seop, et al.
Veröffentlicht: (2026)
Toward Stable World Models: Measuring and Addressing World Instability in Generative Environments
von: Kwon, Soonwoo, et al.
Veröffentlicht: (2025)
von: Kwon, Soonwoo, et al.
Veröffentlicht: (2025)
VDPP: Video Depth Post-Processing for Speed and Scalability
von: Yoon, Daewon, et al.
Veröffentlicht: (2026)
von: Yoon, Daewon, et al.
Veröffentlicht: (2026)
Decompose the model: Mechanistic interpretability in image models with Generalized Integrated Gradients (GIG)
von: Kim, Yearim, et al.
Veröffentlicht: (2024)
von: Kim, Yearim, et al.
Veröffentlicht: (2024)
Unlocking the Potential of Unlabeled Data in Semi-Supervised Domain Generalization
von: Lee, Dongkwan, et al.
Veröffentlicht: (2025)
von: Lee, Dongkwan, et al.
Veröffentlicht: (2025)
ReFlex: Text-Guided Editing of Real Images in Rectified Flow via Mid-Step Feature Extraction and Attention Adaptation
von: Kim, Jimyeong, et al.
Veröffentlicht: (2025)
von: Kim, Jimyeong, et al.
Veröffentlicht: (2025)
Real-Time Intuitive AI Drawing System for Collaboration: Enhancing Human Creativity through Formal and Contextual Intent Integration
von: Song, Jookyung, et al.
Veröffentlicht: (2025)
von: Song, Jookyung, et al.
Veröffentlicht: (2025)
The Role of Teacher Calibration in Knowledge Distillation
von: Kim, Suyoung, et al.
Veröffentlicht: (2025)
von: Kim, Suyoung, et al.
Veröffentlicht: (2025)
CSF: Black-box Fingerprinting via Compositional Semantics for Text-to-Image Models
von: Lee, Junhoo, et al.
Veröffentlicht: (2026)
von: Lee, Junhoo, et al.
Veröffentlicht: (2026)
Harmonizing Visual and Textual Embeddings for Zero-Shot Text-to-Image Customization
von: Song, Yeji, et al.
Veröffentlicht: (2024)
von: Song, Yeji, et al.
Veröffentlicht: (2024)
Coreset Selection for Object Detection
von: Lee, Hojun, et al.
Veröffentlicht: (2024)
von: Lee, Hojun, et al.
Veröffentlicht: (2024)
Semi-Supervised Domain Adaptation for Wildfire Detection
von: Jang, JooYoung, et al.
Veröffentlicht: (2024)
von: Jang, JooYoung, et al.
Veröffentlicht: (2024)
Mitigating the Bias in the Model for Continual Test-Time Adaptation
von: Chung, Inseop, et al.
Veröffentlicht: (2024)
von: Chung, Inseop, et al.
Veröffentlicht: (2024)
LoGoColor: Local-Global 3D Colorization for 360° Scenes
von: Chang, Yeonjin, et al.
Veröffentlicht: (2025)
von: Chang, Yeonjin, et al.
Veröffentlicht: (2025)
ReSpinQuant: Efficient Layer-Wise LLM Quantization via Subspace Residual Rotation Approximation
von: Kim, Suyoung, et al.
Veröffentlicht: (2026)
von: Kim, Suyoung, et al.
Veröffentlicht: (2026)
MSG Score: Automated Video Verification for Reliable Multi-Scene Generation
von: Yoon, Daewon, et al.
Veröffentlicht: (2024)
von: Yoon, Daewon, et al.
Veröffentlicht: (2024)
ROODI: Reconstructing Occluded Objects with Denoising Inpainters
von: Chang, Yeonjin, et al.
Veröffentlicht: (2025)
von: Chang, Yeonjin, et al.
Veröffentlicht: (2025)
Point-to-Point: Sparse Motion Guidance for Controllable Video Editing
von: Song, Yeji, et al.
Veröffentlicht: (2025)
von: Song, Yeji, et al.
Veröffentlicht: (2025)
SplatFlow: Multi-View Rectified Flow Model for 3D Gaussian Splatting Synthesis
von: Go, Hyojun, et al.
Veröffentlicht: (2024)
von: Go, Hyojun, et al.
Veröffentlicht: (2024)
Groupwise Query Specialization and Quality-Aware Multi-Assignment for Transformer-based Visual Relationship Detection
von: Kim, Jongha, et al.
Veröffentlicht: (2024)
von: Kim, Jongha, et al.
Veröffentlicht: (2024)
What's Making That Sound Right Now? Video-centric Audio-Visual Localization
von: Choi, Hahyeon, et al.
Veröffentlicht: (2025)
von: Choi, Hahyeon, et al.
Veröffentlicht: (2025)
TWLV-I: Analysis and Insights from Holistic Evaluation on Video Foundation Models
von: Lee, Hyeongmin, et al.
Veröffentlicht: (2024)
von: Lee, Hyeongmin, et al.
Veröffentlicht: (2024)
ARC-NeRF: Area Ray Casting for Broader Unseen View Coverage in Few-shot Object Rendering
von: Seo, Seunghyeon, et al.
Veröffentlicht: (2024)
von: Seo, Seunghyeon, et al.
Veröffentlicht: (2024)
RefReward-SR: LR-Conditioned Reward Modeling for Preference-Aligned Super-Resolution
von: Song, Yushuai, et al.
Veröffentlicht: (2026)
von: Song, Yushuai, et al.
Veröffentlicht: (2026)
Rethinking Garment Conditioning in Diffusion-based Virtual Try-On
von: Na, Kihyun, et al.
Veröffentlicht: (2025)
von: Na, Kihyun, et al.
Veröffentlicht: (2025)
MERLIN: Multimodal Embedding Refinement via LLM-based Iterative Navigation for Text-Video Retrieval-Rerank Pipeline
von: Han, Donghoon, et al.
Veröffentlicht: (2024)
von: Han, Donghoon, et al.
Veröffentlicht: (2024)
Tempered Self-Similarity Alignment for Physically Plausible Video Generation
von: Kim, Manjin, et al.
Veröffentlicht: (2026)
von: Kim, Manjin, et al.
Veröffentlicht: (2026)
Style Composition within Distinct LoRA modules for Traditional Art
von: Lee, Jaehyun, et al.
Veröffentlicht: (2025)
von: Lee, Jaehyun, et al.
Veröffentlicht: (2025)
A Revisit to the Decoder for Camouflaged Object Detection
von: Ko, Seung Woo, et al.
Veröffentlicht: (2025)
von: Ko, Seung Woo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
MomentMix Augmentation with Length-Aware DETR for Temporally Robust Moment Retrieval
von: Park, Seojeong, et al.
Veröffentlicht: (2024) -
PedaCo-Gen: Scaffolding Pedagogical Agency in Human-AI Collaborative Video Authoring
von: Baek, Injun, et al.
Veröffentlicht: (2026) -
Rethinking Direct Preference Optimization in Diffusion Models
von: Kang, Junyong, et al.
Veröffentlicht: (2025) -
ConcatPlexer: Additional Dim1 Batching for Faster ViTs
von: Han, Donghoon, et al.
Veröffentlicht: (2023) -
From Local to Global to Mechanistic: An iERF-Centered Unified Framework for Interpreting Vision Models
von: Kim, Yearim, et al.
Veröffentlicht: (2026)