Leveraging Image Augmentation for Object Manipulation: Towards Interpretable Controllability in Object-Centric Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Jinwoo, Choi, Janghyuk, Kang, Jaehyun, Lee, Changyeon, Choi, Ho-Jin, Kim, Seon Joo |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ORIDa: Object-centric Real-world Image Composition Dataset
by: Kim, Jinwoo, et al.
Published: (2025)
by: Kim, Jinwoo, et al.
Published: (2025)
SlotVTG: Object-Centric Adapter for Generalizable Video Temporal Grounding
by: Han, Jiwook, et al.
Published: (2026)
by: Han, Jiwook, et al.
Published: (2026)
Accelerating Image Super-Resolution Networks with Pixel-Level Classification
by: Jeong, Jinho, et al.
Published: (2024)
by: Jeong, Jinho, et al.
Published: (2024)
Object Aware Egocentric Online Action Detection
by: An, Joungbin, et al.
Published: (2024)
by: An, Joungbin, et al.
Published: (2024)
Latent Space Super-Resolution for Higher-Resolution Image Generation with Diffusion Models
by: Jeong, Jinho, et al.
Published: (2025)
by: Jeong, Jinho, et al.
Published: (2025)
HAMLET: Switch your Vision-Language-Action Model into a History-Aware Policy
by: Koo, Myungkyu, et al.
Published: (2025)
by: Koo, Myungkyu, et al.
Published: (2025)
UniSkill: Imitating Human Videos via Cross-Embodiment Skill Representations
by: Kim, Hanjung, et al.
Published: (2025)
by: Kim, Hanjung, et al.
Published: (2025)
OCK: Unsupervised Dynamic Video Prediction with Object-Centric Kinematics
by: Song, Yeon-Ji, et al.
Published: (2024)
by: Song, Yeon-Ji, et al.
Published: (2024)
A Review of Image Retrieval Techniques: Data Augmentation and Adversarial Learning Approaches
by: Jinwoo, Kim
Published: (2024)
by: Jinwoo, Kim
Published: (2024)
Attentive Illumination Decomposition Model for Multi-Illuminant White Balancing
by: Kim, Dongyoung, et al.
Published: (2024)
by: Kim, Dongyoung, et al.
Published: (2024)
APT: Improving Diffusion Models for High Resolution Image Generation with Adaptive Path Tracing
by: Han, Sangmin, et al.
Published: (2025)
by: Han, Sangmin, et al.
Published: (2025)
Disentangled Object-Centric Image Representation for Robotic Manipulation
by: Emukpere, David, et al.
Published: (2025)
by: Emukpere, David, et al.
Published: (2025)
Object-Centric Instruction Augmentation for Robotic Manipulation
by: Wen, Junjie, et al.
Published: (2024)
by: Wen, Junjie, et al.
Published: (2024)
Text-Guided 6D Object Pose Rearrangement via Closed-Loop VLM Agents
by: Baik, Sangwon, et al.
Published: (2026)
by: Baik, Sangwon, et al.
Published: (2026)
Hierarchically Structured Neural Bones for Reconstructing Animatable Objects from Casual Videos
by: Jeon, Subin, et al.
Published: (2024)
by: Jeon, Subin, et al.
Published: (2024)
VERIA: Verification-Centric Multimodal Instance Augmentation for Long-Tailed 3D Object Detection
by: Lee, Jumin, et al.
Published: (2026)
by: Lee, Jumin, et al.
Published: (2026)
Object-Centric Representation Learning for Enhanced 3D Semantic Scene Graph Prediction
by: Heo, KunHo, et al.
Published: (2025)
by: Heo, KunHo, et al.
Published: (2025)
Learning to Enhance Aperture Phasor Field for Non-Line-of-Sight Imaging
by: Cho, In, et al.
Published: (2024)
by: Cho, In, et al.
Published: (2024)
Inlier-Centric Post-Training Quantization for Object Detection Models
by: Kim, Minsu, et al.
Published: (2026)
by: Kim, Minsu, et al.
Published: (2026)
OSPO: Object-Centric Self-Improving Preference Optimization for Text-to-Image Generation
by: Oh, Yoonjin, et al.
Published: (2025)
by: Oh, Yoonjin, et al.
Published: (2025)
Open-ended Hierarchical Streaming Video Understanding with Vision Language Models
by: Kang, Hyolim, et al.
Published: (2025)
by: Kang, Hyolim, et al.
Published: (2025)
VISAGE: Video Instance Segmentation with Appearance-Guided Enhancement
by: Kim, Hanjung, et al.
Published: (2023)
by: Kim, Hanjung, et al.
Published: (2023)
Object-Centric World Model for Language-Guided Manipulation
by: Jeong, Youngjoon, et al.
Published: (2025)
by: Jeong, Youngjoon, et al.
Published: (2025)
GTA: Guided Transfer of Spatial Attention from Object-Centric Representations
by: Seo, SeokHyun, et al.
Published: (2024)
by: Seo, SeokHyun, et al.
Published: (2024)
Query-Calibrated Segmental Admission for Descriptor-Agnostic LiDAR Loop Closure in Repetitive Environments
by: Kim, Jaehyun, et al.
Published: (2025)
by: Kim, Jaehyun, et al.
Published: (2025)
Learning Object-Centric Representations in SAR Images with Multi-Level Feature Fusion
by: Jang, Oh-Tae, et al.
Published: (2025)
by: Jang, Oh-Tae, et al.
Published: (2025)
Why Can't I Open My Drawer? Mitigating Object-Driven Shortcuts in Zero-Shot Compositional Action Recognition
by: Ahn, Geo, et al.
Published: (2026)
by: Ahn, Geo, et al.
Published: (2026)
CCMNet: Leveraging Calibrated Color Correction Matrices for Cross-Camera Color Constancy
by: Kim, Dongyoung, et al.
Published: (2025)
by: Kim, Dongyoung, et al.
Published: (2025)
Mine-JEPA: In-Domain Self-Supervised Learning for Mine-Like Object Classification in Side-Scan Sonar
by: Kwon, Taeyoun, et al.
Published: (2026)
by: Kwon, Taeyoun, et al.
Published: (2026)
Beyond Semantics: Disentangling Information Scope in Sparse Autoencoders for CLIP
by: Ro, Yusung, et al.
Published: (2026)
by: Ro, Yusung, et al.
Published: (2026)
Semi-Supervised Domain Adaptation Using Target-Oriented Domain Augmentation for 3D Object Detection
by: Kim, Yecheol, et al.
Published: (2024)
by: Kim, Yecheol, et al.
Published: (2024)
Safety Alignment Backfires: Preventing the Re-emergence of Suppressed Concepts in Fine-tuned Text-to-Image Diffusion Models
by: Kim, Sanghyun, et al.
Published: (2024)
by: Kim, Sanghyun, et al.
Published: (2024)
Layer-Wise Modality Decomposition for Interpretable Multimodal Sensor Fusion
by: Park, Jaehyun, et al.
Published: (2025)
by: Park, Jaehyun, et al.
Published: (2025)
Object Remover Performance Evaluation Methods using Class-wise Object Removal Images
by: Oh, Changsuk, et al.
Published: (2024)
by: Oh, Changsuk, et al.
Published: (2024)
PLOT: Pseudo-Labeling via Video Object Tracking for Scalable Monocular 3D Object Detection
by: Lee, Seokyeong, et al.
Published: (2025)
by: Lee, Seokyeong, et al.
Published: (2025)
Learning Neural Deformation Representation for 4D Dynamic Shape Generation
by: Han, Gyojin, et al.
Published: (2026)
by: Han, Gyojin, et al.
Published: (2026)
3Doodle: Compact Abstraction of Objects with 3D Strokes
by: Choi, Changwoon, et al.
Published: (2024)
by: Choi, Changwoon, et al.
Published: (2024)
Controllable Human Image Generation with Personalized Multi-Garments
by: Choi, Yisol, et al.
Published: (2024)
by: Choi, Yisol, et al.
Published: (2024)
Referring Video Object Segmentation via Language-aligned Track Selection
by: Kim, Seongchan, et al.
Published: (2024)
by: Kim, Seongchan, et al.
Published: (2024)
ParaHome: Parameterizing Everyday Home Activities Towards 3D Generative Modeling of Human-Object Interactions
by: Kim, Jeonghwan, et al.
Published: (2024)
by: Kim, Jeonghwan, et al.
Published: (2024)
Similar Items
-
ORIDa: Object-centric Real-world Image Composition Dataset
by: Kim, Jinwoo, et al.
Published: (2025) -
SlotVTG: Object-Centric Adapter for Generalizable Video Temporal Grounding
by: Han, Jiwook, et al.
Published: (2026) -
Accelerating Image Super-Resolution Networks with Pixel-Level Classification
by: Jeong, Jinho, et al.
Published: (2024) -
Object Aware Egocentric Online Action Detection
by: An, Joungbin, et al.
Published: (2024) -
Latent Space Super-Resolution for Higher-Resolution Image Generation with Diffusion Models
by: Jeong, Jinho, et al.
Published: (2025)