DeLTa: Demonstration and Language-Guided Novel Transparent Object Manipulation
Fuente:
arXiv
Salvato in:
| Autori principali: | Lee, Taeyeop, Kang, Gyuree, Wen, Bowen, Kim, Youngho, Back, Seunghyeok, Kweon, In So, Shim, David Hyunchul, Yoon, Kuk-Jin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Any6D: Model-free 6D Pose Estimation of Novel Objects
di: Lee, Taeyeop, et al.
Pubblicazione: (2025)
di: Lee, Taeyeop, et al.
Pubblicazione: (2025)
Event6D: Event-based Novel Object 6D Pose Tracking
di: Kang, Jae-Young, et al.
Pubblicazione: (2026)
di: Kang, Jae-Young, et al.
Pubblicazione: (2026)
GeoNVS: Geometry Grounded Video Diffusion for Novel View Synthesis
di: Kang, Minjun, et al.
Pubblicazione: (2026)
di: Kang, Minjun, et al.
Pubblicazione: (2026)
Drag4D: Align Your Motion with Text-Driven 3D Scene Generation
di: Kang, Minjun, et al.
Pubblicazione: (2025)
di: Kang, Minjun, et al.
Pubblicazione: (2025)
A Versatile Door Opening System with Mobile Manipulator through Adaptive Position-Force Control and Reinforcement Learning
di: Kang, Gyuree, et al.
Pubblicazione: (2023)
di: Kang, Gyuree, et al.
Pubblicazione: (2023)
DeLTa: A Decoding Strategy based on Logit Trajectory Prediction Improves Factuality and Reasoning Ability
di: He, Yunzhen, et al.
Pubblicazione: (2025)
di: He, Yunzhen, et al.
Pubblicazione: (2025)
SPIBOT: A Drone-Tethered Mobile Gripper for Robust Aerial Object Retrieval in Dynamic Environments
di: Kang, Gyuree, et al.
Pubblicazione: (2024)
di: Kang, Gyuree, et al.
Pubblicazione: (2024)
Ev-3DOD: Pushing the Temporal Boundaries of 3D Object Detection with Event Cameras
di: Cho, Hoonhee, et al.
Pubblicazione: (2025)
di: Cho, Hoonhee, et al.
Pubblicazione: (2025)
Stable Surface Regularization for Fast Few-Shot NeRF
di: Joung, Byeongin, et al.
Pubblicazione: (2024)
di: Joung, Byeongin, et al.
Pubblicazione: (2024)
From Sharp to Blur: Unsupervised Domain Adaptation for 2D Human Pose Estimation Under Extreme Motion Blur Using Event Cameras
di: Kim, Youngho, et al.
Pubblicazione: (2025)
di: Kim, Youngho, et al.
Pubblicazione: (2025)
GMT: Enhancing Generalizable Neural Rendering via Geometry-Driven Multi-Reference Texture Transfer
di: Yoon, Youngho, et al.
Pubblicazione: (2024)
di: Yoon, Youngho, et al.
Pubblicazione: (2024)
A Collaborative Team of UAV-Hexapod for an Autonomous Retrieval System in GNSS-Denied Maritime Environments
di: Lee, Seungwook, et al.
Pubblicazione: (2024)
di: Lee, Seungwook, et al.
Pubblicazione: (2024)
Synchronizing Task Behavior: Aligning Multiple Tasks during Test-Time Training
di: Jeong, Wooseong, et al.
Pubblicazione: (2025)
di: Jeong, Wooseong, et al.
Pubblicazione: (2025)
Robust Adverse Weather Removal via Spectral-based Spatial Grouping
di: Jeong, Yuhwan, et al.
Pubblicazione: (2025)
di: Jeong, Yuhwan, et al.
Pubblicazione: (2025)
Bootstrapping Video Semantic Segmentation Model via Distillation-assisted Test-Time Adaptation
di: Kim, Jihun, et al.
Pubblicazione: (2026)
di: Kim, Jihun, et al.
Pubblicazione: (2026)
TALoS: Enhancing Semantic Scene Completion via Test-time Adaptation on the Line of Sight
di: Jang, Hyun-Kurl, et al.
Pubblicazione: (2024)
di: Jang, Hyun-Kurl, et al.
Pubblicazione: (2024)
Antagonistic Bowden-Cable Actuation of a Lightweight Robotic Hand: Toward Dexterous Manipulation for Payload Constrained Humanoids
di: Min, Sungjae, et al.
Pubblicazione: (2025)
di: Min, Sungjae, et al.
Pubblicazione: (2025)
Finding Meaning in Points: Weakly Supervised Semantic Segmentation for Event Cameras
di: Cho, Hoonhee, et al.
Pubblicazione: (2024)
di: Cho, Hoonhee, et al.
Pubblicazione: (2024)
Learning from Demonstration with Hierarchical Policy Abstractions Toward High-Performance and Courteous Autonomous Racing
di: Chung, Chanyoung, et al.
Pubblicazione: (2024)
di: Chung, Chanyoung, et al.
Pubblicazione: (2024)
ManipForce: Force-Guided Policy Learning with Frequency-Aware Representation for Contact-Rich Manipulation
di: Lee, Geonhyup, et al.
Pubblicazione: (2025)
di: Lee, Geonhyup, et al.
Pubblicazione: (2025)
DC-TTA: Divide-and-Conquer Framework for Test-Time Adaptation of Interactive Segmentation
di: Kim, Jihun, et al.
Pubblicazione: (2025)
di: Kim, Jihun, et al.
Pubblicazione: (2025)
Enhancing Temporal Consistency in Video Editing by Reconstructing Videos with 3D Gaussian Splatting
di: Shin, Inkyu, et al.
Pubblicazione: (2024)
di: Shin, Inkyu, et al.
Pubblicazione: (2024)
Resolving Token-Space Gradient Conflicts: Token Space Manipulation for Transformer-Based Multi-Task Learning
di: Jeong, Wooseong, et al.
Pubblicazione: (2025)
di: Jeong, Wooseong, et al.
Pubblicazione: (2025)
Unleashing the Temporal Potential of Stereo Event Cameras for Continuous-Time 3D Object Detection
di: Kang, Jae-Young, et al.
Pubblicazione: (2025)
di: Kang, Jae-Young, et al.
Pubblicazione: (2025)
DSERT-RoLL: Robust Multi-Modal Perception for Diverse Driving Conditions with Stereo Event-RGB-Thermal Cameras, 4D Radar, and Dual-LiDAR
di: Cho, Hoonhee, et al.
Pubblicazione: (2026)
di: Cho, Hoonhee, et al.
Pubblicazione: (2026)
GraspClutter6D: A Large-scale Real-world Dataset for Robust Perception and Grasping in Cluttered Scenes
di: Back, Seunghyeok, et al.
Pubblicazione: (2025)
di: Back, Seunghyeok, et al.
Pubblicazione: (2025)
MonoDINO-DETR: Depth-Enhanced Monocular 3D Object Detection Using a Vision Foundation Model
di: Kim, Jihyeok, et al.
Pubblicazione: (2025)
di: Kim, Jihyeok, et al.
Pubblicazione: (2025)
Property Estimation in Geotechnical Databases Using Labeled Random Finite Sets
di: Shim, Changbeom, et al.
Pubblicazione: (2024)
di: Shim, Changbeom, et al.
Pubblicazione: (2024)
Weighted isoperimetric ratios and extension problems for fractional conformal Laplacians
di: Jin, Sangdon, et al.
Pubblicazione: (2023)
di: Jin, Sangdon, et al.
Pubblicazione: (2023)
TempFuser: Learning Agile, Tactical, and Acrobatic Flight Maneuvers Using a Long Short-Term Temporal Fusion Transformer
di: Seong, Hyunki, et al.
Pubblicazione: (2023)
di: Seong, Hyunki, et al.
Pubblicazione: (2023)
Self-Supervised Interpretable End-to-End Learning via Latent Functional Modularity
di: Seong, Hyunki, et al.
Pubblicazione: (2024)
di: Seong, Hyunki, et al.
Pubblicazione: (2024)
Skill Q-Network: Learning Adaptive Skill Ensemble for Mapless Navigation in Unknown Environments
di: Seong, Hyunki, et al.
Pubblicazione: (2024)
di: Seong, Hyunki, et al.
Pubblicazione: (2024)
Video Diffusion Models Excel at Tracking Similar-Looking Objects Without Supervision
di: Zhang, Chenshuang, et al.
Pubblicazione: (2025)
di: Zhang, Chenshuang, et al.
Pubblicazione: (2025)
GraspSAM: When Segment Anything Model Meets Grasp Detection
di: Noh, Sangjun, et al.
Pubblicazione: (2024)
di: Noh, Sangjun, et al.
Pubblicazione: (2024)
High-Quality Unknown Object Instance Segmentation via Quadruple Boundary Error Refinement
di: Back, Seunghyeok, et al.
Pubblicazione: (2023)
di: Back, Seunghyeok, et al.
Pubblicazione: (2023)
ImageNet-D: Benchmarking Neural Network Robustness on Diffusion Synthetic Object
di: Zhang, Chenshuang, et al.
Pubblicazione: (2024)
di: Zhang, Chenshuang, et al.
Pubblicazione: (2024)
DINO-VO: A Feature-based Visual Odometry Leveraging a Visual Foundation Model
di: Azhari, Maulana Bisyir, et al.
Pubblicazione: (2025)
di: Azhari, Maulana Bisyir, et al.
Pubblicazione: (2025)
CMTA: Cross-Modal Temporal Alignment for Event-guided Video Deblurring
di: Kim, Taewoo, et al.
Pubblicazione: (2024)
di: Kim, Taewoo, et al.
Pubblicazione: (2024)
Temporal Event Stereo via Joint Learning with Stereoscopic Flow
di: Cho, Hoonhee, et al.
Pubblicazione: (2024)
di: Cho, Hoonhee, et al.
Pubblicazione: (2024)
Selective Task Group Updates for Multi-Task Optimization
di: Jeong, Wooseong, et al.
Pubblicazione: (2025)
di: Jeong, Wooseong, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Any6D: Model-free 6D Pose Estimation of Novel Objects
di: Lee, Taeyeop, et al.
Pubblicazione: (2025) -
Event6D: Event-based Novel Object 6D Pose Tracking
di: Kang, Jae-Young, et al.
Pubblicazione: (2026) -
GeoNVS: Geometry Grounded Video Diffusion for Novel View Synthesis
di: Kang, Minjun, et al.
Pubblicazione: (2026) -
Drag4D: Align Your Motion with Text-Driven 3D Scene Generation
di: Kang, Minjun, et al.
Pubblicazione: (2025) -
A Versatile Door Opening System with Mobile Manipulator through Adaptive Position-Force Control and Reinforcement Learning
di: Kang, Gyuree, et al.
Pubblicazione: (2023)