Robot Synesthesia: A Sound and Emotion Guided AI Painter
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Misra, Vihaan, Schaldenbrand, Peter, Oh, Jean |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ShapeShift: Text-to-Mosaic Synthesis via Semantic Phase-Field Guidance
von: Misra, Vihaan, et al.
Veröffentlicht: (2025)
von: Misra, Vihaan, et al.
Veröffentlicht: (2025)
Robot Synesthesia: In-Hand Manipulation with Visuotactile Sensing
von: Yuan, Ying, et al.
Veröffentlicht: (2023)
von: Yuan, Ying, et al.
Veröffentlicht: (2023)
SCoFT: Self-Contrastive Fine-Tuning for Equitable Image Generation
von: Liu, Zhixuan, et al.
Veröffentlicht: (2024)
von: Liu, Zhixuan, et al.
Veröffentlicht: (2024)
TokenDial: Continuous Attribute Control in Text-to-Video via Spatiotemporal Token Offsets
von: Liu, Zhixuan, et al.
Veröffentlicht: (2026)
von: Liu, Zhixuan, et al.
Veröffentlicht: (2026)
Loomis Painter: Reconstructing the Painting Process
von: Pobitzer, Markus, et al.
Veröffentlicht: (2025)
von: Pobitzer, Markus, et al.
Veröffentlicht: (2025)
HD-Painter: High-Resolution and Prompt-Faithful Text-Guided Image Inpainting with Diffusion Models
von: Manukyan, Hayk, et al.
Veröffentlicht: (2023)
von: Manukyan, Hayk, et al.
Veröffentlicht: (2023)
Token Painter: Training-Free Text-Guided Image Inpainting via Mask Autoregressive Models
von: Jiang, Longtao, et al.
Veröffentlicht: (2025)
von: Jiang, Longtao, et al.
Veröffentlicht: (2025)
GarmentPainter: Efficient 3D Garment Texture Synthesis with Character-Guided Diffusion Model
von: Wu, Jinbo, et al.
Veröffentlicht: (2026)
von: Wu, Jinbo, et al.
Veröffentlicht: (2026)
Optimised ProPainter for Video Diminished Reality Inpainting
von: Li, Pengze, et al.
Veröffentlicht: (2024)
von: Li, Pengze, et al.
Veröffentlicht: (2024)
AttentionPainter: An Efficient and Adaptive Stroke Predictor for Scene Painting
von: Tang, Yizhe, et al.
Veröffentlicht: (2024)
von: Tang, Yizhe, et al.
Veröffentlicht: (2024)
RoadPainter: Points Are Ideal Navigators for Topology transformER
von: Ma, Zhongxing, et al.
Veröffentlicht: (2024)
von: Ma, Zhongxing, et al.
Veröffentlicht: (2024)
AnimatePainter: A Self-Supervised Rendering Framework for Reconstructing Painting Process
von: Hu, Junjie, et al.
Veröffentlicht: (2025)
von: Hu, Junjie, et al.
Veröffentlicht: (2025)
Zero-Painter: Training-Free Layout Control for Text-to-Image Synthesis
von: Ohanyan, Marianna, et al.
Veröffentlicht: (2024)
von: Ohanyan, Marianna, et al.
Veröffentlicht: (2024)
FramePainter: Endowing Interactive Image Editing with Video Diffusion Priors
von: Zhang, Yabo, et al.
Veröffentlicht: (2025)
von: Zhang, Yabo, et al.
Veröffentlicht: (2025)
RoomPainter: View-Integrated Diffusion for Consistent Indoor Scene Texturing
von: Huang, Zhipeng, et al.
Veröffentlicht: (2024)
von: Huang, Zhipeng, et al.
Veröffentlicht: (2024)
PathoPainter: Augmenting Histopathology Segmentation via Tumor-aware Inpainting
von: Liu, Hong, et al.
Veröffentlicht: (2025)
von: Liu, Hong, et al.
Veröffentlicht: (2025)
MambaPainter: Neural Stroke-Based Rendering in a Single Step
von: Sawada, Tomoya, et al.
Veröffentlicht: (2024)
von: Sawada, Tomoya, et al.
Veröffentlicht: (2024)
GaussianPainter: Painting Point Cloud into 3D Gaussians with Normal Guidance
von: Zhou, Jingqiu, et al.
Veröffentlicht: (2024)
von: Zhou, Jingqiu, et al.
Veröffentlicht: (2024)
Applying Medical Imaging Tractography Techniques to Painterly Rendering of Images
von: Di Biase, Alberto
Veröffentlicht: (2025)
von: Di Biase, Alberto
Veröffentlicht: (2025)
TexPainter: Generative Mesh Texturing with Multi-view Consistency
von: Zhang, Hongkun, et al.
Veröffentlicht: (2024)
von: Zhang, Hongkun, et al.
Veröffentlicht: (2024)
FlexPainter: Flexible and Multi-View Consistent Texture Generation
von: Yan, Dongyu, et al.
Veröffentlicht: (2025)
von: Yan, Dongyu, et al.
Veröffentlicht: (2025)
LidarPainter: One-Step Away From Any Lidar View To Novel Guidance
von: Ji, Yuzhou, et al.
Veröffentlicht: (2025)
von: Ji, Yuzhou, et al.
Veröffentlicht: (2025)
VectorPainter: Advanced Stylized Vector Graphics Synthesis Using Stroke-Style Priors
von: Hu, Juncheng, et al.
Veröffentlicht: (2024)
von: Hu, Juncheng, et al.
Veröffentlicht: (2024)
ScenePainter: Semantically Consistent Perpetual 3D Scene Generation with Concept Relation Alignment
von: Xia, Chong, et al.
Veröffentlicht: (2025)
von: Xia, Chong, et al.
Veröffentlicht: (2025)
RePainter: Empowering E-commerce Object Removal via Spatial-matting Reinforcement Learning
von: Guo, Zipeng, et al.
Veröffentlicht: (2025)
von: Guo, Zipeng, et al.
Veröffentlicht: (2025)
ProcessPainter: Learn Painting Process from Sequence Data
von: Song, Yiren, et al.
Veröffentlicht: (2024)
von: Song, Yiren, et al.
Veröffentlicht: (2024)
DreamPainter: Image Background Inpainting for E-commerce Scenarios
von: Zhao, Sijie, et al.
Veröffentlicht: (2025)
von: Zhao, Sijie, et al.
Veröffentlicht: (2025)
AnomalyPainter: Vision-Language-Diffusion Synergy for Zero-Shot Realistic and Diverse Industrial Anomaly Synthesis
von: Lai, Zhangyu, et al.
Veröffentlicht: (2025)
von: Lai, Zhangyu, et al.
Veröffentlicht: (2025)
LLaVA$^3$: Representing 3D Scenes like a Cubist Painter to Boost 3D Scene Understanding of VLMs
von: Petit, Doriand, et al.
Veröffentlicht: (2025)
von: Petit, Doriand, et al.
Veröffentlicht: (2025)
PainterNet: Adaptive Image Inpainting with Actual-Token Attention and Diverse Mask Control
von: Wang, Ruichen, et al.
Veröffentlicht: (2024)
von: Wang, Ruichen, et al.
Veröffentlicht: (2024)
Hearing Hands: Generating Sounds from Physical Interactions in 3D Scenes
von: Dou, Yiming, et al.
Veröffentlicht: (2025)
von: Dou, Yiming, et al.
Veröffentlicht: (2025)
Improving Image De-raining Using Reference-Guided Transformers
von: Ye, Zihao, et al.
Veröffentlicht: (2024)
von: Ye, Zihao, et al.
Veröffentlicht: (2024)
Learning to Hear by Seeing: It's Time for Vision Language Models to Understand Artistic Emotion from Sight and Sound
von: Zhang, Dengming, et al.
Veröffentlicht: (2025)
von: Zhang, Dengming, et al.
Veröffentlicht: (2025)
Draw-In-Mind: Rebalancing Designer-Painter Roles in Unified Multimodal Models Benefits Image Editing
von: Zeng, Ziyun, et al.
Veröffentlicht: (2025)
von: Zeng, Ziyun, et al.
Veröffentlicht: (2025)
SplatPainter: Interactive Authoring of 3D Gaussians from 2D Edits via Test-Time Training
von: Zheng, Yang, et al.
Veröffentlicht: (2025)
von: Zheng, Yang, et al.
Veröffentlicht: (2025)
Unveiling the Cognitive Compass: Theory-of-Mind-Guided Multimodal Emotion Reasoning
von: Luo, Meng, et al.
Veröffentlicht: (2026)
von: Luo, Meng, et al.
Veröffentlicht: (2026)
O1O: Grouping of Known Classes to Identify Unknown Objects as Odd-One-Out
von: Yavuz, Mısra, et al.
Veröffentlicht: (2024)
von: Yavuz, Mısra, et al.
Veröffentlicht: (2024)
ARport: An Augmented Reality System for Markerless Image-Guided Port Placement in Robotic Surgery
von: Han, Zheng, et al.
Veröffentlicht: (2026)
von: Han, Zheng, et al.
Veröffentlicht: (2026)
Dimensional Distribution Emotion State: Leveraging Valence and Arousal as a Common Embedding Space for Visual Emotion Analysis
von: Bergeron, Émile, et al.
Veröffentlicht: (2026)
von: Bergeron, Émile, et al.
Veröffentlicht: (2026)
EASL: Multi-Emotion Guided Semantic Disentanglement for Expressive Sign Language Generation
von: Zhao, Yanchao, et al.
Veröffentlicht: (2025)
von: Zhao, Yanchao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
ShapeShift: Text-to-Mosaic Synthesis via Semantic Phase-Field Guidance
von: Misra, Vihaan, et al.
Veröffentlicht: (2025) -
Robot Synesthesia: In-Hand Manipulation with Visuotactile Sensing
von: Yuan, Ying, et al.
Veröffentlicht: (2023) -
SCoFT: Self-Contrastive Fine-Tuning for Equitable Image Generation
von: Liu, Zhixuan, et al.
Veröffentlicht: (2024) -
TokenDial: Continuous Attribute Control in Text-to-Video via Spatiotemporal Token Offsets
von: Liu, Zhixuan, et al.
Veröffentlicht: (2026) -
Loomis Painter: Reconstructing the Painting Process
von: Pobitzer, Markus, et al.
Veröffentlicht: (2025)