FurniMAS: Language-Guided Furniture Decoration using Multi-Agent System
Fuente:
arXiv
Salvato in:
| Autori principali: | Nguyen, Toan, Le, Tri, Nguyen, Quang, Nguyen, Anh |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
GMAT: Grounded Multi-Agent Clinical Description Generation for Text Encoder in Vision-Language MIL for Whole Slide Image Classification
di: Quang, Ngoc Bui Lam, et al.
Pubblicazione: (2025)
di: Quang, Ngoc Bui Lam, et al.
Pubblicazione: (2025)
CSD-VAR: Content-Style Decomposition in Visual Autoregressive Models
di: Nguyen, Quang-Binh, et al.
Pubblicazione: (2025)
di: Nguyen, Quang-Binh, et al.
Pubblicazione: (2025)
Bidirectional Diffusion Bridge Models
di: Kieu, Duc, et al.
Pubblicazione: (2025)
di: Kieu, Duc, et al.
Pubblicazione: (2025)
Machine Intelligence that Understands Visual and Linguistic Information and Interacts with Humans and Environments
di: Nguyen, Van Quang
Pubblicazione: (2026)
di: Nguyen, Van Quang
Pubblicazione: (2026)
SwiftTry: Fast and Consistent Video Virtual Try-On with Diffusion Models
di: Nguyen, Hung, et al.
Pubblicazione: (2024)
di: Nguyen, Hung, et al.
Pubblicazione: (2024)
MedSteer: Counterfactual Endoscopic Synthesis via Training-Free Activation Steering
di: Pham, Trong-Thang, et al.
Pubblicazione: (2026)
di: Pham, Trong-Thang, et al.
Pubblicazione: (2026)
MADTempo: An Interactive System for Multi-Event Temporal Video Retrieval with Query Augmentation
di: Vu, Huu-An, et al.
Pubblicazione: (2025)
di: Vu, Huu-An, et al.
Pubblicazione: (2025)
WAVER: Writing-style Agnostic Text-Video Retrieval via Distilling Vision-Language Models Through Open-Vocabulary Knowledge
di: Le, Huy, et al.
Pubblicazione: (2023)
di: Le, Huy, et al.
Pubblicazione: (2023)
Unified Interactive Multimodal Moment Retrieval via Cascaded Embedding-Reranking and Temporal-Aware Score Fusion
di: Thanh, Toan Le Ngo, et al.
Pubblicazione: (2025)
di: Thanh, Toan Le Ngo, et al.
Pubblicazione: (2025)
Multimodal Contextualized Support for Enhancing Video Retrieval System
di: Nguyen-Le, Quoc-Bao, et al.
Pubblicazione: (2024)
di: Nguyen-Le, Quoc-Bao, et al.
Pubblicazione: (2024)
Semi-Supervised Semantic Segmentation using Redesigned Self-Training for White Blood Cells
di: Luu, Vinh Quoc, et al.
Pubblicazione: (2024)
di: Luu, Vinh Quoc, et al.
Pubblicazione: (2024)
KTVIC: A Vietnamese Image Captioning Dataset on the Life Domain
di: Pham, Anh-Cuong, et al.
Pubblicazione: (2024)
di: Pham, Anh-Cuong, et al.
Pubblicazione: (2024)
Tracking the Truth: Object-Centric Spatio-Temporal Monitoring for Video Large Language Models
di: Cao, Tri, et al.
Pubblicazione: (2026)
di: Cao, Tri, et al.
Pubblicazione: (2026)
Enhancing Multimodal Entity Linking with Jaccard Distance-based Conditional Contrastive Learning and Contextual Visual Augmentation
di: Nguyen, Cong-Duy, et al.
Pubblicazione: (2025)
di: Nguyen, Cong-Duy, et al.
Pubblicazione: (2025)
InverFill: One-Step Inversion for Enhanced Few-Step Diffusion Inpainting
di: Vu, Duc, et al.
Pubblicazione: (2026)
di: Vu, Duc, et al.
Pubblicazione: (2026)
RT-VLM: Re-Thinking Vision Language Model with 4-Clues for Real-World Object Recognition Robustness
di: Park, Junghyun, et al.
Pubblicazione: (2025)
di: Park, Junghyun, et al.
Pubblicazione: (2025)
V-Math: An Agentic Approach to the Vietnamese National High School Graduation Mathematics Exams
di: Nguyen, Duong Q., et al.
Pubblicazione: (2025)
di: Nguyen, Duong Q., et al.
Pubblicazione: (2025)
FurniScene: A Large-scale 3D Room Dataset with Intricate Furnishing Scenes
di: Zhang, Genghao, et al.
Pubblicazione: (2024)
di: Zhang, Genghao, et al.
Pubblicazione: (2024)
Rethinking Top Probability from Multi-view for Distracted Driver Behaviour Localization
di: Nguyen, Quang Vinh, et al.
Pubblicazione: (2024)
di: Nguyen, Quang Vinh, et al.
Pubblicazione: (2024)
Improving Zero-Shot Object-Level Change Detection by Incorporating Visual Correspondence
di: Nguyen, Hung Huy, et al.
Pubblicazione: (2025)
di: Nguyen, Hung Huy, et al.
Pubblicazione: (2025)
MambaU-Lite: A Lightweight Model based on Mamba and Integrated Channel-Spatial Attention for Skin Lesion Segmentation
di: Nguyen, Thi-Nhu-Quynh, et al.
Pubblicazione: (2024)
di: Nguyen, Thi-Nhu-Quynh, et al.
Pubblicazione: (2024)
STER-VLM: Spatio-Temporal With Enhanced Reference Vision-Language Models
di: Nguyen-Nhu, Tinh-Anh, et al.
Pubblicazione: (2025)
di: Nguyen-Nhu, Tinh-Anh, et al.
Pubblicazione: (2025)
PerspectiveNet: Multi-View Perception for Dynamic Scene Understanding
di: Nguyen, Vinh
Pubblicazione: (2024)
di: Nguyen, Vinh
Pubblicazione: (2024)
Bridging the Training-Deployment Gap: Gated Encoding and Multi-Scale Refinement for Efficient Quantization-Aware Image Enhancement
di: To-Thanh, Dat, et al.
Pubblicazione: (2026)
di: To-Thanh, Dat, et al.
Pubblicazione: (2026)
Phantasia: Context-Adaptive Backdoors in Vision Language Models
di: Tran, Nam Duong, et al.
Pubblicazione: (2026)
di: Tran, Nam Duong, et al.
Pubblicazione: (2026)
Anti-I2V: Safeguarding your photos from malicious image-to-video generation
di: Vu, Duc, et al.
Pubblicazione: (2026)
di: Vu, Duc, et al.
Pubblicazione: (2026)
HDC: Hierarchical Distillation for Multi-level Noisy Consistency in Semi-Supervised Fetal Ultrasound Segmentation
di: Le, Tran Quoc Khanh, et al.
Pubblicazione: (2025)
di: Le, Tran Quoc Khanh, et al.
Pubblicazione: (2025)
Interpreting Radiologist's Intention from Eye Movements in Chest X-ray Diagnosis
di: Pham, Trong-Thang, et al.
Pubblicazione: (2025)
di: Pham, Trong-Thang, et al.
Pubblicazione: (2025)
SwiftBrush v2: Make Your One-step Diffusion Model Better Than Its Teacher
di: Dao, Trung, et al.
Pubblicazione: (2024)
di: Dao, Trung, et al.
Pubblicazione: (2024)
Robustness Evaluation of OCR-based Visual Document Understanding under Multi-Modal Adversarial Attacks
di: Tien, Dong Nguyen, et al.
Pubblicazione: (2025)
di: Tien, Dong Nguyen, et al.
Pubblicazione: (2025)
AutoViVQA: A Large-Scale Automatically Constructed Dataset for Vietnamese Visual Question Answering
di: Tuong, Nguyen Anh, et al.
Pubblicazione: (2026)
di: Tuong, Nguyen Anh, et al.
Pubblicazione: (2026)
Contrastive Integrated Gradients: A Feature Attribution-Based Method for Explaining Whole Slide Image Classification
di: Vu, Anh Mai, et al.
Pubblicazione: (2025)
di: Vu, Anh Mai, et al.
Pubblicazione: (2025)
OE3DIS: Open-Ended 3D Point Cloud Instance Segmentation
di: Nguyen, Phuc D. A., et al.
Pubblicazione: (2024)
di: Nguyen, Phuc D. A., et al.
Pubblicazione: (2024)
Novel 3D Binary Indexed Tree for Volume Computation of 3D Reconstructed Models from Volumetric Data
di: Nguyen-Le, Quoc-Bao, et al.
Pubblicazione: (2024)
di: Nguyen-Le, Quoc-Bao, et al.
Pubblicazione: (2024)
Point Cloud Compression with Bits-back Coding
di: Hieu, Nguyen Quang, et al.
Pubblicazione: (2024)
di: Hieu, Nguyen Quang, et al.
Pubblicazione: (2024)
BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance
di: Le, Huy, et al.
Pubblicazione: (2025)
di: Le, Huy, et al.
Pubblicazione: (2025)
Learning Human Motion with Temporally Conditional Mamba
di: Nguyen, Quang, et al.
Pubblicazione: (2025)
di: Nguyen, Quang, et al.
Pubblicazione: (2025)
PEEB: Part-based Image Classifiers with an Explainable and Editable Language Bottleneck
di: Pham, Thang M., et al.
Pubblicazione: (2024)
di: Pham, Thang M., et al.
Pubblicazione: (2024)
SUGAR: A Sweeter Spot for Generative Unlearning of Many Identities
di: Nguyen, Dung Thuy, et al.
Pubblicazione: (2025)
di: Nguyen, Dung Thuy, et al.
Pubblicazione: (2025)
Multimedia Verification Through Multi-Agent Deep Research Multimodal Large Language Models
di: Le, Huy Hoan, et al.
Pubblicazione: (2025)
di: Le, Huy Hoan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
GMAT: Grounded Multi-Agent Clinical Description Generation for Text Encoder in Vision-Language MIL for Whole Slide Image Classification
di: Quang, Ngoc Bui Lam, et al.
Pubblicazione: (2025) -
CSD-VAR: Content-Style Decomposition in Visual Autoregressive Models
di: Nguyen, Quang-Binh, et al.
Pubblicazione: (2025) -
Bidirectional Diffusion Bridge Models
di: Kieu, Duc, et al.
Pubblicazione: (2025) -
Machine Intelligence that Understands Visual and Linguistic Information and Interacts with Humans and Environments
di: Nguyen, Van Quang
Pubblicazione: (2026) -
SwiftTry: Fast and Consistent Video Virtual Try-On with Diffusion Models
di: Nguyen, Hung, et al.
Pubblicazione: (2024)