SYNTHIA: Novel Concept Design with Affordance Composition
Fuente:
arXiv
Saved in:
| Main Authors: | Ha, Hyeonjeong, Jin, Xiaomeng, Kim, Jeonghwan, Liu, Jiateng, Wang, Zhenhailong, Nguyen, Khanh Duy, Blume, Ansel, Peng, Nanyun, Chang, Kai-Wei, Ji, Heng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PARTONOMY: Large Multimodal Models with Part-Level Visual Understanding
by: Blume, Ansel, et al.
Published: (2025)
by: Blume, Ansel, et al.
Published: (2025)
CreativityBench: Evaluating Agent Creative Reasoning via Affordance-Based Tool Repurposing
by: Qian, Cheng, et al.
Published: (2026)
by: Qian, Cheng, et al.
Published: (2026)
MM-PoisonRAG: Disrupting Multimodal RAG with Local and Global Poisoning Attacks
by: Ha, Hyeonjeong, et al.
Published: (2025)
by: Ha, Hyeonjeong, et al.
Published: (2025)
Search and Detect: Training-Free Long Tail Object Detection via Web-Image Retrieval
by: Sidhu, Mankeerat, et al.
Published: (2024)
by: Sidhu, Mankeerat, et al.
Published: (2024)
Finer: Investigating and Enhancing Fine-Grained Visual Concept Recognition in Large Vision Language Models
by: Kim, Jeonghwan, et al.
Published: (2024)
by: Kim, Jeonghwan, et al.
Published: (2024)
ARMADA: Attribute-Based Multimodal Data Augmentation
by: Jin, Xiaomeng, et al.
Published: (2024)
by: Jin, Xiaomeng, et al.
Published: (2024)
OSExpert: Computer-Use Agents Learning Professional Skills via Exploration
by: Liu, Jiateng, et al.
Published: (2026)
by: Liu, Jiateng, et al.
Published: (2026)
Open Vocabulary Electroencephalography-To-Text Decoding and Zero-shot Sentiment Classification
by: Wang, Zhenhailong, et al.
Published: (2021)
by: Wang, Zhenhailong, et al.
Published: (2021)
Infogent: An Agent-Based Framework for Web Information Aggregation
by: Reddy, Revanth Gangi, et al.
Published: (2024)
by: Reddy, Revanth Gangi, et al.
Published: (2024)
Constructive Distortion: Improving MLLMs with Attention-Guided Image Warping
by: Dalal, Dwip, et al.
Published: (2025)
by: Dalal, Dwip, et al.
Published: (2025)
Scientific Opinion Summarization: Paper Meta-review Generation Dataset, Methods, and Evaluation
by: Zeng, Qi, et al.
Published: (2023)
by: Zeng, Qi, et al.
Published: (2023)
Perception-Aware Policy Optimization for Multimodal Reasoning
by: Wang, Zhenhailong, et al.
Published: (2025)
by: Wang, Zhenhailong, et al.
Published: (2025)
iCONTRA: Toward Thematic Collection Design Via Interactive Concept Transfer
by: Vo, Dinh-Khoi, et al.
Published: (2024)
by: Vo, Dinh-Khoi, et al.
Published: (2024)
Learning to Rank Caption Chains for Video-Text Alignment
by: Blume, Ansel, et al.
Published: (2026)
by: Blume, Ansel, et al.
Published: (2026)
Sustainable development during economic uncertainty: What drives large construction firms to perform corporate social responsibility?
by: Minh Van Nguyen, et al.
Published: (2024)
by: Minh Van Nguyen, et al.
Published: (2024)
MemGuard: Preventing Memory Contamination in Long-Term Memory-Augmented Large Language Models
by: Ha, Hyeonjeong, et al.
Published: (2026)
by: Ha, Hyeonjeong, et al.
Published: (2026)
Multimodal Policy Internalization for Conversational Agents
by: Wang, Zhenhailong, et al.
Published: (2025)
by: Wang, Zhenhailong, et al.
Published: (2025)
Analyzing and Internalizing Complex Policy Documents for LLM Agents
by: Liu, Jiateng, et al.
Published: (2025)
by: Liu, Jiateng, et al.
Published: (2025)
Contrastive Visual Data Augmentation
by: Zhou, Yu, et al.
Published: (2025)
by: Zhou, Yu, et al.
Published: (2025)
Predicting Camera Pose from Perspective Descriptions for Spatial Reasoning
by: Zhang, Xuejun, et al.
Published: (2026)
by: Zhang, Xuejun, et al.
Published: (2026)
SYNTHIA Project Communication, dissemination and engagement materials D10.1
by: MATICAL INNOVATION SL
Published: (2024)
by: MATICAL INNOVATION SL
Published: (2024)
Ensembling Portfolio Strategies for Long-Term Investments: A Distribution-Free Preference Framework for Decision-Making and Algorithms
by: Lam, Duy Khanh
Published: (2024)
by: Lam, Duy Khanh
Published: (2024)
Mean-Variance Portfolio Selection in Long-Term Investments with Unknown Distribution: Online Estimation, Risk Aversion under Ambiguity, and Universality of Algorithms
by: Lam, Duy Khanh
Published: (2024)
by: Lam, Duy Khanh
Published: (2024)
Beating the Best Constant Rebalancing Portfolio in Long-Term Investment: A Generalization of the Kelly Criterion and Universal Learning Algorithm for Markets with Serial Dependence
by: Lam, Duy Khanh
Published: (2025)
by: Lam, Duy Khanh
Published: (2025)
Sequential Portfolio Selection under Latent Side Information-Dependence Structure: Optimality and Universal Learning Algorithms
by: Lam, Duy Khanh
Published: (2025)
by: Lam, Duy Khanh
Published: (2025)
Democratizing LLMs: An Exploration of Cost-Performance Trade-offs in Self-Refined Open-Source Models
by: Shashidhar, Sumuk, et al.
Published: (2023)
by: Shashidhar, Sumuk, et al.
Published: (2023)
Protein Language Models Diverge from Natural Language: Comparative Analysis and Improved Inference
by: Hart, Anna, et al.
Published: (2026)
by: Hart, Anna, et al.
Published: (2026)
A Concept is More Than a Word: Diversified Unlearning in Text-to-Image Diffusion Models
by: Pham, Duc Hao, et al.
Published: (2026)
by: Pham, Duc Hao, et al.
Published: (2026)
Quantum States Seen by a Probe: Partial Trace Over a Region of Space
by: Ansel, Quentin
Published: (2024)
by: Ansel, Quentin
Published: (2024)
Emergent gravity from the correlation of spin-$\tfrac{1}{2}$ systems coupled with a scalar field
by: Ansel, Quentin
Published: (2024)
by: Ansel, Quentin
Published: (2024)
Perspective of a pre‐symptomatic individual with an FTD‐causing MAPT gene variant
by: Ansel Dow
Published: (2025)
by: Ansel Dow
Published: (2025)
Safe Navigation of Bipedal Robots via Koopman Operator-Based Model Predictive Control
by: Kim, Jeonghwan, et al.
Published: (2024)
by: Kim, Jeonghwan, et al.
Published: (2024)
Verbalized Representation Learning for Interpretable Few-Shot Generalization
by: Yang, Cheng-Fu, et al.
Published: (2024)
by: Yang, Cheng-Fu, et al.
Published: (2024)
INTRA: Interaction Relationship-aware Weakly Supervised Affordance Grounding
by: Jang, Ji Ha, et al.
Published: (2024)
by: Jang, Ji Ha, et al.
Published: (2024)
Model Extrapolation Expedites Alignment
by: Zheng, Chujie, et al.
Published: (2024)
by: Zheng, Chujie, et al.
Published: (2024)
Sustainable Supply Chain Practices and the Cost of Equity Capital in Emerging Markets: Do Blockholder Ownership and Geopolitical Risk Matter?
by: Van Ha Nguyen, et al.
Published: (2026)
by: Van Ha Nguyen, et al.
Published: (2026)
Automatic Textual Normalization for Hate Speech Detection
by: Nguyen, Anh Thi-Hoang, et al.
Published: (2023)
by: Nguyen, Anh Thi-Hoang, et al.
Published: (2023)
Interaction-Aware Vehicle Motion Planning with Collision Avoidance Constraints in Highway Traffic
by: Kim, Dongryul, et al.
Published: (2024)
by: Kim, Dongryul, et al.
Published: (2024)
ConTextual: Evaluating Context-Sensitive Text-Rich Visual Reasoning in Large Multimodal Models
by: Wadhawan, Rohan, et al.
Published: (2024)
by: Wadhawan, Rohan, et al.
Published: (2024)
DeepEdit: Knowledge Editing as Decoding with Constraints
by: Wang, Yiwei, et al.
Published: (2024)
by: Wang, Yiwei, et al.
Published: (2024)
Similar Items
-
PARTONOMY: Large Multimodal Models with Part-Level Visual Understanding
by: Blume, Ansel, et al.
Published: (2025) -
CreativityBench: Evaluating Agent Creative Reasoning via Affordance-Based Tool Repurposing
by: Qian, Cheng, et al.
Published: (2026) -
MM-PoisonRAG: Disrupting Multimodal RAG with Local and Global Poisoning Attacks
by: Ha, Hyeonjeong, et al.
Published: (2025) -
Search and Detect: Training-Free Long Tail Object Detection via Web-Image Retrieval
by: Sidhu, Mankeerat, et al.
Published: (2024) -
Finer: Investigating and Enhancing Fine-Grained Visual Concept Recognition in Large Vision Language Models
by: Kim, Jeonghwan, et al.
Published: (2024)