Automated Detection of Visual Attribute Reliance with a Self-Reflective Agent
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Christy, Camuñas, Josep Lopez, Touchet, Jake Thomas, Andreas, Jacob, Lapedriza, Agata, Torralba, Antonio, Shaham, Tamar Rott |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Multimodal Automated Interpretability Agent
di: Shaham, Tamar Rott, et al.
Pubblicazione: (2024)
di: Shaham, Tamar Rott, et al.
Pubblicazione: (2024)
Experimenting with Affective Computing Models in Video Interviews with Spanish-speaking Older Adults
di: Camunas, Josep Lopez, et al.
Pubblicazione: (2025)
di: Camunas, Josep Lopez, et al.
Pubblicazione: (2025)
Vision-Language Binding in In-Context Image Generation
di: Ge, Chris, et al.
Pubblicazione: (2026)
di: Ge, Chris, et al.
Pubblicazione: (2026)
SketchAgent: Language-Driven Sequential Sketch Generation
di: Vinker, Yael, et al.
Pubblicazione: (2024)
di: Vinker, Yael, et al.
Pubblicazione: (2024)
BrainExplore: Large-Scale Discovery of Interpretable Visual Representations in the Human Brain
di: Wasserman, Navve, et al.
Pubblicazione: (2025)
di: Wasserman, Navve, et al.
Pubblicazione: (2025)
The Dual Mechanisms of Spatial Reasoning in Vision-Language Models
di: Cui, Kelly, et al.
Pubblicazione: (2026)
di: Cui, Kelly, et al.
Pubblicazione: (2026)
From Activation to Causality: Discovery of Causal Visual Representations in the Human Brain
di: Golbari, Yuval, et al.
Pubblicazione: (2026)
di: Golbari, Yuval, et al.
Pubblicazione: (2026)
Pitfalls in Evaluating Interpretability Agents
di: Haklay, Tal, et al.
Pubblicazione: (2026)
di: Haklay, Tal, et al.
Pubblicazione: (2026)
A Vision Check-up for Language Models
di: Sharma, Pratyusha, et al.
Pubblicazione: (2024)
di: Sharma, Pratyusha, et al.
Pubblicazione: (2024)
Automatic Discovery of Visual Circuits
di: Rajaram, Achyuta, et al.
Pubblicazione: (2024)
di: Rajaram, Achyuta, et al.
Pubblicazione: (2024)
AdSum: Two-stream Audio-visual Summarization for Automated Video Advertisement Clipping
di: Xie, Wen, et al.
Pubblicazione: (2025)
di: Xie, Wen, et al.
Pubblicazione: (2025)
Understanding Visual Feature Reliance through the Lens of Complexity
di: Fel, Thomas, et al.
Pubblicazione: (2024)
di: Fel, Thomas, et al.
Pubblicazione: (2024)
Letting the neural code speak: Automated characterization of monkey visual neurons through human language
di: Lad, Vedang, et al.
Pubblicazione: (2026)
di: Lad, Vedang, et al.
Pubblicazione: (2026)
RealStats: A Rigorous Real-Only Statistical Framework for Fake Image Detection
di: Zisman, Haim, et al.
Pubblicazione: (2026)
di: Zisman, Haim, et al.
Pubblicazione: (2026)
Dataset Distillation for Pre-Trained Self-Supervised Vision Models
di: Cazenavette, George, et al.
Pubblicazione: (2025)
di: Cazenavette, George, et al.
Pubblicazione: (2025)
Fine-Tuning Enhances Existing Mechanisms: A Case Study on Entity Tracking
di: Prakash, Nikhil, et al.
Pubblicazione: (2024)
di: Prakash, Nikhil, et al.
Pubblicazione: (2024)
MultiModal Action Conditioned Video Generation
di: Li, Yichen, et al.
Pubblicazione: (2025)
di: Li, Yichen, et al.
Pubblicazione: (2025)
PANICL: Mitigating Over-Reliance on Single Prompt in Visual In-Context Learning
di: Zhang, Jiahao, et al.
Pubblicazione: (2025)
di: Zhang, Jiahao, et al.
Pubblicazione: (2025)
Analyzing Cultural Representations of Emotions in LLMs through Mixed Emotion Survey
di: Dudy, Shiran, et al.
Pubblicazione: (2024)
di: Dudy, Shiran, et al.
Pubblicazione: (2024)
Towards Personalized Quantum Federated Learning for Anomaly Detection
di: Rahman, Ratun, et al.
Pubblicazione: (2025)
di: Rahman, Ratun, et al.
Pubblicazione: (2025)
Route, Retrieve, Reflect, Repair: Self-Improving Agentic Framework for Visual Detection and Linguistic Reasoning in Medical Imaging
di: Sayeedi, Md. Faiyaz Abdullah, et al.
Pubblicazione: (2026)
di: Sayeedi, Md. Faiyaz Abdullah, et al.
Pubblicazione: (2026)
Efficient 3D Instance Mapping and Localization with Neural Fields
di: Tang, George, et al.
Pubblicazione: (2024)
di: Tang, George, et al.
Pubblicazione: (2024)
MARS: Memory-Enhanced Agents with Reflective Self-improvement
di: Liang, Xuechen, et al.
Pubblicazione: (2025)
di: Liang, Xuechen, et al.
Pubblicazione: (2025)
Primeros estudios zooarqueológicos en Moreta (Puna de Jujuy, Argentina, S. VII-XVI d.C.)
di: José Luis Camuñas
Pubblicazione: (2021)
di: José Luis Camuñas
Pubblicazione: (2021)
EVLM: Self-Reflective Multimodal Reasoning for Cross-Dimensional Visual Editing
di: Khalid, Umar, et al.
Pubblicazione: (2024)
di: Khalid, Umar, et al.
Pubblicazione: (2024)
Countering the Over-Reliance Trap: Mitigating Object Hallucination for LVLMs via a Self-Validation Framework
di: Liu, Shiyu, et al.
Pubblicazione: (2026)
di: Liu, Shiyu, et al.
Pubblicazione: (2026)
Autonomous and Self-Adapting System for Synthetic Media Detection and Attribution
di: Azizpour, Aref, et al.
Pubblicazione: (2025)
di: Azizpour, Aref, et al.
Pubblicazione: (2025)
Glioma Classification using Multi-sequence MRI and Novel Wavelets-based Feature Fusion
di: Janardhan, Kiranmayee, et al.
Pubblicazione: (2025)
di: Janardhan, Kiranmayee, et al.
Pubblicazione: (2025)
Enhancing Visual Reliance in Text Generation: A Bayesian Perspective on Mitigating Hallucination in Large Vision-Language Models
di: Hu, Nanxing, et al.
Pubblicazione: (2025)
di: Hu, Nanxing, et al.
Pubblicazione: (2025)
Knowledge Detection by Relevant Question and Image Attributes in Visual Question Answering
di: Ahir, Param, et al.
Pubblicazione: (2023)
di: Ahir, Param, et al.
Pubblicazione: (2023)
Self-Supervised Flow Matching for Scalable Multi-Modal Synthesis
di: Chefer, Hila, et al.
Pubblicazione: (2026)
di: Chefer, Hila, et al.
Pubblicazione: (2026)
Motion Attribution for Video Generation
di: Wu, Xindi, et al.
Pubblicazione: (2026)
di: Wu, Xindi, et al.
Pubblicazione: (2026)
Enhancing VICReg: Random-Walk Pairing for Improved Generalization and Better Global Semantics Capturing
di: Simai, Idan, et al.
Pubblicazione: (2025)
di: Simai, Idan, et al.
Pubblicazione: (2025)
Omni-Attribute: Open-vocabulary Attribute Encoder for Visual Concept Personalization
di: Chen, Tsai-Shien, et al.
Pubblicazione: (2025)
di: Chen, Tsai-Shien, et al.
Pubblicazione: (2025)
Seeing Through the PRISM: Compound & Controllable Restoration of Scientific Images
di: Kurinchi-Vendhan, Rupa, et al.
Pubblicazione: (2026)
di: Kurinchi-Vendhan, Rupa, et al.
Pubblicazione: (2026)
OVOD-Agent: A Markov-Bandit Framework for Proactive Visual Reasoning and Self-Evolving Detection
di: Wang, Chujie, et al.
Pubblicazione: (2025)
di: Wang, Chujie, et al.
Pubblicazione: (2025)
Characterizing Model Robustness via Natural Input Gradients
di: Rodríguez-Muñoz, Adrián, et al.
Pubblicazione: (2024)
di: Rodríguez-Muñoz, Adrián, et al.
Pubblicazione: (2024)
SPIRAL: Self-Evolving Action-Conditioned Video Generation via Reflective Planning Agents
di: Yang, Yu, et al.
Pubblicazione: (2026)
di: Yang, Yu, et al.
Pubblicazione: (2026)
ViTA-PAR: Visual and Textual Attribute Alignment with Attribute Prompting for Pedestrian Attribute Recognition
di: Park, Minjeong, et al.
Pubblicazione: (2025)
di: Park, Minjeong, et al.
Pubblicazione: (2025)
Benchmarking and Evolving Reason-Reflect-Rectify for Reflective Visual Generation
di: Wang, Junjie, et al.
Pubblicazione: (2026)
di: Wang, Junjie, et al.
Pubblicazione: (2026)
Documenti analoghi
-
A Multimodal Automated Interpretability Agent
di: Shaham, Tamar Rott, et al.
Pubblicazione: (2024) -
Experimenting with Affective Computing Models in Video Interviews with Spanish-speaking Older Adults
di: Camunas, Josep Lopez, et al.
Pubblicazione: (2025) -
Vision-Language Binding in In-Context Image Generation
di: Ge, Chris, et al.
Pubblicazione: (2026) -
SketchAgent: Language-Driven Sequential Sketch Generation
di: Vinker, Yael, et al.
Pubblicazione: (2024) -
BrainExplore: Large-Scale Discovery of Interpretable Visual Representations in the Human Brain
di: Wasserman, Navve, et al.
Pubblicazione: (2025)