Aligning Characteristic Descriptors with Images for Human-Expert-like Explainability
Fuente:
arXiv
Saved in:
| Main Authors: | Yalavarthi, Bharat Chandra, Ratha, Nalini |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Enhancing Privacy in Face Analytics Using Fully Homomorphic Encryption
by: Yalavarthi, Bharat, et al.
Published: (2024)
by: Yalavarthi, Bharat, et al.
Published: (2024)
Multimodal Privacy-Preserving Entity Resolution with Fully Homomorphic Encryption
by: Roy, Susim, et al.
Published: (2026)
by: Roy, Susim, et al.
Published: (2026)
CROP: Expert-Aligned Image Cropping via Compositional Reasoning and Optimizing Preference
by: Dong, Zhitong, et al.
Published: (2026)
by: Dong, Zhitong, et al.
Published: (2026)
HumanAesExpert: Advancing a Multi-Modality Foundation Model for Human Image Aesthetic Assessment
by: Liao, Zhichao, et al.
Published: (2025)
by: Liao, Zhichao, et al.
Published: (2025)
DescriptorMedSAM: Language-Image Fusion with Multi-Aspect Text Guidance for Medical Image Segmentation
by: Zhang, Wenjie, et al.
Published: (2025)
by: Zhang, Wenjie, et al.
Published: (2025)
Caption-Driven Explorations: Aligning Image and Text Embeddings through Human-Inspired Foveated Vision
by: Zanca, Dario, et al.
Published: (2024)
by: Zanca, Dario, et al.
Published: (2024)
Region in Context: Text-condition Image editing with Human-like semantic reasoning
by: Vu, Thuy Phuong, et al.
Published: (2025)
by: Vu, Thuy Phuong, et al.
Published: (2025)
Cross-Attention Head Position Patterns Can Align with Human Visual Concepts in Text-to-Image Generative Models
by: Park, Jungwon, et al.
Published: (2024)
by: Park, Jungwon, et al.
Published: (2024)
Video-Bench: Human-Aligned Video Generation Benchmark
by: Han, Hui, et al.
Published: (2025)
by: Han, Hui, et al.
Published: (2025)
Expert-Guided Explainable Few-Shot Learning for Medical Image Diagnosis
by: Uddin, Ifrat Ikhtear, et al.
Published: (2025)
by: Uddin, Ifrat Ikhtear, et al.
Published: (2025)
Spatio-Semantic Expert Routing Architecture with Mixture-of-Experts for Referring Image Segmentation
by: Dalaq, Alaa, et al.
Published: (2026)
by: Dalaq, Alaa, et al.
Published: (2026)
Human-Aligned Bench: Fine-Grained Assessment of Reasoning Ability in MLLMs vs. Humans
by: Qiu, Yansheng, et al.
Published: (2025)
by: Qiu, Yansheng, et al.
Published: (2025)
Learning Action Hierarchies via Hybrid Geometric Diffusion
by: Kaushik, Arjun Ramesh, et al.
Published: (2026)
by: Kaushik, Arjun Ramesh, et al.
Published: (2026)
EditReward: A Human-Aligned Reward Model for Instruction-Guided Image Editing
by: Wu, Keming, et al.
Published: (2025)
by: Wu, Keming, et al.
Published: (2025)
Aligning Vision Models with Human Aesthetics in Retrieval: Benchmarks and Algorithms
by: Zhang, Miaosen, et al.
Published: (2024)
by: Zhang, Miaosen, et al.
Published: (2024)
Descriptor: Parasitoid Wasps and Associated Hymenoptera Dataset (DAPWH)
by: Pinheiro, Joao Manoel Herrera, et al.
Published: (2026)
by: Pinheiro, Joao Manoel Herrera, et al.
Published: (2026)
Fixed-length Dense Descriptor for Efficient Fingerprint Matching
by: Pan, Zhiyu, et al.
Published: (2023)
by: Pan, Zhiyu, et al.
Published: (2023)
CytoCLIP: Learning Cytoarchitectural Characteristics in Developing Human Brain Using Contrastive Language Image Pre-Training
by: Ta, Pralaypati, et al.
Published: (2026)
by: Ta, Pralaypati, et al.
Published: (2026)
Representations of Text and Images Align From Layer One
by: Wybitul, Evžen, et al.
Published: (2026)
by: Wybitul, Evžen, et al.
Published: (2026)
Text2Graph VPR: A Text-to-Graph Expert System for Explainable Place Recognition in Changing Environments
by: Yousefzadeh, Saeideh, et al.
Published: (2025)
by: Yousefzadeh, Saeideh, et al.
Published: (2025)
ChartMoE: Mixture of Diversely Aligned Expert Connector for Chart Understanding
by: Xu, Zhengzhuo, et al.
Published: (2024)
by: Xu, Zhengzhuo, et al.
Published: (2024)
Self-Corrected Image Generation with Explainable Latent Rewards
by: Luo, Yinyi, et al.
Published: (2026)
by: Luo, Yinyi, et al.
Published: (2026)
Image-POSER: Reflective RL for Multi-Expert Image Generation and Editing
by: Mohebbi, Hossein, et al.
Published: (2025)
by: Mohebbi, Hossein, et al.
Published: (2025)
Descriptor: Distance-Annotated Traffic Perception Question Answering (DTPQA)
by: Theodoridis, Nikos, et al.
Published: (2025)
by: Theodoridis, Nikos, et al.
Published: (2025)
Belief-Aware VLM Model for Human-like Reasoning
by: Nayak, Anshul, et al.
Published: (2026)
by: Nayak, Anshul, et al.
Published: (2026)
OT-ALD: Aligning Latent Distributions with Optimal Transport for Accelerated Image-to-Image Translation
by: Wang, Zhanpeng, et al.
Published: (2025)
by: Wang, Zhanpeng, et al.
Published: (2025)
InfoDisent: Explainability of Image Classification Models by Information Disentanglement
by: Struski, Łukasz, et al.
Published: (2024)
by: Struski, Łukasz, et al.
Published: (2024)
Explainable Fundus Image Curation and Lesion Detection in Diabetic Retinopathy
by: Mihai, Anca, et al.
Published: (2025)
by: Mihai, Anca, et al.
Published: (2025)
Towards Counterfactual and Contrastive Explainability and Transparency of DCNN Image Classifiers
by: Tariq, Syed Ali, et al.
Published: (2025)
by: Tariq, Syed Ali, et al.
Published: (2025)
Align and Surpass Human Camouflaged Perception: Visual Refocus Reinforcement Fine-Tuning
by: Shen, Ruolin, et al.
Published: (2025)
by: Shen, Ruolin, et al.
Published: (2025)
Local Descriptors Weighted Adaptive Threshold Filtering For Few-Shot Learning
by: Yan, Bingchen
Published: (2024)
by: Yan, Bingchen
Published: (2024)
Human-like Navigation in a World Built for Humans
by: Chandaka, Bhargav, et al.
Published: (2025)
by: Chandaka, Bhargav, et al.
Published: (2025)
Harmonized Tabular-Image Fusion via Gradient-Aligned Alternating Learning
by: Huang, Longfei, et al.
Published: (2026)
by: Huang, Longfei, et al.
Published: (2026)
DiffusionAgent: Navigating Expert Models for Agentic Image Generation
by: Qin, Jie, et al.
Published: (2024)
by: Qin, Jie, et al.
Published: (2024)
SVGauge: Towards Human-Aligned Evaluation for SVG Generation
by: Zini, Leonardo, et al.
Published: (2025)
by: Zini, Leonardo, et al.
Published: (2025)
Simple Agents Outperform Experts in Biomedical Imaging Workflow Optimization
by: Xuefei, et al.
Published: (2025)
by: Xuefei, et al.
Published: (2025)
VidPrism: Heterogeneous Mixture of Experts for Image-to-Video Transfer
by: Lin, Rui, et al.
Published: (2026)
by: Lin, Rui, et al.
Published: (2026)
IntegratedPIFu: Integrated Pixel Aligned Implicit Function for Single-view Human Reconstruction
by: Chan, Kennard Yanting, et al.
Published: (2022)
by: Chan, Kennard Yanting, et al.
Published: (2022)
HEIE: MLLM-Based Hierarchical Explainable AIGC Image Implausibility Evaluator
by: Yang, Fan, et al.
Published: (2024)
by: Yang, Fan, et al.
Published: (2024)
Red Teaming Models for Hyperspectral Image Analysis Using Explainable AI
by: Zaigrajew, Vladimir, et al.
Published: (2024)
by: Zaigrajew, Vladimir, et al.
Published: (2024)
Similar Items
-
Enhancing Privacy in Face Analytics Using Fully Homomorphic Encryption
by: Yalavarthi, Bharat, et al.
Published: (2024) -
Multimodal Privacy-Preserving Entity Resolution with Fully Homomorphic Encryption
by: Roy, Susim, et al.
Published: (2026) -
CROP: Expert-Aligned Image Cropping via Compositional Reasoning and Optimizing Preference
by: Dong, Zhitong, et al.
Published: (2026) -
HumanAesExpert: Advancing a Multi-Modality Foundation Model for Human Image Aesthetic Assessment
by: Liao, Zhichao, et al.
Published: (2025) -
DescriptorMedSAM: Language-Image Fusion with Multi-Aspect Text Guidance for Medical Image Segmentation
by: Zhang, Wenjie, et al.
Published: (2025)