Understanding Cross-Model Perceptual Invariances Through Ensemble Metamers
Fuente:
arXiv
Salvato in:
| Autori principali: | Boehm, Lukas, Mueller, Jonas Leo, Loeffler, Christoffer, Schwinn, Leo, Eskofier, Bjoern, Zanca, Dario |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Don't Get Me Wrong: How to Apply Deep Visual Interpretations to Time Series
di: Loeffler, Christoffer, et al.
Pubblicazione: (2022)
di: Loeffler, Christoffer, et al.
Pubblicazione: (2022)
Caption-Driven Explorations: Aligning Image and Text Embeddings through Human-Inspired Foveated Vision
di: Zanca, Dario, et al.
Pubblicazione: (2024)
di: Zanca, Dario, et al.
Pubblicazione: (2024)
Stratify or Die: Rethinking Data Splits in Image Segmentation
di: Jami, Naga Venkata Sai Jitin, et al.
Pubblicazione: (2025)
di: Jami, Naga Venkata Sai Jitin, et al.
Pubblicazione: (2025)
RadProPoser: Probabilistic Radar Tensor Human Pose Estimation That Knows Its Limits
di: Mueller, Jonas Leo, et al.
Pubblicazione: (2025)
di: Mueller, Jonas Leo, et al.
Pubblicazione: (2025)
Benchmarking Content-Based Puzzle Solvers on Corrupted Jigsaw Puzzles
di: Dirauf, Richard, et al.
Pubblicazione: (2025)
di: Dirauf, Richard, et al.
Pubblicazione: (2025)
FovEx: Human-Inspired Explanations for Vision Transformers and Convolutional Neural Networks
di: Panda, Mahadev Prasad, et al.
Pubblicazione: (2024)
di: Panda, Mahadev Prasad, et al.
Pubblicazione: (2024)
A Unified Approach Towards Active Learning and Out-of-Distribution Detection
di: Schmidt, Sebastian, et al.
Pubblicazione: (2024)
di: Schmidt, Sebastian, et al.
Pubblicazione: (2024)
Enhancing IMU-Based Online Handwriting Recognition via Contrastive Learning with Zero Inference Overhead
di: Li, Jindong, et al.
Pubblicazione: (2026)
di: Li, Jindong, et al.
Pubblicazione: (2026)
Joint Out-of-Distribution Filtering and Data Discovery Active Learning
di: Schmidt, Sebastian, et al.
Pubblicazione: (2025)
di: Schmidt, Sebastian, et al.
Pubblicazione: (2025)
Tokenization vs. Augmentation: A Systematic Study of Writer Variance in IMU-Based Online Handwriting Recognition
di: Li, Jindong, et al.
Pubblicazione: (2026)
di: Li, Jindong, et al.
Pubblicazione: (2026)
Assessing Robustness via Score-Based Adversarial Image Generation
di: Kollovieh, Marcel, et al.
Pubblicazione: (2023)
di: Kollovieh, Marcel, et al.
Pubblicazione: (2023)
FOCUS: Internal MLLM Representations for Efficient Fine-Grained Visual Question Answering
di: Zhong, Liangyu, et al.
Pubblicazione: (2025)
di: Zhong, Liangyu, et al.
Pubblicazione: (2025)
Unexplored flaws in multiple-choice VQA evaluations
di: Rosenthal, Fabio, et al.
Pubblicazione: (2025)
di: Rosenthal, Fabio, et al.
Pubblicazione: (2025)
Exploring Color Invariance through Image-Level Ensemble Learning
di: Gong, Yunpeng, et al.
Pubblicazione: (2024)
di: Gong, Yunpeng, et al.
Pubblicazione: (2024)
PerCo (SD): Open Perceptual Compression
di: Körber, Nikolai, et al.
Pubblicazione: (2024)
di: Körber, Nikolai, et al.
Pubblicazione: (2024)
Multi-View Foundation Models
di: Segre, Leo, et al.
Pubblicazione: (2025)
di: Segre, Leo, et al.
Pubblicazione: (2025)
Vision-Driven Prompt Optimization for Large Language Models in Multimodal Generative Tasks
di: Franklin, Leo, et al.
Pubblicazione: (2025)
di: Franklin, Leo, et al.
Pubblicazione: (2025)
Optimize the Unseen -- Fast NeRF Cleanup with Free Space Prior
di: Segre, Leo, et al.
Pubblicazione: (2024)
di: Segre, Leo, et al.
Pubblicazione: (2024)
VF-NeRF: Viewshed Fields for Rigid NeRF Registration
di: Segre, Leo, et al.
Pubblicazione: (2024)
di: Segre, Leo, et al.
Pubblicazione: (2024)
Towards Video Thinking Test: A Holistic Benchmark for Advanced Video Reasoning and Understanding
di: Zhang, Yuanhan, et al.
Pubblicazione: (2025)
di: Zhang, Yuanhan, et al.
Pubblicazione: (2025)
Trends, Applications, and Challenges in Human Attention Modelling
di: Cartella, Giuseppe, et al.
Pubblicazione: (2024)
di: Cartella, Giuseppe, et al.
Pubblicazione: (2024)
In-Place Panoptic Radiance Field Segmentation with Perceptual Prior for 3D Scene Understanding
di: Li, Shenghao
Pubblicazione: (2024)
di: Li, Shenghao
Pubblicazione: (2024)
Perceptual-GS: Scene-adaptive Perceptual Densification for Gaussian Splatting
di: Zhou, Hongbi, et al.
Pubblicazione: (2025)
di: Zhou, Hongbi, et al.
Pubblicazione: (2025)
Perceptual Classifiers: Detecting Generative Images using Perceptual Features
di: Durbha, Krishna Srikar, et al.
Pubblicazione: (2025)
di: Durbha, Krishna Srikar, et al.
Pubblicazione: (2025)
Susceptibility of Adversarial Attack on Medical Image Segmentation Models
di: Wang, Zhongxuan, et al.
Pubblicazione: (2024)
di: Wang, Zhongxuan, et al.
Pubblicazione: (2024)
MultiMedEval: A Benchmark and a Toolkit for Evaluating Medical Vision-Language Models
di: Royer, Corentin, et al.
Pubblicazione: (2024)
di: Royer, Corentin, et al.
Pubblicazione: (2024)
RouteExtract: A Modular Pipeline for Extracting Routes from Paper Maps
di: Kremser, Bjoern, et al.
Pubblicazione: (2025)
di: Kremser, Bjoern, et al.
Pubblicazione: (2025)
A Decade of You Only Look Once (YOLO) for Object Detection: A Review
di: Ramos, Leo Thomas, et al.
Pubblicazione: (2025)
di: Ramos, Leo Thomas, et al.
Pubblicazione: (2025)
Frequency-Aware Gaussian Splatting Decomposition
di: Lavi, Yishai, et al.
Pubblicazione: (2025)
di: Lavi, Yishai, et al.
Pubblicazione: (2025)
Boosting General Trimap-free Matting in the Real-World Image
di: Zhao, Leo Shan Wenzhang Zhou Grace
Pubblicazione: (2024)
di: Zhao, Leo Shan Wenzhang Zhou Grace
Pubblicazione: (2024)
Can Vision Language Models Learn from Visual Demonstrations of Ambiguous Spatial Reasoning?
di: Zhao, Bowen, et al.
Pubblicazione: (2024)
di: Zhao, Bowen, et al.
Pubblicazione: (2024)
4D-RGPT: Toward Region-level 4D Understanding via Perceptual Distillation
di: Yang, Chiao-An, et al.
Pubblicazione: (2025)
di: Yang, Chiao-An, et al.
Pubblicazione: (2025)
Exploring Facial Biomarkers for Depression through Temporal Analysis of Action Units
di: Parikh, Aditya, et al.
Pubblicazione: (2024)
di: Parikh, Aditya, et al.
Pubblicazione: (2024)
Optimized Learned Image Compression for Facial Expression Recognition
di: Li, Xiumei, et al.
Pubblicazione: (2025)
di: Li, Xiumei, et al.
Pubblicazione: (2025)
Multi-Level Embedding and Alignment Network with Consistency and Invariance Learning for Cross-View Geo-Localization
di: Chen, Zhongwei, et al.
Pubblicazione: (2024)
di: Chen, Zhongwei, et al.
Pubblicazione: (2024)
UniPercept: Towards Unified Perceptual-Level Image Understanding across Aesthetics, Quality, Structure, and Texture
di: Cao, Shuo, et al.
Pubblicazione: (2025)
di: Cao, Shuo, et al.
Pubblicazione: (2025)
UrbanFeel: A Comprehensive Benchmark for Temporal and Perceptual Understanding of City Scenes through Human Perspective
di: He, Jun, et al.
Pubblicazione: (2025)
di: He, Jun, et al.
Pubblicazione: (2025)
VOILA: Evaluation of MLLMs For Perceptual Understanding and Analogical Reasoning
di: Yilmaz, Nilay, et al.
Pubblicazione: (2025)
di: Yilmaz, Nilay, et al.
Pubblicazione: (2025)
Towards Identity-Aware Cross-Modal Retrieval: a Dataset and a Baseline
di: Messina, Nicola, et al.
Pubblicazione: (2024)
di: Messina, Nicola, et al.
Pubblicazione: (2024)
InterAct: Capture and Modelling of Realistic, Expressive and Interactive Activities between Two Persons in Daily Scenarios
di: Huang, Yinghao, et al.
Pubblicazione: (2024)
di: Huang, Yinghao, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Don't Get Me Wrong: How to Apply Deep Visual Interpretations to Time Series
di: Loeffler, Christoffer, et al.
Pubblicazione: (2022) -
Caption-Driven Explorations: Aligning Image and Text Embeddings through Human-Inspired Foveated Vision
di: Zanca, Dario, et al.
Pubblicazione: (2024) -
Stratify or Die: Rethinking Data Splits in Image Segmentation
di: Jami, Naga Venkata Sai Jitin, et al.
Pubblicazione: (2025) -
RadProPoser: Probabilistic Radar Tensor Human Pose Estimation That Knows Its Limits
di: Mueller, Jonas Leo, et al.
Pubblicazione: (2025) -
Benchmarking Content-Based Puzzle Solvers on Corrupted Jigsaw Puzzles
di: Dirauf, Richard, et al.
Pubblicazione: (2025)