Gespeichert in:
| Hauptverfasser: | Passi, Ananya, Robinson, Brian S., Bonner, Michael F. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2605.19155 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
An extremely coarse feedback signal is sufficient for learning human-aligned visual representations
von: Mehta, Yash, et al.
Veröffentlicht: (2026)
von: Mehta, Yash, et al.
Veröffentlicht: (2026)
Universal dimensions of visual representation
von: Chen, Zirui, et al.
Veröffentlicht: (2024)
von: Chen, Zirui, et al.
Veröffentlicht: (2024)
Rapidly deploying on-device eye tracking by distilling visual foundation models
von: Jiang, Cheng, et al.
Veröffentlicht: (2026)
von: Jiang, Cheng, et al.
Veröffentlicht: (2026)
SAGE: Spatial-visual Adaptive Graph Exploration for Efficient Visual Place Recognition
von: Chen, Shunpeng, et al.
Veröffentlicht: (2025)
von: Chen, Shunpeng, et al.
Veröffentlicht: (2025)
Evaluating the Suitability of Different Intraoral Scan Resolutions for Deep Learning-Based Tooth Segmentation
von: Weekley, Daron, et al.
Veröffentlicht: (2025)
von: Weekley, Daron, et al.
Veröffentlicht: (2025)
DIMM: Decoupled Multi-hierarchy Kalman Filter for 3D Object Tracking
von: Zha, Jirong, et al.
Veröffentlicht: (2025)
von: Zha, Jirong, et al.
Veröffentlicht: (2025)
HiERO: understanding the hierarchy of human behavior enhances reasoning on egocentric videos
von: Peirone, Simone Alberto, et al.
Veröffentlicht: (2025)
von: Peirone, Simone Alberto, et al.
Veröffentlicht: (2025)
Towards long-term player tracking with graph hierarchies and domain-specific features
von: Koshkina, Maria, et al.
Veröffentlicht: (2025)
von: Koshkina, Maria, et al.
Veröffentlicht: (2025)
Mining Contextualized Visual Associations from Images for Creativity Understanding
von: Sahu, Ananya, et al.
Veröffentlicht: (2025)
von: Sahu, Ananya, et al.
Veröffentlicht: (2025)
Contrastive Learning-based Multi Modal Architecture for Emoticon Prediction by Employing Image-Text Pairs
von: Pandey, Ananya, et al.
Veröffentlicht: (2024)
von: Pandey, Ananya, et al.
Veröffentlicht: (2024)
Target-Dependent Multimodal Sentiment Analysis Via Employing Visual-to Emotional-Caption Translation Network using Visual-Caption Pairs
von: Pandey, Ananya, et al.
Veröffentlicht: (2024)
von: Pandey, Ananya, et al.
Veröffentlicht: (2024)
FedPartWhole: Federated domain generalization via consistent part-whole hierarchies
von: Radwan, Ahmed, et al.
Veröffentlicht: (2024)
von: Radwan, Ahmed, et al.
Veröffentlicht: (2024)
Characterizing Universal Object Representations Across Vision Models
von: Mahner, Florian P., et al.
Veröffentlicht: (2026)
von: Mahner, Florian P., et al.
Veröffentlicht: (2026)
On the rankability of visual embeddings
von: Sonthalia, Ankit, et al.
Veröffentlicht: (2025)
von: Sonthalia, Ankit, et al.
Veröffentlicht: (2025)
ARMARecon: An ARMA Convolutional Filter based Graph Neural Network for Neurodegenerative Dementias Classification
von: Abburi, VSS Tejaswi, et al.
Veröffentlicht: (2026)
von: Abburi, VSS Tejaswi, et al.
Veröffentlicht: (2026)
Generative Action Tell-Tales: Assessing Human Motion in Synthesized Videos
von: Thomas, Xavier, et al.
Veröffentlicht: (2025)
von: Thomas, Xavier, et al.
Veröffentlicht: (2025)
Characterizing the visual representation of objects from the child's view
von: Yang, Jane, et al.
Veröffentlicht: (2026)
von: Yang, Jane, et al.
Veröffentlicht: (2026)
HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training
von: Wu, Xuecheng, et al.
Veröffentlicht: (2025)
von: Wu, Xuecheng, et al.
Veröffentlicht: (2025)
Neuromorphic visual attention for Sign-language recognition on SpiNNaker
von: Liskova, Sarka, et al.
Veröffentlicht: (2026)
von: Liskova, Sarka, et al.
Veröffentlicht: (2026)
Analyzing Noise Models and Advanced Filtering Algorithms for Image Enhancement
von: Akbar, Sahil Ali, et al.
Veröffentlicht: (2024)
von: Akbar, Sahil Ali, et al.
Veröffentlicht: (2024)
UltrON: Ultrasound Occupancy Networks
von: Wysocki, Magdalena, et al.
Veröffentlicht: (2025)
von: Wysocki, Magdalena, et al.
Veröffentlicht: (2025)
Modelling Visual Semantics via Image Captioning to extract Enhanced Multi-Level Cross-Modal Semantic Incongruity Representation with Attention for Multimodal Sarcasm Detection
von: Aggarwal, Sajal, et al.
Veröffentlicht: (2024)
von: Aggarwal, Sajal, et al.
Veröffentlicht: (2024)
LIT: Large Language Model Driven Intention Tracking for Proactive Human-Robot Collaboration -- A Robot Sous-Chef Application
von: Huang, Zhe, et al.
Veröffentlicht: (2024)
von: Huang, Zhe, et al.
Veröffentlicht: (2024)
TetraSphere: A Neural Descriptor for O(3)-Invariant Point Cloud Analysis
von: Melnyk, Pavlo, et al.
Veröffentlicht: (2022)
von: Melnyk, Pavlo, et al.
Veröffentlicht: (2022)
Wandering around: A bioinspired approach to visual attention through object motion sensitivity
von: D'Angelo, Giulia, et al.
Veröffentlicht: (2025)
von: D'Angelo, Giulia, et al.
Veröffentlicht: (2025)
DualResolution Residual Architecture with Artifact Suppression for Melanocytic Lesion Segmentation
von: Singh, Vikram, et al.
Veröffentlicht: (2025)
von: Singh, Vikram, et al.
Veröffentlicht: (2025)
Spectral Progressive Diffusion for Efficient Image and Video Generation
von: Xiao, Howard, et al.
Veröffentlicht: (2026)
von: Xiao, Howard, et al.
Veröffentlicht: (2026)
Large-scale visual SLAM for in-the-wild videos
von: Sun, Shuo, et al.
Veröffentlicht: (2025)
von: Sun, Shuo, et al.
Veröffentlicht: (2025)
Explaning with trees: interpreting CNNs using hierarchies
von: Rodrigues, Caroline Mazini, et al.
Veröffentlicht: (2024)
von: Rodrigues, Caroline Mazini, et al.
Veröffentlicht: (2024)
Foveated Diffusion: Efficient Spatially Adaptive Image and Video Generation
von: Chao, Brian, et al.
Veröffentlicht: (2026)
von: Chao, Brian, et al.
Veröffentlicht: (2026)
RemEdit: Efficient Diffusion Editing with Riemannian Geometry
von: Adhikarla, Eashan, et al.
Veröffentlicht: (2026)
von: Adhikarla, Eashan, et al.
Veröffentlicht: (2026)
A transition towards virtual representations of visual scenes
von: Pereira, Américo, et al.
Veröffentlicht: (2024)
von: Pereira, Américo, et al.
Veröffentlicht: (2024)
AI-driven visual monitoring of industrial assembly tasks
von: Nardon, Mattia, et al.
Veröffentlicht: (2025)
von: Nardon, Mattia, et al.
Veröffentlicht: (2025)
Perception Encoder: The best visual embeddings are not at the output of the network
von: Bolya, Daniel, et al.
Veröffentlicht: (2025)
von: Bolya, Daniel, et al.
Veröffentlicht: (2025)
Can visual language models resolve textual ambiguity with visual cues? Let visual puns tell you!
von: Chung, Jiwan, et al.
Veröffentlicht: (2024)
von: Chung, Jiwan, et al.
Veröffentlicht: (2024)
Automated mapping of virtual environments with visual predictive coding
von: Gornet, James, et al.
Veröffentlicht: (2023)
von: Gornet, James, et al.
Veröffentlicht: (2023)
Block-Sparse Global Attention for Efficient Multi-View Geometry Transformers
von: Wang, Chung-Shien Brian, et al.
Veröffentlicht: (2025)
von: Wang, Chung-Shien Brian, et al.
Veröffentlicht: (2025)
Incremental dimension reduction for efficient and accurate visual anomaly detection
von: Lee, Teng-Yok
Veröffentlicht: (2026)
von: Lee, Teng-Yok
Veröffentlicht: (2026)
SCHIGAND: A Synthetic Facial Generation Mode Pipeline
von: Kadali, Ananya, et al.
Veröffentlicht: (2026)
von: Kadali, Ananya, et al.
Veröffentlicht: (2026)
Affine transformation estimation improves visual self-supervised learning
von: Torpey, David, et al.
Veröffentlicht: (2024)
von: Torpey, David, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
An extremely coarse feedback signal is sufficient for learning human-aligned visual representations
von: Mehta, Yash, et al.
Veröffentlicht: (2026) -
Universal dimensions of visual representation
von: Chen, Zirui, et al.
Veröffentlicht: (2024) -
Rapidly deploying on-device eye tracking by distilling visual foundation models
von: Jiang, Cheng, et al.
Veröffentlicht: (2026) -
SAGE: Spatial-visual Adaptive Graph Exploration for Efficient Visual Place Recognition
von: Chen, Shunpeng, et al.
Veröffentlicht: (2025) -
Evaluating the Suitability of Different Intraoral Scan Resolutions for Deep Learning-Based Tooth Segmentation
von: Weekley, Daron, et al.
Veröffentlicht: (2025)