Compositional Entailment Learning for Hyperbolic Vision-Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Pal, Avik, van Spengler, Max, di Melendugno, Guido Maria D'Amely, Flaborea, Alessandro, Galasso, Fabio, Mettes, Pascal |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Hyp2Nav: Hyperbolic Planning and Curiosity for Crowd Navigation
von: di Melendugno, Guido Maria D'Amely, et al.
Veröffentlicht: (2024)
von: di Melendugno, Guido Maria D'Amely, et al.
Veröffentlicht: (2024)
PREGO: online mistake detection in PRocedural EGOcentric videos
von: Flaborea, Alessandro, et al.
Veröffentlicht: (2024)
von: Flaborea, Alessandro, et al.
Veröffentlicht: (2024)
TI-PREGO: Chain of Thought and In-Context Learning for Online Mistake Detection in PRocedural EGOcentric Videos
von: Plini, Leonardo, et al.
Veröffentlicht: (2024)
von: Plini, Leonardo, et al.
Veröffentlicht: (2024)
Contracting Skeletal Kinematics for Human-Related Video Anomaly Detection
von: Flaborea, Alessandro, et al.
Veröffentlicht: (2023)
von: Flaborea, Alessandro, et al.
Veröffentlicht: (2023)
ANTHROPOS-V: benchmarking the novel task of Crowd Volume Estimation
von: Collorone, Luca, et al.
Veröffentlicht: (2025)
von: Collorone, Luca, et al.
Veröffentlicht: (2025)
Balanced Hyperbolic Embeddings Are Natural Out-of-Distribution Detectors
von: Kasarla, Tejaswi, et al.
Veröffentlicht: (2025)
von: Kasarla, Tejaswi, et al.
Veröffentlicht: (2025)
Low-distortion and GPU-compatible Tree Embeddings in Hyperbolic Space
von: van Spengler, Max, et al.
Veröffentlicht: (2025)
von: van Spengler, Max, et al.
Veröffentlicht: (2025)
Not All Latent Spaces Are Flat: Hyperbolic Concept Control
von: Briglia, Maria Rosaria, et al.
Veröffentlicht: (2026)
von: Briglia, Maria Rosaria, et al.
Veröffentlicht: (2026)
Adversarial Attacks on Hyperbolic Networks
von: van Spengler, Max, et al.
Veröffentlicht: (2024)
von: van Spengler, Max, et al.
Veröffentlicht: (2024)
Hyperbolic Safety-Aware Vision-Language Models
von: Poppi, Tobia, et al.
Veröffentlicht: (2025)
von: Poppi, Tobia, et al.
Veröffentlicht: (2025)
Hyperbolic Concept Bottleneck Models
von: Uyterlinde, Daniel, et al.
Veröffentlicht: (2026)
von: Uyterlinde, Daniel, et al.
Veröffentlicht: (2026)
In-Context Learning Improves Compositional Understanding of Vision-Language Models
von: Nulli, Matteo, et al.
Veröffentlicht: (2024)
von: Nulli, Matteo, et al.
Veröffentlicht: (2024)
Continual Hyperbolic Learning of Instances and Classes
von: Ayoughi, Melika, et al.
Veröffentlicht: (2025)
von: Ayoughi, Melika, et al.
Veröffentlicht: (2025)
HYPE: Hyperbolic Entailment Filtering for Underspecified Images and Texts
von: Kim, Wonjae, et al.
Veröffentlicht: (2024)
von: Kim, Wonjae, et al.
Veröffentlicht: (2024)
Looking Beyond the Obvious: A Survey on Abstract Concept Recognition for Video Understanding
von: Mago, Gowreesh, et al.
Veröffentlicht: (2025)
von: Mago, Gowreesh, et al.
Veröffentlicht: (2025)
DELST: Dual Entailment Learning for Hyperbolic Image-Gene Pretraining in Spatial Transcriptomics
von: Chen, Xulin, et al.
Veröffentlicht: (2025)
von: Chen, Xulin, et al.
Veröffentlicht: (2025)
Hyperbolic Active Learning for Semantic Segmentation under Domain Shift
von: Franco, Luca, et al.
Veröffentlicht: (2023)
von: Franco, Luca, et al.
Veröffentlicht: (2023)
HyperAlign: Hyperbolic Entailment Cones for Adaptive Text-to-Image Alignment Assessment
von: Chen, Wenzhi, et al.
Veröffentlicht: (2026)
von: Chen, Wenzhi, et al.
Veröffentlicht: (2026)
VELOCITI: Benchmarking Video-Language Compositional Reasoning with Strict Entailment
von: Saravanan, Darshana, et al.
Veröffentlicht: (2024)
von: Saravanan, Darshana, et al.
Veröffentlicht: (2024)
Find the Cliffhanger: Multi-Modal Trailerness in Soap Operas
von: Bretti, Carlo, et al.
Veröffentlicht: (2024)
von: Bretti, Carlo, et al.
Veröffentlicht: (2024)
Probing Vision-Language Understanding through the Visual Entailment Task: promises and pitfalls
von: Pitta, Elena, et al.
Veröffentlicht: (2025)
von: Pitta, Elena, et al.
Veröffentlicht: (2025)
Global and Local Entailment Learning for Natural World Imagery
von: Sastry, Srikumar, et al.
Veröffentlicht: (2025)
von: Sastry, Srikumar, et al.
Veröffentlicht: (2025)
PHyCLIP: $\ell_1$-Product of Hyperbolic Factors Unifies Hierarchy and Compositionality in Vision-Language Representation Learning
von: Yoshikawa, Daiki, et al.
Veröffentlicht: (2025)
von: Yoshikawa, Daiki, et al.
Veröffentlicht: (2025)
About latent roles in forecasting players in team sports
von: Scofano, Luca, et al.
Veröffentlicht: (2023)
von: Scofano, Luca, et al.
Veröffentlicht: (2023)
PhysTalk: Language-driven Real-time Physics in 3D Gaussian Scenes
von: Collorone, Luca, et al.
Veröffentlicht: (2025)
von: Collorone, Luca, et al.
Veröffentlicht: (2025)
Union-over-Intersections: Object Detection beyond Winner-Takes-All
von: Bhowmik, Aritra, et al.
Veröffentlicht: (2023)
von: Bhowmik, Aritra, et al.
Veröffentlicht: (2023)
Length-Aware Motion Synthesis via Latent Diffusion
von: Sampieri, Alessio, et al.
Veröffentlicht: (2024)
von: Sampieri, Alessio, et al.
Veröffentlicht: (2024)
Uncertainty-guided Compositional Alignment with Part-to-Whole Semantic Representativeness in Hyperbolic Vision-Language Models
von: Kim, Hayeon, et al.
Veröffentlicht: (2026)
von: Kim, Hayeon, et al.
Veröffentlicht: (2026)
MoDiPO: text-to-motion alignment via AI-feedback-driven Direct Preference Optimization
von: Pappa, Massimiliano, et al.
Veröffentlicht: (2024)
von: Pappa, Massimiliano, et al.
Veröffentlicht: (2024)
Lightweight Uncertainty Quantification with Simplex Semantic Segmentation for Terrain Traversability
von: Dijk, Judith, et al.
Veröffentlicht: (2024)
von: Dijk, Judith, et al.
Veröffentlicht: (2024)
Hyperbolic and Evidence-Prioritized Experts for Large Vision-Language Models
von: Zhou, Zijie, et al.
Veröffentlicht: (2026)
von: Zhou, Zijie, et al.
Veröffentlicht: (2026)
Iterated Learning Improves Compositionality in Large Vision-Language Models
von: Zheng, Chenhao, et al.
Veröffentlicht: (2024)
von: Zheng, Chenhao, et al.
Veröffentlicht: (2024)
PiercingEye: Dual-Space Video Violence Detection with Hyperbolic Vision-Language Guidance
von: Leng, Jiaxu, et al.
Veröffentlicht: (2025)
von: Leng, Jiaxu, et al.
Veröffentlicht: (2025)
Human Motion Unlearning
von: De Matteis, Edoardo, et al.
Veröffentlicht: (2025)
von: De Matteis, Edoardo, et al.
Veröffentlicht: (2025)
Social EgoMesh Estimation
von: Scofano, Luca, et al.
Veröffentlicht: (2024)
von: Scofano, Luca, et al.
Veröffentlicht: (2024)
Joint Analysis of Optical and SAR Vegetation Indices for Vineyard Monitoring: Assessing Biomass Dynamics and Phenological Stages over Po Valley, Italy
von: Bergamaschi, Andrea, et al.
Veröffentlicht: (2025)
von: Bergamaschi, Andrea, et al.
Veröffentlicht: (2025)
$\text{H}^2$em: Learning Hierarchical Hyperbolic Embeddings for Compositional Zero-Shot Learning
von: Li, Lin, et al.
Veröffentlicht: (2025)
von: Li, Lin, et al.
Veröffentlicht: (2025)
HexFormer: Hyperbolic Vision Transformer with Exponential Map Aggregation
von: Alyoussef, Haya, et al.
Veröffentlicht: (2026)
von: Alyoussef, Haya, et al.
Veröffentlicht: (2026)
OVOSE: Open-Vocabulary Semantic Segmentation in Event-Based Cameras
von: Rahman, Muhammad Rameez Ur, et al.
Veröffentlicht: (2024)
von: Rahman, Muhammad Rameez Ur, et al.
Veröffentlicht: (2024)
CaTS-Bench: Can Language Models Describe Time Series?
von: Zhou, Luca, et al.
Veröffentlicht: (2025)
von: Zhou, Luca, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Hyp2Nav: Hyperbolic Planning and Curiosity for Crowd Navigation
von: di Melendugno, Guido Maria D'Amely, et al.
Veröffentlicht: (2024) -
PREGO: online mistake detection in PRocedural EGOcentric videos
von: Flaborea, Alessandro, et al.
Veröffentlicht: (2024) -
TI-PREGO: Chain of Thought and In-Context Learning for Online Mistake Detection in PRocedural EGOcentric Videos
von: Plini, Leonardo, et al.
Veröffentlicht: (2024) -
Contracting Skeletal Kinematics for Human-Related Video Anomaly Detection
von: Flaborea, Alessandro, et al.
Veröffentlicht: (2023) -
ANTHROPOS-V: benchmarking the novel task of Crowd Volume Estimation
von: Collorone, Luca, et al.
Veröffentlicht: (2025)