Improving Contrastive Learning for Referring Expression Counting
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Triaridis, Kostas, Kaliosis, Panagiotis, Nguyen, E-Ro, Xu, Jingyi, Le, Hieu, Samaras, Dimitris |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Phrase-Instance Alignment for Generalized Referring Segmentation
von: Nguyen, E-Ro, et al.
Veröffentlicht: (2024)
von: Nguyen, E-Ro, et al.
Veröffentlicht: (2024)
Assessing Sample Quality via the Latent Space of Generative Models
von: Xu, Jingyi, et al.
Veröffentlicht: (2024)
von: Xu, Jingyi, et al.
Veröffentlicht: (2024)
Importance-Based Token Merging for Efficient Image and Video Generation
von: Wu, Haoyu, et al.
Veröffentlicht: (2024)
von: Wu, Haoyu, et al.
Veröffentlicht: (2024)
Talking Head Generation via AU-Guided Landmark Prediction
von: Chang, Shao-Yu, et al.
Veröffentlicht: (2025)
von: Chang, Shao-Yu, et al.
Veröffentlicht: (2025)
Mitigating Diffusion Model Hallucinations with Dynamic Guidance
von: Triaridis, Kostas, et al.
Veröffentlicht: (2025)
von: Triaridis, Kostas, et al.
Veröffentlicht: (2025)
One Attention, One Scale: Phase-Aligned Rotary Positional Embeddings for Mixed-Resolution Diffusion Transformer
von: Wu, Haoyu, et al.
Veröffentlicht: (2025)
von: Wu, Haoyu, et al.
Veröffentlicht: (2025)
Embedding Physical Reasoning into Diffusion-Based Shadow Generation
von: Hu, Shilin, et al.
Veröffentlicht: (2025)
von: Hu, Shilin, et al.
Veröffentlicht: (2025)
Cast and Attached Shadow Detection via Iterative Light and Geometry Reasoning
von: Hu, Shilin, et al.
Veröffentlicht: (2025)
von: Hu, Shilin, et al.
Veröffentlicht: (2025)
Weighting Pseudo-Labels via High-Activation Feature Index Similarity and Object Detection for Semi-Supervised Segmentation
von: Howlader, Prantik, et al.
Veröffentlicht: (2024)
von: Howlader, Prantik, et al.
Veröffentlicht: (2024)
CORA: Consistency-Guided Semi-Supervised Framework for Reasoning Segmentation
von: Howlader, Prantik, et al.
Veröffentlicht: (2025)
von: Howlader, Prantik, et al.
Veröffentlicht: (2025)
ZoomLDM: Latent Diffusion Model for multi-scale image generation
von: Yellapragada, Srikar, et al.
Veröffentlicht: (2024)
von: Yellapragada, Srikar, et al.
Veröffentlicht: (2024)
Learning to Align: Addressing Character Frequency Distribution Shifts in Handwritten Text Recognition
von: Kaliosis, Panagiotis, et al.
Veröffentlicht: (2025)
von: Kaliosis, Panagiotis, et al.
Veröffentlicht: (2025)
Beyond Pixels: Semi-Supervised Semantic Segmentation with a Multi-scale Patch-based Multi-Label Classifier
von: Howlader, Prantik, et al.
Veröffentlicht: (2024)
von: Howlader, Prantik, et al.
Veröffentlicht: (2024)
Few-shot Personalized Scanpath Prediction
von: Xue, Ruoyu, et al.
Veröffentlicht: (2025)
von: Xue, Ruoyu, et al.
Veröffentlicht: (2025)
MMFusion: Combining Image Forensic Filters for Visual Manipulation Detection and Localization
von: Triaridis, Kostas, et al.
Veröffentlicht: (2023)
von: Triaridis, Kostas, et al.
Veröffentlicht: (2023)
Personalized Image Descriptions from Attention Sequences
von: Xue, Ruoyu, et al.
Veröffentlicht: (2025)
von: Xue, Ruoyu, et al.
Veröffentlicht: (2025)
Shadow Removal Refinement via Material-Consistent Shadow Edges
von: Hu, Shilin, et al.
Veröffentlicht: (2024)
von: Hu, Shilin, et al.
Veröffentlicht: (2024)
JEAN: Joint Expression and Audio-guided NeRF-based Talking Face Generation
von: Chakkera, Sai Tanmay Reddy, et al.
Veröffentlicht: (2024)
von: Chakkera, Sai Tanmay Reddy, et al.
Veröffentlicht: (2024)
Counting Stacked Objects
von: Dumery, Corentin, et al.
Veröffentlicht: (2024)
von: Dumery, Corentin, et al.
Veröffentlicht: (2024)
Decoupling What to Count and Where to See for Referring Expression Counting
von: Zou, Yuda, et al.
Veröffentlicht: (2025)
von: Zou, Yuda, et al.
Veröffentlicht: (2025)
Self-supervised co-salient object detection via feature correspondence at multiple scales
von: Chakraborty, Souradeep, et al.
Veröffentlicht: (2024)
von: Chakraborty, Souradeep, et al.
Veröffentlicht: (2024)
Pathology Image Compression with Pre-trained Autoencoders
von: Yellapragada, Srikar, et al.
Veröffentlicht: (2025)
von: Yellapragada, Srikar, et al.
Veröffentlicht: (2025)
Automated Counting of Stacked Objects in Industrial Inspection
von: Dumery, Corentin, et al.
Veröffentlicht: (2026)
von: Dumery, Corentin, et al.
Veröffentlicht: (2026)
MI-NeRF: Learning a Single Face NeRF from Multiple Identities
von: Chatziagapi, Aggelina, et al.
Veröffentlicht: (2024)
von: Chatziagapi, Aggelina, et al.
Veröffentlicht: (2024)
What about gravity in video generation? Post-Training Newton's Laws with Verifiable Rewards
von: Le, Minh-Quan, et al.
Veröffentlicht: (2025)
von: Le, Minh-Quan, et al.
Veröffentlicht: (2025)
Learning 3D Reconstruction with Priors in Test Time
von: Zhou, Lei, et al.
Veröffentlicht: (2026)
von: Zhou, Lei, et al.
Veröffentlicht: (2026)
TopoDiffusionNet: A Topology-aware Diffusion Model
von: Gupta, Saumya, et al.
Veröffentlicht: (2024)
von: Gupta, Saumya, et al.
Veröffentlicht: (2024)
Multi-view Gaze Target Estimation
von: Miao, Qiaomu, et al.
Veröffentlicht: (2025)
von: Miao, Qiaomu, et al.
Veröffentlicht: (2025)
Exploring Contextual Attribute Density in Referring Expression Counting
von: Wang, Zhicheng, et al.
Veröffentlicht: (2025)
von: Wang, Zhicheng, et al.
Veröffentlicht: (2025)
Learning to Weight Parameters for Training Data Attribution
von: Li, Shuangqi, et al.
Veröffentlicht: (2025)
von: Li, Shuangqi, et al.
Veröffentlicht: (2025)
MIGS: Multi-Identity Gaussian Splatting via Tensor Decomposition
von: Chatziagapi, Aggelina, et al.
Veröffentlicht: (2024)
von: Chatziagapi, Aggelina, et al.
Veröffentlicht: (2024)
CAKE: Real-time Action Detection via Motion Distillation and Background-aware Contrastive Learning
von: Hoang, Hieu, et al.
Veröffentlicht: (2026)
von: Hoang, Hieu, et al.
Veröffentlicht: (2026)
GECKO: Gigapixel Vision-Concept Contrastive Pretraining in Histopathology
von: Kapse, Saarthak, et al.
Veröffentlicht: (2025)
von: Kapse, Saarthak, et al.
Veröffentlicht: (2025)
MonoLoss: A Training Objective for Interpretable Monosemantic Representations
von: Nasiri-Sarvi, Ali, et al.
Veröffentlicht: (2026)
von: Nasiri-Sarvi, Ali, et al.
Veröffentlicht: (2026)
Fast constrained sampling in pre-trained diffusion models
von: Graikos, Alexandros, et al.
Veröffentlicht: (2024)
von: Graikos, Alexandros, et al.
Veröffentlicht: (2024)
Learned representation-guided diffusion models for large-image generation
von: Graikos, Alexandros, et al.
Veröffentlicht: (2023)
von: Graikos, Alexandros, et al.
Veröffentlicht: (2023)
Pairwise-Constrained Implicit Functions for 3D Human Heart Modelling
von: Le, Hieu, et al.
Veröffentlicht: (2023)
von: Le, Hieu, et al.
Veröffentlicht: (2023)
All Seeds Are Not Equal: Enhancing Compositional Text-to-Image Generation with Reliable Random Seeds
von: Li, Shuangqi, et al.
Veröffentlicht: (2024)
von: Li, Shuangqi, et al.
Veröffentlicht: (2024)
Learning Relighting and Intrinsic Decomposition in Neural Radiance Fields
von: Yang, Yixiong, et al.
Veröffentlicht: (2024)
von: Yang, Yixiong, et al.
Veröffentlicht: (2024)
PathSegDiff: Pathology Segmentation using Diffusion model representations
von: Danisetty, Sachin Kumar, et al.
Veröffentlicht: (2025)
von: Danisetty, Sachin Kumar, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Phrase-Instance Alignment for Generalized Referring Segmentation
von: Nguyen, E-Ro, et al.
Veröffentlicht: (2024) -
Assessing Sample Quality via the Latent Space of Generative Models
von: Xu, Jingyi, et al.
Veröffentlicht: (2024) -
Importance-Based Token Merging for Efficient Image and Video Generation
von: Wu, Haoyu, et al.
Veröffentlicht: (2024) -
Talking Head Generation via AU-Guided Landmark Prediction
von: Chang, Shao-Yu, et al.
Veröffentlicht: (2025) -
Mitigating Diffusion Model Hallucinations with Dynamic Guidance
von: Triaridis, Kostas, et al.
Veröffentlicht: (2025)