CATVis: Context-Aware Thought Visualization
Fuente:
arXiv
Guardado en:
| Autores principales: | Mehmood, Tariq, Ahmad, Hamza, Shakeel, Muhammad Haroon, Taj, Murtaza |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
DFA-CON: A Contrastive Learning Approach for Detecting Copyright Infringement in DeepFake Art
por: Wahab, Haroon, et al.
Publicado: (2025)
por: Wahab, Haroon, et al.
Publicado: (2025)
Camera Calibration through Geometric Constraints from Rotation and Projection Matrices
por: Waleed, Muhammad, et al.
Publicado: (2024)
por: Waleed, Muhammad, et al.
Publicado: (2024)
A Deep Features-Based Approach Using Modified ResNet50 and Gradient Boosting for Visual Sentiments Classification
por: Arslan, Muhammad, et al.
Publicado: (2024)
por: Arslan, Muhammad, et al.
Publicado: (2024)
DocVXQA: Context-Aware Visual Explanations for Document Question Answering
por: Souibgui, Mohamed Ali, et al.
Publicado: (2025)
por: Souibgui, Mohamed Ali, et al.
Publicado: (2025)
S-Chain: Structured Visual Chain-of-Thought For Medicine
por: Le-Duc, Khai, et al.
Publicado: (2025)
por: Le-Duc, Khai, et al.
Publicado: (2025)
Thermal Vision: Pioneering Non-Invasive Temperature Tracking in Congested Spaces
por: Samal, Arijit, et al.
Publicado: (2024)
por: Samal, Arijit, et al.
Publicado: (2024)
PromptHub: Enhancing Multi-Prompt Visual In-Context Learning with Locality-Aware Fusion, Concentration and Alignment
por: Luo, Tianci, et al.
Publicado: (2026)
por: Luo, Tianci, et al.
Publicado: (2026)
Context-Aware Meta-Learning
por: Fifty, Christopher, et al.
Publicado: (2023)
por: Fifty, Christopher, et al.
Publicado: (2023)
Bias Redistribution in Visual Machine Unlearning: Does Forgetting One Group Harm Another?
por: Haruna, Yunusa, et al.
Publicado: (2026)
por: Haruna, Yunusa, et al.
Publicado: (2026)
VisReason: A Large-Scale Dataset for Visual Chain-of-Thought Reasoning
por: Li, Lingxiao, et al.
Publicado: (2025)
por: Li, Lingxiao, et al.
Publicado: (2025)
Test-Time Visual In-Context Tuning
por: Xie, Jiahao, et al.
Publicado: (2025)
por: Xie, Jiahao, et al.
Publicado: (2025)
Unlocking Generalization for Robotics via Modularity and Scale
por: Dalal, Murtaza
Publicado: (2025)
por: Dalal, Murtaza
Publicado: (2025)
Imagine while Reasoning in Space: Multimodal Visualization-of-Thought
por: Li, Chengzu, et al.
Publicado: (2025)
por: Li, Chengzu, et al.
Publicado: (2025)
A Realistic Protocol for Evaluation of Weakly Supervised Object Localization
por: Murtaza, Shakeeb, et al.
Publicado: (2024)
por: Murtaza, Shakeeb, et al.
Publicado: (2024)
Leveraging Transformers for Weakly Supervised Object Localization in Unconstrained Videos
por: Murtaza, Shakeeb, et al.
Publicado: (2024)
por: Murtaza, Shakeeb, et al.
Publicado: (2024)
Personalized Vision via Visual In-Context Learning
por: Jiang, Yuxin, et al.
Publicado: (2025)
por: Jiang, Yuxin, et al.
Publicado: (2025)
Multi-Task Learning for Visually Grounded Reasoning in Gastrointestinal VQA
por: Safwan, Itbaan, et al.
Publicado: (2025)
por: Safwan, Itbaan, et al.
Publicado: (2025)
CASHG: Context-Aware Stylized Online Handwriting Generation
por: Shin, Jinsu, et al.
Publicado: (2026)
por: Shin, Jinsu, et al.
Publicado: (2026)
Context Sensitivity Improves Human-Machine Visual Alignment
por: Born, Frieda, et al.
Publicado: (2026)
por: Born, Frieda, et al.
Publicado: (2026)
Context-Aware Multimodal Pretraining
por: Roth, Karsten, et al.
Publicado: (2024)
por: Roth, Karsten, et al.
Publicado: (2024)
Stable Diffusion Models are Secretly Good at Visual In-Context Learning
por: Oorloff, Trevine, et al.
Publicado: (2025)
por: Oorloff, Trevine, et al.
Publicado: (2025)
Spatio-Temporal driven Attention Graph Neural Network with Block Adjacency matrix (STAG-NN-BA) for Remote Land-use Change Detection
por: Nazir, Usman, et al.
Publicado: (2023)
por: Nazir, Usman, et al.
Publicado: (2023)
TeD-Loc: Text Distillation for Weakly Supervised Object Localization
por: Murtaza, Shakeeb, et al.
Publicado: (2025)
por: Murtaza, Shakeeb, et al.
Publicado: (2025)
Capturing Context-Aware Route Choice Semantics for Trajectory Representation Learning
por: Cao, Ji, et al.
Publicado: (2025)
por: Cao, Ji, et al.
Publicado: (2025)
Exploring State-of-the-art models for Early Detection of Forest Fires
por: Ahmed, Sharjeel, et al.
Publicado: (2025)
por: Ahmed, Sharjeel, et al.
Publicado: (2025)
Architecture-Aware Explanation Auditing for Industrial Visual Inspection
por: Jia, Sibo, et al.
Publicado: (2026)
por: Jia, Sibo, et al.
Publicado: (2026)
Learning to Compress Contexts for Efficient Knowledge-based Visual Question Answering
por: Weng, Weixi, et al.
Publicado: (2024)
por: Weng, Weixi, et al.
Publicado: (2024)
Neural Language of Thought Models
por: Wu, Yi-Fu, et al.
Publicado: (2024)
por: Wu, Yi-Fu, et al.
Publicado: (2024)
LFRA-Net: A Lightweight Focal and Region-Aware Attention Network for Retinal Vessel Segmentatio
por: Mehmood, Mehwish, et al.
Publicado: (2025)
por: Mehmood, Mehwish, et al.
Publicado: (2025)
Chain-of-Visual-Thought: Teaching VLMs to See and Think Better with Continuous Visual Tokens
por: Qin, Yiming, et al.
Publicado: (2025)
por: Qin, Yiming, et al.
Publicado: (2025)
Context-Based Semantic-Aware Alignment for Semi-Supervised Multi-Label Learning
por: Fan, Heng-Bo, et al.
Publicado: (2024)
por: Fan, Heng-Bo, et al.
Publicado: (2024)
Context-Aware Aerial Object Detection: Leveraging Inter-Object and Background Relationships
por: Ren, Botao, et al.
Publicado: (2024)
por: Ren, Botao, et al.
Publicado: (2024)
Gated-Attention Feature-Fusion Based Framework for Poverty Prediction
por: Ramzan, Muhammad Umer, et al.
Publicado: (2024)
por: Ramzan, Muhammad Umer, et al.
Publicado: (2024)
Trend-Aware Fashion Recommendation with Visual Segmentation and Semantic Similarity
por: Djilani, Mohamed, et al.
Publicado: (2025)
por: Djilani, Mohamed, et al.
Publicado: (2025)
Visual Analysis of Prediction Uncertainty in Neural Networks for Deep Image Synthesis
por: Dutta, Soumya, et al.
Publicado: (2024)
por: Dutta, Soumya, et al.
Publicado: (2024)
CapCLIP: A Vision-Language Representation Alignment Approach for Wireless Capsule Endoscopy Analysis
por: Wahab, Haroon, et al.
Publicado: (2026)
por: Wahab, Haroon, et al.
Publicado: (2026)
Improved Crop and Weed Detection with Diverse Data Ensemble Learning
por: Asad, Muhammad Hamza, et al.
Publicado: (2023)
por: Asad, Muhammad Hamza, et al.
Publicado: (2023)
The Effects of Grouped Structural Global Pruning of Vision Transformers on Domain Generalisation
por: Riaz, Hamza, et al.
Publicado: (2025)
por: Riaz, Hamza, et al.
Publicado: (2025)
Channel-Aware Probing for Multi-Channel Imaging
por: Marikkar, Umar, et al.
Publicado: (2026)
por: Marikkar, Umar, et al.
Publicado: (2026)
CAMELTrack: Context-Aware Multi-cue ExpLoitation for Online Multi-Object Tracking
por: Somers, Vladimir, et al.
Publicado: (2025)
por: Somers, Vladimir, et al.
Publicado: (2025)
Ejemplares similares
-
DFA-CON: A Contrastive Learning Approach for Detecting Copyright Infringement in DeepFake Art
por: Wahab, Haroon, et al.
Publicado: (2025) -
Camera Calibration through Geometric Constraints from Rotation and Projection Matrices
por: Waleed, Muhammad, et al.
Publicado: (2024) -
A Deep Features-Based Approach Using Modified ResNet50 and Gradient Boosting for Visual Sentiments Classification
por: Arslan, Muhammad, et al.
Publicado: (2024) -
DocVXQA: Context-Aware Visual Explanations for Document Question Answering
por: Souibgui, Mohamed Ali, et al.
Publicado: (2025) -
S-Chain: Structured Visual Chain-of-Thought For Medicine
por: Le-Duc, Khai, et al.
Publicado: (2025)