Glance-MCMT: A General MCMT Framework with Glance Initialization and Progressive Association
Fuente:
arXiv
Guardado en:
| Autor principal: | Hashempoor, Hamidreza |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
GlanceVAD: Exploring Glance Supervision for Label-efficient Video Anomaly Detection
por: Zhang, Huaxin, et al.
Publicado: (2024)
por: Zhang, Huaxin, et al.
Publicado: (2024)
FastTracker: Real-Time and Accurate Visual Tracking
por: Hashempoor, Hamidreza, et al.
Publicado: (2025)
por: Hashempoor, Hamidreza, et al.
Publicado: (2025)
Glance: Accelerating Diffusion Models with 1 Sample
por: Dong, Zhuobai, et al.
Publicado: (2025)
por: Dong, Zhuobai, et al.
Publicado: (2025)
Glance and Focus Reinforcement for Pan-cancer Screening
por: Wu, Linshan, et al.
Publicado: (2026)
por: Wu, Linshan, et al.
Publicado: (2026)
FeatureSORT: Essential Features for Effective Tracking
por: Hashempoor, Hamidreza, et al.
Publicado: (2024)
por: Hashempoor, Hamidreza, et al.
Publicado: (2024)
Glance and Focus: Memory Prompting for Multi-Event Video Question Answering
por: Bai, Ziyi, et al.
Publicado: (2024)
por: Bai, Ziyi, et al.
Publicado: (2024)
Hallucination at a Glance: Controlled Visual Edits and Fine-Grained Multimodal Learning
por: Bai, Tianyi, et al.
Publicado: (2025)
por: Bai, Tianyi, et al.
Publicado: (2025)
Glance-or-Gaze: Incentivizing LMMs to Adaptively Focus Search via Reinforcement Learning
por: Bai, Hongbo, et al.
Publicado: (2026)
por: Bai, Hongbo, et al.
Publicado: (2026)
Insights from Visual Cognition: Understanding Human Action Dynamics with Overall Glance and Refined Gaze Transformer
por: Xing, Bohao, et al.
Publicado: (2026)
por: Xing, Bohao, et al.
Publicado: (2026)
Learning at a Glance: Towards Interpretable Data-limited Continual Semantic Segmentation via Semantic-Invariance Modelling
por: Yuan, Bo, et al.
Publicado: (2024)
por: Yuan, Bo, et al.
Publicado: (2024)
Every Angle Is Worth A Second Glance: Mining Kinematic Skeletal Structures from Multi-view Joint Cloud
por: Jiang, Junkun, et al.
Publicado: (2025)
por: Jiang, Junkun, et al.
Publicado: (2025)
Just a Few Glances: Open-Set Visual Perception with Image Prompt Paradigm
por: Zhang, Jinrong, et al.
Publicado: (2024)
por: Zhang, Jinrong, et al.
Publicado: (2024)
MS-Glance: Bio-Insipred Non-semantic Context Vectors and their Applications in Supervising Image Reconstruction
por: Gao, Ziqi, et al.
Publicado: (2024)
por: Gao, Ziqi, et al.
Publicado: (2024)
DistillNeRF: Perceiving 3D Scenes from Single-Glance Images by Distilling Neural Fields and Foundation Model Features
por: Wang, Letian, et al.
Publicado: (2024)
por: Wang, Letian, et al.
Publicado: (2024)
Initialize to Generalize: A Stronger Initialization Pipeline for Sparse-View 3DGS
por: Zhou, Feng, et al.
Publicado: (2025)
por: Zhou, Feng, et al.
Publicado: (2025)
A General Framework to Boost 3D GS Initialization for Text-to-3D Generation by Lexical Richness
por: Jiang, Lutao, et al.
Publicado: (2024)
por: Jiang, Lutao, et al.
Publicado: (2024)
LM-MCVT: A Lightweight Multi-modal Multi-view Convolutional-Vision Transformer Approach for 3D Object Recognition
por: Xiong, Songsong, et al.
Publicado: (2025)
por: Xiong, Songsong, et al.
Publicado: (2025)
X-Oscar: A Progressive Framework for High-quality Text-guided 3D Animatable Avatar Generation
por: Ma, Yiwei, et al.
Publicado: (2024)
por: Ma, Yiwei, et al.
Publicado: (2024)
Asymmetric Dual Self-Distillation for 3D Self-Supervised Representation Learning
por: Leijenaar, Remco F., et al.
Publicado: (2025)
por: Leijenaar, Remco F., et al.
Publicado: (2025)
ETTA: Efficient Test-Time Adaptation for Vision-Language Models through Dynamic Embedding Updates
por: Dastmalchi, Hamidreza, et al.
Publicado: (2025)
por: Dastmalchi, Hamidreza, et al.
Publicado: (2025)
Towards Open-World Grasping with Large Vision-Language Models
por: Tziafas, Georgios, et al.
Publicado: (2024)
por: Tziafas, Georgios, et al.
Publicado: (2024)
A Progressive Framework of Vision-language Knowledge Distillation and Alignment for Multilingual Scene
por: Zhang, Wenbo, et al.
Publicado: (2024)
por: Zhang, Wenbo, et al.
Publicado: (2024)
Text2Graph VPR: A Text-to-Graph Expert System for Explainable Place Recognition in Changing Environments
por: Yousefzadeh, Saeideh, et al.
Publicado: (2025)
por: Yousefzadeh, Saeideh, et al.
Publicado: (2025)
Progressive Checkerboards for Autoregressive Multiscale Image Generation
por: Eigen, David
Publicado: (2026)
por: Eigen, David
Publicado: (2026)
Fighting Hallucinations with Counterfactuals: Diffusion-Guided Perturbations for LVLM Hallucination Suppression
por: Dastmalchi, Hamidreza, et al.
Publicado: (2026)
por: Dastmalchi, Hamidreza, et al.
Publicado: (2026)
GLip: A Global-Local Integrated Progressive Framework for Robust Visual Speech Recognition
por: Wang, Tianyue, et al.
Publicado: (2025)
por: Wang, Tianyue, et al.
Publicado: (2025)
Twin Co-Adaptive Dialogue for Progressive Image Generation
por: Wang, Jianhui, et al.
Publicado: (2025)
por: Wang, Jianhui, et al.
Publicado: (2025)
Enhancing Image Generation Fidelity via Progressive Prompts
por: Xiong, Zhen, et al.
Publicado: (2025)
por: Xiong, Zhen, et al.
Publicado: (2025)
Spectral Progressive Diffusion for Efficient Image and Video Generation
por: Xiao, Howard, et al.
Publicado: (2026)
por: Xiao, Howard, et al.
Publicado: (2026)
PSF-4D: A Progressive Sampling Framework for View Consistent 4D Editing
por: Iqbal, Hasan, et al.
Publicado: (2025)
por: Iqbal, Hasan, et al.
Publicado: (2025)
Structured Initialization for Vision Transformers
por: Zheng, Jianqiao, et al.
Publicado: (2025)
por: Zheng, Jianqiao, et al.
Publicado: (2025)
DrivePTS: A Progressive Learning Framework with Textual and Structural Enhancement for Driving Scene Generation
por: Wang, Zhechao, et al.
Publicado: (2026)
por: Wang, Zhechao, et al.
Publicado: (2026)
HABIT: Chrono-Synergia Robust Progressive Learning Framework for Composed Image Retrieval
por: Li, Zixu, et al.
Publicado: (2026)
por: Li, Zixu, et al.
Publicado: (2026)
U-Turn Diffusion
por: Behjoo, Hamidreza, et al.
Publicado: (2023)
por: Behjoo, Hamidreza, et al.
Publicado: (2023)
Video Analysis and Generation via a Semantic Progress Function
por: Metzer, Gal, et al.
Publicado: (2026)
por: Metzer, Gal, et al.
Publicado: (2026)
Recurrent Generic Contour-based Instance Segmentation with Progressive Learning
por: Feng, Hao, et al.
Publicado: (2023)
por: Feng, Hao, et al.
Publicado: (2023)
Longitudinal NSCLC Treatment Progression via Multimodal Generative Models
por: Mantegna, Massimiliano, et al.
Publicado: (2026)
por: Mantegna, Massimiliano, et al.
Publicado: (2026)
Gaussian On-the-Fly Splatting: A Progressive Framework for Robust Near Real-Time 3DGS Optimization
por: Xu, Yiwei, et al.
Publicado: (2025)
por: Xu, Yiwei, et al.
Publicado: (2025)
A Progressive Single-Modality to Multi-Modality Classification Framework for Alzheimer's Disease Sub-type Diagnosis
por: Liu, Yuxiao, et al.
Publicado: (2024)
por: Liu, Yuxiao, et al.
Publicado: (2024)
Hypothesis Testing for Progressive Kernel Estimation and VCM Framework
por: Lin, Zehui, et al.
Publicado: (2025)
por: Lin, Zehui, et al.
Publicado: (2025)
Ejemplares similares
-
GlanceVAD: Exploring Glance Supervision for Label-efficient Video Anomaly Detection
por: Zhang, Huaxin, et al.
Publicado: (2024) -
FastTracker: Real-Time and Accurate Visual Tracking
por: Hashempoor, Hamidreza, et al.
Publicado: (2025) -
Glance: Accelerating Diffusion Models with 1 Sample
por: Dong, Zhuobai, et al.
Publicado: (2025) -
Glance and Focus Reinforcement for Pan-cancer Screening
por: Wu, Linshan, et al.
Publicado: (2026) -
FeatureSORT: Essential Features for Effective Tracking
por: Hashempoor, Hamidreza, et al.
Publicado: (2024)