Adding Thermal Awareness to Visual Systems in Real-Time via Distilled Diffusion Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Guo, Yuchen, Gong, Junli, Dong, Wenjun, Cheung, Yiuming, Su, Weifeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Bringing Multimodal Large Language Models to Infrared-Visible Image Fusion Quality Assessment
von: Guo, Yuchen, et al.
Veröffentlicht: (2026)
von: Guo, Yuchen, et al.
Veröffentlicht: (2026)
Can Segmentation Models Understand the World? Towards Proactive Affordance Reasoning via Visual Chain-of-Thought
von: Guo, Yuchen, et al.
Veröffentlicht: (2026)
von: Guo, Yuchen, et al.
Veröffentlicht: (2026)
LumiVideo: An Intelligent Agentic System for Video Color Grading
von: Guo, Yuchen, et al.
Veröffentlicht: (2026)
von: Guo, Yuchen, et al.
Veröffentlicht: (2026)
Fuse4Seg: Image Fusion for Multi-Modal Medical Segmentation via Bi-level Optimization
von: Guo, Yuchen, et al.
Veröffentlicht: (2024)
von: Guo, Yuchen, et al.
Veröffentlicht: (2024)
DAE-Fuse: An Adaptive Discriminative Autoencoder for Multi-Modality Image Fusion
von: Guo, Yuchen, et al.
Veröffentlicht: (2024)
von: Guo, Yuchen, et al.
Veröffentlicht: (2024)
Inference-Time Diffusion Model Distillation
von: Park, Geon Yeong, et al.
Veröffentlicht: (2024)
von: Park, Geon Yeong, et al.
Veröffentlicht: (2024)
X-Adapter: Adding Universal Compatibility of Plugins for Upgraded Diffusion Model
von: Ran, Lingmin, et al.
Veröffentlicht: (2023)
von: Ran, Lingmin, et al.
Veröffentlicht: (2023)
Segment Any RGB-Thermal Model with Language-aided Distillation
von: Xing, Dong, et al.
Veröffentlicht: (2025)
von: Xing, Dong, et al.
Veröffentlicht: (2025)
Real-Time Visual Attribution Streaming in Thinking Model
von: Kang, Seil, et al.
Veröffentlicht: (2026)
von: Kang, Seil, et al.
Veröffentlicht: (2026)
ChangeDiff: A Multi-Temporal Change Detection Data Generator with Flexible Text Prompts via Diffusion Model
von: Zang, Qi, et al.
Veröffentlicht: (2024)
von: Zang, Qi, et al.
Veröffentlicht: (2024)
DiffusionAgent: Navigating Expert Models for Agentic Image Generation
von: Qin, Jie, et al.
Veröffentlicht: (2024)
von: Qin, Jie, et al.
Veröffentlicht: (2024)
LiveTalk: Real-Time Multimodal Interactive Video Diffusion via Improved On-Policy Distillation
von: Chern, Ethan, et al.
Veröffentlicht: (2025)
von: Chern, Ethan, et al.
Veröffentlicht: (2025)
Collaborative Few-Step Distillation and Low-Bit Quantization for Wan2.2 Dual-Expert Video Diffusion Models
von: Du, Jinyang, et al.
Veröffentlicht: (2026)
von: Du, Jinyang, et al.
Veröffentlicht: (2026)
HyperAlign: Hypernetwork for Efficient Test-Time Alignment of Diffusion Models
von: Xie, Xin, et al.
Veröffentlicht: (2026)
von: Xie, Xin, et al.
Veröffentlicht: (2026)
Generative Dataset Distillation Based on Diffusion Model
von: Su, Duo, et al.
Veröffentlicht: (2024)
von: Su, Duo, et al.
Veröffentlicht: (2024)
Continuous-Time Distribution Matching for Few-Step Diffusion Distillation
von: Liu, Tao, et al.
Veröffentlicht: (2026)
von: Liu, Tao, et al.
Veröffentlicht: (2026)
Robust MLLM Unlearning via Visual Knowledge Distillation
von: Wang, Yuhang, et al.
Veröffentlicht: (2025)
von: Wang, Yuhang, et al.
Veröffentlicht: (2025)
GreenEye: Development of Real-Time Traffic Signal Recognition System for Visual Impairments
von: Kim, Danu
Veröffentlicht: (2024)
von: Kim, Danu
Veröffentlicht: (2024)
Enabling Real-Time Colonoscopic Polyp Segmentation on Commodity CPUs via Ultra-Lightweight Architecture
von: Gao, Weihao, et al.
Veröffentlicht: (2026)
von: Gao, Weihao, et al.
Veröffentlicht: (2026)
Time-Aware One Step Diffusion Network for Real-World Image Super-Resolution
von: Zhang, Tianyi, et al.
Veröffentlicht: (2025)
von: Zhang, Tianyi, et al.
Veröffentlicht: (2025)
Foreground-Aware Dataset Distillation via Dynamic Patch Selection
von: Li, Longzhen, et al.
Veröffentlicht: (2026)
von: Li, Longzhen, et al.
Veröffentlicht: (2026)
Diffusion Models Are Real-Time Game Engines
von: Valevski, Dani, et al.
Veröffentlicht: (2024)
von: Valevski, Dani, et al.
Veröffentlicht: (2024)
IDA-VLM: Towards Movie Understanding via ID-Aware Large Vision-Language Model
von: Ji, Yatai, et al.
Veröffentlicht: (2024)
von: Ji, Yatai, et al.
Veröffentlicht: (2024)
PRISM: Precision-Recall Informed Data-Free Knowledge Distillation via Generative Diffusion
von: He, Xuewan, et al.
Veröffentlicht: (2025)
von: He, Xuewan, et al.
Veröffentlicht: (2025)
DiffAttn: Diffusion-Based Drivers' Visual Attention Prediction with LLM-Enhanced Semantic Reasoning
von: Liu, Weimin, et al.
Veröffentlicht: (2026)
von: Liu, Weimin, et al.
Veröffentlicht: (2026)
Accelerating Diffusion Models with One-to-Many Knowledge Distillation
von: Zhang, Linfeng, et al.
Veröffentlicht: (2024)
von: Zhang, Linfeng, et al.
Veröffentlicht: (2024)
AnimateDiff-Lightning: Cross-Model Diffusion Distillation
von: Lin, Shanchuan, et al.
Veröffentlicht: (2024)
von: Lin, Shanchuan, et al.
Veröffentlicht: (2024)
One Step Diffusion-based Super-Resolution with Time-Aware Distillation
von: He, Xiao, et al.
Veröffentlicht: (2024)
von: He, Xiao, et al.
Veröffentlicht: (2024)
Adapting VACE for Real-Time Autoregressive Video Diffusion
von: Fosdick, Ryan
Veröffentlicht: (2026)
von: Fosdick, Ryan
Veröffentlicht: (2026)
DynaSplat: Dynamic-Static Gaussian Splatting with Hierarchical Motion Decomposition for Scene Reconstruction
von: Deng, Junli, et al.
Veröffentlicht: (2025)
von: Deng, Junli, et al.
Veröffentlicht: (2025)
Robotic System with AI for Real Time Weed Detection, Canopy Aware Spraying, and Droplet Pattern Evaluation
von: Rasool, Inayat, et al.
Veröffentlicht: (2025)
von: Rasool, Inayat, et al.
Veröffentlicht: (2025)
SoulX-FlashTalk: Real-Time Infinite Streaming of Audio-Driven Avatars via Self-Correcting Bidirectional Distillation
von: Shen, Le, et al.
Veröffentlicht: (2025)
von: Shen, Le, et al.
Veröffentlicht: (2025)
Dynamic Eraser for Guided Concept Erasure in Diffusion Models
von: Gong, Qinghui
Veröffentlicht: (2026)
von: Gong, Qinghui
Veröffentlicht: (2026)
FDBPL: Faster Distillation-Based Prompt Learning for Region-Aware Vision-Language Models Adaptation
von: Zhang, Zherui, et al.
Veröffentlicht: (2025)
von: Zhang, Zherui, et al.
Veröffentlicht: (2025)
Illusion-Aware Visual Preprocessing and Anti-Illusion Prompting for Classic Illusion Understanding in Vision-Language Models
von: Zha, Junli, et al.
Veröffentlicht: (2026)
von: Zha, Junli, et al.
Veröffentlicht: (2026)
LP-LLM: End-to-End Real-World Degraded License Plate Text Recognition via Large Multimodal Models
von: Gong, Haoyan, et al.
Veröffentlicht: (2026)
von: Gong, Haoyan, et al.
Veröffentlicht: (2026)
From Structure to Detail: Hierarchical Distillation for Efficient Diffusion Model
von: Cheng, Hanbo, et al.
Veröffentlicht: (2025)
von: Cheng, Hanbo, et al.
Veröffentlicht: (2025)
GVD: Guiding Video Diffusion Model for Scalable Video Distillation
von: Li, Kunyang, et al.
Veröffentlicht: (2025)
von: Li, Kunyang, et al.
Veröffentlicht: (2025)
Distilling Multi-view Diffusion Models into 3D Generators
von: Qin, Hao, et al.
Veröffentlicht: (2025)
von: Qin, Hao, et al.
Veröffentlicht: (2025)
See the past: Time-Reversed Scene Reconstruction from Thermal Traces Using Visual Language Models
von: Contreras, Kebin, et al.
Veröffentlicht: (2025)
von: Contreras, Kebin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Bringing Multimodal Large Language Models to Infrared-Visible Image Fusion Quality Assessment
von: Guo, Yuchen, et al.
Veröffentlicht: (2026) -
Can Segmentation Models Understand the World? Towards Proactive Affordance Reasoning via Visual Chain-of-Thought
von: Guo, Yuchen, et al.
Veröffentlicht: (2026) -
LumiVideo: An Intelligent Agentic System for Video Color Grading
von: Guo, Yuchen, et al.
Veröffentlicht: (2026) -
Fuse4Seg: Image Fusion for Multi-Modal Medical Segmentation via Bi-level Optimization
von: Guo, Yuchen, et al.
Veröffentlicht: (2024) -
DAE-Fuse: An Adaptive Discriminative Autoencoder for Multi-Modality Image Fusion
von: Guo, Yuchen, et al.
Veröffentlicht: (2024)