Lattice Boltzmann Model for Learning Real-World Pixel Dynamicity
Fuente:
arXiv
Saved in:
| Main Authors: | Zheng, Guangze, Lin, Shijie, Zuo, Haobo, Si, Si, Wang, Ming-Shan, Fu, Changhong, Pan, Jia |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
NetTrack: Tracking Highly Dynamic Objects with a Net
by: Zheng, Guangze, et al.
Published: (2024)
by: Zheng, Guangze, et al.
Published: (2024)
Progressive Representation Learning for Real-Time UAV Tracking
by: Fu, Changhong, et al.
Published: (2024)
by: Fu, Changhong, et al.
Published: (2024)
SAM-DA: UAV Tracks Anything at Night with SAM-Powered Domain Adaptation
by: Fu, Changhong, et al.
Published: (2023)
by: Fu, Changhong, et al.
Published: (2023)
DaDiff: Domain-aware Diffusion Model for Nighttime UAV Tracking
by: Zuo, Haobo, et al.
Published: (2024)
by: Zuo, Haobo, et al.
Published: (2024)
Prompt-Driven Temporal Domain Adaptation for Nighttime UAV Tracking
by: Fu, Changhong, et al.
Published: (2024)
by: Fu, Changhong, et al.
Published: (2024)
Dual Prompt-Driven Feature Encoding for Nighttime UAV Tracking
by: Wang, Yiheng, et al.
Published: (2026)
by: Wang, Yiheng, et al.
Published: (2026)
AnyTSR: Any-Scale Thermal Super-Resolution for UAV
by: Li, Mengyuan, et al.
Published: (2025)
by: Li, Mengyuan, et al.
Published: (2025)
EdgeSpotter: Multi-Scale Dense Text Spotting for Industrial Panel Monitoring
by: Fu, Changhong, et al.
Published: (2025)
by: Fu, Changhong, et al.
Published: (2025)
RealDPO: Real or Not Real, that is the Preference
by: Cheng, Guo, et al.
Published: (2025)
by: Cheng, Guo, et al.
Published: (2025)
Dreamweaver: Learning Compositional World Models from Pixels
by: Baek, Junyeob, et al.
Published: (2025)
by: Baek, Junyeob, et al.
Published: (2025)
Enhancing Nighttime UAV Tracking with Light Distribution Suppression
by: Yao, Liangliang, et al.
Published: (2024)
by: Yao, Liangliang, et al.
Published: (2024)
Object Attribute Matters in Visual Question Answering
by: Li, Peize, et al.
Published: (2023)
by: Li, Peize, et al.
Published: (2023)
Real-World Efficient Blind Motion Deblurring via Blur Pixel Discretization
by: Kim, Insoo, et al.
Published: (2024)
by: Kim, Insoo, et al.
Published: (2024)
Contrastive Learning Guided Latent Diffusion Model for Image-to-Image Translation
by: Si, Qi, et al.
Published: (2025)
by: Si, Qi, et al.
Published: (2025)
Beyond Pixels: Vector-to-Graph Transformation for Reliable Schematic Auditing
by: Ma, Chengwei, et al.
Published: (2026)
by: Ma, Chengwei, et al.
Published: (2026)
mAVE: A Watermark for Joint Audio-Visual Generation Models
by: Si, Luyang, et al.
Published: (2026)
by: Si, Luyang, et al.
Published: (2026)
DiLA: Disentangled Latent Action World Models
by: Zhang, Tianqiu, et al.
Published: (2026)
by: Zhang, Tianqiu, et al.
Published: (2026)
ORL-LDM: Offline Reinforcement Learning Guided Latent Diffusion Model Super-Resolution Reconstruction
by: Lyu, Shijie
Published: (2025)
by: Lyu, Shijie
Published: (2025)
UniPixel: Unified Object Referring and Segmentation for Pixel-Level Visual Reasoning
by: Liu, Ye, et al.
Published: (2025)
by: Liu, Ye, et al.
Published: (2025)
PhysWorld: From Real Videos to World Models of Deformable Objects via Physics-Aware Demonstration Synthesis
by: Yang, Yu, et al.
Published: (2025)
by: Yang, Yu, et al.
Published: (2025)
The Language of Touch: Translating Vibrations into Text with Dual-Branch Learning
by: Chen, Jin, et al.
Published: (2026)
by: Chen, Jin, et al.
Published: (2026)
DexWorldModel: Causal Latent World Modeling towards Automated Learning of Embodied Tasks
by: Deng, Yueci, et al.
Published: (2026)
by: Deng, Yueci, et al.
Published: (2026)
LatticeWorld: A Multimodal Large Language Model-Empowered Framework for Interactive Complex World Generation
by: Duan, Yinglin, et al.
Published: (2025)
by: Duan, Yinglin, et al.
Published: (2025)
Beyond Words and Pixels: A Benchmark for Implicit World Knowledge Reasoning in Generative Models
by: Han, Tianyang, et al.
Published: (2025)
by: Han, Tianyang, et al.
Published: (2025)
Towards Visual Discrimination and Reasoning of Real-World Physical Dynamics: Physics-Grounded Anomaly Detection
by: Li, Wenqiao, et al.
Published: (2025)
by: Li, Wenqiao, et al.
Published: (2025)
Structure Abstraction and Generalization in a Hippocampal-Entorhinal Inspired World Model
by: Zhang, Tianqiu, et al.
Published: (2026)
by: Zhang, Tianqiu, et al.
Published: (2026)
Pixels, Patterns, but No Poetry: To See The World like Humans
by: Gao, Hongcheng, et al.
Published: (2025)
by: Gao, Hongcheng, et al.
Published: (2025)
ClearSight: Visual Signal Enhancement for Object Hallucination Mitigation in Multimodal Large language Models
by: Yin, Hao, et al.
Published: (2025)
by: Yin, Hao, et al.
Published: (2025)
Restoring Real-World Images with an Internal Detail Enhancement Diffusion Model
by: Xiao, Peng, et al.
Published: (2025)
by: Xiao, Peng, et al.
Published: (2025)
Diagnosing and Repairing Unsafe Channels in Vision-Language Models via Causal Discovery and Dual-Modal Safety Subspace Projection
by: Fu, Jinhu, et al.
Published: (2026)
by: Fu, Jinhu, et al.
Published: (2026)
GaussianWorld: Gaussian World Model for Streaming 3D Occupancy Prediction
by: Zuo, Sicheng, et al.
Published: (2024)
by: Zuo, Sicheng, et al.
Published: (2024)
GLaMM: Pixel Grounding Large Multimodal Model
by: Rasheed, Hanoona, et al.
Published: (2023)
by: Rasheed, Hanoona, et al.
Published: (2023)
Thinking Ahead: Foresight Intelligence in MLLMs and World Models
by: Gong, Zhantao, et al.
Published: (2025)
by: Gong, Zhantao, et al.
Published: (2025)
EgoPlan-Bench2: A Benchmark for Multimodal Large Language Model Planning in Real-World Scenarios
by: Qiu, Lu, et al.
Published: (2024)
by: Qiu, Lu, et al.
Published: (2024)
Pixel Reasoner: Incentivizing Pixel-Space Reasoning with Curiosity-Driven Reinforcement Learning
by: Wang, Haozhe, et al.
Published: (2025)
by: Wang, Haozhe, et al.
Published: (2025)
Optimization of Autonomous Driving Image Detection Based on RFAConv and Triplet Attention
by: Ling, Zhipeng, et al.
Published: (2024)
by: Ling, Zhipeng, et al.
Published: (2024)
Layer-Wise Feature Metric of Semantic-Pixel Matching for Few-Shot Learning
by: Tang, Hao, et al.
Published: (2024)
by: Tang, Hao, et al.
Published: (2024)
MSConv: Multiplicative and Subtractive Convolution for Face Recognition
by: Zhou, Si, et al.
Published: (2025)
by: Zhou, Si, et al.
Published: (2025)
Not All Pixels Are Equal: Pixel-wise Meta-Learning for Medical Segmentation with Noisy Labels
by: Mu, Chenyu, et al.
Published: (2025)
by: Mu, Chenyu, et al.
Published: (2025)
From Pixels to Predicates: Learning Symbolic World Models via Pretrained Vision-Language Models
by: Athalye, Ashay, et al.
Published: (2024)
by: Athalye, Ashay, et al.
Published: (2024)
Similar Items
-
NetTrack: Tracking Highly Dynamic Objects with a Net
by: Zheng, Guangze, et al.
Published: (2024) -
Progressive Representation Learning for Real-Time UAV Tracking
by: Fu, Changhong, et al.
Published: (2024) -
SAM-DA: UAV Tracks Anything at Night with SAM-Powered Domain Adaptation
by: Fu, Changhong, et al.
Published: (2023) -
DaDiff: Domain-aware Diffusion Model for Nighttime UAV Tracking
by: Zuo, Haobo, et al.
Published: (2024) -
Prompt-Driven Temporal Domain Adaptation for Nighttime UAV Tracking
by: Fu, Changhong, et al.
Published: (2024)