Gespeichert in:
| Hauptverfasser: | Ma, Yeyao, Li, Chen, Zhang, Xiaosong, Hu, Han, Xie, Weidi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2602.12155 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
X-Omni: Reinforcement Learning Makes Discrete Autoregressive Image Generative Models Great Again
von: Geng, Zigang, et al.
Veröffentlicht: (2025)
von: Geng, Zigang, et al.
Veröffentlicht: (2025)
PSALM: Pixelwise SegmentAtion with Large Multi-Modal Model
von: Zhang, Zheng, et al.
Veröffentlicht: (2024)
von: Zhang, Zheng, et al.
Veröffentlicht: (2024)
MatchTime: Towards Automatic Soccer Game Commentary Generation
von: Rao, Jiayuan, et al.
Veröffentlicht: (2024)
von: Rao, Jiayuan, et al.
Veröffentlicht: (2024)
Moving Object Segmentation: All You Need Is SAM (and Flow)
von: Xie, Junyu, et al.
Veröffentlicht: (2024)
von: Xie, Junyu, et al.
Veröffentlicht: (2024)
MedFlowSeg: Flow Matching for Medical Image Segmentation with Frequency-Aware Attention
von: Chen, Zhi, et al.
Veröffentlicht: (2026)
von: Chen, Zhi, et al.
Veröffentlicht: (2026)
SceneGen: Single-Image 3D Scene Generation in One Feedforward Pass
von: Meng, Yanxu, et al.
Veröffentlicht: (2025)
von: Meng, Yanxu, et al.
Veröffentlicht: (2025)
MT-EditFlow: Reinforcement Learning for Multi-Turn Image Editing with Flow Matching
von: Huang, Jiahui, et al.
Veröffentlicht: (2026)
von: Huang, Jiahui, et al.
Veröffentlicht: (2026)
Beyond Imitation: Constraint-Aware Trajectory Generation with Flow Matching For End-to-End Autonomous Driving
von: Liu, Lin, et al.
Veröffentlicht: (2025)
von: Liu, Lin, et al.
Veröffentlicht: (2025)
ELIP: Enhanced Visual-Language Foundation Models for Image Retrieval
von: Zhan, Guanqi, et al.
Veröffentlicht: (2025)
von: Zhan, Guanqi, et al.
Veröffentlicht: (2025)
GMOS: Grounding Moving Object Segmentation in 3D Space and Time
von: Xie, Junyu, et al.
Veröffentlicht: (2026)
von: Xie, Junyu, et al.
Veröffentlicht: (2026)
FMVP: Masked Flow Matching for Adversarial Video Purification
von: Tang, Duoxun, et al.
Veröffentlicht: (2026)
von: Tang, Duoxun, et al.
Veröffentlicht: (2026)
Revisiting Multi-Task Visual Representation Learning
von: Di, Shangzhe, et al.
Veröffentlicht: (2026)
von: Di, Shangzhe, et al.
Veröffentlicht: (2026)
Grounded Question-Answering in Long Egocentric Videos
von: Di, Shangzhe, et al.
Veröffentlicht: (2023)
von: Di, Shangzhe, et al.
Veröffentlicht: (2023)
EchoSight: Advancing Visual-Language Models with Wiki Knowledge
von: Yan, Yibin, et al.
Veröffentlicht: (2024)
von: Yan, Yibin, et al.
Veröffentlicht: (2024)
A Sanity Check on Composed Image Retrieval
von: Liu, Yikun, et al.
Veröffentlicht: (2026)
von: Liu, Yikun, et al.
Veröffentlicht: (2026)
Multi-Sentence Grounding for Long-term Instructional Video
von: Li, Zeqian, et al.
Veröffentlicht: (2023)
von: Li, Zeqian, et al.
Veröffentlicht: (2023)
Aligning Latent Geometry for Spherical Flow Matching in Image Generation
von: Meral, Tuna Han Salih, et al.
Veröffentlicht: (2026)
von: Meral, Tuna Han Salih, et al.
Veröffentlicht: (2026)
Zero-shot Composed Text-Image Retrieval
von: Liu, Yikun, et al.
Veröffentlicht: (2023)
von: Liu, Yikun, et al.
Veröffentlicht: (2023)
Grounded Multi-Hop VideoQA in Long-Form Egocentric Videos
von: Chen, Qirui, et al.
Veröffentlicht: (2024)
von: Chen, Qirui, et al.
Veröffentlicht: (2024)
FlowLUT: Efficient Image Enhancement via Differentiable LUTs and Iterative Flow Matching
von: Hu, Liubing, et al.
Veröffentlicht: (2025)
von: Hu, Liubing, et al.
Veröffentlicht: (2025)
FlowInOne:Unifying Multimodal Generation as Image-in, Image-out Flow Matching
von: Yi, Junchao, et al.
Veröffentlicht: (2026)
von: Yi, Junchao, et al.
Veröffentlicht: (2026)
A Sanity Check for AI-generated Image Detection
von: Yan, Shilin, et al.
Veröffentlicht: (2024)
von: Yan, Shilin, et al.
Veröffentlicht: (2024)
FREPix: Frequency-Heterogeneous Flow Matching for Pixel-Space Image Generation
von: Lin, Mingfeng, et al.
Veröffentlicht: (2026)
von: Lin, Mingfeng, et al.
Veröffentlicht: (2026)
DenseStep2M: A Scalable, Training-Free Pipeline for Dense Instructional Video Annotation
von: Ge, Mingji, et al.
Veröffentlicht: (2026)
von: Ge, Mingji, et al.
Veröffentlicht: (2026)
Learning Patient-Specific Disease Dynamics with Latent Flow Matching for Longitudinal Imaging Generation
von: Chen, Hao, et al.
Veröffentlicht: (2025)
von: Chen, Hao, et al.
Veröffentlicht: (2025)
WaterMamba: Visual State Space Model for Underwater Image Enhancement
von: Guan, Meisheng, et al.
Veröffentlicht: (2024)
von: Guan, Meisheng, et al.
Veröffentlicht: (2024)
Appearance-Based Refinement for Object-Centric Motion Segmentation
von: Xie, Junyu, et al.
Veröffentlicht: (2023)
von: Xie, Junyu, et al.
Veröffentlicht: (2023)
Kernel Adversarial Learning for Real-world Image Super-resolution
von: Wang, Hu, et al.
Veröffentlicht: (2021)
von: Wang, Hu, et al.
Veröffentlicht: (2021)
Adversarial Distribution Matching for Diffusion Distillation Towards Efficient Image and Video Synthesis
von: Lu, Yanzuo, et al.
Veröffentlicht: (2025)
von: Lu, Yanzuo, et al.
Veröffentlicht: (2025)
FlowAR: Scale-wise Autoregressive Image Generation Meets Flow Matching
von: Ren, Sucheng, et al.
Veröffentlicht: (2024)
von: Ren, Sucheng, et al.
Veröffentlicht: (2024)
CurveFlow: Curvature-Guided Flow Matching for Image Generation
von: Luo, Yan, et al.
Veröffentlicht: (2025)
von: Luo, Yan, et al.
Veröffentlicht: (2025)
Character-Centric Understanding of Animated Movies
von: Gui, Zhongrui, et al.
Veröffentlicht: (2025)
von: Gui, Zhongrui, et al.
Veröffentlicht: (2025)
Aerial Monocular 3D Object Detection
von: Hu, Yue, et al.
Veröffentlicht: (2022)
von: Hu, Yue, et al.
Veröffentlicht: (2022)
Flow of Truth: Proactive Temporal Forensics for Image-to-Video Generation
von: Chen, Yuzhuo, et al.
Veröffentlicht: (2026)
von: Chen, Yuzhuo, et al.
Veröffentlicht: (2026)
Frequency-Aware Flow Matching for High-Quality Image Generation
von: Ren, Sucheng, et al.
Veröffentlicht: (2026)
von: Ren, Sucheng, et al.
Veröffentlicht: (2026)
A General Protocol to Probe Large Vision Models for 3D Physical Understanding
von: Zhan, Guanqi, et al.
Veröffentlicht: (2023)
von: Zhan, Guanqi, et al.
Veröffentlicht: (2023)
Few-Shot Distribution-Aligned Flow Matching for Data Synthesis in Medical Image Segmentation
von: Yang, Jie, et al.
Veröffentlicht: (2026)
von: Yang, Jie, et al.
Veröffentlicht: (2026)
Learning Straight Flows: Variational Flow Matching for Efficient Generation
von: Ma, Chenrui, et al.
Veröffentlicht: (2025)
von: Ma, Chenrui, et al.
Veröffentlicht: (2025)
Fine-grained Spatiotemporal Grounding on Egocentric Videos
von: Liang, Shuo, et al.
Veröffentlicht: (2025)
von: Liang, Shuo, et al.
Veröffentlicht: (2025)
Can We Build Scene Graphs, Not Classify Them? FlowSG: Progressive Image-Conditioned Scene Graph Generation with Flow Matching
von: Hu, Xin, et al.
Veröffentlicht: (2026)
von: Hu, Xin, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
X-Omni: Reinforcement Learning Makes Discrete Autoregressive Image Generative Models Great Again
von: Geng, Zigang, et al.
Veröffentlicht: (2025) -
PSALM: Pixelwise SegmentAtion with Large Multi-Modal Model
von: Zhang, Zheng, et al.
Veröffentlicht: (2024) -
MatchTime: Towards Automatic Soccer Game Commentary Generation
von: Rao, Jiayuan, et al.
Veröffentlicht: (2024) -
Moving Object Segmentation: All You Need Is SAM (and Flow)
von: Xie, Junyu, et al.
Veröffentlicht: (2024) -
MedFlowSeg: Flow Matching for Medical Image Segmentation with Frequency-Aware Attention
von: Chen, Zhi, et al.
Veröffentlicht: (2026)