Integration Flow Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Jingjing, Zhang, Dan, Luo, Joshua, Yang, Yin, Luo, Feng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning Discrete Autoregressive Priors with Wasserstein Gradient Flow
von: Zheng, Bowen, et al.
Veröffentlicht: (2026)
von: Zheng, Bowen, et al.
Veröffentlicht: (2026)
milliFlow: Scene Flow Estimation on mmWave Radar Point Cloud for Human Motion Sensing
von: Ding, Fangqiang, et al.
Veröffentlicht: (2023)
von: Ding, Fangqiang, et al.
Veröffentlicht: (2023)
Diff-Instruct++: Training One-step Text-to-image Generator Model to Align with Human Preferences
von: Luo, Weijian
Veröffentlicht: (2024)
von: Luo, Weijian
Veröffentlicht: (2024)
Flow Generator Matching
von: Huang, Zemin, et al.
Veröffentlicht: (2024)
von: Huang, Zemin, et al.
Veröffentlicht: (2024)
Score Distillation of Flow Matching Models
von: Zhou, Mingyuan, et al.
Veröffentlicht: (2025)
von: Zhou, Mingyuan, et al.
Veröffentlicht: (2025)
Towards High-Order Mean Flow Generative Models: Feasibility, Expressivity, and Provably Efficient Criteria
von: Cao, Yang, et al.
Veröffentlicht: (2025)
von: Cao, Yang, et al.
Veröffentlicht: (2025)
World Modeling with Probabilistic Structure Integration
von: Kotar, Klemen, et al.
Veröffentlicht: (2025)
von: Kotar, Klemen, et al.
Veröffentlicht: (2025)
Diagnosing and Improving Diffusion Models by Estimating the Optimal Loss Value
von: Xu, Yixian, et al.
Veröffentlicht: (2025)
von: Xu, Yixian, et al.
Veröffentlicht: (2025)
Advancing Multimodal Large Language Models with Quantization-Aware Scale Learning for Efficient Adaptation
von: Xie, Jingjing, et al.
Veröffentlicht: (2024)
von: Xie, Jingjing, et al.
Veröffentlicht: (2024)
Taming the Entropy Cliff: Variable Codebook Size Quantization for Autoregressive Visual Generation
von: Zheng, Bowen, et al.
Veröffentlicht: (2026)
von: Zheng, Bowen, et al.
Veröffentlicht: (2026)
Temporal Test-Time Adaptation with State-Space Models
von: Schirmer, Mona, et al.
Veröffentlicht: (2024)
von: Schirmer, Mona, et al.
Veröffentlicht: (2024)
PhysGaussian: Physics-Integrated 3D Gaussians for Generative Dynamics
von: Xie, Tianyi, et al.
Veröffentlicht: (2023)
von: Xie, Tianyi, et al.
Veröffentlicht: (2023)
DiffRaman: A Conditional Latent Denoising Diffusion Probabilistic Model for Bacterial Raman Spectroscopy Identification Under Limited Data Conditions
von: Yao, Haiming, et al.
Veröffentlicht: (2024)
von: Yao, Haiming, et al.
Veröffentlicht: (2024)
AEGPO: Adaptive Entropy-Guided Policy Optimization for Diffusion Models
von: Li, Yuming, et al.
Veröffentlicht: (2026)
von: Li, Yuming, et al.
Veröffentlicht: (2026)
MageBench: Bridging Large Multimodal Models to Agents
von: Zhang, Miaosen, et al.
Veröffentlicht: (2024)
von: Zhang, Miaosen, et al.
Veröffentlicht: (2024)
Multi-scale Quaternion CNN and BiGRU with Cross Self-attention Feature Fusion for Fault Diagnosis of Bearing
von: Liu, Huanbai, et al.
Veröffentlicht: (2024)
von: Liu, Huanbai, et al.
Veröffentlicht: (2024)
E-GRPO: High Entropy Steps Drive Effective Reinforcement Learning for Flow Models
von: Zhang, Shengjun, et al.
Veröffentlicht: (2026)
von: Zhang, Shengjun, et al.
Veröffentlicht: (2026)
Grounding Video Models to Actions through Goal Conditioned Exploration
von: Luo, Yunhao, et al.
Veröffentlicht: (2024)
von: Luo, Yunhao, et al.
Veröffentlicht: (2024)
SCFlow: Implicitly Learning Style and Content Disentanglement with Flow Models
von: Ma, Pingchuan, et al.
Veröffentlicht: (2025)
von: Ma, Pingchuan, et al.
Veröffentlicht: (2025)
What Helps---and What Hurts: Bidirectional Explanations for Vision Transformers
von: Su, Qin, et al.
Veröffentlicht: (2026)
von: Su, Qin, et al.
Veröffentlicht: (2026)
CrossVL: Complexity-Aware Feature Routing and Paired Curriculum for Cross-View Vision-Language Detection
von: Liu, Zhipeng, et al.
Veröffentlicht: (2026)
von: Liu, Zhipeng, et al.
Veröffentlicht: (2026)
David and Goliath: Small One-step Model Beats Large Diffusion with Score Post-training
von: Luo, Weijian, et al.
Veröffentlicht: (2024)
von: Luo, Weijian, et al.
Veröffentlicht: (2024)
AdvantageFlow: Advantage-Weighted Least Squares for RL in Flow Models
von: Kveton, Branislav, et al.
Veröffentlicht: (2026)
von: Kveton, Branislav, et al.
Veröffentlicht: (2026)
Co-PLNet: A Collaborative Point-Line Network for Prompt-Guided Wireframe Parsing
von: Wang, Chao, et al.
Veröffentlicht: (2026)
von: Wang, Chao, et al.
Veröffentlicht: (2026)
Stepwise Credit Assignment for GRPO on Flow-Matching Models
von: Savani, Yash, et al.
Veröffentlicht: (2026)
von: Savani, Yash, et al.
Veröffentlicht: (2026)
SurgLLM: A Versatile Large Multimodal Model with Spatial Focus and Temporal Awareness for Surgical Video Understanding
von: Chen, Zhen, et al.
Veröffentlicht: (2025)
von: Chen, Zhen, et al.
Veröffentlicht: (2025)
MVR: Multi-view Video Reward Shaping for Reinforcement Learning
von: Luo, Lirui, et al.
Veröffentlicht: (2026)
von: Luo, Lirui, et al.
Veröffentlicht: (2026)
NinA: Normalizing Flows in Action. Training VLA Models with Normalizing Flows
von: Tarasov, Denis, et al.
Veröffentlicht: (2025)
von: Tarasov, Denis, et al.
Veröffentlicht: (2025)
A Cascading Cooperative Multi-agent Framework for On-ramp Merging Control Integrating Large Language Models
von: Zhang, Miao, et al.
Veröffentlicht: (2025)
von: Zhang, Miao, et al.
Veröffentlicht: (2025)
Adversarial Supervision Makes Layout-to-Image Diffusion Models Thrive
von: Li, Yumeng, et al.
Veröffentlicht: (2024)
von: Li, Yumeng, et al.
Veröffentlicht: (2024)
Fine-Tuning a Large Vision-Language Model for Artwork's Scoring and Critique
von: Zhang, Zhehan, et al.
Veröffentlicht: (2026)
von: Zhang, Zhehan, et al.
Veröffentlicht: (2026)
Longitudinal Flow Matching for Trajectory Modeling
von: Islam, Mohammad Mohaiminul, et al.
Veröffentlicht: (2025)
von: Islam, Mohammad Mohaiminul, et al.
Veröffentlicht: (2025)
LSPT: Long-term Spatial Prompt Tuning for Visual Representation Learning
von: Mo, Shentong, et al.
Veröffentlicht: (2024)
von: Mo, Shentong, et al.
Veröffentlicht: (2024)
A Large-scale Medical Visual Task Adaptation Benchmark
von: Mo, Shentong, et al.
Veröffentlicht: (2024)
von: Mo, Shentong, et al.
Veröffentlicht: (2024)
Denoising Fisher Training For Neural Implicit Samplers
von: Luo, Weijian, et al.
Veröffentlicht: (2024)
von: Luo, Weijian, et al.
Veröffentlicht: (2024)
Entropy-Aware Structural Alignment for Zero-Shot Handwritten Chinese Character Recognition
von: Luo, Qiuming, et al.
Veröffentlicht: (2026)
von: Luo, Qiuming, et al.
Veröffentlicht: (2026)
HiPath: Hierarchical Vision-Language Alignment for Structured Pathology Report Prediction
von: Yuan, Ruicheng, et al.
Veröffentlicht: (2026)
von: Yuan, Ruicheng, et al.
Veröffentlicht: (2026)
Benchmarking Feature Upsampling Methods for Vision Foundation Models using Interactive Segmentation
von: Havrylov, Volodymyr, et al.
Veröffentlicht: (2025)
von: Havrylov, Volodymyr, et al.
Veröffentlicht: (2025)
CMT: Mid-Training for Efficient Learning of Consistency, Mean Flow, and Flow Map Models
von: Hu, Zheyuan, et al.
Veröffentlicht: (2025)
von: Hu, Zheyuan, et al.
Veröffentlicht: (2025)
MetaFormer Baselines for Vision
von: Yu, Weihao, et al.
Veröffentlicht: (2022)
von: Yu, Weihao, et al.
Veröffentlicht: (2022)
Ähnliche Einträge
-
Learning Discrete Autoregressive Priors with Wasserstein Gradient Flow
von: Zheng, Bowen, et al.
Veröffentlicht: (2026) -
milliFlow: Scene Flow Estimation on mmWave Radar Point Cloud for Human Motion Sensing
von: Ding, Fangqiang, et al.
Veröffentlicht: (2023) -
Diff-Instruct++: Training One-step Text-to-image Generator Model to Align with Human Preferences
von: Luo, Weijian
Veröffentlicht: (2024) -
Flow Generator Matching
von: Huang, Zemin, et al.
Veröffentlicht: (2024) -
Score Distillation of Flow Matching Models
von: Zhou, Mingyuan, et al.
Veröffentlicht: (2025)