RealDPO: Real or Not Real, that is the Preference
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cheng, Guo, Yang, Danni, Huang, Ziqi, Si, Jianlou, Si, Chenyang, Liu, Ziwei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Lattice Boltzmann Model for Learning Real-World Pixel Dynamicity
von: Zheng, Guangze, et al.
Veröffentlicht: (2025)
von: Zheng, Guangze, et al.
Veröffentlicht: (2025)
When Preferences Diverge: Aligning Diffusion Models with Minority-Aware Adaptive DPO
von: Zhang, Lingfan, et al.
Veröffentlicht: (2025)
von: Zhang, Lingfan, et al.
Veröffentlicht: (2025)
Reg-DPO: SFT-Regularized Direct Preference Optimization with GT-Pair for Improving Video Generation
von: Du, Jie, et al.
Veröffentlicht: (2025)
von: Du, Jie, et al.
Veröffentlicht: (2025)
FreeInit: Bridging Initialization Gap in Video Diffusion Models
von: Wu, Tianxing, et al.
Veröffentlicht: (2023)
von: Wu, Tianxing, et al.
Veröffentlicht: (2023)
VL-DPO: Vision-Language-Guided Finetuning for Preference-Aligned Autonomous Driving
von: Xu, Zhefan, et al.
Veröffentlicht: (2026)
von: Xu, Zhefan, et al.
Veröffentlicht: (2026)
Density-guided Translator Boosts Synthetic-to-Real Unsupervised Domain Adaptive Segmentation of 3D Point Clouds
von: Yuan, Zhimin, et al.
Veröffentlicht: (2024)
von: Yuan, Zhimin, et al.
Veröffentlicht: (2024)
MIA-DPO: Multi-Image Augmented Direct Preference Optimization For Large Vision-Language Models
von: Liu, Ziyu, et al.
Veröffentlicht: (2024)
von: Liu, Ziyu, et al.
Veröffentlicht: (2024)
RealWonder: Real-Time Physical Action-Conditioned Video Generation
von: Liu, Wei, et al.
Veröffentlicht: (2026)
von: Liu, Wei, et al.
Veröffentlicht: (2026)
VideoDPO: Omni-Preference Alignment for Video Diffusion Generation
von: Liu, Runtao, et al.
Veröffentlicht: (2024)
von: Liu, Runtao, et al.
Veröffentlicht: (2024)
FocusDPO: Dynamic Preference Optimization for Multi-Subject Personalized Image Generation via Adaptive Focus
von: Jin, Qiaoqiao, et al.
Veröffentlicht: (2025)
von: Jin, Qiaoqiao, et al.
Veröffentlicht: (2025)
Sim2Real-AD: A Modular Sim-to-Real Framework for Deploying VLM-Guided Reinforcement Learning in Real-World Autonomous Driving
von: Huang, Zilin, et al.
Veröffentlicht: (2026)
von: Huang, Zilin, et al.
Veröffentlicht: (2026)
GARF: Learning Generalizable 3D Reassembly for Real-World Fractures
von: Li, Sihang, et al.
Veröffentlicht: (2025)
von: Li, Sihang, et al.
Veröffentlicht: (2025)
DP$^2$O-SR: Direct Perceptual Preference Optimization for Real-World Image Super-Resolution
von: Wu, Rongyuan, et al.
Veröffentlicht: (2025)
von: Wu, Rongyuan, et al.
Veröffentlicht: (2025)
On the Real-World Adversarial Robustness of Real-Time Semantic Segmentation Models for Autonomous Driving
von: Rossolini, Giulio, et al.
Veröffentlicht: (2022)
von: Rossolini, Giulio, et al.
Veröffentlicht: (2022)
EmbodieDreamer: Advancing Real2Sim2Real Transfer for Policy Training via Embodied World Modeling
von: Wang, Boyuan, et al.
Veröffentlicht: (2025)
von: Wang, Boyuan, et al.
Veröffentlicht: (2025)
TARAC: Mitigating Hallucination in LVLMs via Temporal Attention Real-time Accumulative Connection
von: Jiang, Lei, et al.
Veröffentlicht: (2025)
von: Jiang, Lei, et al.
Veröffentlicht: (2025)
Real-Time Crowd Counting for Embedded Systems with Lightweight Architecture
von: Zhao, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Zhao, Zhiyuan, et al.
Veröffentlicht: (2025)
Make-it-Real: Unleashing Large Multimodal Model for Painting 3D Objects with Realistic Materials
von: Fang, Ye, et al.
Veröffentlicht: (2024)
von: Fang, Ye, et al.
Veröffentlicht: (2024)
MDPBench: A Benchmark for Multilingual Document Parsing in Real-World Scenarios
von: Li, Zhang, et al.
Veröffentlicht: (2026)
von: Li, Zhang, et al.
Veröffentlicht: (2026)
DART: Depth-Enhanced Accurate and Real-Time Background Matting
von: Li, Hanxi, et al.
Veröffentlicht: (2024)
von: Li, Hanxi, et al.
Veröffentlicht: (2024)
RealAppliance: Let High-fidelity Appliance Assets Controllable and Workable as Aligned Real Manuals
von: Gao, Yuzheng, et al.
Veröffentlicht: (2025)
von: Gao, Yuzheng, et al.
Veröffentlicht: (2025)
A Real-to-Sim-to-Real Approach to Robotic Manipulation with VLM-Generated Iterative Keypoint Rewards
von: Patel, Shivansh, et al.
Veröffentlicht: (2025)
von: Patel, Shivansh, et al.
Veröffentlicht: (2025)
YCDa: YCbCr Decoupled Attention for Real-time Realistic Camouflaged Object Detection
von: Zheng, PeiHuang, et al.
Veröffentlicht: (2026)
von: Zheng, PeiHuang, et al.
Veröffentlicht: (2026)
Real2SAM2Real: Generative 3D Caches as Complementary Context for Video Diffusion
von: Wu, Jiayi, et al.
Veröffentlicht: (2026)
von: Wu, Jiayi, et al.
Veröffentlicht: (2026)
V-IRL: Grounding Virtual Intelligence in Real Life
von: Yang, Jihan, et al.
Veröffentlicht: (2024)
von: Yang, Jihan, et al.
Veröffentlicht: (2024)
Inversion-DPO: Precise and Efficient Post-Training for Diffusion Models
von: Li, Zejian, et al.
Veröffentlicht: (2025)
von: Li, Zejian, et al.
Veröffentlicht: (2025)
An Overall Real-Time Mechanism for Classification and Quality Evaluation of Rice
von: Xia, Wanke, et al.
Veröffentlicht: (2025)
von: Xia, Wanke, et al.
Veröffentlicht: (2025)
RealGen: Photorealistic Text-to-Image Generation via Detector-Guided Rewards
von: Ye, Junyan, et al.
Veröffentlicht: (2025)
von: Ye, Junyan, et al.
Veröffentlicht: (2025)
No Need For Real Anomaly: MLLM Empowered Zero-Shot Video Anomaly Detection
von: Dai, Zunkai, et al.
Veröffentlicht: (2026)
von: Dai, Zunkai, et al.
Veröffentlicht: (2026)
Tell me Habibi, is it Real or Fake?
von: Kuckreja, Kartik, et al.
Veröffentlicht: (2025)
von: Kuckreja, Kartik, et al.
Veröffentlicht: (2025)
Real-time Yemeni Currency Detection
von: AL-Edreesi, Edrees, et al.
Veröffentlicht: (2024)
von: AL-Edreesi, Edrees, et al.
Veröffentlicht: (2024)
MineWorld: a Real-Time and Open-Source Interactive World Model on Minecraft
von: Guo, Junliang, et al.
Veröffentlicht: (2025)
von: Guo, Junliang, et al.
Veröffentlicht: (2025)
Video as the New Language for Real-World Decision Making
von: Yang, Sherry, et al.
Veröffentlicht: (2024)
von: Yang, Sherry, et al.
Veröffentlicht: (2024)
mDPO: Conditional Preference Optimization for Multimodal Large Language Models
von: Wang, Fei, et al.
Veröffentlicht: (2024)
von: Wang, Fei, et al.
Veröffentlicht: (2024)
GOOD: Training-Free Guided Diffusion Sampling for Out-of-Distribution Detection
von: Gao, Xin, et al.
Veröffentlicht: (2025)
von: Gao, Xin, et al.
Veröffentlicht: (2025)
FocusDD: Real-World Scene Infusion for Robust Dataset Distillation
von: Hu, Youbing, et al.
Veröffentlicht: (2025)
von: Hu, Youbing, et al.
Veröffentlicht: (2025)
SA-Occ: Satellite-Assisted 3D Occupancy Prediction in Real World
von: Chen, Chen, et al.
Veröffentlicht: (2025)
von: Chen, Chen, et al.
Veröffentlicht: (2025)
V-DPO: Mitigating Hallucination in Large Vision Language Models via Vision-Guided Direct Preference Optimization
von: Xie, Yuxi, et al.
Veröffentlicht: (2024)
von: Xie, Yuxi, et al.
Veröffentlicht: (2024)
SceneTAP: Scene-Coherent Typographic Adversarial Planner against Vision-Language Models in Real-World Environments
von: Cao, Yue, et al.
Veröffentlicht: (2024)
von: Cao, Yue, et al.
Veröffentlicht: (2024)
Real-Time Human Action Recognition on Embedded Platforms
von: Wang, Ruiqi, et al.
Veröffentlicht: (2024)
von: Wang, Ruiqi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Lattice Boltzmann Model for Learning Real-World Pixel Dynamicity
von: Zheng, Guangze, et al.
Veröffentlicht: (2025) -
When Preferences Diverge: Aligning Diffusion Models with Minority-Aware Adaptive DPO
von: Zhang, Lingfan, et al.
Veröffentlicht: (2025) -
Reg-DPO: SFT-Regularized Direct Preference Optimization with GT-Pair for Improving Video Generation
von: Du, Jie, et al.
Veröffentlicht: (2025) -
FreeInit: Bridging Initialization Gap in Video Diffusion Models
von: Wu, Tianxing, et al.
Veröffentlicht: (2023) -
VL-DPO: Vision-Language-Guided Finetuning for Preference-Aligned Autonomous Driving
von: Xu, Zhefan, et al.
Veröffentlicht: (2026)