Saved in:
| Main Authors: | Nam, Jisu, Son, Soowon, Xu, Zhan, Shi, Jing, Liu, Difan, Liu, Feng, Misraa, Aashish, Kim, Seungryong, Zhou, Yang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2503.15406 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TrackCraft3R: Repurposing Video Diffusion Transformers for Dense 3D Tracking
by: Nam, Jisu, et al.
Published: (2026)
by: Nam, Jisu, et al.
Published: (2026)
Emergent Temporal Correspondences from Video Diffusion Transformers
by: Nam, Jisu, et al.
Published: (2025)
by: Nam, Jisu, et al.
Published: (2025)
Repurposing Video Diffusion Transformers for Robust Point Tracking
by: Son, Soowon, et al.
Published: (2025)
by: Son, Soowon, et al.
Published: (2025)
Waymo Driverless Car Data Analysis and Driving Modeling using CNN and LSTM
by: Misraa, Aashish Kumar, et al.
Published: (2025)
by: Misraa, Aashish Kumar, et al.
Published: (2025)
Customize-A-Video: One-Shot Motion Customization of Text-to-Video Diffusion Models
by: Ren, Yixuan, et al.
Published: (2024)
by: Ren, Yixuan, et al.
Published: (2024)
TETO: Tracking Events with Teacher Observation for Motion Estimation and Frame Interpolation
by: Yang, Jini, et al.
Published: (2026)
by: Yang, Jini, et al.
Published: (2026)
Diffusion Model for Dense Matching
by: Nam, Jisu, et al.
Published: (2023)
by: Nam, Jisu, et al.
Published: (2023)
MoDiTalker: Motion-Disentangled Diffusion Model for High-Fidelity Talking Head Generation
by: Kim, Seyeon, et al.
Published: (2024)
by: Kim, Seyeon, et al.
Published: (2024)
Modeling the effects of insecticides and external effortson crop production
by: A.K. Misraa
Published: (2021)
by: A.K. Misraa
Published: (2021)
Local All-Pair Correspondence for Point Tracking
by: Cho, Seokju, et al.
Published: (2024)
by: Cho, Seokju, et al.
Published: (2024)
DreamMatcher: Appearance Matching Self-Attention for Semantically-Consistent Text-to-Image Personalization
by: Nam, Jisu, et al.
Published: (2024)
by: Nam, Jisu, et al.
Published: (2024)
Squeezing Capacity from Multimodal Large Language Models for Subject-driven Generation
by: Zheng, Shuhong, et al.
Published: (2026)
by: Zheng, Shuhong, et al.
Published: (2026)
Deep Forcing: Training-Free Long Video Generation with Deep Sink and Participative Compression
by: Yi, Jung, et al.
Published: (2025)
by: Yi, Jung, et al.
Published: (2025)
Pose-dIVE: Pose-Diversified Augmentation with Diffusion Model for Person Re-Identification
by: Kim, Inès Hyeonsu, et al.
Published: (2024)
by: Kim, Inès Hyeonsu, et al.
Published: (2024)
Move-in-2D: 2D-Conditioned Human Motion Generation
by: Huang, Hsin-Ping, et al.
Published: (2024)
by: Huang, Hsin-Ping, et al.
Published: (2024)
DGFM: Full Body Dance Generation Driven by Music Foundation Models
by: Liu, Xinran, et al.
Published: (2025)
by: Liu, Xinran, et al.
Published: (2025)
MATRIX: Mask Track Alignment for Interaction-aware Video Generation
by: Jin, Siyoon, et al.
Published: (2025)
by: Jin, Siyoon, et al.
Published: (2025)
Appearance Matching Adapter for Exemplar-based Semantic Image Synthesis in-the-Wild
by: Jin, Siyoon, et al.
Published: (2024)
by: Jin, Siyoon, et al.
Published: (2024)
Progressive Autoregressive Video Diffusion Models
by: Xie, Desai, et al.
Published: (2024)
by: Xie, Desai, et al.
Published: (2024)
CORAL: Correspondence Alignment for Improved Virtual Try-On
by: Kim, Jiyoung, et al.
Published: (2026)
by: Kim, Jiyoung, et al.
Published: (2026)
PerMix-RLVR: Preserving Persona Expressivity under Verifiable-Reward Alignment
by: Oh, Jihwan, et al.
Published: (2026)
by: Oh, Jihwan, et al.
Published: (2026)
WorldCam: Interactive Autoregressive 3D Gaming Worlds with Camera Pose as a Unifying Geometric Representation
by: Nam, Jisu, et al.
Published: (2026)
by: Nam, Jisu, et al.
Published: (2026)
Fast SAM 3D Body: Accelerating SAM 3D Body for Real-Time Full-Body Human Mesh Recovery
by: Yang, Timing, et al.
Published: (2026)
by: Yang, Timing, et al.
Published: (2026)
CorGi: Contribution-Guided Block-Wise Interval Caching for Training-Free Acceleration of Diffusion Transformers
by: Son, Yonglak, et al.
Published: (2025)
by: Son, Yonglak, et al.
Published: (2025)
Confidence Regularized Masked Language Modeling using Text Length
by: Ji, Seunghyun, et al.
Published: (2025)
by: Ji, Seunghyun, et al.
Published: (2025)
PersonaCraft: Personalized and Controllable Full-Body Multi-Human Scene Generation Using Occlusion-Aware 3D-Conditioned Diffusion
by: Kim, Gwanghyun, et al.
Published: (2024)
by: Kim, Gwanghyun, et al.
Published: (2024)
EEG-Based Speech Decoding: A Novel Approach Using Multi-Kernel Ensemble Diffusion Models
by: Kim, Soowon, et al.
Published: (2024)
by: Kim, Soowon, et al.
Published: (2024)
PersonaMagic: Stage-Regulated High-Fidelity Face Customization with Tandem Equilibrium
by: Li, Xinzhe, et al.
Published: (2024)
by: Li, Xinzhe, et al.
Published: (2024)
Interact-Custom: Customized Human Object Interaction Image Generation
by: Xu, Zhu, et al.
Published: (2025)
by: Xu, Zhu, et al.
Published: (2025)
Spotting Out-of-Character Behavior: Atomic-Level Evaluation of Persona Fidelity in Open-Ended Generation
by: Shin, Jisu, et al.
Published: (2025)
by: Shin, Jisu, et al.
Published: (2025)
DuetSVG: Unified Multimodal SVG Generation with Internal Visual Guidance
by: Zhang, Peiying, et al.
Published: (2025)
by: Zhang, Peiying, et al.
Published: (2025)
Repurposing Geometric Foundation Models for Multi-view Diffusion
by: Jang, Wooseok, et al.
Published: (2026)
by: Jang, Wooseok, et al.
Published: (2026)
Customized FinGPT Search Agents Using Foundation Models
by: Tian, Felix, et al.
Published: (2024)
by: Tian, Felix, et al.
Published: (2024)
HIMO: A New Benchmark for Full-Body Human Interacting with Multiple Objects
by: Lv, Xintao, et al.
Published: (2024)
by: Lv, Xintao, et al.
Published: (2024)
SAM 3D Body: Robust Full-Body Human Mesh Recovery
by: Yang, Xitong, et al.
Published: (2026)
by: Yang, Xitong, et al.
Published: (2026)
You Are What You Bought: Generating Customer Personas for E-commerce Applications
by: Shi, Yimin, et al.
Published: (2025)
by: Shi, Yimin, et al.
Published: (2025)
PersonaBOT: Bringing Customer Personas to Life with LLMs and RAG
by: Rizwan, Muhammed, et al.
Published: (2025)
by: Rizwan, Muhammed, et al.
Published: (2025)
Lightweight Diffusion-based Framework for Online Imagined Speech Decoding in Aphasia
by: Ko, Eunyeong, et al.
Published: (2025)
by: Ko, Eunyeong, et al.
Published: (2025)
SapiensID: Foundation for Human Recognition
by: Kim, Minchul, et al.
Published: (2025)
by: Kim, Minchul, et al.
Published: (2025)
Strategic Data Ordering: Enhancing Large Language Model Performance through Curriculum Learning
by: Kim, Jisu, et al.
Published: (2024)
by: Kim, Jisu, et al.
Published: (2024)
Similar Items
-
TrackCraft3R: Repurposing Video Diffusion Transformers for Dense 3D Tracking
by: Nam, Jisu, et al.
Published: (2026) -
Emergent Temporal Correspondences from Video Diffusion Transformers
by: Nam, Jisu, et al.
Published: (2025) -
Repurposing Video Diffusion Transformers for Robust Point Tracking
by: Son, Soowon, et al.
Published: (2025) -
Waymo Driverless Car Data Analysis and Driving Modeling using CNN and LSTM
by: Misraa, Aashish Kumar, et al.
Published: (2025) -
Customize-A-Video: One-Shot Motion Customization of Text-to-Video Diffusion Models
by: Ren, Yixuan, et al.
Published: (2024)