DisCoRD: Discrete Tokens to Continuous Motion via Rectified Flow Decoding
Fuente:
arXiv
Saved in:
| Main Authors: | Cho, Jungbin, Kim, Junwan, Kim, Jisoo, Kim, Minseo, Kang, Mingu, Hong, Sungeun, Oh, Tae-Hyun, Yu, Youngjae |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SceneAdapt: Scene-aware Adaptation of Human Motion Diffusion
by: Cho, Jungbin, et al.
Published: (2025)
by: Cho, Jungbin, et al.
Published: (2025)
DEEPTalk: Dynamic Emotion Embedding for Probabilistic Speech-Driven 3D Face Animation
by: Kim, Jisoo, et al.
Published: (2024)
by: Kim, Jisoo, et al.
Published: (2024)
V.I.P. : Iterative Online Preference Distillation for Efficient Video Diffusion Models
by: Kim, Jisoo, et al.
Published: (2025)
by: Kim, Jisoo, et al.
Published: (2025)
EgoSpeak: Learning When to Speak for Egocentric Conversational Agents in the Wild
by: Kim, Junhyeok, et al.
Published: (2025)
by: Kim, Junhyeok, et al.
Published: (2025)
AVIN-Chat: An Audio-Visual Interactive Chatbot System with Emotional State Tuning
by: Park, Chanhyuk, et al.
Published: (2024)
by: Park, Chanhyuk, et al.
Published: (2024)
ReDi: Rectified Discrete Flow
by: Yoo, Jaehoon, et al.
Published: (2025)
by: Yoo, Jaehoon, et al.
Published: (2025)
A Proof of the Exact Convergence Rate of Gradient Descent
by: Kim, Jungbin
Published: (2024)
by: Kim, Jungbin
Published: (2024)
A Proof of Exact Convergence Rate of Gradient Descent. Part I. Performance Criterion $\Vert \nabla f(x_N)\Vert^2/(f(x_0)-f_*)$
by: Kim, Jungbin
Published: (2024)
by: Kim, Jungbin
Published: (2024)
Continuous Degradation Modeling via Latent Flow Matching for Real-World Super-Resolution
by: Kim, Hyeonjae, et al.
Published: (2026)
by: Kim, Hyeonjae, et al.
Published: (2026)
SMILE: Multimodal Dataset for Understanding Laughter in Video with Language Models
by: Hyun, Lee, et al.
Published: (2023)
by: Hyun, Lee, et al.
Published: (2023)
Self‐Rectifying Volatile Memristor for Highly Dynamic Functions
by: Dongyeol Ju, et al.
Published: (2025)
by: Dongyeol Ju, et al.
Published: (2025)
Reducing Peak Memory Usage for Modern Multimodal Large Language Model Pipelines
by: Kim, Junwan, et al.
Published: (2026)
by: Kim, Junwan, et al.
Published: (2026)
SLIP & ETHICS: Graduated Intervention for AI Emotional Companions
by: Kim, Minseo
Published: (2026)
by: Kim, Minseo
Published: (2026)
Kinodynamic Task and Motion Planning using VLM-guided and Interleaved Sampling
by: Kwon, Minseo, et al.
Published: (2025)
by: Kwon, Minseo, et al.
Published: (2025)
Accelerated Gradient Methods for Geodesically Convex Optimization: Tractable Algorithms and Convergence Analysis
by: Kim, Jungbin, et al.
Published: (2022)
by: Kim, Jungbin, et al.
Published: (2022)
Horospherically Convex Optimization on Hadamard Manifolds Part I: Analysis and Algorithms
by: Criscitiello, Christopher, et al.
Published: (2025)
by: Criscitiello, Christopher, et al.
Published: (2025)
Learning-based Axial Video Motion Magnification
by: Byung-Ki, Kwon, et al.
Published: (2023)
by: Byung-Ki, Kwon, et al.
Published: (2023)
On the size of universal graphs for spanning trees
by: Kim, Jaehoon, et al.
Published: (2025)
by: Kim, Jaehoon, et al.
Published: (2025)
AFLL: Real-time Load Stabilization for MMO Game Servers Based on Circular Causality Learning
by: Kang, Shinsuk, et al.
Published: (2026)
by: Kang, Shinsuk, et al.
Published: (2026)
Robust 3D Shape Reconstruction in Zero-Shot from a Single Image in the Wild
by: Cho, Junhyeong, et al.
Published: (2024)
by: Cho, Junhyeong, et al.
Published: (2024)
SafeFlow: Real-Time Text-Driven Humanoid Whole-Body Control via Physics-Guided Rectified Flow and Selective Safety Gating
by: Cho, Hanbyel, et al.
Published: (2026)
by: Cho, Hanbyel, et al.
Published: (2026)
Learning Where It Matters: Geometric Anchoring for Robust Preference Alignment
by: Cho, Youngjae, et al.
Published: (2026)
by: Cho, Youngjae, et al.
Published: (2026)
High Fidelity Text-to-Speech Via Discrete Tokens Using Token Transducer and Group Masked Language Model
by: Lee, Joun Yeop, et al.
Published: (2024)
by: Lee, Joun Yeop, et al.
Published: (2024)
FPRF: Feed-Forward Photorealistic Style Transfer of Large-Scale 3D Neural Radiance Fields
by: Kim, GeonU, et al.
Published: (2024)
by: Kim, GeonU, et al.
Published: (2024)
Light-Wave Engineering for Selective Polarization of a Single $\mathbf{Q}$ Valley in Transition Metal Dichalcogenides
by: Kim, Youngjae
Published: (2025)
by: Kim, Youngjae
Published: (2025)
Pseudospins revealed through the giant dynamical Franz-Keldysh effect in massless Dirac materials
by: Kim, Youngjae
Published: (2024)
by: Kim, Youngjae
Published: (2024)
Improving Visual Token Reduction via Rectifying Distortions for Efficient Multimodal LLM Inference
by: Cho, Hyeonwoo, et al.
Published: (2026)
by: Cho, Hyeonwoo, et al.
Published: (2026)
FacEDiT: Unified Talking Face Editing and Generation via Facial Motion Infilling
by: Sung-Bin, Kim, et al.
Published: (2025)
by: Sung-Bin, Kim, et al.
Published: (2025)
Revisit What You See: Revealing Visual Semantics in Vision Tokens to Guide LVLM Decoding
by: Cho, Beomsik, et al.
Published: (2025)
by: Cho, Beomsik, et al.
Published: (2025)
TripleSumm: Adaptive Triple-Modality Fusion for Video Summarization
by: Kim, Sumin, et al.
Published: (2026)
by: Kim, Sumin, et al.
Published: (2026)
UCMNet: Uncertainty-Aware Context Memory Network for Under-Display Camera Image Restoration
by: Kim, Daehyun, et al.
Published: (2026)
by: Kim, Daehyun, et al.
Published: (2026)
ConceptPrism: Concept Disentanglement in Personalized Diffusion Models via Residual Token Optimization
by: Kim, Minseo, et al.
Published: (2026)
by: Kim, Minseo, et al.
Published: (2026)
RA-Touch: Retrieval-Augmented Touch Understanding with Enriched Visual Data
by: Cho, Yoorhim, et al.
Published: (2025)
by: Cho, Yoorhim, et al.
Published: (2025)
Track-centric Iterative Learning for Global Trajectory Optimization in Autonomous Racing
by: Nam, Youngim, et al.
Published: (2026)
by: Nam, Youngim, et al.
Published: (2026)
MIRRAMS: Learning Robust Tabular Models under Unseen Missingness Shifts
by: Lee, Jihye, et al.
Published: (2025)
by: Lee, Jihye, et al.
Published: (2025)
Memorize Early, Then Query: Inlier-Memorization-Guided Active Outlier Detection
by: Kang, Minseo, et al.
Published: (2026)
by: Kang, Minseo, et al.
Published: (2026)
Identifiable Token Correspondence for World Models
by: Kim, Youngin, et al.
Published: (2026)
by: Kim, Youngin, et al.
Published: (2026)
Self-Rectifying Diffusion Sampling with Perturbed-Attention Guidance
by: Ahn, Donghoon, et al.
Published: (2024)
by: Ahn, Donghoon, et al.
Published: (2024)
Subtle Risks, Critical Failures: A Framework for Diagnosing Physical Safety of LLMs for Embodied Decision Making
by: Son, Yejin, et al.
Published: (2025)
by: Son, Yejin, et al.
Published: (2025)
DisCo-Diff: Enhancing Continuous Diffusion Models with Discrete Latents
by: Xu, Yilun, et al.
Published: (2024)
by: Xu, Yilun, et al.
Published: (2024)
Similar Items
-
SceneAdapt: Scene-aware Adaptation of Human Motion Diffusion
by: Cho, Jungbin, et al.
Published: (2025) -
DEEPTalk: Dynamic Emotion Embedding for Probabilistic Speech-Driven 3D Face Animation
by: Kim, Jisoo, et al.
Published: (2024) -
V.I.P. : Iterative Online Preference Distillation for Efficient Video Diffusion Models
by: Kim, Jisoo, et al.
Published: (2025) -
EgoSpeak: Learning When to Speak for Egocentric Conversational Agents in the Wild
by: Kim, Junhyeok, et al.
Published: (2025) -
AVIN-Chat: An Audio-Visual Interactive Chatbot System with Emotional State Tuning
by: Park, Chanhyuk, et al.
Published: (2024)