Making Video Models Adhere to User Intent with Minor Adjustments
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ajisafe, Daniel, Hedlin, Eric, Rhodin, Helge, Yi, Kwang Moo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Unsupervised Keypoints from Pretrained Diffusion Models
von: Hedlin, Eric, et al.
Veröffentlicht: (2023)
von: Hedlin, Eric, et al.
Veröffentlicht: (2023)
CasCalib: Cascaded Calibration for Motion Capture from Sparse Unsynchronized Cameras
von: Tang, James, et al.
Veröffentlicht: (2024)
von: Tang, James, et al.
Veröffentlicht: (2024)
Mirror-Aware Neural Humans
von: Ajisafe, Daniel, et al.
Veröffentlicht: (2023)
von: Ajisafe, Daniel, et al.
Veröffentlicht: (2023)
NESI: Shape Representation via Neural Explicit Surface Intersection
von: Zhang, Congyi, et al.
Veröffentlicht: (2024)
von: Zhang, Congyi, et al.
Veröffentlicht: (2024)
LatentKeypointGAN: Controlling Images via Latent Keypoints
von: He, Xingzhe, et al.
Veröffentlicht: (2021)
von: He, Xingzhe, et al.
Veröffentlicht: (2021)
Gaussian Shadow Casting for Neural Characters
von: Bolanos, Luis, et al.
Veröffentlicht: (2024)
von: Bolanos, Luis, et al.
Veröffentlicht: (2024)
Radiant Foam: Real-Time Differentiable Ray Tracing
von: Govindarajan, Shrisudhan, et al.
Veröffentlicht: (2025)
von: Govindarajan, Shrisudhan, et al.
Veröffentlicht: (2025)
Evaluating Alternatives to SFM Point Cloud Initialization for Gaussian Splatting
von: Foroutan, Yalda, et al.
Veröffentlicht: (2024)
von: Foroutan, Yalda, et al.
Veröffentlicht: (2024)
LSE-NeRF: Learning Sensor Modeling Errors for Deblured Neural Radiance Fields with RGB-Event Stereo
von: Tang, Wei Zhi, et al.
Veröffentlicht: (2024)
von: Tang, Wei Zhi, et al.
Veröffentlicht: (2024)
Volumetric Rendering with Baked Quadrature Fields
von: Sharma, Gopal, et al.
Veröffentlicht: (2023)
von: Sharma, Gopal, et al.
Veröffentlicht: (2023)
ReDepth Anything: Test-Time Depth Refinement via Self-Supervised Re-lighting
von: Bhattarai, Ananta R., et al.
Veröffentlicht: (2025)
von: Bhattarai, Ananta R., et al.
Veröffentlicht: (2025)
Follow My Hold: Hand-Object Interaction Reconstruction through Geometric Guidance
von: Aytekin, Ayce Idil, et al.
Veröffentlicht: (2025)
von: Aytekin, Ayce Idil, et al.
Veröffentlicht: (2025)
Object and Contact Point Tracking in Demonstrations Using 3D Gaussian Splatting
von: Büttner, Michael, et al.
Veröffentlicht: (2024)
von: Büttner, Michael, et al.
Veröffentlicht: (2024)
E-3DPSM: A State Machine for Event-Based Egocentric 3D Human Pose Estimation
von: Deshmukh, Mayur, et al.
Veröffentlicht: (2026)
von: Deshmukh, Mayur, et al.
Veröffentlicht: (2026)
Salience-Based Adaptive Masking: Revisiting Token Dynamics for Enhanced Pre-training
von: Choi, Hyesong, et al.
Veröffentlicht: (2024)
von: Choi, Hyesong, et al.
Veröffentlicht: (2024)
ROODI: Reconstructing Occluded Objects with Denoising Inpainters
von: Chang, Yeonjin, et al.
Veröffentlicht: (2025)
von: Chang, Yeonjin, et al.
Veröffentlicht: (2025)
Improving Generative Pre-Training: An In-depth Study of Masked Image Modeling and Denoising Models
von: Choi, Hyesong, et al.
Veröffentlicht: (2024)
von: Choi, Hyesong, et al.
Veröffentlicht: (2024)
Audio-Driven Universal Gaussian Head Avatars
von: Teotia, Kartik, et al.
Veröffentlicht: (2025)
von: Teotia, Kartik, et al.
Veröffentlicht: (2025)
SONIC: Spectral Optimization of Noise for Inpainting with Consistency
von: Baek, Seungyeon, et al.
Veröffentlicht: (2025)
von: Baek, Seungyeon, et al.
Veröffentlicht: (2025)
Power Foam: Unifying Real-Time Differentiable Ray Tracing and Rasterization
von: Govindarajan, Shrisudhan, et al.
Veröffentlicht: (2026)
von: Govindarajan, Shrisudhan, et al.
Veröffentlicht: (2026)
Semantic Foam: Unifying Spatial and Semantic Scene Decomposition
von: Sharafeldin, Amr, et al.
Veröffentlicht: (2026)
von: Sharafeldin, Amr, et al.
Veröffentlicht: (2026)
FullCircle: Effortless 3D Reconstruction from Casual 360$^\circ$ Captures
von: Foroutan, Yalda, et al.
Veröffentlicht: (2026)
von: Foroutan, Yalda, et al.
Veröffentlicht: (2026)
NoKSR: Kernel-Free Neural Surface Reconstruction via Point Cloud Serialization
von: Li, Zhen, et al.
Veröffentlicht: (2025)
von: Li, Zhen, et al.
Veröffentlicht: (2025)
A Data Perspective on Enhanced Identity Preservation for Diffusion Personalization
von: He, Xingzhe, et al.
Veröffentlicht: (2023)
von: He, Xingzhe, et al.
Veröffentlicht: (2023)
Representing Animatable Avatar via Factorized Neural Fields
von: Song, Chunjin, et al.
Veröffentlicht: (2024)
von: Song, Chunjin, et al.
Veröffentlicht: (2024)
DreamTexture: Shape from Virtual Texture with Analysis by Augmentation
von: Bhattarai, Ananta R., et al.
Veröffentlicht: (2025)
von: Bhattarai, Ananta R., et al.
Veröffentlicht: (2025)
Grasp in Gaussians: Fast Monocular Reconstruction of Dynamic Hand-Object Interactions
von: Aytekin, Ayce Idil, et al.
Veröffentlicht: (2026)
von: Aytekin, Ayce Idil, et al.
Veröffentlicht: (2026)
3D Gaussian Splatting as Markov Chain Monte Carlo
von: Kheradmand, Shakiba, et al.
Veröffentlicht: (2024)
von: Kheradmand, Shakiba, et al.
Veröffentlicht: (2024)
Lagrangian Hashing for Compressed Neural Field Representations
von: Govindarajan, Shrisudhan, et al.
Veröffentlicht: (2024)
von: Govindarajan, Shrisudhan, et al.
Veröffentlicht: (2024)
VideoUFO: A Million-Scale User-Focused Dataset for Text-to-Video Generation
von: Wang, Wenhao, et al.
Veröffentlicht: (2025)
von: Wang, Wenhao, et al.
Veröffentlicht: (2025)
StochasticSplats: Stochastic Rasterization for Sorting-Free 3D Gaussian Splatting
von: Kheradmand, Shakiba, et al.
Veröffentlicht: (2025)
von: Kheradmand, Shakiba, et al.
Veröffentlicht: (2025)
Make Your Training Flexible: Towards Deployment-Efficient Video Models
von: Wang, Chenting, et al.
Veröffentlicht: (2025)
von: Wang, Chenting, et al.
Veröffentlicht: (2025)
AnyUser: Translating Sketched User Intent into Domestic Robots
von: Yang, Songyuan, et al.
Veröffentlicht: (2026)
von: Yang, Songyuan, et al.
Veröffentlicht: (2026)
PointNeRF++: A multi-scale, point-based Neural Radiance Field
von: Sun, Weiwei, et al.
Veröffentlicht: (2023)
von: Sun, Weiwei, et al.
Veröffentlicht: (2023)
BANF: Band-limited Neural Fields for Levels of Detail Reconstruction
von: Shabanov, Ahan, et al.
Veröffentlicht: (2024)
von: Shabanov, Ahan, et al.
Veröffentlicht: (2024)
IntRec: Intent-based Retrieval with Contrastive Refinement
von: Shamsolmoali, Pourya, et al.
Veröffentlicht: (2026)
von: Shamsolmoali, Pourya, et al.
Veröffentlicht: (2026)
Beyond Real versus Fake Towards Intent-Aware Video Analysis
von: Atreya, Saurabh, et al.
Veröffentlicht: (2025)
von: Atreya, Saurabh, et al.
Veröffentlicht: (2025)
Digital Twin Generation from Visual Data: A Survey
von: Melnik, Andrew, et al.
Veröffentlicht: (2025)
von: Melnik, Andrew, et al.
Veröffentlicht: (2025)
Weakly Supervised Point Cloud Segmentation via Conservative Propagation of Scene-level Labels
von: Xia, Shaobo, et al.
Veröffentlicht: (2023)
von: Xia, Shaobo, et al.
Veröffentlicht: (2023)
ProBA: Probabilistic Bundle Adjustment with the Bhattacharyya Coefficient
von: Chui, Jason, et al.
Veröffentlicht: (2025)
von: Chui, Jason, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Unsupervised Keypoints from Pretrained Diffusion Models
von: Hedlin, Eric, et al.
Veröffentlicht: (2023) -
CasCalib: Cascaded Calibration for Motion Capture from Sparse Unsynchronized Cameras
von: Tang, James, et al.
Veröffentlicht: (2024) -
Mirror-Aware Neural Humans
von: Ajisafe, Daniel, et al.
Veröffentlicht: (2023) -
NESI: Shape Representation via Neural Explicit Surface Intersection
von: Zhang, Congyi, et al.
Veröffentlicht: (2024) -
LatentKeypointGAN: Controlling Images via Latent Keypoints
von: He, Xingzhe, et al.
Veröffentlicht: (2021)