E.T. the Exceptional Trajectories: Text-to-camera-trajectory generation with character awareness
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Courant, Robin, Dufour, Nicolas, Wang, Xi, Christie, Marc, Kalogeiton, Vicky |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Pulp Motion: Framing-aware multimodal camera and human motion generation
von: Courant, Robin, et al.
Veröffentlicht: (2025)
von: Courant, Robin, et al.
Veröffentlicht: (2025)
AKiRa: Augmentation Kit on Rays for optical video generation
von: Wang, Xi, et al.
Veröffentlicht: (2024)
von: Wang, Xi, et al.
Veröffentlicht: (2024)
FunnyNet-W: Multimodal Learning of Funny Moments in Videos in the Wild
von: Liu, Zhi-Song, et al.
Veröffentlicht: (2024)
von: Liu, Zhi-Song, et al.
Veröffentlicht: (2024)
Don't drop your samples! Coherence-aware training benefits Conditional diffusion
von: Dufour, Nicolas, et al.
Veröffentlicht: (2024)
von: Dufour, Nicolas, et al.
Veröffentlicht: (2024)
Around the World in 80 Timesteps: A Generative Approach to Global Visual Geolocation
von: Dufour, Nicolas, et al.
Veröffentlicht: (2024)
von: Dufour, Nicolas, et al.
Veröffentlicht: (2024)
How far can we go with ImageNet for Text-to-Image generation?
von: Degeorge, L., et al.
Veröffentlicht: (2025)
von: Degeorge, L., et al.
Veröffentlicht: (2025)
Training-Free Synthetic Data Generation with Dual IP-Adapter Guidance
von: Boudier, Luc, et al.
Veröffentlicht: (2025)
von: Boudier, Luc, et al.
Veröffentlicht: (2025)
MIRO: MultI-Reward cOnditioned pretraining improves T2I quality and efficiency
von: Dufour, Nicolas, et al.
Veröffentlicht: (2025)
von: Dufour, Nicolas, et al.
Veröffentlicht: (2025)
SF20K Competition 2025: Summary and findings
von: Ghermi, Ridouane, et al.
Veröffentlicht: (2026)
von: Ghermi, Ridouane, et al.
Veröffentlicht: (2026)
What about gravity in video generation? Post-Training Newton's Laws with Verifiable Rewards
von: Le, Minh-Quan, et al.
Veröffentlicht: (2025)
von: Le, Minh-Quan, et al.
Veröffentlicht: (2025)
Analysis of Classifier-Free Guidance Weight Schedulers
von: Wang, Xi, et al.
Veröffentlicht: (2024)
von: Wang, Xi, et al.
Veröffentlicht: (2024)
MUSE: Manipulating Unified Framework for Synthesizing Emotions in Images via Test-Time Optimization
von: Xia, Yingjie, et al.
Veröffentlicht: (2025)
von: Xia, Yingjie, et al.
Veröffentlicht: (2025)
Long Story Short: Story-level Video Understanding from 20K Short Films
von: Ghermi, Ridouane, et al.
Veröffentlicht: (2024)
von: Ghermi, Ridouane, et al.
Veröffentlicht: (2024)
Bridging Text and Image for Artist Style Transfer via Contrastive Learning
von: Liu, Zhi-Song, et al.
Veröffentlicht: (2024)
von: Liu, Zhi-Song, et al.
Veröffentlicht: (2024)
Studying Image Diffusion Features for Zero-Shot Video Object Segmentation
von: Delatolas, Thanos, et al.
Veröffentlicht: (2025)
von: Delatolas, Thanos, et al.
Veröffentlicht: (2025)
Soft-Di[M]O: Improving One-Step Discrete Image Generation with Soft Embeddings
von: Zhu, Yuanzhi, et al.
Veröffentlicht: (2025)
von: Zhu, Yuanzhi, et al.
Veröffentlicht: (2025)
Di$\mathtt{[M]}$O: Distilling Masked Diffusion Models into One-step Generator
von: Zhu, Yuanzhi, et al.
Veröffentlicht: (2025)
von: Zhu, Yuanzhi, et al.
Veröffentlicht: (2025)
Diffusion Reinforcement Learning via Centered Reward Distillation
von: Zhu, Yuanzhi, et al.
Veröffentlicht: (2026)
von: Zhu, Yuanzhi, et al.
Veröffentlicht: (2026)
Name Your Style: An Arbitrary Artist-aware Image Style Transfer
von: Liu, Zhi-Song, et al.
Veröffentlicht: (2022)
von: Liu, Zhi-Song, et al.
Veröffentlicht: (2022)
T-REGS: Minimum Spanning Tree Regularization for Self-Supervised Learning
von: Mordacq, Julie, et al.
Veröffentlicht: (2025)
von: Mordacq, Julie, et al.
Veröffentlicht: (2025)
One-step Diffusion Models with Bregman Density Ratio Matching
von: Zhu, Yuanzhi, et al.
Veröffentlicht: (2025)
von: Zhu, Yuanzhi, et al.
Veröffentlicht: (2025)
ADAPT: Multimodal Learning for Detecting Physiological Changes under Missing Modalities
von: Mordacq, Julie, et al.
Veröffentlicht: (2024)
von: Mordacq, Julie, et al.
Veröffentlicht: (2024)
LEAD: Latent Realignment for Human Motion Diffusion
von: Andreou, Nefeli, et al.
Veröffentlicht: (2024)
von: Andreou, Nefeli, et al.
Veröffentlicht: (2024)
Make me an Expert: Distilling from Generalist Black-Box Models into Specialized Models for Semantic Segmentation
von: Benigmim, Yasser, et al.
Veröffentlicht: (2025)
von: Benigmim, Yasser, et al.
Veröffentlicht: (2025)
SIGHT: Synthesizing Image-Text Conditioned and Geometry-Guided 3D Hand-Object Trajectories
von: Gavryushin, Alexey, et al.
Veröffentlicht: (2025)
von: Gavryushin, Alexey, et al.
Veröffentlicht: (2025)
Collaborating Foundation Models for Domain Generalized Semantic Segmentation
von: Benigmim, Yasser, et al.
Veröffentlicht: (2023)
von: Benigmim, Yasser, et al.
Veröffentlicht: (2023)
FakeParts: a New Family of AI-Generated DeepFakes
von: Liu, Ziyi, et al.
Veröffentlicht: (2025)
von: Liu, Ziyi, et al.
Veröffentlicht: (2025)
EgoNav: Egocentric Scene-aware Human Trajectory Prediction
von: Wang, Weizhuo, et al.
Veröffentlicht: (2024)
von: Wang, Weizhuo, et al.
Veröffentlicht: (2024)
Leveraging Text Localization for Scene Text Removal via Text-aware Masked Image Modeling
von: Wang, Zixiao, et al.
Veröffentlicht: (2024)
von: Wang, Zixiao, et al.
Veröffentlicht: (2024)
M${^2}$Depth: Self-supervised Two-Frame Multi-camera Metric Depth Estimation
von: Zou, Yingshuang, et al.
Veröffentlicht: (2024)
von: Zou, Yingshuang, et al.
Veröffentlicht: (2024)
WoVoGen: World Volume-aware Diffusion for Controllable Multi-camera Driving Scene Generation
von: Lu, Jiachen, et al.
Veröffentlicht: (2023)
von: Lu, Jiachen, et al.
Veröffentlicht: (2023)
Attention-aware Social Graph Transformer Networks for Stochastic Trajectory Prediction
von: Liu, Yao, et al.
Veröffentlicht: (2023)
von: Liu, Yao, et al.
Veröffentlicht: (2023)
FAST: Foreground-aware Diffusion with Accelerated Sampling Trajectory for Segmentation-oriented Anomaly Synthesis
von: Xu, Xichen, et al.
Veröffentlicht: (2025)
von: Xu, Xichen, et al.
Veröffentlicht: (2025)
Physical Plausibility-aware Trajectory Prediction via Locomotion Embodiment
von: Taketsugu, Hiromu, et al.
Veröffentlicht: (2025)
von: Taketsugu, Hiromu, et al.
Veröffentlicht: (2025)
Device-aware Optical Adversarial Attack for a Portable Projector-camera System
von: Jiang, Ning, et al.
Veröffentlicht: (2025)
von: Jiang, Ning, et al.
Veröffentlicht: (2025)
Trajectory-aware Shifted State Space Models for Online Video Super-Resolution
von: Zhu, Qiang, et al.
Veröffentlicht: (2025)
von: Zhu, Qiang, et al.
Veröffentlicht: (2025)
Geometry aware 3D generation from in-the-wild images in ImageNet
von: Shen, Qijia, et al.
Veröffentlicht: (2024)
von: Shen, Qijia, et al.
Veröffentlicht: (2024)
Towards Foundation Models for Cryo-ET Subtomogram Analysis
von: Jiang, Runmin, et al.
Veröffentlicht: (2025)
von: Jiang, Runmin, et al.
Veröffentlicht: (2025)
Text2Place: Affordance-aware Text Guided Human Placement
von: Parihar, Rishubh, et al.
Veröffentlicht: (2024)
von: Parihar, Rishubh, et al.
Veröffentlicht: (2024)
TextIM: Part-aware Interactive Motion Synthesis from Text
von: Fan, Siyuan, et al.
Veröffentlicht: (2024)
von: Fan, Siyuan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Pulp Motion: Framing-aware multimodal camera and human motion generation
von: Courant, Robin, et al.
Veröffentlicht: (2025) -
AKiRa: Augmentation Kit on Rays for optical video generation
von: Wang, Xi, et al.
Veröffentlicht: (2024) -
FunnyNet-W: Multimodal Learning of Funny Moments in Videos in the Wild
von: Liu, Zhi-Song, et al.
Veröffentlicht: (2024) -
Don't drop your samples! Coherence-aware training benefits Conditional diffusion
von: Dufour, Nicolas, et al.
Veröffentlicht: (2024) -
Around the World in 80 Timesteps: A Generative Approach to Global Visual Geolocation
von: Dufour, Nicolas, et al.
Veröffentlicht: (2024)