One Trajectory, One Token: Grounded Video Tokenization via Panoptic Sub-object Trajectory
Fuente:
arXiv
Saved in:
| Main Authors: | Zheng, Chenhao, Zhang, Jieyu, Salehi, Mohammadreza, Gao, Ziqi, Iyengar, Vishnu, Kobori, Norimasa, Kong, Quan, Krishna, Ranjay |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TrajTok: Learning Trajectory Tokens enables better Video Understanding
by: Zheng, Chenhao, et al.
Published: (2026)
by: Zheng, Chenhao, et al.
Published: (2026)
Negative Token Merging: Image-based Adversarial Feature Guidance
by: Singh, Jaskirat, et al.
Published: (2024)
by: Singh, Jaskirat, et al.
Published: (2024)
TrajectoryCrafter: Redirecting Camera Trajectory for Monocular Videos via Diffusion Models
by: YU, Mark, et al.
Published: (2025)
by: YU, Mark, et al.
Published: (2025)
Scaling Mesh Generation via Compressive Tokenization
by: Weng, Haohan, et al.
Published: (2024)
by: Weng, Haohan, et al.
Published: (2024)
Token Perturbation Guidance for Diffusion Models
by: Rajabi, Javad, et al.
Published: (2025)
by: Rajabi, Javad, et al.
Published: (2025)
TokenLight: Precise Lighting Control in Images using Attribute Tokens
by: Chaturvedi, Sumit, et al.
Published: (2026)
by: Chaturvedi, Sumit, et al.
Published: (2026)
MolmoPoint: Better Pointing for VLMs with Grounding Tokens
by: Clark, Christopher, et al.
Published: (2026)
by: Clark, Christopher, et al.
Published: (2026)
3D Shape Tokenization via Latent Flow Matching
by: Chang, Jen-Hao Rick, et al.
Published: (2024)
by: Chang, Jen-Hao Rick, et al.
Published: (2024)
One-Shot Method for Computing Generalized Winding Numbers
by: Martens, Cedric, et al.
Published: (2024)
by: Martens, Cedric, et al.
Published: (2024)
Temporally Smooth Mesh Extraction for Procedural Scenes with Long-Range Camera Trajectories using Spacetime Octrees
by: Ma, Zeyu, et al.
Published: (2025)
by: Ma, Zeyu, et al.
Published: (2025)
Symbol as Points: Panoptic Symbol Spotting via Point-based Representation
by: Liu, Wenlong, et al.
Published: (2024)
by: Liu, Wenlong, et al.
Published: (2024)
SANA-Sprint: One-Step Diffusion with Continuous-Time Consistency Distillation
by: Chen, Junsong, et al.
Published: (2025)
by: Chen, Junsong, et al.
Published: (2025)
Skin Tokens: A Learned Compact Representation for Unified Autoregressive Rigging
by: Zhang, Jia-peng, et al.
Published: (2026)
by: Zhang, Jia-peng, et al.
Published: (2026)
Conformal Slit Mapping Based Spiral Tool Trajectory Planning for Ball-end Milling on Complex Freeform Surfaces
by: Shen, Changqing, et al.
Published: (2025)
by: Shen, Changqing, et al.
Published: (2025)
SuperVoxelGPT: Adaptive and Ordered 3D Tokenization for Autoregressive Shape Generation
by: Li, Yuan, et al.
Published: (2026)
by: Li, Yuan, et al.
Published: (2026)
E-VLC: A Real-World Dataset for Event-based Visible Light Communication And Localization
by: Shiba, Shintaro, et al.
Published: (2025)
by: Shiba, Shintaro, et al.
Published: (2025)
One Model to Rig Them All: Diverse Skeleton Rigging with UniRig
by: Zhang, Jia-Peng, et al.
Published: (2025)
by: Zhang, Jia-Peng, et al.
Published: (2025)
BrickAnything: Geometry-Conditioned Buildable Brick Generation with Structure-Aware Tokenization
by: Ni, Zhengyang, et al.
Published: (2026)
by: Ni, Zhengyang, et al.
Published: (2026)
Good Token Hunting: A Hitchhiker's Guide to Token Selection for Visual Geometry Transformers
by: Zheng, Shuhong, et al.
Published: (2026)
by: Zheng, Shuhong, et al.
Published: (2026)
Transforming a Non-Differentiable Rasterizer into a Differentiable One with Stochastic Gradient Estimation
by: Deliot, Thomas, et al.
Published: (2024)
by: Deliot, Thomas, et al.
Published: (2024)
Meshing of High-Dimensional Toroidal Manifolds from Quasi-Periodic Three-Body Problem Dynamics using Parameterization via Discrete One-Forms
by: Basile, Dante, et al.
Published: (2025)
by: Basile, Dante, et al.
Published: (2025)
Implicit Swept Volume SDF: Enabling Continuous Collision-Free Trajectory Generation for Arbitrary Shapes
by: Wang, Jingping, et al.
Published: (2024)
by: Wang, Jingping, et al.
Published: (2024)
Automatic Camera Trajectory Control with Enhanced Immersion for Virtual Cinematography
by: Wu, Xinyi, et al.
Published: (2023)
by: Wu, Xinyi, et al.
Published: (2023)
Taming Real-World Space-Time Video Super-Resolution with One-Step Diffusion
by: Wei, Shuoyan, et al.
Published: (2026)
by: Wei, Shuoyan, et al.
Published: (2026)
One algebra for all : Geometric Algebra methods for neurosymbolic XR scene authoring, animation and neural rendering
by: Kamarianakis, Manos, et al.
Published: (2025)
by: Kamarianakis, Manos, et al.
Published: (2025)
LiTo: Surface Light Field Tokenization
by: Chang, Jen-Hao Rick, et al.
Published: (2026)
by: Chang, Jen-Hao Rick, et al.
Published: (2026)
TLControl: Trajectory and Language Control for Human Motion Synthesis
by: Wan, Weilin, et al.
Published: (2023)
by: Wan, Weilin, et al.
Published: (2023)
One-shot Embroidery Customization via Contrastive LoRA Modulation
by: Ma, Jun, et al.
Published: (2025)
by: Ma, Jun, et al.
Published: (2025)
One Shot, One Talk: Whole-body Talking Avatar from a Single Image
by: Xiang, Jun, et al.
Published: (2024)
by: Xiang, Jun, et al.
Published: (2024)
Tokenizing Buildings: A Transformer for Layout Synthesis
by: de Guevara, Manuel Ladron, et al.
Published: (2025)
by: de Guevara, Manuel Ladron, et al.
Published: (2025)
Split&Splat: Zero-Shot Panoptic Segmentation via Explicit Instance Modeling and 3D Gaussian Splatting
by: Monchieri, Leonardo, et al.
Published: (2026)
by: Monchieri, Leonardo, et al.
Published: (2026)
Learning to Importance Sample in Primary Sample Space
by: Zheng, Quan, et al.
Published: (2018)
by: Zheng, Quan, et al.
Published: (2018)
VideoFrom3D: 3D Scene Video Generation via Complementary Image and Video Diffusion Models
by: Kim, Geonung, et al.
Published: (2025)
by: Kim, Geonung, et al.
Published: (2025)
SAEdit: Token-level control for continuous image editing via Sparse AutoEncoder
by: Kamenetsky, Ronen, et al.
Published: (2025)
by: Kamenetsky, Ronen, et al.
Published: (2025)
Commercial Vehicle Braking Optimization: A Robust SIFT-Trajectory Approach
by: Li, Zhe, et al.
Published: (2025)
by: Li, Zhe, et al.
Published: (2025)
Neural Network-Based Tracking and 3D Reconstruction of Baseball Pitch Trajectories from Single-View 2D Video
by: Hsieh, Jhen
Published: (2024)
by: Hsieh, Jhen
Published: (2024)
Scaling Transformer-Based Novel View Synthesis Models with Token Disentanglement and Synthetic Data
by: Nair, Nithin Gopalakrishnan, et al.
Published: (2025)
by: Nair, Nithin Gopalakrishnan, et al.
Published: (2025)
3D Skew Gaussian Splatting with Any Camera Trajectory Visualization Engine
by: Zhao, Beizhen, et al.
Published: (2026)
by: Zhao, Beizhen, et al.
Published: (2026)
Learning 3D Garment Animation from Trajectories of A Piece of Cloth
by: Shao, Yidi, et al.
Published: (2025)
by: Shao, Yidi, et al.
Published: (2025)
ACT-R: Adaptive Camera Trajectories for Single View 3D Reconstruction
by: Wang, Yizhi, et al.
Published: (2025)
by: Wang, Yizhi, et al.
Published: (2025)
Similar Items
-
TrajTok: Learning Trajectory Tokens enables better Video Understanding
by: Zheng, Chenhao, et al.
Published: (2026) -
Negative Token Merging: Image-based Adversarial Feature Guidance
by: Singh, Jaskirat, et al.
Published: (2024) -
TrajectoryCrafter: Redirecting Camera Trajectory for Monocular Videos via Diffusion Models
by: YU, Mark, et al.
Published: (2025) -
Scaling Mesh Generation via Compressive Tokenization
by: Weng, Haohan, et al.
Published: (2024) -
Token Perturbation Guidance for Diffusion Models
by: Rajabi, Javad, et al.
Published: (2025)