Poly-Autoregressive Prediction for Modeling Interactions
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Thakkar, Neerja, Sadjadpour, Tara, Rajasegaran, Jathushan, Ginosar, Shiry, Malik, Jitendra |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Gaussian Masked Autoencoders
von: Rajasegaran, Jathushan, et al.
Veröffentlicht: (2025)
von: Rajasegaran, Jathushan, et al.
Veröffentlicht: (2025)
Synergy and Synchrony in Couple Dances
von: Maluleke, Vongani, et al.
Veröffentlicht: (2024)
von: Maluleke, Vongani, et al.
Veröffentlicht: (2024)
Forecasting Motion in the Wild
von: Thakkar, Neerja, et al.
Veröffentlicht: (2026)
von: Thakkar, Neerja, et al.
Veröffentlicht: (2026)
Tracking by Predicting 3-D Gaussians Over Time
von: Baranwal, Tanish, et al.
Veröffentlicht: (2025)
von: Baranwal, Tanish, et al.
Veröffentlicht: (2025)
Scaling Properties of Diffusion Models for Perceptual Tasks
von: Ravishankar, Rahul, et al.
Veröffentlicht: (2024)
von: Ravishankar, Rahul, et al.
Veröffentlicht: (2024)
An Empirical Study of Autoregressive Pre-training from Videos
von: Rajasegaran, Jathushan, et al.
Veröffentlicht: (2025)
von: Rajasegaran, Jathushan, et al.
Veröffentlicht: (2025)
Adaptive Human Trajectory Prediction via Latent Corridors
von: Thakkar, Neerja, et al.
Veröffentlicht: (2023)
von: Thakkar, Neerja, et al.
Veröffentlicht: (2023)
Synthesizing Moving People with 3D Control
von: Li, Boyi, et al.
Veröffentlicht: (2024)
von: Li, Boyi, et al.
Veröffentlicht: (2024)
Humanoid Locomotion as Next Token Prediction
von: Radosavovic, Ilija, et al.
Veröffentlicht: (2024)
von: Radosavovic, Ilija, et al.
Veröffentlicht: (2024)
FewShotNeRF: Meta-Learning-based Novel View Synthesis for Rapid Scene-Specific Adaptation
von: Sivakumar, Piraveen, et al.
Veröffentlicht: (2024)
von: Sivakumar, Piraveen, et al.
Veröffentlicht: (2024)
Pose Priors from Language Models
von: Subramanian, Sanjay, et al.
Veröffentlicht: (2024)
von: Subramanian, Sanjay, et al.
Veröffentlicht: (2024)
Diffusion Models as Data Mining Tools
von: Siglidis, Ioannis, et al.
Veröffentlicht: (2024)
von: Siglidis, Ioannis, et al.
Veröffentlicht: (2024)
EgoPet: Egomotion and Interaction Data from an Animal's Perspective
von: Bar, Amir, et al.
Veröffentlicht: (2024)
von: Bar, Amir, et al.
Veröffentlicht: (2024)
Back into Plato's Cave: Examining Cross-modal Representational Convergence at Scale
von: Koepke, A. Sophia, et al.
Veröffentlicht: (2026)
von: Koepke, A. Sophia, et al.
Veröffentlicht: (2026)
KiVA: Kid-inspired Visual Analogies for Testing Large Multimodal Models
von: Yiu, Eunice, et al.
Veröffentlicht: (2024)
von: Yiu, Eunice, et al.
Veröffentlicht: (2024)
Diffusion Forcing for Multi-Agent Interaction Sequence Modeling
von: Maluleke, Vongani H., et al.
Veröffentlicht: (2025)
von: Maluleke, Vongani H., et al.
Veröffentlicht: (2025)
Human-level 3D shape perception emerges from multi-view learning
von: Bonnen, Tyler, et al.
Veröffentlicht: (2026)
von: Bonnen, Tyler, et al.
Veröffentlicht: (2026)
Reconstructing Hand-Held Objects in 3D from Images and Videos
von: Wu, Jane, et al.
Veröffentlicht: (2024)
von: Wu, Jane, et al.
Veröffentlicht: (2024)
HINT: Hierarchical Interaction Modeling for Autoregressive Multi-Human Motion Generation
von: Liu, Mengge, et al.
Veröffentlicht: (2026)
von: Liu, Mengge, et al.
Veröffentlicht: (2026)
Long-Context Autoregressive Video Modeling with Next-Frame Prediction
von: Gu, Yuchao, et al.
Veröffentlicht: (2025)
von: Gu, Yuchao, et al.
Veröffentlicht: (2025)
FVAR: Visual Autoregressive Modeling via Next Focus Prediction
von: Li, Xiaofan, et al.
Veröffentlicht: (2025)
von: Li, Xiaofan, et al.
Veröffentlicht: (2025)
Frozen Forecasting: A Unified Evaluation
von: Walker, Jacob C, et al.
Veröffentlicht: (2025)
von: Walker, Jacob C, et al.
Veröffentlicht: (2025)
Autoregressive Flow Matching for Motion Prediction
von: Xie, Johnathan, et al.
Veröffentlicht: (2025)
von: Xie, Johnathan, et al.
Veröffentlicht: (2025)
Interact2Ar: Full-Body Human-Human Interaction Generation via Autoregressive Diffusion Models
von: Ruiz-Ponce, Pablo, et al.
Veröffentlicht: (2025)
von: Ruiz-Ponce, Pablo, et al.
Veröffentlicht: (2025)
FlexVAR: Flexible Visual Autoregressive Modeling without Residual Prediction
von: Jiao, Siyu, et al.
Veröffentlicht: (2025)
von: Jiao, Siyu, et al.
Veröffentlicht: (2025)
Next Patch Prediction for Autoregressive Visual Generation
von: Pang, Yatian, et al.
Veröffentlicht: (2024)
von: Pang, Yatian, et al.
Veröffentlicht: (2024)
Reconstructing People, Places, and Cameras
von: Müller, Lea, et al.
Veröffentlicht: (2024)
von: Müller, Lea, et al.
Veröffentlicht: (2024)
Customize Your Visual Autoregressive Recipe with Set Autoregressive Modeling
von: Liu, Wenze, et al.
Veröffentlicht: (2024)
von: Liu, Wenze, et al.
Veröffentlicht: (2024)
ARIG: Autoregressive Interactive Head Generation for Real-time Conversations
von: Guo, Ying, et al.
Veröffentlicht: (2025)
von: Guo, Ying, et al.
Veröffentlicht: (2025)
Hand-Object Interaction Pretraining from Videos
von: Singh, Himanshu Gaurav, et al.
Veröffentlicht: (2024)
von: Singh, Himanshu Gaurav, et al.
Veröffentlicht: (2024)
Perception Encoder: The best visual embeddings are not at the output of the network
von: Bolya, Daniel, et al.
Veröffentlicht: (2025)
von: Bolya, Daniel, et al.
Veröffentlicht: (2025)
Autoregressive Video Generation beyond Next Frames Prediction
von: Ren, Sucheng, et al.
Veröffentlicht: (2025)
von: Ren, Sucheng, et al.
Veröffentlicht: (2025)
Learning from Next-Frame Prediction: Autoregressive Video Modeling Encodes Effective Representations
von: Li, Jinghan, et al.
Veröffentlicht: (2025)
von: Li, Jinghan, et al.
Veröffentlicht: (2025)
Automating Deformable Gasket Assembly
von: Adebola, Simeon, et al.
Veröffentlicht: (2024)
von: Adebola, Simeon, et al.
Veröffentlicht: (2024)
Visual Implicit Autoregressive Modeling
von: Jiang, Pengfei, et al.
Veröffentlicht: (2026)
von: Jiang, Pengfei, et al.
Veröffentlicht: (2026)
Knot Forcing: Taming Autoregressive Video Diffusion Models for Real-time Infinite Interactive Portrait Animation
von: Xiao, Steven, et al.
Veröffentlicht: (2025)
von: Xiao, Steven, et al.
Veröffentlicht: (2025)
Improving Autoregressive Visual Generation with Cluster-Oriented Token Prediction
von: Hu, Teng, et al.
Veröffentlicht: (2025)
von: Hu, Teng, et al.
Veröffentlicht: (2025)
From Prediction to Perfection: Introducing Refinement to Autoregressive Image Generation
von: Cheng, Cheng, et al.
Veröffentlicht: (2025)
von: Cheng, Cheng, et al.
Veröffentlicht: (2025)
Autoregressive Universal Video Segmentation Model
von: Heo, Miran, et al.
Veröffentlicht: (2025)
von: Heo, Miran, et al.
Veröffentlicht: (2025)
Visual Self-Refinement for Autoregressive Models
von: Wang, Jiamian, et al.
Veröffentlicht: (2025)
von: Wang, Jiamian, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Gaussian Masked Autoencoders
von: Rajasegaran, Jathushan, et al.
Veröffentlicht: (2025) -
Synergy and Synchrony in Couple Dances
von: Maluleke, Vongani, et al.
Veröffentlicht: (2024) -
Forecasting Motion in the Wild
von: Thakkar, Neerja, et al.
Veröffentlicht: (2026) -
Tracking by Predicting 3-D Gaussians Over Time
von: Baranwal, Tanish, et al.
Veröffentlicht: (2025) -
Scaling Properties of Diffusion Models for Perceptual Tasks
von: Ravishankar, Rahul, et al.
Veröffentlicht: (2024)