Object-Centric Temporal Consistency via Conditional Autoregressive Inductive Biases
Fuente:
arXiv
Saved in:
| Main Authors: | Meo, Cristian, Nakano, Akihiro, Lică, Mircea, Didolkar, Aniket, Suzuki, Masahiro, Goyal, Anirudh, Zhang, Mengmi, Dauwels, Justin, Matsuo, Yutaka, Bengio, Yoshua |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Masked Generative Priors Improve World Models Sequence Modelling Capabilities
by: Meo, Cristian, et al.
Published: (2024)
by: Meo, Cristian, et al.
Published: (2024)
Zero-Shot Object-Centric Representation Learning
by: Didolkar, Aniket, et al.
Published: (2024)
by: Didolkar, Aniket, et al.
Published: (2024)
Bayesian-LoRA: LoRA based Parameter Efficient Fine-Tuning using Optimal Quantization levels and Rank Values trough Differentiable Bayesian Gates
by: Meo, Cristian, et al.
Published: (2024)
by: Meo, Cristian, et al.
Published: (2024)
$α$-TCVAE: On the relationship between Disentanglement and Diversity
by: Meo, Cristian, et al.
Published: (2024)
by: Meo, Cristian, et al.
Published: (2024)
When Object-Centric World Models Meet Policy Learning: From Pixels to Policies, and Where It Breaks
by: Ferraro, Stefano, et al.
Published: (2025)
by: Ferraro, Stefano, et al.
Published: (2025)
Metacognitive Reuse: Turning Recurring LLM Reasoning Into Concise Behaviors
by: Didolkar, Aniket, et al.
Published: (2025)
by: Didolkar, Aniket, et al.
Published: (2025)
Extreme Precipitation Nowcasting using Transformer-based Generative Models
by: Meo, Cristian, et al.
Published: (2024)
by: Meo, Cristian, et al.
Published: (2024)
Metacognitive Capabilities of LLMs: An Exploration in Mathematical Problem Solving
by: Didolkar, Aniket, et al.
Published: (2024)
by: Didolkar, Aniket, et al.
Published: (2024)
BlockGPT: Spatio-Temporal Modelling of Rainfall via Frame-Level Autoregression
by: Meo, Cristian, et al.
Published: (2025)
by: Meo, Cristian, et al.
Published: (2025)
Evaluation of Vision-LLMs in Surveillance Video
by: Benschop, Pascal, et al.
Published: (2025)
by: Benschop, Pascal, et al.
Published: (2025)
EfficientTDMPC: Improved MPC Objectives for Sample-Efficient Continuous Control
by: Evers, Thomas, et al.
Published: (2026)
by: Evers, Thomas, et al.
Published: (2026)
Learning Beyond Pattern Matching? Assaying Mathematical Understanding in LLMs
by: Guo, Siyuan, et al.
Published: (2024)
by: Guo, Siyuan, et al.
Published: (2024)
Enhancing Unimodal Latent Representations in Multimodal VAEs through Iterative Amortized Inference
by: Oshima, Yuta, et al.
Published: (2024)
by: Oshima, Yuta, et al.
Published: (2024)
Slot-VAE: Object-Centric Scene Generation with Slot Attention
by: Wang, Yanbo, et al.
Published: (2023)
by: Wang, Yanbo, et al.
Published: (2023)
WorldPack: Compressed Memory Improves Spatial Consistency in Video World Modeling
by: Oshima, Yuta, et al.
Published: (2025)
by: Oshima, Yuta, et al.
Published: (2025)
CTRL-O: Language-Controllable Object-Centric Visual Representation Learning
by: Didolkar, Aniket, et al.
Published: (2025)
by: Didolkar, Aniket, et al.
Published: (2025)
Shaping Inductive Bias in Diffusion Models through Frequency-Based Noise Control
by: Jiralerspong, Thomas, et al.
Published: (2025)
by: Jiralerspong, Thomas, et al.
Published: (2025)
Assessing the Geographic Generalization and Physical Consistency of Generative Models for Climate Downscaling
by: Saccardi, Carlo, et al.
Published: (2025)
by: Saccardi, Carlo, et al.
Published: (2025)
SSM Meets Video Diffusion Models: Efficient Long-Term Video Generation with Structured State Spaces
by: Oshima, Yuta, et al.
Published: (2024)
by: Oshima, Yuta, et al.
Published: (2024)
Inference-Time Text-to-Video Alignment with Diffusion Latent Beam Search
by: Oshima, Yuta, et al.
Published: (2025)
by: Oshima, Yuta, et al.
Published: (2025)
The Embodied World Model Based on LLM with Visual Information and Prediction-Oriented Prompts
by: Haijima, Wakana, et al.
Published: (2024)
by: Haijima, Wakana, et al.
Published: (2024)
Machine learning and information theory concepts towards an AI Mathematician
by: Bengio, Yoshua, et al.
Published: (2024)
by: Bengio, Yoshua, et al.
Published: (2024)
Rethinking Thinking Tokens: LLMs as Improvement Operators
by: Madaan, Lovish, et al.
Published: (2025)
by: Madaan, Lovish, et al.
Published: (2025)
Unlearning via Sparse Representations
by: Shah, Vedant, et al.
Published: (2023)
by: Shah, Vedant, et al.
Published: (2023)
Baking Symmetry into GFlowNets
by: Ma, George, et al.
Published: (2024)
by: Ma, George, et al.
Published: (2024)
Precipitation Nowcasting Using Physics Informed Discriminator Generative Models
by: Yin, Junzhe, et al.
Published: (2024)
by: Yin, Junzhe, et al.
Published: (2024)
MindForge: Empowering Embodied Agents with Theory of Mind for Lifelong Cultural Learning
by: Lică, Mircea, et al.
Published: (2024)
by: Lică, Mircea, et al.
Published: (2024)
Unveiling Divergent Inductive Biases of LLMs on Temporal Data
by: Kishore, Sindhu, et al.
Published: (2024)
by: Kishore, Sindhu, et al.
Published: (2024)
Noticing the Watcher: LLM Agents Can Infer CoT Monitoring from Blocking Feedback
by: Jiralerspong, Thomas, et al.
Published: (2026)
by: Jiralerspong, Thomas, et al.
Published: (2026)
Bio-Inspired Artificial Neural Networks based on Predictive Coding
by: Casnici, Davide, et al.
Published: (2025)
by: Casnici, Davide, et al.
Published: (2025)
Compositional Scene Understanding through Inverse Generative Modeling
by: Wang, Yanbo, et al.
Published: (2025)
by: Wang, Yanbo, et al.
Published: (2025)
Comments on resolution of nonassociativity in SFT- an example from axioms of BCFT-
by: Matsuo Yutaka
Published: (2002)
by: Matsuo Yutaka
Published: (2002)
Sliding Window Recurrences for Sequence Models
by: Secrieru, Dragos, et al.
Published: (2025)
by: Secrieru, Dragos, et al.
Published: (2025)
Language Models Need Inductive Biases to Count Inductively
by: Chang, Yingshan, et al.
Published: (2024)
by: Chang, Yingshan, et al.
Published: (2024)
Consciousness-Inspired Spatio-Temporal Abstractions for Better Generalization in Reinforcement Learning
by: Zhao, Mingde, et al.
Published: (2023)
by: Zhao, Mingde, et al.
Published: (2023)
Instilling Inductive Biases with Subnetworks
by: Zhang, Enyan, et al.
Published: (2023)
by: Zhang, Enyan, et al.
Published: (2023)
Assessing Situational and Spatial Awareness of VLMs with Synthetically Generated Video
by: Benschop, Pascal, et al.
Published: (2026)
by: Benschop, Pascal, et al.
Published: (2026)
The Good, The Efficient and the Inductive Biases: Exploring Efficiency in Deep Learning Through the Use of Inductive Biases
by: Romero, David W.
Published: (2024)
by: Romero, David W.
Published: (2024)
On Generalization for Generative Flow Networks
by: Krichel, Anas, et al.
Published: (2024)
by: Krichel, Anas, et al.
Published: (2024)
Interventional Causal Representation Learning
by: Ahuja, Kartik, et al.
Published: (2022)
by: Ahuja, Kartik, et al.
Published: (2022)
Similar Items
-
Masked Generative Priors Improve World Models Sequence Modelling Capabilities
by: Meo, Cristian, et al.
Published: (2024) -
Zero-Shot Object-Centric Representation Learning
by: Didolkar, Aniket, et al.
Published: (2024) -
Bayesian-LoRA: LoRA based Parameter Efficient Fine-Tuning using Optimal Quantization levels and Rank Values trough Differentiable Bayesian Gates
by: Meo, Cristian, et al.
Published: (2024) -
$α$-TCVAE: On the relationship between Disentanglement and Diversity
by: Meo, Cristian, et al.
Published: (2024) -
When Object-Centric World Models Meet Policy Learning: From Pixels to Policies, and Where It Breaks
by: Ferraro, Stefano, et al.
Published: (2025)