Zero-shot World Models Are Developmentally Efficient Learners
Fuente:
arXiv
Saved in:
| Main Authors: | Aw, Khai Loong, Kotar, Klemen, Lee, Wanhee, Kim, Seungwoo, Jedoui, Khaled, Venkatesh, Rahul, Chen, Lilian Naing, Frank, Michael C., Yamins, Daniel L. K. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Unified 3D Scene Understanding Through Physical World Modeling
by: Lee, Wanhee, et al.
Published: (2026)
by: Lee, Wanhee, et al.
Published: (2026)
3D Scene Understanding Through Local Random Access Sequence Modeling
by: Lee, Wanhee, et al.
Published: (2025)
by: Lee, Wanhee, et al.
Published: (2025)
World Modeling with Probabilistic Structure Integration
by: Kotar, Klemen, et al.
Published: (2025)
by: Kotar, Klemen, et al.
Published: (2025)
Taming generative video models for zero-shot optical flow extraction
by: Kim, Seungwoo, et al.
Published: (2025)
by: Kim, Seungwoo, et al.
Published: (2025)
Physical Object Understanding with a Physically Controllable World Model
by: Venkatesh, Rahul, et al.
Published: (2026)
by: Venkatesh, Rahul, et al.
Published: (2026)
Understanding Physical Dynamics with Counterfactual World Modeling
by: Venkatesh, Rahul, et al.
Published: (2023)
by: Venkatesh, Rahul, et al.
Published: (2023)
Discovering and using Spelke segments
by: Venkatesh, Rahul, et al.
Published: (2025)
by: Venkatesh, Rahul, et al.
Published: (2025)
Representing Speech Through Autoregressive Prediction of Cochlear Tokens
by: Tuckute, Greta, et al.
Published: (2025)
by: Tuckute, Greta, et al.
Published: (2025)
Model Connectomes: A Generational Approach to Data-Efficient Language Models
by: Kotar, Klemen, et al.
Published: (2025)
by: Kotar, Klemen, et al.
Published: (2025)
Self-Supervised Learning of Motion Concepts by Optimizing Counterfactuals
by: Stojanov, Stefan, et al.
Published: (2025)
by: Stojanov, Stefan, et al.
Published: (2025)
Zero-shot Interactive Perception
by: Sripada, Venkatesh, et al.
Published: (2026)
by: Sripada, Venkatesh, et al.
Published: (2026)
Instruction-tuning Aligns LLMs to the Human Brain
by: Aw, Khai Loong, et al.
Published: (2023)
by: Aw, Khai Loong, et al.
Published: (2023)
Characterizing the visual representation of objects from the child's view
by: Yang, Jane, et al.
Published: (2026)
by: Yang, Jane, et al.
Published: (2026)
Noise is an Efficient Learner for Zero-Shot Vision-Language Models
by: Imam, Raza, et al.
Published: (2025)
by: Imam, Raza, et al.
Published: (2025)
World Action Models are Zero-shot Policies
by: Ye, Seonghyeon, et al.
Published: (2026)
by: Ye, Seonghyeon, et al.
Published: (2026)
Efficient and Versatile Robust Fine-Tuning of Zero-shot Models
by: Kim, Sungyeon, et al.
Published: (2024)
by: Kim, Sungyeon, et al.
Published: (2024)
Assessing the alignment between infants' visual and linguistic experience using multimodal language models
by: Tan, Alvin Wei Ming, et al.
Published: (2025)
by: Tan, Alvin Wei Ming, et al.
Published: (2025)
Improved Detection Performance of Cognitive Radio Networks in AWGN and Rayleigh Fading Environments
by: Ying Loong Lee
Published: (2013)
by: Ying Loong Lee
Published: (2013)
FEATHer: Fourier-Efficient Adaptive Temporal Hierarchy Forecaster for Time-Series Forecasting
by: Lee, Jaehoon, et al.
Published: (2026)
by: Lee, Jaehoon, et al.
Published: (2026)
Fundamental Efficiency Limits of Transition-Metal Dichalcogenide Solar Cells with Carrier Multiplication and Hot-Carrier Effects
by: Lee, Seungwoo
Published: (2026)
by: Lee, Seungwoo
Published: (2026)
Transition Metal Dichalcogenides Multijunction Solar Cells Toward the Multicolor Limit
by: Lee, Seungwoo
Published: (2026)
by: Lee, Seungwoo
Published: (2026)
Validating Generative Agent-Based Models of Social Norm Enforcement: From Replication to Novel Predictions
by: Cross, Logan, et al.
Published: (2025)
by: Cross, Logan, et al.
Published: (2025)
Contextrast: Contextual Contrastive Learning for Semantic Segmentation
by: Sung, Changki, et al.
Published: (2024)
by: Sung, Changki, et al.
Published: (2024)
Zero-shot Quantization: A Comprehensive Survey
by: Kim, Minjun, et al.
Published: (2025)
by: Kim, Minjun, et al.
Published: (2025)
Emergence of Fluctuation Relations in UNO
by: Sidajaya, Peter, et al.
Published: (2024)
by: Sidajaya, Peter, et al.
Published: (2024)
LLMs as Zero-shot Graph Learners: Alignment of GNN Representations with LLM Token Embeddings
by: Wang, Duo, et al.
Published: (2024)
by: Wang, Duo, et al.
Published: (2024)
Effectiveness of Zero-shot-CoT in Japanese Prompts
by: Takayama, Shusuke, et al.
Published: (2025)
by: Takayama, Shusuke, et al.
Published: (2025)
Reverso: Efficient Time Series Foundation Models for Zero-shot Forecasting
by: Fu, Xinghong, et al.
Published: (2026)
by: Fu, Xinghong, et al.
Published: (2026)
Zero-shot Text-guided Infinite Image Synthesis with LLM guidance
by: Kwon, Soyeong, et al.
Published: (2024)
by: Kwon, Soyeong, et al.
Published: (2024)
Zero-shot Commonsense Reasoning over Machine Imagination
by: Park, Hyuntae, et al.
Published: (2024)
by: Park, Hyuntae, et al.
Published: (2024)
Modality-Aware Representation Learning for Zero-shot Sketch-based Image Retrieval
by: Lyou, Eunyi, et al.
Published: (2024)
by: Lyou, Eunyi, et al.
Published: (2024)
Zero-shot World Models via Search in Memory
by: Malato, Federico, et al.
Published: (2025)
by: Malato, Federico, et al.
Published: (2025)
Language-only Efficient Training of Zero-shot Composed Image Retrieval
by: Gu, Geonmo, et al.
Published: (2023)
by: Gu, Geonmo, et al.
Published: (2023)
Flashback: Memory-Driven Zero-shot, Real-time Video Anomaly Detection
by: Lee, Hyogun, et al.
Published: (2025)
by: Lee, Hyogun, et al.
Published: (2025)
IFCap: Image-like Retrieval and Frequency-based Entity Filtering for Zero-shot Captioning
by: Lee, Soeun, et al.
Published: (2024)
by: Lee, Soeun, et al.
Published: (2024)
Enhancing Spatio-Temporal Zero-shot Action Recognition with Language-driven Description Attributes
by: Kim, Yehna, et al.
Published: (2025)
by: Kim, Yehna, et al.
Published: (2025)
U-Net for crab image semantic segmentation with PyTorch : UAV imaging and deep learning approach can id entify Brachyura in tidal flats
by: Dongwoo Kim, et al.
Published: (2024)
by: Dongwoo Kim, et al.
Published: (2024)
ARC-NeRF: Area Ray Casting for Broader Unseen View Coverage in Few-shot Object Rendering
by: Seo, Seunghyeon, et al.
Published: (2024)
by: Seo, Seunghyeon, et al.
Published: (2024)
The Role of Masking for Efficient Supervised Knowledge Distillation of Vision Transformers
by: Son, Seungwoo, et al.
Published: (2023)
by: Son, Seungwoo, et al.
Published: (2023)
InstantFamily: Masked Attention for Zero-shot Multi-ID Image Generation
by: Kim, Chanran, et al.
Published: (2024)
by: Kim, Chanran, et al.
Published: (2024)
Similar Items
-
Unified 3D Scene Understanding Through Physical World Modeling
by: Lee, Wanhee, et al.
Published: (2026) -
3D Scene Understanding Through Local Random Access Sequence Modeling
by: Lee, Wanhee, et al.
Published: (2025) -
World Modeling with Probabilistic Structure Integration
by: Kotar, Klemen, et al.
Published: (2025) -
Taming generative video models for zero-shot optical flow extraction
by: Kim, Seungwoo, et al.
Published: (2025) -
Physical Object Understanding with a Physically Controllable World Model
by: Venkatesh, Rahul, et al.
Published: (2026)