AI Must Embrace Specialization via Superhuman Adaptable Intelligence
Fuente:
arXiv
Salvato in:
| Autori principali: | Goldfeder, Judah, Wyder, Philippe, LeCun, Yann, Ziv, Ravid Shwartz |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Video Representation Learning with Joint-Embedding Predictive Architectures
di: Drozdov, Katrina, et al.
Pubblicazione: (2024)
di: Drozdov, Katrina, et al.
Pubblicazione: (2024)
Soft Clustering Anchors for Self-Supervised Speech Representation Learning in Joint Embedding Prediction Architectures
di: Ioannides, Georgios, et al.
Pubblicazione: (2026)
di: Ioannides, Georgios, et al.
Pubblicazione: (2026)
A superpersuasive autonomous policy debating system
di: Roush, Allen, et al.
Pubblicazione: (2025)
di: Roush, Allen, et al.
Pubblicazione: (2025)
From Tokens to Thoughts: How LLMs and Humans Trade Compression for Meaning
di: Shani, Chen, et al.
Pubblicazione: (2025)
di: Shani, Chen, et al.
Pubblicazione: (2025)
Variance-Covariance Regularization Improves Representation Learning
di: Zhu, Jiachen, et al.
Pubblicazione: (2023)
di: Zhu, Jiachen, et al.
Pubblicazione: (2023)
An Information-Theoretic Perspective on Variance-Invariance-Covariance Regularization
di: Shwartz-Ziv, Ravid, et al.
Pubblicazione: (2023)
di: Shwartz-Ziv, Ravid, et al.
Pubblicazione: (2023)
JEPA as a Neural Tokenizer: Learning Robust Speech Representations with Density Adaptive Attention
di: Ioannides, Georgios, et al.
Pubblicazione: (2025)
di: Ioannides, Georgios, et al.
Pubblicazione: (2025)
Layer by Layer: Uncovering Hidden Representations in Language Models
di: Skean, Oscar, et al.
Pubblicazione: (2025)
di: Skean, Oscar, et al.
Pubblicazione: (2025)
Attention Sinks and Compression Valleys in LLMs are Two Sides of the Same Coin
di: Queipo-de-Llano, Enrique, et al.
Pubblicazione: (2025)
di: Queipo-de-Llano, Enrique, et al.
Pubblicazione: (2025)
The Entropy Enigma: Success and Failure of Entropy Minimization
di: Press, Ori, et al.
Pubblicazione: (2024)
di: Press, Ori, et al.
Pubblicazione: (2024)
Does Representation Matter? Exploring Intermediate Layers in Large Language Models
di: Skean, Oscar, et al.
Pubblicazione: (2024)
di: Skean, Oscar, et al.
Pubblicazione: (2024)
Antislop: A Comprehensive Framework for Identifying and Eliminating Repetitive Patterns in Language Models
di: Paech, Samuel, et al.
Pubblicazione: (2025)
di: Paech, Samuel, et al.
Pubblicazione: (2025)
Why AI systems don't learn and what to do about it: Lessons on autonomous learning from cognitive science
di: Dupoux, Emmanuel, et al.
Pubblicazione: (2026)
di: Dupoux, Emmanuel, et al.
Pubblicazione: (2026)
Fast and Exact Enumeration of Deep Networks Partitions Regions
di: Balestriero, Randall, et al.
Pubblicazione: (2024)
di: Balestriero, Randall, et al.
Pubblicazione: (2024)
Do Multi-Agents Dream of Electric Screens? Achieving Perfect Accuracy on AndroidWorld Through Task Decomposition
di: Favreau, Pierre-Louis, et al.
Pubblicazione: (2026)
di: Favreau, Pierre-Louis, et al.
Pubblicazione: (2026)
On Training in Imagination
di: Timor, Nadav, et al.
Pubblicazione: (2026)
di: Timor, Nadav, et al.
Pubblicazione: (2026)
LeJEPA: Provable and Scalable Self-Supervised Learning Without the Heuristics
di: Balestriero, Randall, et al.
Pubblicazione: (2025)
di: Balestriero, Randall, et al.
Pubblicazione: (2025)
Learning by Reconstruction Produces Uninformative Features For Perception
di: Balestriero, Randall, et al.
Pubblicazione: (2024)
di: Balestriero, Randall, et al.
Pubblicazione: (2024)
LLM-JEPA: Large Language Models Meet Joint Embedding Predictive Architectures
di: Huang, Hai, et al.
Pubblicazione: (2025)
di: Huang, Hai, et al.
Pubblicazione: (2025)
NdLinear: Preserving Multi-Dimensional Structure for Parameter-Efficient Neural Networks
di: Reneau, Alex, et al.
Pubblicazione: (2025)
di: Reneau, Alex, et al.
Pubblicazione: (2025)
Rate-In: Information-Driven Adaptive Dropout Rates for Improved Inference-Time Uncertainty Estimation
di: Zeevi, Tal, et al.
Pubblicazione: (2024)
di: Zeevi, Tal, et al.
Pubblicazione: (2024)
DINO-WM: World Models on Pre-trained Visual Features enable Zero-shot Planning
di: Zhou, Gaoyue, et al.
Pubblicazione: (2024)
di: Zhou, Gaoyue, et al.
Pubblicazione: (2024)
The Spike, the Sparse and the Sink: Anatomy of Massive Activations and Attention Sinks
di: Sun, Shangwen, et al.
Pubblicazione: (2026)
di: Sun, Shangwen, et al.
Pubblicazione: (2026)
Position: AI Safety Must Embrace an Antifragile Perspective
di: Jin, Ming, et al.
Pubblicazione: (2025)
di: Jin, Ming, et al.
Pubblicazione: (2025)
Knolling Bot: Teaching Robots the Human Notion of Tidiness
di: Hu, Yuhang, et al.
Pubblicazione: (2023)
di: Hu, Yuhang, et al.
Pubblicazione: (2023)
Beyond Cropping and Rotation: Automated Evolution of Powerful Task-Specific Augmentations with Generative Models
di: Goldfeder, Judah, et al.
Pubblicazione: (2026)
di: Goldfeder, Judah, et al.
Pubblicazione: (2026)
Causal-JEPA: Learning World Models through Object-Level Latent Masking
di: Nam, Heejeong, et al.
Pubblicazione: (2026)
di: Nam, Heejeong, et al.
Pubblicazione: (2026)
When Attention Collapses: How Degenerate Layers in LLMs Enable Smaller, Stronger Models
di: Sanyal, Sunny, et al.
Pubblicazione: (2024)
di: Sanyal, Sunny, et al.
Pubblicazione: (2024)
Evidence of an Emergent "Self" in Continual Robot Learning
di: Jhunjhunwala, Adidev, et al.
Pubblicazione: (2026)
di: Jhunjhunwala, Adidev, et al.
Pubblicazione: (2026)
Gaussian Embeddings: How JEPAs Secretly Learn Your Data Density
di: Balestriero, Randall, et al.
Pubblicazione: (2025)
di: Balestriero, Randall, et al.
Pubblicazione: (2025)
LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels
di: Maes, Lucas, et al.
Pubblicazione: (2026)
di: Maes, Lucas, et al.
Pubblicazione: (2026)
Just How Flexible are Neural Networks in Practice?
di: Shwartz-Ziv, Ravid, et al.
Pubblicazione: (2024)
di: Shwartz-Ziv, Ravid, et al.
Pubblicazione: (2024)
Value-guided action planning with JEPA world models
di: Destrade, Matthieu, et al.
Pubblicazione: (2025)
di: Destrade, Matthieu, et al.
Pubblicazione: (2025)
Open-Endedness is Essential for Artificial Superhuman Intelligence
di: Hughes, Edward, et al.
Pubblicazione: (2024)
di: Hughes, Edward, et al.
Pubblicazione: (2024)
Blockwise Self-Supervised Learning at Scale
di: Siddiqui, Shoaib Ahmed, et al.
Pubblicazione: (2023)
di: Siddiqui, Shoaib Ahmed, et al.
Pubblicazione: (2023)
Light-weight probing of unsupervised representations for Reinforcement Learning
di: Zhang, Wancong, et al.
Pubblicazione: (2022)
di: Zhang, Wancong, et al.
Pubblicazione: (2022)
Seq-VCR: Preventing Collapse in Intermediate Transformer Representations for Enhanced Reasoning
di: Arefin, Md Rifat, et al.
Pubblicazione: (2024)
di: Arefin, Md Rifat, et al.
Pubblicazione: (2024)
What Drives Success in Physical Planning with Joint-Embedding Predictive World Models?
di: Terver, Basile, et al.
Pubblicazione: (2025)
di: Terver, Basile, et al.
Pubblicazione: (2025)
Language Games as the Pathway to Artificial Superhuman Intelligence
di: Wen, Ying, et al.
Pubblicazione: (2025)
di: Wen, Ying, et al.
Pubblicazione: (2025)
The Illusion of AI Expertise Under Uncertainty: Navigating Elusive Ground Truth via a Probabilistic Paradigm
di: Elangovan, Aparna, et al.
Pubblicazione: (2026)
di: Elangovan, Aparna, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Video Representation Learning with Joint-Embedding Predictive Architectures
di: Drozdov, Katrina, et al.
Pubblicazione: (2024) -
Soft Clustering Anchors for Self-Supervised Speech Representation Learning in Joint Embedding Prediction Architectures
di: Ioannides, Georgios, et al.
Pubblicazione: (2026) -
A superpersuasive autonomous policy debating system
di: Roush, Allen, et al.
Pubblicazione: (2025) -
From Tokens to Thoughts: How LLMs and Humans Trade Compression for Meaning
di: Shani, Chen, et al.
Pubblicazione: (2025) -
Variance-Covariance Regularization Improves Representation Learning
di: Zhu, Jiachen, et al.
Pubblicazione: (2023)