Learning to (Learn at Test Time): RNNs with Expressive Hidden States
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Yu, Li, Xinhao, Dalal, Karan, Xu, Jiarui, Vikram, Arjun, Zhang, Genghan, Dubois, Yann, Chen, Xinlei, Wang, Xiaolong, Koyejo, Sanmi, Hashimoto, Tatsunori, Guestrin, Carlos |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning to (Learn at Test Time)
von: Sun, Yu, et al.
Veröffentlicht: (2023)
von: Sun, Yu, et al.
Veröffentlicht: (2023)
End-to-End Test-Time Training for Long Context
von: Tandon, Arnuv, et al.
Veröffentlicht: (2025)
von: Tandon, Arnuv, et al.
Veröffentlicht: (2025)
One-Minute Video Generation with Test-Time Training
von: Dalal, Karan, et al.
Veröffentlicht: (2025)
von: Dalal, Karan, et al.
Veröffentlicht: (2025)
Discovering Implicit Large Language Model Alignment Objectives
von: Chen, Edward, et al.
Veröffentlicht: (2026)
von: Chen, Edward, et al.
Veröffentlicht: (2026)
Evaluating Self-Supervised Learning via Risk Decomposition
von: Dubois, Yann, et al.
Veröffentlicht: (2023)
von: Dubois, Yann, et al.
Veröffentlicht: (2023)
Benchmarking Distributional Alignment of Large Language Models
von: Meister, Nicole, et al.
Veröffentlicht: (2024)
von: Meister, Nicole, et al.
Veröffentlicht: (2024)
The Extractive-Abstractive Spectrum: Uncovering Verifiability Trade-offs in LLM Generations
von: Worledge, Theodora, et al.
Veröffentlicht: (2024)
von: Worledge, Theodora, et al.
Veröffentlicht: (2024)
Interactive Multi-Objective Probabilistic Preference Learning with Soft and Hard Bounds
von: Chen, Edward, et al.
Veröffentlicht: (2025)
von: Chen, Edward, et al.
Veröffentlicht: (2025)
AlpacaFarm: A Simulation Framework for Methods that Learn from Human Feedback
von: Dubois, Yann, et al.
Veröffentlicht: (2023)
von: Dubois, Yann, et al.
Veröffentlicht: (2023)
Learning to Discover at Test Time
von: Yuksekgonul, Mert, et al.
Veröffentlicht: (2026)
von: Yuksekgonul, Mert, et al.
Veröffentlicht: (2026)
Length-Controlled AlpacaEval: A Simple Way to Debias Automatic Evaluators
von: Dubois, Yann, et al.
Veröffentlicht: (2024)
von: Dubois, Yann, et al.
Veröffentlicht: (2024)
In-Context Learning of Energy Functions
von: Schaeffer, Rylan, et al.
Veröffentlicht: (2024)
von: Schaeffer, Rylan, et al.
Veröffentlicht: (2024)
On the Role of Depth in the Expressivity of RNNs
von: Lizaire, Maude, et al.
Veröffentlicht: (2026)
von: Lizaire, Maude, et al.
Veröffentlicht: (2026)
ALMo: Interactive Aim-Limit-Defined, Multi-Objective System for Personalized High-Dose-Rate Brachytherapy Treatment Planning and Visualization for Cervical Cancer
von: Chen, Edward, et al.
Veröffentlicht: (2026)
von: Chen, Edward, et al.
Veröffentlicht: (2026)
Causally Inspired Regularization Enables Domain General Representations
von: Salaudeen, Olawale, et al.
Veröffentlicht: (2024)
von: Salaudeen, Olawale, et al.
Veröffentlicht: (2024)
Let's Measure Information Step-by-Step: AI-Based Evaluation Beyond Vibes
von: Robertson, Zachary, et al.
Veröffentlicht: (2025)
von: Robertson, Zachary, et al.
Veröffentlicht: (2025)
CURE: Cultural Understanding and Reasoning Evaluation - A Framework for "Thick" Culture Alignment Evaluation in LLMs
von: Vo, Truong, et al.
Veröffentlicht: (2025)
von: Vo, Truong, et al.
Veröffentlicht: (2025)
A Framework for Objective-Driven Dynamical Stochastic Fields
von: Zhang, Yibo Jacky, et al.
Veröffentlicht: (2025)
von: Zhang, Yibo Jacky, et al.
Veröffentlicht: (2025)
Modeling Multi-Objective Tradeoffs with Monotonic Utility Functions
von: Chen, Edward, et al.
Veröffentlicht: (2024)
von: Chen, Edward, et al.
Veröffentlicht: (2024)
Implicit Language Models are RNNs: Balancing Parallelization and Expressivity
von: Schöne, Mark, et al.
Veröffentlicht: (2025)
von: Schöne, Mark, et al.
Veröffentlicht: (2025)
High-Dimensional Markov-switching Ordinary Differential Processes
von: Tsai, Katherine, et al.
Veröffentlicht: (2024)
von: Tsai, Katherine, et al.
Veröffentlicht: (2024)
Distributional Machine Unlearning via Selective Data Removal
von: Allouah, Youssef, et al.
Veröffentlicht: (2025)
von: Allouah, Youssef, et al.
Veröffentlicht: (2025)
SCENEBench: An Audio Understanding Benchmark Grounded in Assistive and Industrial Use Cases
von: Iyer, Laya, et al.
Veröffentlicht: (2026)
von: Iyer, Laya, et al.
Veröffentlicht: (2026)
HiFA: High-fidelity Text-to-3D Generation with Advanced Diffusion Guidance
von: Zhu, Junzhe, et al.
Veröffentlicht: (2023)
von: Zhu, Junzhe, et al.
Veröffentlicht: (2023)
Is Pre-training Truly Better Than Meta-Learning?
von: Miranda, Brando, et al.
Veröffentlicht: (2023)
von: Miranda, Brando, et al.
Veröffentlicht: (2023)
Language Models with Conformal Factuality Guarantees
von: Mohri, Christopher, et al.
Veröffentlicht: (2024)
von: Mohri, Christopher, et al.
Veröffentlicht: (2024)
Model Equality Testing: Which Model Is This API Serving?
von: Gao, Irena, et al.
Veröffentlicht: (2024)
von: Gao, Irena, et al.
Veröffentlicht: (2024)
Scaling Laws for the Value of Individual Data Points in Machine Learning
von: Covert, Ian, et al.
Veröffentlicht: (2024)
von: Covert, Ian, et al.
Veröffentlicht: (2024)
Reasoning Models Don't Just Think Longer, They Move Differently
von: Gjølbye, Anders, et al.
Veröffentlicht: (2026)
von: Gjølbye, Anders, et al.
Veröffentlicht: (2026)
Is Backpropagation Optimal? When Synthetic Gradients Improve Sample Efficiency
von: Zhang, Yibo Jacky, et al.
Veröffentlicht: (2026)
von: Zhang, Yibo Jacky, et al.
Veröffentlicht: (2026)
The Inadequacy of Offline LLM Evaluations: A Need to Account for Personalization in Model Behavior
von: Wang, Angelina, et al.
Veröffentlicht: (2025)
von: Wang, Angelina, et al.
Veröffentlicht: (2025)
Principled Federated Domain Adaptation: Gradient Projection and Auto-Weighting
von: Jiang, Enyi, et al.
Veröffentlicht: (2023)
von: Jiang, Enyi, et al.
Veröffentlicht: (2023)
Steering Away from Memorization: Reachability-Constrained Reinforcement Learning for Text-to-Image Diffusion
von: Karnik, Sathwik, et al.
Veröffentlicht: (2026)
von: Karnik, Sathwik, et al.
Veröffentlicht: (2026)
Reasoning to Learn from Latent Thoughts
von: Ruan, Yangjun, et al.
Veröffentlicht: (2025)
von: Ruan, Yangjun, et al.
Veröffentlicht: (2025)
Adaptive Compression in Federated Learning via Side Information
von: Isik, Berivan, et al.
Veröffentlicht: (2023)
von: Isik, Berivan, et al.
Veröffentlicht: (2023)
Mechanistic Interpretability of RNNs emulating Hidden Markov Models
von: Torre, Elia, et al.
Veröffentlicht: (2025)
von: Torre, Elia, et al.
Veröffentlicht: (2025)
Learning State-Tracking from Code Using Linear RNNs
von: Siems, Julien, et al.
Veröffentlicht: (2026)
von: Siems, Julien, et al.
Veröffentlicht: (2026)
Test-Time Training on Video Streams
von: Wang, Renhao, et al.
Veröffentlicht: (2023)
von: Wang, Renhao, et al.
Veröffentlicht: (2023)
Frontières et mobilité au quotidien
von: Dubois, Yann
Veröffentlicht: (2020)
von: Dubois, Yann
Veröffentlicht: (2020)
Pantograph: A Machine-to-Machine Interaction Interface for Advanced Theorem Proving, High Level Reasoning, and Data Extraction in Lean 4
von: Aniva, Leni, et al.
Veröffentlicht: (2024)
von: Aniva, Leni, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Learning to (Learn at Test Time)
von: Sun, Yu, et al.
Veröffentlicht: (2023) -
End-to-End Test-Time Training for Long Context
von: Tandon, Arnuv, et al.
Veröffentlicht: (2025) -
One-Minute Video Generation with Test-Time Training
von: Dalal, Karan, et al.
Veröffentlicht: (2025) -
Discovering Implicit Large Language Model Alignment Objectives
von: Chen, Edward, et al.
Veröffentlicht: (2026) -
Evaluating Self-Supervised Learning via Risk Decomposition
von: Dubois, Yann, et al.
Veröffentlicht: (2023)