Reinforced Fast Weights with Next-Sequence Prediction
Fuente:
arXiv
Saved in:
| Main Authors: | Hwang, Hee Seung, Wu, Xindi, Chun, Sanghyuk, Russakovsky, Olga |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Visual Compositional Tuning
by: Wu, Xindi, et al.
Published: (2025)
by: Wu, Xindi, et al.
Published: (2025)
ICONS: Influence Consensus for Vision-Language Data Selection
by: Wu, Xindi, et al.
Published: (2024)
by: Wu, Xindi, et al.
Published: (2024)
Vision-Language Dataset Distillation
by: Wu, Xindi, et al.
Published: (2023)
by: Wu, Xindi, et al.
Published: (2023)
Toward Interactive Regional Understanding in Vision-Large Language Models
by: Lee, Jungbeom, et al.
Published: (2024)
by: Lee, Jungbeom, et al.
Published: (2024)
Beyond Objects: Contextual Synthetic Data Generation for Fine-Grained Classification
by: Yang, William, et al.
Published: (2025)
by: Yang, William, et al.
Published: (2025)
ConceptMix: A Compositional Image Generation Benchmark with Controllable Difficulty
by: Wu, Xindi, et al.
Published: (2024)
by: Wu, Xindi, et al.
Published: (2024)
BOW: Reinforcement Learning for Bottlenecked Next Word Prediction
by: Shen, Ming, et al.
Published: (2025)
by: Shen, Ming, et al.
Published: (2025)
Analyzing the Roles of Language and Vision in Learning from Limited Data
by: Chen, Allison, et al.
Published: (2024)
by: Chen, Allison, et al.
Published: (2024)
CREFT: Sequential Multi-Agent LLM for Character Relation Extraction
by: Chun, Ye Eun, et al.
Published: (2025)
by: Chun, Ye Eun, et al.
Published: (2025)
On Support Samples of Next Word Prediction
by: Li, Yuqian, et al.
Published: (2025)
by: Li, Yuqian, et al.
Published: (2025)
Dual-Scale World Models for LLM Agents Towards Hard-Exploration Problems
by: Kim, Minsoo, et al.
Published: (2025)
by: Kim, Minsoo, et al.
Published: (2025)
From Weights to Activations: Is Steering the Next Frontier of Adaptation?
by: Ostermann, Simon, et al.
Published: (2026)
by: Ostermann, Simon, et al.
Published: (2026)
Diversity or Precision? A Deep Dive into Next Token Prediction
by: Wu, Haoyuan, et al.
Published: (2025)
by: Wu, Haoyuan, et al.
Published: (2025)
Lossless Compression of Large Language Model-Generated Text via Next-Token Prediction
by: Mao, Yu, et al.
Published: (2025)
by: Mao, Yu, et al.
Published: (2025)
Counterfactual-Consistency Prompting for Relative Temporal Understanding in Large Language Models
by: Kim, Jongho, et al.
Published: (2025)
by: Kim, Jongho, et al.
Published: (2025)
CoEx -- Co-evolving World-model and Exploration
by: Kim, Minsoo, et al.
Published: (2025)
by: Kim, Minsoo, et al.
Published: (2025)
Improved Probabilistic Image-Text Representations
by: Chun, Sanghyuk
Published: (2023)
by: Chun, Sanghyuk
Published: (2023)
Multiplicity is an Inevitable and Inherent Challenge in Multimodal Learning
by: Chun, Sanghyuk
Published: (2025)
by: Chun, Sanghyuk
Published: (2025)
BIBERT-Pipe on Biomedical Nested Named Entity Linking at BioASQ 2025
by: Li, Chunyu, et al.
Published: (2025)
by: Li, Chunyu, et al.
Published: (2025)
Intended Target Identification for Anomia Patients with Gradient-based Selective Augmentation
by: Kim, Jongho, et al.
Published: (2025)
by: Kim, Jongho, et al.
Published: (2025)
Disentangling Questions from Query Generation for Task-Adaptive Retrieval
by: Lee, Yoonsang, et al.
Published: (2024)
by: Lee, Yoonsang, et al.
Published: (2024)
Benchmarking Testing in Automated Theorem Proving
by: Kim, Jongyoon, et al.
Published: (2026)
by: Kim, Jongyoon, et al.
Published: (2026)
Unifying Specialized Visual Encoders for Video Language Models
by: Chung, Jihoon, et al.
Published: (2025)
by: Chung, Jihoon, et al.
Published: (2025)
Can You Learn Semantics Through Next-Word Prediction? The Case of Entailment
by: Merrill, William, et al.
Published: (2024)
by: Merrill, William, et al.
Published: (2024)
DuET: Dual Execution for Test Output Prediction with Generated Code and Pseudocode
by: Han, Hojae, et al.
Published: (2026)
by: Han, Hojae, et al.
Published: (2026)
Auxiliary Knowledge-Induced Learning for Automatic Multi-Label Medical Document Classification
by: Wang, Xindi, et al.
Published: (2024)
by: Wang, Xindi, et al.
Published: (2024)
Chain of Grounded Objectives: Bridging Process and Goal-oriented Prompting for Code Generation
by: Yeo, Sangyeop, et al.
Published: (2025)
by: Yeo, Sangyeop, et al.
Published: (2025)
OffsetBias: Leveraging Debiased Data for Tuning Evaluators
by: Park, Junsoo, et al.
Published: (2024)
by: Park, Junsoo, et al.
Published: (2024)
Training LLMs Beyond Next Token Prediction -- Filling the Mutual Information Gap
by: Yang, Chun-Hao, et al.
Published: (2025)
by: Yang, Chun-Hao, et al.
Published: (2025)
Simulating Weighted Automata over Sequences and Trees with Transformers
by: Rizvi, Michael, et al.
Published: (2024)
by: Rizvi, Michael, et al.
Published: (2024)
HARP: Hesitation-Aware Reframing in Transformer Inference Pass
by: Storaï, Romain, et al.
Published: (2024)
by: Storaï, Romain, et al.
Published: (2024)
Can David Beat Goliath? On Multi-Hop Reasoning with Resource-Constrained Agents
by: Han, Hojae, et al.
Published: (2026)
by: Han, Hojae, et al.
Published: (2026)
Chaining Event Spans for Temporal Relation Grounding
by: Kim, Jongho, et al.
Published: (2025)
by: Kim, Jongho, et al.
Published: (2025)
SAFE: Stepwise Atomic Feedback for Error correction in Multi-hop Reasoning
by: Kwon, Daeyong, et al.
Published: (2026)
by: Kwon, Daeyong, et al.
Published: (2026)
NITP: Next Implicit Token Prediction for LLM Pre-training
by: Zhang, Xiangdong, et al.
Published: (2026)
by: Zhang, Xiangdong, et al.
Published: (2026)
ARES: Alternating Reinforcement Learning and Supervised Fine-Tuning for Enhanced Multi-Modal Chain-of-Thought Reasoning Through Diverse AI Feedback
by: Byun, Ju-Seung, et al.
Published: (2024)
by: Byun, Ju-Seung, et al.
Published: (2024)
Textless Dependency Parsing by Labeled Sequence Prediction
by: Kando, Shunsuke, et al.
Published: (2024)
by: Kando, Shunsuke, et al.
Published: (2024)
Agent-as-Judge for Factual Summarization of Long Narratives
by: Jeong, Yeonseok, et al.
Published: (2025)
by: Jeong, Yeonseok, et al.
Published: (2025)
ConfClip: Confidence-Weighted and Clipped Reward for Reinforcement Learning in LLMs
by: Zhang, Bonan, et al.
Published: (2025)
by: Zhang, Bonan, et al.
Published: (2025)
Making Qwen3 Think in Korean with Reinforcement Learning
by: Lee, Jungyup, et al.
Published: (2025)
by: Lee, Jungyup, et al.
Published: (2025)
Similar Items
-
Visual Compositional Tuning
by: Wu, Xindi, et al.
Published: (2025) -
ICONS: Influence Consensus for Vision-Language Data Selection
by: Wu, Xindi, et al.
Published: (2024) -
Vision-Language Dataset Distillation
by: Wu, Xindi, et al.
Published: (2023) -
Toward Interactive Regional Understanding in Vision-Large Language Models
by: Lee, Jungbeom, et al.
Published: (2024) -
Beyond Objects: Contextual Synthetic Data Generation for Fine-Grained Classification
by: Yang, William, et al.
Published: (2025)