Bifrost: Steering Strategic Trajectories to Bridge Contextual Gaps for Self-Improving Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Tran, Quan M., Huang, Zhuo, Zhang, Wenbin, Han, Bo, Yatani, Koji, Sugiyama, Masashi, Liu, Tongliang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BrokenBind: Universal Modality Exploration beyond Dataset Boundaries
by: Huang, Zhuo, et al.
Published: (2026)
by: Huang, Zhuo, et al.
Published: (2026)
Towards Effective Evaluations and Comparisons for LLM Unlearning Methods
by: Wang, Qizhou, et al.
Published: (2024)
by: Wang, Qizhou, et al.
Published: (2024)
BadLabel: A Robust Perspective on Evaluating and Enhancing Label-noise Learning
by: Zhang, Jingfeng, et al.
Published: (2023)
by: Zhang, Jingfeng, et al.
Published: (2023)
Towards Scalable Oversight with Collaborative Multi-Agent Debate in Error Detection
by: Chen, Yongqiang, et al.
Published: (2025)
by: Chen, Yongqiang, et al.
Published: (2025)
On the Thinking-Language Modeling Gap in Large Language Models
by: Liu, Chenxi, et al.
Published: (2025)
by: Liu, Chenxi, et al.
Published: (2025)
Non-stationary Online Learning for Curved Losses: Improved Dynamic Regret via Mixability
by: Zhang, Yu-Jie, et al.
Published: (2025)
by: Zhang, Yu-Jie, et al.
Published: (2025)
Reinforcement Learning with Verifiable yet Noisy Rewards under Imperfect Verifiers
by: Cai, Xin-Qiang, et al.
Published: (2025)
by: Cai, Xin-Qiang, et al.
Published: (2025)
Exploring Criteria of Loss Reweighting to Enhance LLM Unlearning
by: Yang, Puning, et al.
Published: (2025)
by: Yang, Puning, et al.
Published: (2025)
Mind the Gap Between Prototypes and Images in Cross-domain Finetuning
by: Tian, Hongduan, et al.
Published: (2024)
by: Tian, Hongduan, et al.
Published: (2024)
Is Gradient Ascent Really Necessary? Memorize to Forget for Machine Unlearning
by: Huang, Zhuo, et al.
Published: (2026)
by: Huang, Zhuo, et al.
Published: (2026)
MeGU: Machine-Guided Unlearning with Target Feature Disentanglement
by: Wang, Haoyu, et al.
Published: (2026)
by: Wang, Haoyu, et al.
Published: (2026)
Enriching Disentanglement: From Logical Definitions to Quantitative Metrics
by: Zhang, Yivan, et al.
Published: (2023)
by: Zhang, Yivan, et al.
Published: (2023)
A Category-theoretical Meta-analysis of Definitions of Disentanglement
by: Zhang, Yivan, et al.
Published: (2023)
by: Zhang, Yivan, et al.
Published: (2023)
A Fast Algorithm for the Real-Valued Combinatorial Pure Exploration of Multi-Armed Bandit
by: Nakamura, Shintaro, et al.
Published: (2023)
by: Nakamura, Shintaro, et al.
Published: (2023)
What Is Preference Optimization Doing, and Why?
by: Wang, Yue, et al.
Published: (2025)
by: Wang, Yue, et al.
Published: (2025)
VEC-SBM: Optimal Community Detection with Vectorial Edges Covariates
by: Braun, Guillaume, et al.
Published: (2024)
by: Braun, Guillaume, et al.
Published: (2024)
Riemannian Langevin Dynamics: Strong Convergence of Geometric Euler-Maruyama Scheme
by: Zhan, Zhiyuan, et al.
Published: (2026)
by: Zhan, Zhiyuan, et al.
Published: (2026)
Research as Resistance: Recognizing and Reconsidering HCI's Role in Technology Hype Cycles
by: Sramek, Zefan, et al.
Published: (2025)
by: Sramek, Zefan, et al.
Published: (2025)
Multi-Player Approaches for Dueling Bandits
by: Raveh, Or, et al.
Published: (2024)
by: Raveh, Or, et al.
Published: (2024)
VI-CuRL: Stabilizing Verifier-Independent RL Reasoning via Confidence-Guided Variance Reduction
by: Cai, Xin-Qiang, et al.
Published: (2026)
by: Cai, Xin-Qiang, et al.
Published: (2026)
Decoupling the Class Label and the Target Concept in Machine Unlearning
by: Zhu, Jianing, et al.
Published: (2024)
by: Zhu, Jianing, et al.
Published: (2024)
Enhancing Sample Selection Against Label Noise by Cutting Mislabeled Easy Examples
by: Yuan, Suqin, et al.
Published: (2025)
by: Yuan, Suqin, et al.
Published: (2025)
On the Over-Memorization During Natural, Robust and Catastrophic Overfitting
by: Lin, Runqi, et al.
Published: (2023)
by: Lin, Runqi, et al.
Published: (2023)
Label Distribution Learning with Biased Annotations by Learning Multi-Label Representation
by: Kou, Zhiqiang, et al.
Published: (2025)
by: Kou, Zhiqiang, et al.
Published: (2025)
Parallel Simulation for Log-concave Sampling and Score-based Diffusion Models
by: Zhou, Huanjian, et al.
Published: (2024)
by: Zhou, Huanjian, et al.
Published: (2024)
COBRA: Contextual Bandit Algorithm for Ensuring Truthful Strategic Agents
by: Verma, Arun, et al.
Published: (2025)
by: Verma, Arun, et al.
Published: (2025)
Practical estimation of the optimal classification error with soft labels and calibration
by: Ushio, Ryota, et al.
Published: (2025)
by: Ushio, Ryota, et al.
Published: (2025)
The Survival Bandit Problem
by: Riou, Charles, et al.
Published: (2022)
by: Riou, Charles, et al.
Published: (2022)
Thompson Exploration with Best Challenger Rule in Best Arm Identification
by: Lee, Jongyeong, et al.
Published: (2023)
by: Lee, Jongyeong, et al.
Published: (2023)
Bridging the Gap between Chemical Reaction Pretraining and Conditional Molecule Generation with a Unified Model
by: Qiang, Bo, et al.
Published: (2023)
by: Qiang, Bo, et al.
Published: (2023)
Towards Understanding Valuable Preference Data for Large Language Model Alignment
by: Zhang, Zizhuo, et al.
Published: (2025)
by: Zhang, Zizhuo, et al.
Published: (2025)
Context-Enhanced Multi-View Trajectory Representation Learning: Bridging the Gap through Self-Supervised Models
by: Qian, Tangwen, et al.
Published: (2024)
by: Qian, Tangwen, et al.
Published: (2024)
Learnability Gaps of Strategic Classification
by: Cohen, Lee, et al.
Published: (2024)
by: Cohen, Lee, et al.
Published: (2024)
From Debate to Equilibrium: Belief-Driven Multi-Agent LLM Reasoning via Bayesian Nash Equilibrium
by: Yi, Xie, et al.
Published: (2025)
by: Yi, Xie, et al.
Published: (2025)
Reinforcement Learning with Options and State Representation
by: Ghriss, Ayoub, et al.
Published: (2024)
by: Ghriss, Ayoub, et al.
Published: (2024)
What If the Input is Expanded in OOD Detection?
by: Zhang, Boxuan, et al.
Published: (2024)
by: Zhang, Boxuan, et al.
Published: (2024)
Offline Reinforcement Learning from Datasets with Structured Non-Stationarity
by: Ackermann, Johannes, et al.
Published: (2024)
by: Ackermann, Johannes, et al.
Published: (2024)
FedImpro: Measuring and Improving Client Update in Federated Learning
by: Tang, Zhenheng, et al.
Published: (2024)
by: Tang, Zhenheng, et al.
Published: (2024)
Bifrost-1: Bridging Multimodal LLMs and Diffusion Models with Patch-level CLIP Latents
by: Lin, Han, et al.
Published: (2025)
by: Lin, Han, et al.
Published: (2025)
LLM Routing with Dueling Feedback
by: Chiang, Chao-Kai, et al.
Published: (2025)
by: Chiang, Chao-Kai, et al.
Published: (2025)
Similar Items
-
BrokenBind: Universal Modality Exploration beyond Dataset Boundaries
by: Huang, Zhuo, et al.
Published: (2026) -
Towards Effective Evaluations and Comparisons for LLM Unlearning Methods
by: Wang, Qizhou, et al.
Published: (2024) -
BadLabel: A Robust Perspective on Evaluating and Enhancing Label-noise Learning
by: Zhang, Jingfeng, et al.
Published: (2023) -
Towards Scalable Oversight with Collaborative Multi-Agent Debate in Error Detection
by: Chen, Yongqiang, et al.
Published: (2025) -
On the Thinking-Language Modeling Gap in Large Language Models
by: Liu, Chenxi, et al.
Published: (2025)