Guardado en:
| Autores principales: | Cao, Xianwei, Quan, Dou, Zhang, Zhenliang, Wang, Shuang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2603.22813 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
CLNet: Cross-View Correspondence Makes a Stronger Geo-Localizationer
por: Cao, Xianwei, et al.
Publicado: (2025)
por: Cao, Xianwei, et al.
Publicado: (2025)
Dynamics-Aligned Shared Hypernetworks for Contextual RL under Discontinuous Shifts
por: Benad, Jan, et al.
Publicado: (2026)
por: Benad, Jan, et al.
Publicado: (2026)
Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain
por: Panagopoulos, Dimitris, et al.
Publicado: (2025)
por: Panagopoulos, Dimitris, et al.
Publicado: (2025)
Exploring Human-Machine Coexistence in Symmetrical Reality
por: Zhang, Zhenliang
Publicado: (2026)
por: Zhang, Zhenliang
Publicado: (2026)
Time-Scaling Is What Agents Need Now
por: Liu, Zhi, et al.
Publicado: (2026)
por: Liu, Zhi, et al.
Publicado: (2026)
Spectral Invariant Learning for Dynamic Graphs under Distribution Shifts
por: Zhang, Zeyang, et al.
Publicado: (2024)
por: Zhang, Zeyang, et al.
Publicado: (2024)
Contrastive Learning of Preferences with a Contextual InfoNCE Loss
por: Bertram, Timo, et al.
Publicado: (2024)
por: Bertram, Timo, et al.
Publicado: (2024)
Understand What LLM Needs: Dual Preference Alignment for Retrieval-Augmented Generation
por: Dong, Guanting, et al.
Publicado: (2024)
por: Dong, Guanting, et al.
Publicado: (2024)
Contextual Position Encoding: Learning to Count What's Important
por: Golovneva, Olga, et al.
Publicado: (2024)
por: Golovneva, Olga, et al.
Publicado: (2024)
Reinforcement Learning-based Sequential Route Recommendation for System-Optimal Traffic Assignment
por: Wang, Leizhen, et al.
Publicado: (2025)
por: Wang, Leizhen, et al.
Publicado: (2025)
Silencing the Guardrails: Inference-Time Jailbreaking via Dynamic Contextual Representation Ablation
por: Xing, Wenpeng, et al.
Publicado: (2026)
por: Xing, Wenpeng, et al.
Publicado: (2026)
Human-in-the-Loop Multi-Agent Ventilator Decision Support with Contextual Bandit Preference Learning
por: Li, Sijia, et al.
Publicado: (2026)
por: Li, Sijia, et al.
Publicado: (2026)
Where and What: Reasoning Dynamic and Implicit Preferences in Situated Conversational Recommendation
por: Lin, Dongding, et al.
Publicado: (2026)
por: Lin, Dongding, et al.
Publicado: (2026)
Toward Effective Tool-Integrated Reasoning via Self-Evolved Preference Learning
por: Chen, Yifei, et al.
Publicado: (2025)
por: Chen, Yifei, et al.
Publicado: (2025)
ICR Probe: Tracking Hidden State Dynamics for Reliable Hallucination Detection in LLMs
por: Zhang, Zhenliang, et al.
Publicado: (2025)
por: Zhang, Zhenliang, et al.
Publicado: (2025)
ATIR: Towards Audio-Text Interleaved Contextual Retrieval
por: Zhao, Tong, et al.
Publicado: (2026)
por: Zhao, Tong, et al.
Publicado: (2026)
From Abstract to Contextual: What LLMs Still Cannot Do in Mathematics
por: Cao, Bowen, et al.
Publicado: (2026)
por: Cao, Bowen, et al.
Publicado: (2026)
Lightweight Adapter Learning for More Generalized Remote Sensing Change Detection
por: Quan, Dou, et al.
Publicado: (2025)
por: Quan, Dou, et al.
Publicado: (2025)
TeachAnything: A Multimodal Crowdsourcing Platform for Training Embodied AI Agents in Symmetrical Reality
por: Liu, Zidong, et al.
Publicado: (2026)
por: Liu, Zidong, et al.
Publicado: (2026)
Attention Basin: Why Contextual Position Matters in Large Language Models
por: Yi, Zihao, et al.
Publicado: (2025)
por: Yi, Zihao, et al.
Publicado: (2025)
Adaptive Shielding for Safe Reinforcement Learning under Hidden-Parameter Dynamics Shifts
por: Kwon, Minjae, et al.
Publicado: (2025)
por: Kwon, Minjae, et al.
Publicado: (2025)
Graphs Generalization under Distribution Shifts
por: Tian, Qin, et al.
Publicado: (2024)
por: Tian, Qin, et al.
Publicado: (2024)
Focus on What Matters: Fisher-Guided Adaptive Multimodal Fusion for Vulnerability Detection
por: Bian, Yun, et al.
Publicado: (2026)
por: Bian, Yun, et al.
Publicado: (2026)
Small-Margin Preferences Still Matter-If You Train Them Right
por: Pang, Jinlong, et al.
Publicado: (2026)
por: Pang, Jinlong, et al.
Publicado: (2026)
Via Negativa for AI Alignment: Why Negative Constraints Are Structurally Superior to Positive Preferences
por: Cheng, Quan
Publicado: (2026)
por: Cheng, Quan
Publicado: (2026)
$\boldsymbol{f}$-OPD: Stabilizing Long-Horizon On-Policy Distillation with Freshness-Aware Control
por: Chen, Xianwei, et al.
Publicado: (2026)
por: Chen, Xianwei, et al.
Publicado: (2026)
OMPO: A Unified Framework for RL under Policy and Dynamics Shifts
por: Luo, Yu, et al.
Publicado: (2024)
por: Luo, Yu, et al.
Publicado: (2024)
Statistical Inference for Misspecified Contextual Bandits
por: Guo, Yongyi, et al.
Publicado: (2025)
por: Guo, Yongyi, et al.
Publicado: (2025)
DynamicPO: Dynamic Preference Optimization for Recommendation
por: Hu, Xingyu, et al.
Publicado: (2026)
por: Hu, Xingyu, et al.
Publicado: (2026)
Rule Learning for Knowledge Graph Reasoning under Agnostic Distribution Shift
por: Liu, Shixuan, et al.
Publicado: (2025)
por: Liu, Shixuan, et al.
Publicado: (2025)
Now It Sounds Like You: Learning Personalized Vocabulary On Device
por: Wang, Sid, et al.
Publicado: (2023)
por: Wang, Sid, et al.
Publicado: (2023)
An Enhanced Federated Prototype Learning Method under Domain Shift
por: Kuang, Liang, et al.
Publicado: (2024)
por: Kuang, Liang, et al.
Publicado: (2024)
What Matters in Data for DPO?
por: Pan, Yu, et al.
Publicado: (2025)
por: Pan, Yu, et al.
Publicado: (2025)
Rewarding What Matters: Step-by-Step Reinforcement Learning for Task-Oriented Dialogue
por: Du, Huifang, et al.
Publicado: (2024)
por: Du, Huifang, et al.
Publicado: (2024)
What Matters for Batch Online Reinforcement Learning in Robotics?
por: Dong, Perry, et al.
Publicado: (2025)
por: Dong, Perry, et al.
Publicado: (2025)
Domain-Contextualized Inference: A Computable Graph Architecture for Explicit-Domain Reasoning
por: Li, Chao, et al.
Publicado: (2026)
por: Li, Chao, et al.
Publicado: (2026)
Multi-Objective Planning with Contextual Lexicographic Reward Preferences
por: Rustagi, Pulkit, et al.
Publicado: (2025)
por: Rustagi, Pulkit, et al.
Publicado: (2025)
Graph Fairness Learning under Distribution Shifts
por: Li, Yibo, et al.
Publicado: (2024)
por: Li, Yibo, et al.
Publicado: (2024)
Preference Consistency Matters: Enhancing Preference Learning in Language Models with Automated Self-Curation of Training Corpora
por: Lee, JoonHo, et al.
Publicado: (2024)
por: Lee, JoonHo, et al.
Publicado: (2024)
Contextual Preference Collaborative Measure Framework Based on Belief System
por: Yu, Hang, et al.
Publicado: (2025)
por: Yu, Hang, et al.
Publicado: (2025)
Ejemplares similares
-
CLNet: Cross-View Correspondence Makes a Stronger Geo-Localizationer
por: Cao, Xianwei, et al.
Publicado: (2025) -
Dynamics-Aligned Shared Hypernetworks for Contextual RL under Discontinuous Shifts
por: Benad, Jan, et al.
Publicado: (2026) -
Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain
por: Panagopoulos, Dimitris, et al.
Publicado: (2025) -
Exploring Human-Machine Coexistence in Symmetrical Reality
por: Zhang, Zhenliang
Publicado: (2026) -
Time-Scaling Is What Agents Need Now
por: Liu, Zhi, et al.
Publicado: (2026)