OSCAR: Operating System Control via State-Aware Reasoning and Re-Planning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Xiaoqiang, Liu, Bang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
R$^3$Mem: Bridging Memory Retention and Retrieval via Reversible Compression
von: Wang, Xiaoqiang, et al.
Veröffentlicht: (2025)
von: Wang, Xiaoqiang, et al.
Veröffentlicht: (2025)
ReCode: Unify Plan and Action for Universal Granularity Control
von: Yu, Zhaoyang, et al.
Veröffentlicht: (2025)
von: Yu, Zhaoyang, et al.
Veröffentlicht: (2025)
ReFeed: Multi-dimensional Summarization Refinement with Reflective Reasoning on Feedback
von: Yun, Taewon, et al.
Veröffentlicht: (2025)
von: Yun, Taewon, et al.
Veröffentlicht: (2025)
AdaCtrl: Towards Adaptive and Controllable Reasoning via Difficulty-Aware Budgeting
von: Huang, Shijue, et al.
Veröffentlicht: (2025)
von: Huang, Shijue, et al.
Veröffentlicht: (2025)
IMAGINE: Integrating Multi-Agent System into One Model for Complex Reasoning and Planning
von: Zhang, Xikai, et al.
Veröffentlicht: (2025)
von: Zhang, Xikai, et al.
Veröffentlicht: (2025)
Improving Context Fidelity via Native Retrieval-Augmented Reasoning
von: Wang, Suyuchen, et al.
Veröffentlicht: (2025)
von: Wang, Suyuchen, et al.
Veröffentlicht: (2025)
OSCAR: Orchestrated Self-verification and Cross-path Refinement
von: Shah, Yash, et al.
Veröffentlicht: (2026)
von: Shah, Yash, et al.
Veröffentlicht: (2026)
Mem-$π$: Adaptive Memory through Learning When and What to Generate
von: Wang, Xiaoqiang, et al.
Veröffentlicht: (2026)
von: Wang, Xiaoqiang, et al.
Veröffentlicht: (2026)
ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning
von: Chen, Mingyang, et al.
Veröffentlicht: (2025)
von: Chen, Mingyang, et al.
Veröffentlicht: (2025)
Explainable Chain-of-Thought Reasoning: An Empirical Analysis on State-Aware Reasoning Dynamics
von: Yu, Sheldon, et al.
Veröffentlicht: (2025)
von: Yu, Sheldon, et al.
Veröffentlicht: (2025)
Latent Chain-of-Thought as Planning: Decoupling Reasoning from Verbalization
von: Wang, Jiecong, et al.
Veröffentlicht: (2026)
von: Wang, Jiecong, et al.
Veröffentlicht: (2026)
OSCAR: Online Soft Compression And Reranking
von: Louis, Maxime, et al.
Veröffentlicht: (2025)
von: Louis, Maxime, et al.
Veröffentlicht: (2025)
Controllable LLM Reasoning via Sparse Autoencoder-Based Steering
von: Fang, Yi, et al.
Veröffentlicht: (2026)
von: Fang, Yi, et al.
Veröffentlicht: (2026)
DiReCT: Diagnostic Reasoning for Clinical Notes via Large Language Models
von: Wang, Bowen, et al.
Veröffentlicht: (2024)
von: Wang, Bowen, et al.
Veröffentlicht: (2024)
ReFlect: An Effective Harness System for Complex Long-Horizon LLM Reasoning
von: Huang, Fan
Veröffentlicht: (2026)
von: Huang, Fan
Veröffentlicht: (2026)
GraCoRe: Benchmarking Graph Comprehension and Complex Reasoning in Large Language Models
von: Yuan, Zike, et al.
Veröffentlicht: (2024)
von: Yuan, Zike, et al.
Veröffentlicht: (2024)
ReCrit: Transition-Aware Reinforcement Learning for Scientific Critic Reasoning
von: Xu, Wanghan, et al.
Veröffentlicht: (2026)
von: Xu, Wanghan, et al.
Veröffentlicht: (2026)
PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization
von: Cao, Zouying, et al.
Veröffentlicht: (2025)
von: Cao, Zouying, et al.
Veröffentlicht: (2025)
RAG+: Enhancing Retrieval-Augmented Generation with Application-Aware Reasoning
von: Wang, Yu, et al.
Veröffentlicht: (2025)
von: Wang, Yu, et al.
Veröffentlicht: (2025)
Stop Before You Fail: Operational Capability Boundaries for Mitigating Unproductive Reasoning in Large Reasoning Models
von: Zhang, Qingjie, et al.
Veröffentlicht: (2025)
von: Zhang, Qingjie, et al.
Veröffentlicht: (2025)
ReIn: Conversational Error Recovery with Reasoning Inception
von: Kim, Takyoung, et al.
Veröffentlicht: (2026)
von: Kim, Takyoung, et al.
Veröffentlicht: (2026)
ReasonOps: Operator Segmentation for LLM Reasoning Traces
von: Lee, Daniel, et al.
Veröffentlicht: (2026)
von: Lee, Daniel, et al.
Veröffentlicht: (2026)
ReCo: Reliable Causal Chain Reasoning via Structural Causal Recurrent Neural Networks
von: Xiong, Kai, et al.
Veröffentlicht: (2022)
von: Xiong, Kai, et al.
Veröffentlicht: (2022)
Reinforced Context Order Recovery for Adaptive Reasoning and Planning
von: Ma, Long, et al.
Veröffentlicht: (2025)
von: Ma, Long, et al.
Veröffentlicht: (2025)
DS@GT at CheckThat! 2025: Detecting Subjectivity via Transfer-Learning and Corrective Data Augmentation
von: Heil, Maximilian, et al.
Veröffentlicht: (2025)
von: Heil, Maximilian, et al.
Veröffentlicht: (2025)
MSCoRe: A Benchmark for Multi-Stage Collaborative Reasoning in LLM Agents
von: Lei, Yuzhen, et al.
Veröffentlicht: (2025)
von: Lei, Yuzhen, et al.
Veröffentlicht: (2025)
ReProbe: Efficient Test-Time Scaling of Multi-Step Reasoning by Probing Internal States of Large Language Models
von: Ni, Jingwei, et al.
Veröffentlicht: (2025)
von: Ni, Jingwei, et al.
Veröffentlicht: (2025)
Mitigating Shortcut Reasoning in Language Models: A Gradient-Aware Training Approach
von: Cao, Hongyu, et al.
Veröffentlicht: (2026)
von: Cao, Hongyu, et al.
Veröffentlicht: (2026)
Outdated Issue Aware Decoding for Reasoning Questions on Edited Knowledge
von: Sun, Zengkui, et al.
Veröffentlicht: (2024)
von: Sun, Zengkui, et al.
Veröffentlicht: (2024)
How Reliable are LLMs for Reasoning on the Re-ranking task?
von: Islam, Nafis Tanveer, et al.
Veröffentlicht: (2025)
von: Islam, Nafis Tanveer, et al.
Veröffentlicht: (2025)
Plan Then Action:High-Level Planning Guidance Reinforcement Learning for LLM Reasoning
von: Dou, Zhihao, et al.
Veröffentlicht: (2025)
von: Dou, Zhihao, et al.
Veröffentlicht: (2025)
AstroReason-Bench: Evaluating Unified Agentic Planning across Heterogeneous Space Planning Problems
von: Wang, Weiyi, et al.
Veröffentlicht: (2026)
von: Wang, Weiyi, et al.
Veröffentlicht: (2026)
OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning
von: Yu, Fei, et al.
Veröffentlicht: (2023)
von: Yu, Fei, et al.
Veröffentlicht: (2023)
PlanGEN: A Multi-Agent Framework for Generating Planning and Reasoning Trajectories for Complex Problem Solving
von: Parmar, Mihir, et al.
Veröffentlicht: (2025)
von: Parmar, Mihir, et al.
Veröffentlicht: (2025)
ReAttn: Improving Attention-based Re-ranking via Attention Re-weighting
von: Tian, Yuxing, et al.
Veröffentlicht: (2026)
von: Tian, Yuxing, et al.
Veröffentlicht: (2026)
OraPlan-SQL: A Planning-Centric Framework for Complex Bilingual NL2SQL Reasoning
von: Liu, Marianne Menglin, et al.
Veröffentlicht: (2025)
von: Liu, Marianne Menglin, et al.
Veröffentlicht: (2025)
Harder Is Better: Boosting Mathematical Reasoning via Difficulty-Aware GRPO and Multi-Aspect Question Reformulation
von: Dai, Yanqi, et al.
Veröffentlicht: (2026)
von: Dai, Yanqi, et al.
Veröffentlicht: (2026)
Learning to Plan & Reason for Evaluation with Thinking-LLM-as-a-Judge
von: Saha, Swarnadeep, et al.
Veröffentlicht: (2025)
von: Saha, Swarnadeep, et al.
Veröffentlicht: (2025)
BoostTaxo: Zero-Shot Taxonomy Induction via Boosting-Style Agentic Reasoning and Constraint-Aware Calibration
von: Ling, Yancheng, et al.
Veröffentlicht: (2026)
von: Ling, Yancheng, et al.
Veröffentlicht: (2026)
Matrix as Plan: Structured Logical Reasoning with Feedback-Driven Replanning
von: Chen, Ke, et al.
Veröffentlicht: (2026)
von: Chen, Ke, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
R$^3$Mem: Bridging Memory Retention and Retrieval via Reversible Compression
von: Wang, Xiaoqiang, et al.
Veröffentlicht: (2025) -
ReCode: Unify Plan and Action for Universal Granularity Control
von: Yu, Zhaoyang, et al.
Veröffentlicht: (2025) -
ReFeed: Multi-dimensional Summarization Refinement with Reflective Reasoning on Feedback
von: Yun, Taewon, et al.
Veröffentlicht: (2025) -
AdaCtrl: Towards Adaptive and Controllable Reasoning via Difficulty-Aware Budgeting
von: Huang, Shijue, et al.
Veröffentlicht: (2025) -
IMAGINE: Integrating Multi-Agent System into One Model for Complex Reasoning and Planning
von: Zhang, Xikai, et al.
Veröffentlicht: (2025)