Taming Actor-Observer Asymmetry in Agents via Dialectical Alignment
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Bobo, Wu, Rui, Ji, Zibo, Zhang, Meishan, Fei, Hao, Zhang, Min, Lee, Mong-Li, Hsu, Wynne |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Dialect Normalization using Large Language Models and Morphological Rules
di: Dimakis, Antonios, et al.
Pubblicazione: (2025)
di: Dimakis, Antonios, et al.
Pubblicazione: (2025)
Dialect Matters: Cross-Lingual ASR Transfer for Low-Resource Indic Language Varieties
di: Dhasmana, Akriti, et al.
Pubblicazione: (2026)
di: Dhasmana, Akriti, et al.
Pubblicazione: (2026)
EnDive: A Cross-Dialect Benchmark for Fairness and Performance in Large Language Models
di: Gupta, Abhay, et al.
Pubblicazione: (2025)
di: Gupta, Abhay, et al.
Pubblicazione: (2025)
Strategy Adaptation in Large Language Model Werewolf Agents
di: Nakamori, Fuya, et al.
Pubblicazione: (2025)
di: Nakamori, Fuya, et al.
Pubblicazione: (2025)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
di: Peters, Sydney, et al.
Pubblicazione: (2025)
di: Peters, Sydney, et al.
Pubblicazione: (2025)
Homogeneous Keys, Heterogeneous Values: Exploiting Local KV Cache Asymmetry for Long-Context LLMs
di: Cui, Wanyun, et al.
Pubblicazione: (2025)
di: Cui, Wanyun, et al.
Pubblicazione: (2025)
Learning From Failure: Integrating Negative Examples when Fine-tuning Large Language Models as Agents
di: Wang, Renxi, et al.
Pubblicazione: (2024)
di: Wang, Renxi, et al.
Pubblicazione: (2024)
HACHIMI: Scalable and Controllable Student Persona Generation via Orchestrated Agents
di: Jiang, Yilin, et al.
Pubblicazione: (2026)
di: Jiang, Yilin, et al.
Pubblicazione: (2026)
An ethical study of generative AI from the Actor-Network Theory perspective
di: Li, Yuying, et al.
Pubblicazione: (2024)
di: Li, Yuying, et al.
Pubblicazione: (2024)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
di: Ashuach, Tomer, et al.
Pubblicazione: (2025)
di: Ashuach, Tomer, et al.
Pubblicazione: (2025)
Test-Time Scaling of Reasoning Models for Machine Translation
di: Li, Zihao, et al.
Pubblicazione: (2025)
di: Li, Zihao, et al.
Pubblicazione: (2025)
Cultural Benchmarking of LLMs in Standard and Dialectal Arabic Dialogues
di: Kautsar, Muhammad Dehan Al, et al.
Pubblicazione: (2026)
di: Kautsar, Muhammad Dehan Al, et al.
Pubblicazione: (2026)
MT-Ranker: Reference-free machine translation evaluation by inter-system ranking
di: Moosa, Ibraheem Muhammad, et al.
Pubblicazione: (2024)
di: Moosa, Ibraheem Muhammad, et al.
Pubblicazione: (2024)
Improving Retrospective Language Agents via Joint Policy Gradient Optimization
di: Feng, Xueyang, et al.
Pubblicazione: (2025)
di: Feng, Xueyang, et al.
Pubblicazione: (2025)
Unstructured Text Enhanced Open-domain Dialogue System: A Systematic Survey
di: Ma, Longxuan, et al.
Pubblicazione: (2024)
di: Ma, Longxuan, et al.
Pubblicazione: (2024)
Policy-driven Knowledge Selection and Response Generation for Document-grounded Dialogue
di: Ma, Longxuan, et al.
Pubblicazione: (2024)
di: Ma, Longxuan, et al.
Pubblicazione: (2024)
R-Genie: Reasoning-Guided Generative Image Editing
di: Zhang, Dong, et al.
Pubblicazione: (2025)
di: Zhang, Dong, et al.
Pubblicazione: (2025)
Towards Alignment-Centric Paradigm: A Survey of Instruction Tuning in Large Language Models
di: Han, Xudong, et al.
Pubblicazione: (2025)
di: Han, Xudong, et al.
Pubblicazione: (2025)
RUQuant: Towards Refining Uniform Quantization for Large Language Models
di: Liu, Han, et al.
Pubblicazione: (2026)
di: Liu, Han, et al.
Pubblicazione: (2026)
ToolGen: Unified Tool Retrieval and Calling via Generation
di: Wang, Renxi, et al.
Pubblicazione: (2024)
di: Wang, Renxi, et al.
Pubblicazione: (2024)
Cross-lingual Human-Preference Alignment for Neural Machine Translation with Direct Quality Optimization
di: Uhlig, Kaden, et al.
Pubblicazione: (2024)
di: Uhlig, Kaden, et al.
Pubblicazione: (2024)
RIDE: Enhancing Large Language Model Alignment through Restyled In-Context Learning Demonstration Exemplars
di: Hua, Yuncheng, et al.
Pubblicazione: (2025)
di: Hua, Yuncheng, et al.
Pubblicazione: (2025)
PaperAudit-Bench: Benchmarking Error Detection in Research Papers for Critical Automated Peer Review
di: Tu, Songjun, et al.
Pubblicazione: (2026)
di: Tu, Songjun, et al.
Pubblicazione: (2026)
OPOR-Bench: Evaluating Large Language Models on Online Public Opinion Report Generation
di: Yu, Jinzheng, et al.
Pubblicazione: (2025)
di: Yu, Jinzheng, et al.
Pubblicazione: (2025)
When Self-Reference Fails to Close: Matrix-Level Dynamics in Large Language Models
di: Bae, Ji Ho
Pubblicazione: (2026)
di: Bae, Ji Ho
Pubblicazione: (2026)
SEPTQ: A Simple and Effective Post-Training Quantization Paradigm for Large Language Models
di: Liu, Han, et al.
Pubblicazione: (2026)
di: Liu, Han, et al.
Pubblicazione: (2026)
BabelDOC: Better Layout-Preserving PDF Translation via Intermediate Representation
di: Yang, Qi, et al.
Pubblicazione: (2026)
di: Yang, Qi, et al.
Pubblicazione: (2026)
LoRS: Efficient Low-Rank Adaptation for Sparse Large Language Model
di: Hu, Yuxuan, et al.
Pubblicazione: (2025)
di: Hu, Yuxuan, et al.
Pubblicazione: (2025)
Adaptive Schema-aware Event Extraction with Retrieval-Augmented Generation
di: Liang, Sheng, et al.
Pubblicazione: (2025)
di: Liang, Sheng, et al.
Pubblicazione: (2025)
ML-Promise: A Multilingual Dataset for Corporate Promise Verification
di: Seki, Yohei, et al.
Pubblicazione: (2024)
di: Seki, Yohei, et al.
Pubblicazione: (2024)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
di: Saji, Alan, et al.
Pubblicazione: (2025)
di: Saji, Alan, et al.
Pubblicazione: (2025)
Partially Recentralization Softmax Loss for Vision-Language Models Robustness
di: Wang, Hao, et al.
Pubblicazione: (2024)
di: Wang, Hao, et al.
Pubblicazione: (2024)
myNER: Contextualized Burmese Named Entity Recognition with Bidirectional LSTM and fastText Embeddings via Joint Training with POS Tagging
di: Thant, Kaung Lwin, et al.
Pubblicazione: (2025)
di: Thant, Kaung Lwin, et al.
Pubblicazione: (2025)
Towards Fundamental Language Models: Does Linguistic Competence Scale with Model Size?
di: Collado-Montañez, Jaime, et al.
Pubblicazione: (2025)
di: Collado-Montañez, Jaime, et al.
Pubblicazione: (2025)
SeLeRoSa: Sentence-Level Romanian Satire Detection Dataset
di: Smădu, Răzvan-Alexandru, et al.
Pubblicazione: (2025)
di: Smădu, Răzvan-Alexandru, et al.
Pubblicazione: (2025)
Enhancing Paraphrase Type Generation: The Impact of DPO and RLHF Evaluated with Human-Ranked Data
di: Lübbers, Christopher Lee
Pubblicazione: (2025)
di: Lübbers, Christopher Lee
Pubblicazione: (2025)
Temporal Knowledge Question Answering via Abstract Reasoning Induction
di: Chen, Ziyang, et al.
Pubblicazione: (2023)
di: Chen, Ziyang, et al.
Pubblicazione: (2023)
Learning to Generate Structured Output with Schema Reinforcement Learning
di: Lu, Yaxi, et al.
Pubblicazione: (2025)
di: Lu, Yaxi, et al.
Pubblicazione: (2025)
Alignment Backfire: Language-Dependent Reversal of Safety Interventions Across 16 Languages in LLM Multi-Agent Systems
di: Fukui, Hiroki
Pubblicazione: (2026)
di: Fukui, Hiroki
Pubblicazione: (2026)
HumanLLM: Benchmarking and Improving LLM Anthropomorphism via Human Cognitive Patterns
di: Wang, Xintao, et al.
Pubblicazione: (2026)
di: Wang, Xintao, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Dialect Normalization using Large Language Models and Morphological Rules
di: Dimakis, Antonios, et al.
Pubblicazione: (2025) -
Dialect Matters: Cross-Lingual ASR Transfer for Low-Resource Indic Language Varieties
di: Dhasmana, Akriti, et al.
Pubblicazione: (2026) -
EnDive: A Cross-Dialect Benchmark for Fairness and Performance in Large Language Models
di: Gupta, Abhay, et al.
Pubblicazione: (2025) -
Strategy Adaptation in Large Language Model Werewolf Agents
di: Nakamori, Fuya, et al.
Pubblicazione: (2025) -
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
di: Peters, Sydney, et al.
Pubblicazione: (2025)