Frictional Agent Alignment Framework: Slow Down and Don't Break Things
Fuente:
arXiv
Saved in:
| Main Authors: | Nath, Abhijnan, Graff, Carine, Bachinin, Andrei, Krishnaswamy, Nikhil |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Collaborate, Deliberate, Evaluate: How LLM Alignment Affects Coordinated Multi-Agent Outcomes
by: Nath, Abhijnan, et al.
Published: (2025)
by: Nath, Abhijnan, et al.
Published: (2025)
Dynamic Epistemic Friction in Dialogue
by: Obiso, Timothy, et al.
Published: (2025)
by: Obiso, Timothy, et al.
Published: (2025)
CRAFT: Grounded Multi-Agent Coordination Under Partial Information
by: Nath, Abhijnan, et al.
Published: (2026)
by: Nath, Abhijnan, et al.
Published: (2026)
Okay, Let's Do This! Modeling Event Coreference with Generated Rationales and Knowledge Distillation
by: Nath, Abhijnan, et al.
Published: (2024)
by: Nath, Abhijnan, et al.
Published: (2024)
Simultaneous Reward Distillation and Preference Learning: Get You a Language Model Who Can Do Both
by: Nath, Abhijnan, et al.
Published: (2024)
by: Nath, Abhijnan, et al.
Published: (2024)
Learning "Partner-Aware" Collaborators in Multi-Party Collaboration
by: Nath, Abhijnan, et al.
Published: (2025)
by: Nath, Abhijnan, et al.
Published: (2025)
Any Other Thoughts, Hedgehog? Linking Deliberation Chains in Collaborative Dialogues
by: Nath, Abhijnan, et al.
Published: (2024)
by: Nath, Abhijnan, et al.
Published: (2024)
Multimodal Cross-Document Event Coreference Resolution Using Linear Semantic Transfer and Mixed-Modality Ensembles
by: Nath, Abhijnan, et al.
Published: (2024)
by: Nath, Abhijnan, et al.
Published: (2024)
Frictive Policy Optimization for LLMs: Epistemic Intervention, Risk-Sensitive Control, and Reflective Alignment
by: Pustejovsky, James, et al.
Published: (2026)
by: Pustejovsky, James, et al.
Published: (2026)
Don't Command, Cultivate: An Exploratory Study of System-2 Alignment
by: Wang, Yuhang, et al.
Published: (2024)
by: Wang, Yuhang, et al.
Published: (2024)
Exploring Failure Cases in Multimodal Reasoning About Physical Dynamics
by: Ghaffari, Sadaf, et al.
Published: (2024)
by: Ghaffari, Sadaf, et al.
Published: (2024)
Cross-Lingual Transfer Robustness to Lower-Resource Languages on Adversarial Datasets
by: Manafi, Shadi, et al.
Published: (2024)
by: Manafi, Shadi, et al.
Published: (2024)
Don't Break the Cache: An Evaluation of Prompt Caching for Long-Horizon Agentic Tasks
by: Lumer, Elias, et al.
Published: (2026)
by: Lumer, Elias, et al.
Published: (2026)
Don't Rank, Combine! Combining Machine Translation Hypotheses Using Quality Estimation
by: Vernikos, Giorgos, et al.
Published: (2024)
by: Vernikos, Giorgos, et al.
Published: (2024)
Sample, Don't Search: Rethinking Test-Time Alignment for Language Models
by: Faria, Gonçalo, et al.
Published: (2025)
by: Faria, Gonçalo, et al.
Published: (2025)
Don't Touch My Diacritics
by: Gorman, Kyle, et al.
Published: (2024)
by: Gorman, Kyle, et al.
Published: (2024)
Don't Pay Attention
by: Hammoud, Mohammad, et al.
Published: (2025)
by: Hammoud, Mohammad, et al.
Published: (2025)
The Impact of Background Speech on Interruption Detection in Collaborative Groups
by: Bradford, Mariah, et al.
Published: (2025)
by: Bradford, Mariah, et al.
Published: (2025)
Don't Take Things Out of Context: Attention Intervention for Enhancing Chain-of-Thought Reasoning in Large Language Models
by: Yan, Shaotian, et al.
Published: (2025)
by: Yan, Shaotian, et al.
Published: (2025)
If You Don't Understand It, Don't Use It: Eliminating Trojans with Filters Between Layers
by: Hernandez, Adriano
Published: (2024)
by: Hernandez, Adriano
Published: (2024)
Don't Throw Away Your Pretrained Model
by: Feng, Shangbin, et al.
Published: (2025)
by: Feng, Shangbin, et al.
Published: (2025)
Don't Say No: Jailbreaking LLM by Suppressing Refusal
by: Zhou, Yukai, et al.
Published: (2024)
by: Zhou, Yukai, et al.
Published: (2024)
Don't Forget Your Reward Values: Language Model Alignment via Value-based Calibration
by: Mao, Xin, et al.
Published: (2024)
by: Mao, Xin, et al.
Published: (2024)
Hatevolution: What Static Benchmarks Don't Tell Us
by: Di Bonaventura, Chiara, et al.
Published: (2025)
by: Di Bonaventura, Chiara, et al.
Published: (2025)
Think, But Don't Overthink: Reproducing Recursive Language Models
by: Wang, Daren
Published: (2026)
by: Wang, Daren
Published: (2026)
Don't Trust Generative Agents to Mimic Communication on Social Networks Unless You Benchmarked their Empirical Realism
by: Münker, Simon, et al.
Published: (2025)
by: Münker, Simon, et al.
Published: (2025)
Reasoning Models Reason Well, Until They Don't
by: Rameshkumar, Revanth, et al.
Published: (2025)
by: Rameshkumar, Revanth, et al.
Published: (2025)
Don't Throw Away Data: Better Sequence Knowledge Distillation
by: Wang, Jun, et al.
Published: (2024)
by: Wang, Jun, et al.
Published: (2024)
Mind the Gap: Pitfalls of LLM Alignment with Asian Public Opinion
by: Shankar, Hari, et al.
Published: (2026)
by: Shankar, Hari, et al.
Published: (2026)
s3: You Don't Need That Much Data to Train a Search Agent via RL
by: Jiang, Pengcheng, et al.
Published: (2025)
by: Jiang, Pengcheng, et al.
Published: (2025)
Don't Walk the Line: Boundary Guidance for Filtered Generation
by: Ball, Sarah, et al.
Published: (2025)
by: Ball, Sarah, et al.
Published: (2025)
Language Models Don't Learn the Physical Manifestation of Language
by: Lee, Bruce W., et al.
Published: (2024)
by: Lee, Bruce W., et al.
Published: (2024)
Can AI Assistants Know What They Don't Know?
by: Cheng, Qinyuan, et al.
Published: (2024)
by: Cheng, Qinyuan, et al.
Published: (2024)
Better Slow than Sorry: Introducing Positive Friction for Reliable Dialogue Systems
by: İnan, Mert, et al.
Published: (2025)
by: İnan, Mert, et al.
Published: (2025)
Don't Lose Focus: Activation Steering via Key-Orthogonal Projections
by: Luo, Haoyan, et al.
Published: (2026)
by: Luo, Haoyan, et al.
Published: (2026)
Pointer-Generator Networks for Low-Resource Machine Translation: Don't Copy That!
by: Bafna, Niyati, et al.
Published: (2024)
by: Bafna, Niyati, et al.
Published: (2024)
Reuse, Don't Retrain: A Recipe for Continued Pretraining of Language Models
by: Parmar, Jupinder, et al.
Published: (2024)
by: Parmar, Jupinder, et al.
Published: (2024)
Be like a Goldfish, Don't Memorize! Mitigating Memorization in Generative LLMs
by: Hans, Abhimanyu, et al.
Published: (2024)
by: Hans, Abhimanyu, et al.
Published: (2024)
Don't Start Over: A Cost-Effective Framework for Migrating Personalized Prompts Between LLMs
by: Zhao, Ziyi, et al.
Published: (2026)
by: Zhao, Ziyi, et al.
Published: (2026)
Don't Adapt Small Language Models for Tools; Adapt Tool Schemas to the Models
by: Lee, Jonggeun, et al.
Published: (2025)
by: Lee, Jonggeun, et al.
Published: (2025)
Similar Items
-
Collaborate, Deliberate, Evaluate: How LLM Alignment Affects Coordinated Multi-Agent Outcomes
by: Nath, Abhijnan, et al.
Published: (2025) -
Dynamic Epistemic Friction in Dialogue
by: Obiso, Timothy, et al.
Published: (2025) -
CRAFT: Grounded Multi-Agent Coordination Under Partial Information
by: Nath, Abhijnan, et al.
Published: (2026) -
Okay, Let's Do This! Modeling Event Coreference with Generated Rationales and Knowledge Distillation
by: Nath, Abhijnan, et al.
Published: (2024) -
Simultaneous Reward Distillation and Preference Learning: Get You a Language Model Who Can Do Both
by: Nath, Abhijnan, et al.
Published: (2024)