Saved in:
| Main Authors: | Ngweta, Lilian, Kate, Kiran, Tsay, Jason, Rizk, Yara |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2504.06969 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
When Agents go Astray: Course-Correcting SWE Agents with PRMs
by: Gandhi, Shubham, et al.
Published: (2025)
by: Gandhi, Shubham, et al.
Published: (2025)
Aligners: Decoupling LLMs and Alignment
by: Ngweta, Lilian, et al.
Published: (2024)
by: Ngweta, Lilian, et al.
Published: (2024)
Repairing Tool Calls Using Post-tool Execution Reflection and RAG
by: Tsay, Jason, et al.
Published: (2025)
by: Tsay, Jason, et al.
Published: (2025)
NESTFUL: A Benchmark for Evaluating LLMs on Nested Sequences of API Calls
by: Basu, Kinjal, et al.
Published: (2024)
by: Basu, Kinjal, et al.
Published: (2024)
Intuitive or Dependent? Investigating LLMs' Behavior Style to Conflicting Prompts
by: Ying, Jiahao, et al.
Published: (2023)
by: Ying, Jiahao, et al.
Published: (2023)
Agent Lifecycle Toolkit (ALTK): Reusable Middleware Components for Robust AI Agents
by: Wright, Zidane, et al.
Published: (2026)
by: Wright, Zidane, et al.
Published: (2026)
Better Call Claude: Can LLMs Detect Changes of Writing Style?
by: Römisch, Johannes, et al.
Published: (2025)
by: Römisch, Johannes, et al.
Published: (2025)
ToolRM: Outcome Reward Models for Tool-Calling Large Language Models
by: Agarwal, Mayank, et al.
Published: (2025)
by: Agarwal, Mayank, et al.
Published: (2025)
Harmful Prompt Laundering: Jailbreaking LLMs with Abductive Styles and Symbolic Encoding
by: Joo, Seongho, et al.
Published: (2025)
by: Joo, Seongho, et al.
Published: (2025)
LLMs Are Biased Towards Output Formats! Systematically Evaluating and Mitigating Output Format Bias of LLMs
by: Long, Do Xuan, et al.
Published: (2024)
by: Long, Do Xuan, et al.
Published: (2024)
Prompt Decorators: A Declarative and Composable Syntax for Reasoning, Formatting, and Control in LLMs
by: Heris, Mostapha Kalami
Published: (2025)
by: Heris, Mostapha Kalami
Published: (2025)
Selective Prompting Tuning for Personalized Conversations with LLMs
by: Huang, Qiushi, et al.
Published: (2024)
by: Huang, Qiushi, et al.
Published: (2024)
Style-Specific Neurons for Steering LLMs in Text Style Transfer
by: Lai, Wen, et al.
Published: (2024)
by: Lai, Wen, et al.
Published: (2024)
How Good Are LLMs at Processing Tool Outputs?
by: Kate, Kiran, et al.
Published: (2025)
by: Kate, Kiran, et al.
Published: (2025)
Style-Compress: An LLM-Based Prompt Compression Framework Considering Task-Specific Styles
by: Pu, Xiao, et al.
Published: (2024)
by: Pu, Xiao, et al.
Published: (2024)
Towards Self-Robust LLMs: Intrinsic Prompt Noise Resistance via CoIPO
by: Yang, Xin, et al.
Published: (2026)
by: Yang, Xin, et al.
Published: (2026)
PromptRobust: Towards Evaluating the Robustness of Large Language Models on Adversarial Prompts
by: Zhu, Kaijie, et al.
Published: (2023)
by: Zhu, Kaijie, et al.
Published: (2023)
Towards Self-Referential Analytic Assessment: A Profile-Based Approach to L2 Writing Evaluation with LLMs
by: Bannò, Stefano, et al.
Published: (2026)
by: Bannò, Stefano, et al.
Published: (2026)
StyleDubber: Towards Multi-Scale Style Learning for Movie Dubbing
by: Cong, Gaoxiang, et al.
Published: (2024)
by: Cong, Gaoxiang, et al.
Published: (2024)
A Chain-of-Thought Prompting Approach with LLMs for Evaluating Students' Formative Assessment Responses in Science
by: Cohn, Clayton, et al.
Published: (2024)
by: Cohn, Clayton, et al.
Published: (2024)
Enhancing Function-Calling Capabilities in LLMs: Strategies for Prompt Formats, Data Integration, and Multilingual Translation
by: Chen, Yi-Chang, et al.
Published: (2024)
by: Chen, Yi-Chang, et al.
Published: (2024)
Morphemes Without Borders: Evaluating Root-Pattern Morphology in Arabic Tokenizers and LLMs
by: Alakeel, Yara, et al.
Published: (2026)
by: Alakeel, Yara, et al.
Published: (2026)
Short-Path Prompting in LLMs: Analyzing Reasoning Instability and Solutions for Robust Performance
by: Tang, Zuoli, et al.
Published: (2025)
by: Tang, Zuoli, et al.
Published: (2025)
Format Matters: The Robustness of Multimodal LLMs in Reviewing Evidence from Tables and Charts
by: Ho, Xanh, et al.
Published: (2025)
by: Ho, Xanh, et al.
Published: (2025)
Surface Reading LLMs: Synthetic Text and its Styles
by: Bajohr, Hannes
Published: (2025)
by: Bajohr, Hannes
Published: (2025)
StyleRec: A Benchmark Dataset for Prompt Recovery in Writing Style Transformation
by: Liu, Shenyang, et al.
Published: (2025)
by: Liu, Shenyang, et al.
Published: (2025)
Flip-Flop Consistency: Unsupervised Training for Robustness to Prompt Perturbations in LLMs
by: Hejabi, Parsa, et al.
Published: (2025)
by: Hejabi, Parsa, et al.
Published: (2025)
Towards Style Alignment in Cross-Cultural Translation
by: Havaldar, Shreya, et al.
Published: (2025)
by: Havaldar, Shreya, et al.
Published: (2025)
Fragile Reasoning: A Mechanistic Analysis of LLM Sensitivity to Meaning-Preserving Perturbations
by: Han, Shou-Tzu, et al.
Published: (2026)
by: Han, Shou-Tzu, et al.
Published: (2026)
SC2: Towards Enhancing Content Preservation and Style Consistency in Long Text Style Transfer
by: Zhao, Jie, et al.
Published: (2024)
by: Zhao, Jie, et al.
Published: (2024)
When Style Breaks Safety: Defending LLMs Against Superficial Style Alignment
by: Xiao, Yuxin, et al.
Published: (2025)
by: Xiao, Yuxin, et al.
Published: (2025)
APM: Evaluating Style Personalization in LLMs with Arbitrary Preference Mappings
by: Spohn, Philipp, et al.
Published: (2026)
by: Spohn, Philipp, et al.
Published: (2026)
Generation, Evaluation, and Explanation of Novelists' Styles with Single-Token Prompts
by: Rezaei, Mosab, et al.
Published: (2025)
by: Rezaei, Mosab, et al.
Published: (2025)
One Model Transfer to All: On Robust Jailbreak Prompts Generation against LLMs
by: Li, Linbao, et al.
Published: (2025)
by: Li, Linbao, et al.
Published: (2025)
Towards Robust and Parameter-Efficient Knowledge Unlearning for LLMs
by: Cha, Sungmin, et al.
Published: (2024)
by: Cha, Sungmin, et al.
Published: (2024)
LakotaBERT: A Transformer-based Model for Low Resource Lakota Language
by: Parankusham, Kanishka, et al.
Published: (2025)
by: Parankusham, Kanishka, et al.
Published: (2025)
Towards Robust Real-World Spreadsheet Understanding with Multi-Agent Multi-Format Reasoning
by: Ren, Houxing, et al.
Published: (2026)
by: Ren, Houxing, et al.
Published: (2026)
Towards a Benchmark for Large Language Models for Business Process Management Tasks
by: Busch, Kiran, et al.
Published: (2024)
by: Busch, Kiran, et al.
Published: (2024)
SPARC: Subspace-Aware Prompt Adaptation for Robust Continual Learning in LLMs
by: Jayasuriya, Dinithi, et al.
Published: (2025)
by: Jayasuriya, Dinithi, et al.
Published: (2025)
Study on LLMs for Promptagator-Style Dense Retriever Training
by: Gwon, Daniel, et al.
Published: (2025)
by: Gwon, Daniel, et al.
Published: (2025)
Similar Items
-
When Agents go Astray: Course-Correcting SWE Agents with PRMs
by: Gandhi, Shubham, et al.
Published: (2025) -
Aligners: Decoupling LLMs and Alignment
by: Ngweta, Lilian, et al.
Published: (2024) -
Repairing Tool Calls Using Post-tool Execution Reflection and RAG
by: Tsay, Jason, et al.
Published: (2025) -
NESTFUL: A Benchmark for Evaluating LLMs on Nested Sequences of API Calls
by: Basu, Kinjal, et al.
Published: (2024) -
Intuitive or Dependent? Investigating LLMs' Behavior Style to Conflicting Prompts
by: Ying, Jiahao, et al.
Published: (2023)