Saved in:
| Main Authors: | Arnould, Daniel, Aziz, Rashad, Kang, Zixuan, Changal, Tanav, Zhu, Kevin, Dev, Sunishchal, Grand, Gabriel, Kulkarni, Shreyas Sunil |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2606.01182 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BED-LLM: Intelligent Information Gathering with LLMs and Bayesian Experimental Design
by: Choudhury, Deepro, et al.
Published: (2025)
by: Choudhury, Deepro, et al.
Published: (2025)
Visualizing and Benchmarking LLM Factual Hallucination Tendencies via Internal State Analysis and Clustering
by: Mao, Nathan, et al.
Published: (2026)
by: Mao, Nathan, et al.
Published: (2026)
COMPASS: Context-Modulated PID Attention Steering System for Hallucination Mitigation
by: Sahay, Kenji, et al.
Published: (2025)
by: Sahay, Kenji, et al.
Published: (2025)
ProMoral-Bench: Evaluating Prompting Strategies for Moral Reasoning and Safety in LLMs
by: Thomas, Rohan Subramanian, et al.
Published: (2026)
by: Thomas, Rohan Subramanian, et al.
Published: (2026)
Modeling and Predicting Multi-Turn Answer Instability in Large Language Models
by: He, Jiahang, et al.
Published: (2025)
by: He, Jiahang, et al.
Published: (2025)
DuoLens: A Framework for Robust Detection of Machine-Generated Multilingual Text and Code
by: Agrawal, Shriyansh, et al.
Published: (2025)
by: Agrawal, Shriyansh, et al.
Published: (2025)
REAMS: Reasoning Enhanced Algorithm for Maths Solving
by: Singh, Eishkaran, et al.
Published: (2025)
by: Singh, Eishkaran, et al.
Published: (2025)
MoEMoE: Question Guided Dense and Scalable Sparse Mixture-of-Expert for Multi-source Multi-modal Answering
by: Verma, Vinay Kumar, et al.
Published: (2025)
by: Verma, Vinay Kumar, et al.
Published: (2025)
Alignment-Constrained Dynamic Pruning for LLMs: Identifying and Preserving Alignment-Critical Circuits
by: Patel, Dev, et al.
Published: (2025)
by: Patel, Dev, et al.
Published: (2025)
Low-Resource Safety Failures Are Action Failures, Not Representation Failures
by: Aziz, Rashad, et al.
Published: (2026)
by: Aziz, Rashad, et al.
Published: (2026)
SALT: Steering Activations towards Leakage-free Thinking in Chain of Thought
by: Batra, Shourya, et al.
Published: (2025)
by: Batra, Shourya, et al.
Published: (2025)
BED: Bi-Encoder-Based Detectors for Out-of-Distribution Detection
by: Owen, Louis, et al.
Published: (2023)
by: Owen, Louis, et al.
Published: (2023)
Emergent Misalignment via In-Context Learning: Narrow in-context examples can produce broadly misaligned LLMs
by: Afonin, Nikita, et al.
Published: (2025)
by: Afonin, Nikita, et al.
Published: (2025)
Sumudu Neural Operator for ODEs and PDEs
by: Zelenskiy, Ben, et al.
Published: (2025)
by: Zelenskiy, Ben, et al.
Published: (2025)
SEA-BED: How Do Embedding Models Represent Southeast Asian Languages?
by: Ponwitayarat, Wuttikorn, et al.
Published: (2025)
by: Ponwitayarat, Wuttikorn, et al.
Published: (2025)
INSURE-Dial: A Phase-Aware Conversational Dataset & Benchmark for Compliance Verification and Phase Detection
by: Kulkarni, Shubham, et al.
Published: (2026)
by: Kulkarni, Shubham, et al.
Published: (2026)
Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction
by: Rashad, Mohamed
Published: (2024)
by: Rashad, Mohamed
Published: (2024)
Twisted Malle's Conjecture
by: Choudhary, Tanav
Published: (2025)
by: Choudhary, Tanav
Published: (2025)
On the codimension 1 PGL(3) orbit closures in $\text{Gr}(3,6)$
by: Choudhary, Tanav
Published: (2025)
by: Choudhary, Tanav
Published: (2025)
Cognis: Context-Aware Memory for Conversational AI Agents
by: Daftari, Parshva, et al.
Published: (2026)
by: Daftari, Parshva, et al.
Published: (2026)
Chopping Trees: Semantic Similarity Based Dynamic Pruning for Tree-of-Thought Reasoning
by: Kim, Joongho, et al.
Published: (2025)
by: Kim, Joongho, et al.
Published: (2025)
Understanding Figurative Meaning through Explainable Visual Entailment
by: Saakyan, Arkadiy, et al.
Published: (2024)
by: Saakyan, Arkadiy, et al.
Published: (2024)
Limits of Emergent Reasoning of Large Language Models in Agentic Frameworks for Deterministic Games
by: Su, Chris, et al.
Published: (2025)
by: Su, Chris, et al.
Published: (2025)
Shoot First, Ask Questions Later? Building Rational Agents that Explore and Act Like People
by: Grand, Gabriel, et al.
Published: (2025)
by: Grand, Gabriel, et al.
Published: (2025)
Loose LIPS Sink Ships: Asking Questions in Battleship with Language-Informed Program Sampling
by: Grand, Gabriel, et al.
Published: (2024)
by: Grand, Gabriel, et al.
Published: (2024)
BYO-Eval: Build Your Own Dataset for Fine-Grained Visual Assessment of Multimodal Language Models
by: Arnould, Ludovic, et al.
Published: (2025)
by: Arnould, Ludovic, et al.
Published: (2025)
AgentChangeBench: A Multi-Dimensional Evaluation Framework for Goal-Shift Robustness in Conversational AI
by: Rana, Manik, et al.
Published: (2025)
by: Rana, Manik, et al.
Published: (2025)
Advanced LPeg techniques: A dual case study approach
by: Zhu, Zixuan
Published: (2025)
by: Zhu, Zixuan
Published: (2025)
Build, Borrow, or Just Fine-Tune? A Political Scientist's Guide to Choosing NLP Models
by: Meher, Shreyas
Published: (2026)
by: Meher, Shreyas
Published: (2026)
Tool-Aware Planning in Contact Center AI: Evaluating LLMs through Lineage-Guided Query Decomposition
by: Nathan, Varun, et al.
Published: (2026)
by: Nathan, Varun, et al.
Published: (2026)
Profiling checkpointing schedules in adjoint ST-AD
by: Hascoët, Laurent, et al.
Published: (2024)
by: Hascoët, Laurent, et al.
Published: (2024)
History-Aware Conversational Dense Retrieval
by: Mo, Fengran, et al.
Published: (2024)
by: Mo, Fengran, et al.
Published: (2024)
Emergent Persuasion: Will LLMs Persuade Without Being Prompted?
by: Chang, Vincent, et al.
Published: (2025)
by: Chang, Vincent, et al.
Published: (2025)
Inference-Time Chain-of-Thought Pruning with Latent Informativeness Signals
by: Li, Sophie, et al.
Published: (2025)
by: Li, Sophie, et al.
Published: (2025)
Uncertainty-Aware Budget Allocation for Adaptive Test-Time Reasoning
by: Nguyen, Manh, et al.
Published: (2026)
by: Nguyen, Manh, et al.
Published: (2026)
CA-BERT: Leveraging Context Awareness for Enhanced Multi-Turn Chat Interaction
by: Liu, Minghao, et al.
Published: (2024)
by: Liu, Minghao, et al.
Published: (2024)
CA*: Addressing Evaluation Pitfalls in Computation-Aware Latency for Simultaneous Speech Translation
by: Xu, Xi, et al.
Published: (2024)
by: Xu, Xi, et al.
Published: (2024)
BiCA: Effective Biomedical Dense Retrieval with Citation-Aware Hard Negatives
by: Sinha, Aarush, et al.
Published: (2025)
by: Sinha, Aarush, et al.
Published: (2025)
Self-Steering Language Models
by: Grand, Gabriel, et al.
Published: (2025)
by: Grand, Gabriel, et al.
Published: (2025)
Clarify, Abstain or Answer? Strategising in Conversation with Belief-Augmented Generation
by: Baan, Joris, et al.
Published: (2026)
by: Baan, Joris, et al.
Published: (2026)
Similar Items
-
BED-LLM: Intelligent Information Gathering with LLMs and Bayesian Experimental Design
by: Choudhury, Deepro, et al.
Published: (2025) -
Visualizing and Benchmarking LLM Factual Hallucination Tendencies via Internal State Analysis and Clustering
by: Mao, Nathan, et al.
Published: (2026) -
COMPASS: Context-Modulated PID Attention Steering System for Hallucination Mitigation
by: Sahay, Kenji, et al.
Published: (2025) -
ProMoral-Bench: Evaluating Prompting Strategies for Moral Reasoning and Safety in LLMs
by: Thomas, Rohan Subramanian, et al.
Published: (2026) -
Modeling and Predicting Multi-Turn Answer Instability in Large Language Models
by: He, Jiahang, et al.
Published: (2025)