KnowThyself: An Agentic Assistant for LLM Interpretability
Fuente:
arXiv
Saved in:
| Main Authors: | Prasai, Suraj, Du, Mengnan, Zhang, Ying, Yang, Fan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MATH-PT: A Math Reasoning Benchmark for European and Brazilian Portuguese
by: Teixeira, Tiago, et al.
Published: (2026)
by: Teixeira, Tiago, et al.
Published: (2026)
ARAG: Agentic Retrieval Augmented Generation for Personalized Recommendation
by: Maragheh, Reza Yousefi, et al.
Published: (2025)
by: Maragheh, Reza Yousefi, et al.
Published: (2025)
Knowledge-Aware Iterative Retrieval for Multi-Agent Systems
by: Song, Seyoung
Published: (2025)
by: Song, Seyoung
Published: (2025)
MALTopic: Multi-Agent LLM Topic Modeling Framework
by: Sharma, Yash
Published: (2026)
by: Sharma, Yash
Published: (2026)
PestMA: LLM-based Multi-Agent System for Informed Pest Management
by: Shi, Hongrui, et al.
Published: (2025)
by: Shi, Hongrui, et al.
Published: (2025)
Review of Case-Based Reasoning for LLM Agents: Theoretical Foundations, Architectural Components, and Cognitive Integration
by: Hatalis, Kostas, et al.
Published: (2025)
by: Hatalis, Kostas, et al.
Published: (2025)
Pioneer Agent: Continual Improvement of Small Language Models in Production
by: Atreja, Dhruv, et al.
Published: (2026)
by: Atreja, Dhruv, et al.
Published: (2026)
Deciphering Digital Detectives: Understanding LLM Behaviors and Capabilities in Multi-Agent Mystery Games
by: Wu, Dekun, et al.
Published: (2023)
by: Wu, Dekun, et al.
Published: (2023)
ATANT v1.1: Positioning Continuity Evaluation Against Memory, Long-Context, and Agentic-Memory Benchmarks
by: Tanguturi, Samuel Sameer
Published: (2026)
by: Tanguturi, Samuel Sameer
Published: (2026)
Pareto-Optimized Open-Source LLMs for Healthcare via Context Retrieval
by: Bayarri-Planas, Jordi, et al.
Published: (2024)
by: Bayarri-Planas, Jordi, et al.
Published: (2024)
DecisionBench: A Benchmark for Emergent Delegation in Long-Horizon Agentic Workflows
by: Gao, Yuxuan, et al.
Published: (2026)
by: Gao, Yuxuan, et al.
Published: (2026)
GSAR: Typed Grounding for Hallucination Detection and Recovery in Multi-Agent LLMs
by: Kamelhar, Federico A.
Published: (2026)
by: Kamelhar, Federico A.
Published: (2026)
Agent WARPP: Workflow Adherence via Runtime Parallel Personalization
by: Mazzolenis, Maria Emilia, et al.
Published: (2025)
by: Mazzolenis, Maria Emilia, et al.
Published: (2025)
The Station: An Open-World Environment for AI-Driven Discovery
by: Chung, Stephen, et al.
Published: (2025)
by: Chung, Stephen, et al.
Published: (2025)
Personality-Driven Decision-Making in LLM-Based Autonomous Agents
by: Newsham, Lewis, et al.
Published: (2025)
by: Newsham, Lewis, et al.
Published: (2025)
Literature Review Of Multi-Agent Debate For Problem-Solving
by: Tillmann, Arne
Published: (2025)
by: Tillmann, Arne
Published: (2025)
Talk is Cheap, Communication is Hard: Dynamic Grounding Failures and Repair in Multi-Agent Negotiation
by: Yao, Yiheng, et al.
Published: (2026)
by: Yao, Yiheng, et al.
Published: (2026)
Tool-RoCo: An Agent-as-Tool Self-organization Large Language Model Benchmark in Multi-robot Cooperation
by: Zhang, Ke, et al.
Published: (2025)
by: Zhang, Ke, et al.
Published: (2025)
Using LLM-Based Approaches to Enhance and Automate Topic Labeling
by: Khandelwal, Trishia
Published: (2025)
by: Khandelwal, Trishia
Published: (2025)
Equip Pre-ranking with Target Attention by Residual Quantization
by: Li, Yutong, et al.
Published: (2025)
by: Li, Yutong, et al.
Published: (2025)
QuickSilver -- Speeding up LLM Inference through Dynamic Token Halting, KV Skipping, Contextual Token Fusion, and Adaptive Matryoshka Quantization
by: Khanna, Danush, et al.
Published: (2025)
by: Khanna, Danush, et al.
Published: (2025)
GraphCompliance: Aligning Policy and Context Graphs for LLM-Based Regulatory Compliance
by: Chung, Jiseong, et al.
Published: (2025)
by: Chung, Jiseong, et al.
Published: (2025)
Multi-Agent Systems Powered by Large Language Models: Applications in Swarm Intelligence
by: Jimenez-Romero, Cristian, et al.
Published: (2025)
by: Jimenez-Romero, Cristian, et al.
Published: (2025)
Exploring Design of Multi-Agent LLM Dialogues for Research Ideation
by: Ueda, Keisuke, et al.
Published: (2025)
by: Ueda, Keisuke, et al.
Published: (2025)
CAG: Chunked Augmented Generation for Google Chrome's Built-in Gemini Nano
by: Surulimuthu, Vivek Vellaiyappan, et al.
Published: (2024)
by: Surulimuthu, Vivek Vellaiyappan, et al.
Published: (2024)
TRIZ Agents: A Multi-Agent LLM Approach for TRIZ-Based Innovation
by: Szczepanik, Kamil, et al.
Published: (2025)
by: Szczepanik, Kamil, et al.
Published: (2025)
EMOS: Embodiment-aware Heterogeneous Multi-robot Operating System with LLM Agents
by: Chen, Junting, et al.
Published: (2024)
by: Chen, Junting, et al.
Published: (2024)
TwinVoice: A Multi-dimensional Benchmark Towards Digital Twins via LLM Persona Simulation
by: Du, Bangde, et al.
Published: (2025)
by: Du, Bangde, et al.
Published: (2025)
MIRA: Empowering One-Touch AI Services on Smartphones with MLLM-based Instruction Recommendation
by: Bian, Zhipeng, et al.
Published: (2025)
by: Bian, Zhipeng, et al.
Published: (2025)
ATANT: An Evaluation Framework for AI Continuity
by: Tanguturi, Samuel Sameer
Published: (2026)
by: Tanguturi, Samuel Sameer
Published: (2026)
Mixture of Experts Approaches in Dense Retrieval Tasks
by: Sokli, Effrosyni, et al.
Published: (2025)
by: Sokli, Effrosyni, et al.
Published: (2025)
A Language for Describing Agentic LLM Contexts
by: Pelc, Noga Peleg, et al.
Published: (2026)
by: Pelc, Noga Peleg, et al.
Published: (2026)
FlexStructRAG: Flexible Structure-Aware Multi-Granular Relational Retrieval for RAG
by: Chen, Mengzhu, et al.
Published: (2026)
by: Chen, Mengzhu, et al.
Published: (2026)
Anchor-and-Resume Concession Under Dynamic Pricing for LLM-Augmented Freight Negotiation
by: Nguyen, Hoang, et al.
Published: (2026)
by: Nguyen, Hoang, et al.
Published: (2026)
LLM Reasoning for Cold-Start Item Recommendation
by: Li, Shijun, et al.
Published: (2025)
by: Li, Shijun, et al.
Published: (2025)
A Domain-Agnostic Neurosymbolic Approach for Big Social Data Analysis: Evaluating Mental Health Sentiment on Social Media during COVID-19
by: Khandelwal, Vedant, et al.
Published: (2024)
by: Khandelwal, Vedant, et al.
Published: (2024)
Quo Vadis ChatGPT? From Large Language Models to Large Knowledge Models
by: Venkatasubramanian, Venkat, et al.
Published: (2024)
by: Venkatasubramanian, Venkat, et al.
Published: (2024)
Social Learning through Interactions with Other Agents: A Survey
by: Hillier, Dylan, et al.
Published: (2024)
by: Hillier, Dylan, et al.
Published: (2024)
OpenAI Cribbed Our Tax Example, But Can GPT-4 Really Do Tax?
by: Blair-Stanek, Andrew, et al.
Published: (2023)
by: Blair-Stanek, Andrew, et al.
Published: (2023)
ReFoRCE: A Text-to-SQL Agent with Self-Refinement, Consensus Enforcement, and Column Exploration
by: Deng, Minghang, et al.
Published: (2025)
by: Deng, Minghang, et al.
Published: (2025)
Similar Items
-
MATH-PT: A Math Reasoning Benchmark for European and Brazilian Portuguese
by: Teixeira, Tiago, et al.
Published: (2026) -
ARAG: Agentic Retrieval Augmented Generation for Personalized Recommendation
by: Maragheh, Reza Yousefi, et al.
Published: (2025) -
Knowledge-Aware Iterative Retrieval for Multi-Agent Systems
by: Song, Seyoung
Published: (2025) -
MALTopic: Multi-Agent LLM Topic Modeling Framework
by: Sharma, Yash
Published: (2026) -
PestMA: LLM-based Multi-Agent System for Informed Pest Management
by: Shi, Hongrui, et al.
Published: (2025)