Interpretable Emergent Language Using Inter-Agent Transformers
Fuente:
arXiv
Salvato in:
| Autore principale: | Bhardwaj, Mannan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Talking Heads: Understanding Inter-layer Communication in Transformer Language Models
di: Merullo, Jack, et al.
Pubblicazione: (2024)
di: Merullo, Jack, et al.
Pubblicazione: (2024)
Emergent Languages in Populations of Language Model Agents: From Token Efficiency to Oversight Evasion
di: Beltoft, Stine Lyngsø, et al.
Pubblicazione: (2026)
di: Beltoft, Stine Lyngsø, et al.
Pubblicazione: (2026)
Speaking Your Language: Spatial Relationships in Interpretable Emergent Communication
di: Lipinski, Olaf, et al.
Pubblicazione: (2024)
di: Lipinski, Olaf, et al.
Pubblicazione: (2024)
Emergent Convergence in Multi-Agent LLM Annotation
di: Parfenova, Angelina, et al.
Pubblicazione: (2025)
di: Parfenova, Angelina, et al.
Pubblicazione: (2025)
SignAttention: On the Interpretability of Transformer Models for Sign Language Translation
di: Bianco, Pedro Alejandro Dal, et al.
Pubblicazione: (2024)
di: Bianco, Pedro Alejandro Dal, et al.
Pubblicazione: (2024)
VQEL: Enabling Self-Play in Emergent Language Games via Agent-Internal Vector Quantization
di: Paqaleh, Mohammad Mahdi Samiei, et al.
Pubblicazione: (2025)
di: Paqaleh, Mohammad Mahdi Samiei, et al.
Pubblicazione: (2025)
Using Large Language Models for the Interpretation of Building Regulations
di: Fuchs, Stefan, et al.
Pubblicazione: (2024)
di: Fuchs, Stefan, et al.
Pubblicazione: (2024)
Language Models are Homer Simpson! Safety Re-Alignment of Fine-tuned Language Models through Task Arithmetic
di: Bhardwaj, Rishabh, et al.
Pubblicazione: (2024)
di: Bhardwaj, Rishabh, et al.
Pubblicazione: (2024)
Emergent Introspective Awareness in Large Language Models
di: Lindsey, Jack
Pubblicazione: (2026)
di: Lindsey, Jack
Pubblicazione: (2026)
The Grammar of Transformers: A Systematic Review of Interpretability Research on Syntactic Knowledge in Language Models
di: Graichen, Nora, et al.
Pubblicazione: (2026)
di: Graichen, Nora, et al.
Pubblicazione: (2026)
SuperLocalMemory V3.3: The Living Brain -- Biologically-Inspired Forgetting, Cognitive Quantization, and Multi-Channel Retrieval for Zero-LLM Agent Memory Systems
di: Bhardwaj, Varun Pratap
Pubblicazione: (2026)
di: Bhardwaj, Varun Pratap
Pubblicazione: (2026)
Scaled and Inter-token Relation Enhanced Transformer for Sample-restricted Residential NILM
di: Rahman, Minhajur, et al.
Pubblicazione: (2024)
di: Rahman, Minhajur, et al.
Pubblicazione: (2024)
Unleashing the Emergent Cognitive Synergy in Large Language Models: A Task-Solving Agent through Multi-Persona Self-Collaboration
di: Wang, Zhenhailong, et al.
Pubblicazione: (2023)
di: Wang, Zhenhailong, et al.
Pubblicazione: (2023)
Automated Interpretability and Feature Discovery in Language Models with Agents
di: Marin-Llobet, Arnau, et al.
Pubblicazione: (2026)
di: Marin-Llobet, Arnau, et al.
Pubblicazione: (2026)
Attention-MoA: Enhancing Mixture-of-Agents via Inter-Agent Semantic Attention and Deep Residual Synthesis
di: Wen, Jianyu, et al.
Pubblicazione: (2026)
di: Wen, Jianyu, et al.
Pubblicazione: (2026)
A Monosemantic Attribution Framework for Stable Interpretability in Clinical Neuroscience Transformer-Based Language Models
di: Mamalakis, Michail, et al.
Pubblicazione: (2026)
di: Mamalakis, Michail, et al.
Pubblicazione: (2026)
Emergent Abilities of Large Language Models under Continued Pretraining for Language Adaptation
di: Elhady, Ahmed, et al.
Pubblicazione: (2025)
di: Elhady, Ahmed, et al.
Pubblicazione: (2025)
Investigate-Consolidate-Exploit: A General Strategy for Inter-Task Agent Self-Evolution
di: Qian, Cheng, et al.
Pubblicazione: (2024)
di: Qian, Cheng, et al.
Pubblicazione: (2024)
From Tokens to Lattices: Emergent Lattice Structures in Language Models
di: Xiong, Bo, et al.
Pubblicazione: (2025)
di: Xiong, Bo, et al.
Pubblicazione: (2025)
Humanlike Cognitive Patterns as Emergent Phenomena in Large Language Models
di: Tang, Zhisheng, et al.
Pubblicazione: (2024)
di: Tang, Zhisheng, et al.
Pubblicazione: (2024)
Emergent Semantics Beyond Token Embeddings: Transformer LMs with Frozen Visual Unicode Representations
di: Bochkov, A.
Pubblicazione: (2025)
di: Bochkov, A.
Pubblicazione: (2025)
The Dual-Stream Transformer: Channelized Architecture for Interpretable Language Modeling
di: Kerce, J. Clayton, et al.
Pubblicazione: (2026)
di: Kerce, J. Clayton, et al.
Pubblicazione: (2026)
Emergent Symbolic Mechanisms Support Abstract Reasoning in Large Language Models
di: Yang, Yukang, et al.
Pubblicazione: (2025)
di: Yang, Yukang, et al.
Pubblicazione: (2025)
Multi-Agent Causal Discovery Using Large Language Models
di: Le, Hao Duong, et al.
Pubblicazione: (2024)
di: Le, Hao Duong, et al.
Pubblicazione: (2024)
Enhancing Answer Reliability Through Inter-Model Consensus of Large Language Models
di: Amiri-Margavi, Alireza, et al.
Pubblicazione: (2024)
di: Amiri-Margavi, Alireza, et al.
Pubblicazione: (2024)
Transforming Wearable Data into Personal Health Insights using Large Language Model Agents
di: Merrill, Mike A., et al.
Pubblicazione: (2024)
di: Merrill, Mike A., et al.
Pubblicazione: (2024)
Generative Emergent Communication: Large Language Model is a Collective World Model
di: Taniguchi, Tadahiro, et al.
Pubblicazione: (2024)
di: Taniguchi, Tadahiro, et al.
Pubblicazione: (2024)
Emergent Structured Representations Support Flexible In-Context Inference in Large Language Models
di: Xu, Ningyu, et al.
Pubblicazione: (2026)
di: Xu, Ningyu, et al.
Pubblicazione: (2026)
Your Agent May Misevolve: Emergent Risks in Self-evolving LLM Agents
di: Shao, Shuai, et al.
Pubblicazione: (2025)
di: Shao, Shuai, et al.
Pubblicazione: (2025)
Interpreting Key Mechanisms of Factual Recall in Transformer-Based Language Models
di: Lv, Ang, et al.
Pubblicazione: (2024)
di: Lv, Ang, et al.
Pubblicazione: (2024)
MobileVLM: A Vision-Language Model for Better Intra- and Inter-UI Understanding
di: Wu, Qinzhuo, et al.
Pubblicazione: (2024)
di: Wu, Qinzhuo, et al.
Pubblicazione: (2024)
Examining Inter-Consistency of Large Language Models Collaboration: An In-depth Analysis via Debate
di: Xiong, Kai, et al.
Pubblicazione: (2023)
di: Xiong, Kai, et al.
Pubblicazione: (2023)
ARCANE: A Multi-Agent Framework for Interpretable and Configurable Alignment
di: Masters, Charlie, et al.
Pubblicazione: (2025)
di: Masters, Charlie, et al.
Pubblicazione: (2025)
AgentSHAP: Interpreting LLM Agent Tool Importance with Monte Carlo Shapley Value Estimation
di: Horovicz, Miriam
Pubblicazione: (2025)
di: Horovicz, Miriam
Pubblicazione: (2025)
U-shaped and Inverted-U Scaling behind Emergent Abilities of Large Language Models
di: Wu, Tung-Yu, et al.
Pubblicazione: (2024)
di: Wu, Tung-Yu, et al.
Pubblicazione: (2024)
AgentGroupChat: An Interactive Group Chat Simulacra For Better Eliciting Emergent Behavior
di: Gu, Zhouhong, et al.
Pubblicazione: (2024)
di: Gu, Zhouhong, et al.
Pubblicazione: (2024)
AIOS Compiler: LLM as Interpreter for Natural Language Programming and Flow Programming of AI Agents
di: Xu, Shuyuan, et al.
Pubblicazione: (2024)
di: Xu, Shuyuan, et al.
Pubblicazione: (2024)
Enhanced Sentiment Interpretation via a Lexicon-Fuzzy-Transformer Framework
di: Rokhva, Shayan, et al.
Pubblicazione: (2025)
di: Rokhva, Shayan, et al.
Pubblicazione: (2025)
Interpretability of Language Models via Task Spaces
di: Weber, Lucas, et al.
Pubblicazione: (2024)
di: Weber, Lucas, et al.
Pubblicazione: (2024)
Representations as Language: An Information-Theoretic Framework for Interpretability
di: Conklin, Henry, et al.
Pubblicazione: (2024)
di: Conklin, Henry, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Talking Heads: Understanding Inter-layer Communication in Transformer Language Models
di: Merullo, Jack, et al.
Pubblicazione: (2024) -
Emergent Languages in Populations of Language Model Agents: From Token Efficiency to Oversight Evasion
di: Beltoft, Stine Lyngsø, et al.
Pubblicazione: (2026) -
Speaking Your Language: Spatial Relationships in Interpretable Emergent Communication
di: Lipinski, Olaf, et al.
Pubblicazione: (2024) -
Emergent Convergence in Multi-Agent LLM Annotation
di: Parfenova, Angelina, et al.
Pubblicazione: (2025) -
SignAttention: On the Interpretability of Transformer Models for Sign Language Translation
di: Bianco, Pedro Alejandro Dal, et al.
Pubblicazione: (2024)