Exploring How LLMs Capture and Represent Domain-Specific Knowledge
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Garcia, Mirian Hipolito, Couturier, Camille, Diaz, Daniel Madrigal, Mallick, Ankur, Kyrillidis, Anastasios, Sim, Robert, Ruhle, Victor, Rajmohan, Saravan |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
LEGOMem: Modular Procedural Memory for Multi-agent LLM Systems for Workflow Automation
par: Han, Dongge, et autres
Publié: (2025)
par: Han, Dongge, et autres
Publié: (2025)
Revisiting Transformer Layer Parameterization Through Causal Energy Minimization
par: Xu, Jin, et autres
Publié: (2026)
par: Xu, Jin, et autres
Publié: (2026)
Semantic Caching of Contextual Summaries for Efficient Question-Answering with Language Models
par: Couturier, Camille, et autres
Publié: (2025)
par: Couturier, Camille, et autres
Publié: (2025)
Enhancing Reasoning Capabilities of Small Language Models with Blueprints and Prompt Template Search
par: Han, Dongge, et autres
Publié: (2025)
par: Han, Dongge, et autres
Publié: (2025)
Hybrid-RACA: Hybrid Retrieval-Augmented Composition Assistance for Real-time Text Prediction
par: Xia, Menglin, et autres
Publié: (2023)
par: Xia, Menglin, et autres
Publié: (2023)
Cost-Aware Retrieval-Augmentation Reasoning Models with Adaptive Retrieval Depth
par: Hashemi, Helia, et autres
Publié: (2025)
par: Hashemi, Helia, et autres
Publié: (2025)
OdysseyBench: Evaluating LLM Agents on Long-Horizon Complex Office Application Workflows
par: Wang, Weixuan, et autres
Publié: (2025)
par: Wang, Weixuan, et autres
Publié: (2025)
Memora: A Harmonic Memory Representation Balancing Abstraction and Specificity
par: Xia, Menglin, et autres
Publié: (2026)
par: Xia, Menglin, et autres
Publié: (2026)
Sweeping Heterogeneity with Smart MoPs: Mixture of Prompts for LLM Task Adaptation
par: Dun, Chen, et autres
Publié: (2023)
par: Dun, Chen, et autres
Publié: (2023)
TurboAttention: Efficient Attention Approximation For High Throughputs LLMs
par: Kang, Hao, et autres
Publié: (2024)
par: Kang, Hao, et autres
Publié: (2024)
Budget-Aware Agentic Routing via Boundary-Guided Training
par: Zhang, Caiqi, et autres
Publié: (2026)
par: Zhang, Caiqi, et autres
Publié: (2026)
EcoAct: Economic Agent Determines When to Register What Action
par: Zhang, Shaokun, et autres
Publié: (2024)
par: Zhang, Shaokun, et autres
Publié: (2024)
Lean Attention: Hardware-Aware Scalable Attention Mechanism for the Decode-Phase of Transformers
par: Sanovar, Rya, et autres
Publié: (2024)
par: Sanovar, Rya, et autres
Publié: (2024)
BEST-Route: Adaptive LLM Routing with Test-Time Optimal Compute
par: Ding, Dujian, et autres
Publié: (2025)
par: Ding, Dujian, et autres
Publié: (2025)
SageServe: Optimizing LLM Serving on Cloud Data Centers with Forecast Aware Auto-Scaling
par: Jaiswal, Shashwat, et autres
Publié: (2025)
par: Jaiswal, Shashwat, et autres
Publié: (2025)
Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing
par: Ding, Dujian, et autres
Publié: (2024)
par: Ding, Dujian, et autres
Publié: (2024)
Intelligent Router for LLM Workloads: Improving Performance Through Workload-Aware Load Balancing
par: Jain, Kunal, et autres
Publié: (2024)
par: Jain, Kunal, et autres
Publié: (2024)
AutoAdapt: An Automated Domain Adaptation Framework for LLMs
par: Sinha, Sidharth, et autres
Publié: (2026)
par: Sinha, Sidharth, et autres
Publié: (2026)
Guided by the Experts: Provable Feature Learning Dynamic of Soft-Routed Mixture-of-Experts
par: Liao, Fangshuo, et autres
Publié: (2025)
par: Liao, Fangshuo, et autres
Publié: (2025)
GHOST: Unmasking Phantom States in Mamba2 via Grouped Hidden-state Output-aware Selection & Truncation
par: Menezes, Michael, et autres
Publié: (2026)
par: Menezes, Michael, et autres
Publié: (2026)
Provable Accelerated Convergence of Nesterov's Momentum for Deep ReLU Neural Networks
par: Liao, Fangshuo, et autres
Publié: (2023)
par: Liao, Fangshuo, et autres
Publié: (2023)
A Tale of Two Graphs: Separating Knowledge Exploration from Outline Structure for Open-Ended Deep Research
par: Shi, Zhuofan, et autres
Publié: (2026)
par: Shi, Zhuofan, et autres
Publié: (2026)
Learning to Specialize: Joint Gating-Expert Training for Adaptive MoEs in Decentralized Settings
par: Farhat, Yehya, et autres
Publié: (2023)
par: Farhat, Yehya, et autres
Publié: (2023)
Using non-convex optimization in quantum process tomography: Factored gradient descent is tough to beat
par: Quiroga, David A., et autres
Publié: (2023)
par: Quiroga, David A., et autres
Publié: (2023)
Better Schedules for Low Precision Training of Deep Neural Networks
par: Wolfe, Cameron R., et autres
Publié: (2024)
par: Wolfe, Cameron R., et autres
Publié: (2024)
Towards Active Synthetic Data Generation for Finetuning Language Models
par: Kessler, Samuel, et autres
Publié: (2025)
par: Kessler, Samuel, et autres
Publié: (2025)
Minerva: A Programmable Memory Test Benchmark for Language Models
par: Xia, Menglin, et autres
Publié: (2025)
par: Xia, Menglin, et autres
Publié: (2025)
One Model, Two Roles: Emergent Specialization in a Shared Recurrent Transformer
par: Shen, Jucheng, et autres
Publié: (2026)
par: Shen, Jucheng, et autres
Publié: (2026)
Quantum EigenGame for excited state calculation
par: Quiroga, David, et autres
Publié: (2025)
par: Quiroga, David, et autres
Publié: (2025)
Provable Model-Parallel Distributed Principal Component Analysis with Parallel Deflation
par: Liao, Fangshuo, et autres
Publié: (2025)
par: Liao, Fangshuo, et autres
Publié: (2025)
AdaPaD: Adaptive Parallel Deflation for PEFT with Self-Correcting Rank Discovery
par: Su, Barbara, et autres
Publié: (2026)
par: Su, Barbara, et autres
Publié: (2026)
SGD at the Edge of Stability: The Stochastic Sharpness Gap
par: Liao, Fangshuo, et autres
Publié: (2026)
par: Liao, Fangshuo, et autres
Publié: (2026)
Producción de Lecanicillium (= Verticillium) lecanii en diferentes sustratos y patogenicidad
par: Hipólito Cortez Madrigal
Publié: (2007)
par: Hipólito Cortez Madrigal
Publié: (2007)
LLM Reasoning as Trajectories: Step-Specific Representation Geometry and Correctness Signals
par: Sun, Lihao, et autres
Publié: (2026)
par: Sun, Lihao, et autres
Publié: (2026)
One Rank at a Time: Cascading Error Dynamics in Sequential Learning
par: Vandchali, Mahtab Alizadeh, et autres
Publié: (2025)
par: Vandchali, Mahtab Alizadeh, et autres
Publié: (2025)
Unveiling Hidden Pivotal Players with GoalNet: A GNN-Based Soccer Player Evaluation System
par: Jiang, Jacky Hao, et autres
Publié: (2025)
par: Jiang, Jacky Hao, et autres
Publié: (2025)
Ensuring Fair LLM Serving Amid Diverse Applications
par: Khan, Redwan Ibne Seraj, et autres
Publié: (2024)
par: Khan, Redwan Ibne Seraj, et autres
Publié: (2024)
CI-Work: Benchmarking Contextual Integrity in Enterprise LLM Agents
par: Fu, Wenjie, et autres
Publié: (2026)
par: Fu, Wenjie, et autres
Publié: (2026)
Three Birds with One Stone: Improving Performance, Convergence, and System Throughput with Nest
par: Huo, Yuqian, et autres
Publié: (2025)
par: Huo, Yuqian, et autres
Publié: (2025)
A Catalyst Framework for the Quantum Linear System Problem via the Proximal Point Algorithm
par: Kim, Junhyung Lyle, et autres
Publié: (2024)
par: Kim, Junhyung Lyle, et autres
Publié: (2024)
Documents similaires
-
LEGOMem: Modular Procedural Memory for Multi-agent LLM Systems for Workflow Automation
par: Han, Dongge, et autres
Publié: (2025) -
Revisiting Transformer Layer Parameterization Through Causal Energy Minimization
par: Xu, Jin, et autres
Publié: (2026) -
Semantic Caching of Contextual Summaries for Efficient Question-Answering with Language Models
par: Couturier, Camille, et autres
Publié: (2025) -
Enhancing Reasoning Capabilities of Small Language Models with Blueprints and Prompt Template Search
par: Han, Dongge, et autres
Publié: (2025) -
Hybrid-RACA: Hybrid Retrieval-Augmented Composition Assistance for Real-time Text Prediction
par: Xia, Menglin, et autres
Publié: (2023)