Memora: A Harmonic Memory Representation Balancing Abstraction and Specificity
Fuente:
arXiv
Guardado en:
| Autores principales: | Xia, Menglin, Zhang, Xuchao, Dixit, Shantanu, Harimurugan, Paramaguru, Wang, Rujia, Ruhle, Victor, Sim, Robert, Bansal, Chetan, Rajmohan, Saravan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Hybrid-RACA: Hybrid Retrieval-Augmented Composition Assistance for Real-time Text Prediction
por: Xia, Menglin, et al.
Publicado: (2023)
por: Xia, Menglin, et al.
Publicado: (2023)
AMPO: Active Multi-Preference Optimization for Self-play Preference Selection
por: Gupta, Taneesh, et al.
Publicado: (2025)
por: Gupta, Taneesh, et al.
Publicado: (2025)
REFA: Reference Free Alignment for multi-preference optimization
por: Gupta, Taneesh, et al.
Publicado: (2024)
por: Gupta, Taneesh, et al.
Publicado: (2024)
LEGOMem: Modular Procedural Memory for Multi-agent LLM Systems for Workflow Automation
por: Han, Dongge, et al.
Publicado: (2025)
por: Han, Dongge, et al.
Publicado: (2025)
Multi-Preference Optimization: Generalizing DPO via Set-Level Contrasts
por: Gupta, Taneesh, et al.
Publicado: (2024)
por: Gupta, Taneesh, et al.
Publicado: (2024)
Synergistic Weak-Strong Collaboration by Aligning Preferences
por: Jiao, Yizhu, et al.
Publicado: (2025)
por: Jiao, Yizhu, et al.
Publicado: (2025)
Budget-Aware Agentic Routing via Boundary-Guided Training
por: Zhang, Caiqi, et al.
Publicado: (2026)
por: Zhang, Caiqi, et al.
Publicado: (2026)
Revisiting Transformer Layer Parameterization Through Causal Energy Minimization
por: Xu, Jin, et al.
Publicado: (2026)
por: Xu, Jin, et al.
Publicado: (2026)
X-lifecycle Learning for Cloud Incident Management using LLMs
por: Goel, Drishti, et al.
Publicado: (2024)
por: Goel, Drishti, et al.
Publicado: (2024)
Enhancing Reasoning Capabilities of Small Language Models with Blueprints and Prompt Template Search
por: Han, Dongge, et al.
Publicado: (2025)
por: Han, Dongge, et al.
Publicado: (2025)
SynthAgent: Adapting Web Agents with Synthetic Supervision
por: Wang, Zhaoyang, et al.
Publicado: (2025)
por: Wang, Zhaoyang, et al.
Publicado: (2025)
Semantic Caching of Contextual Summaries for Efficient Question-Answering with Language Models
por: Couturier, Camille, et al.
Publicado: (2025)
por: Couturier, Camille, et al.
Publicado: (2025)
Automated Root Causing of Cloud Incidents using In-Context Learning with GPT-4
por: Zhang, Xuchao, et al.
Publicado: (2024)
por: Zhang, Xuchao, et al.
Publicado: (2024)
TurboAttention: Efficient Attention Approximation For High Throughputs LLMs
por: Kang, Hao, et al.
Publicado: (2024)
por: Kang, Hao, et al.
Publicado: (2024)
AIOpsLab: A Holistic Framework to Evaluate AI Agents for Enabling Autonomous Clouds
por: Chen, Yinfang, et al.
Publicado: (2025)
por: Chen, Yinfang, et al.
Publicado: (2025)
LLM Reasoning as Trajectories: Step-Specific Representation Geometry and Correctness Signals
por: Sun, Lihao, et al.
Publicado: (2026)
por: Sun, Lihao, et al.
Publicado: (2026)
Minerva: A Programmable Memory Test Benchmark for Language Models
por: Xia, Menglin, et al.
Publicado: (2025)
por: Xia, Menglin, et al.
Publicado: (2025)
Cost-Aware Retrieval-Augmentation Reasoning Models with Adaptive Retrieval Depth
por: Hashemi, Helia, et al.
Publicado: (2025)
por: Hashemi, Helia, et al.
Publicado: (2025)
Generative Caching for Structurally Similar Prompts and Responses
por: Chakraborty, Sarthak, et al.
Publicado: (2025)
por: Chakraborty, Sarthak, et al.
Publicado: (2025)
Ensuring Fair LLM Serving Amid Diverse Applications
por: Khan, Redwan Ibne Seraj, et al.
Publicado: (2024)
por: Khan, Redwan Ibne Seraj, et al.
Publicado: (2024)
AutoAdapt: An Automated Domain Adaptation Framework for LLMs
por: Sinha, Sidharth, et al.
Publicado: (2026)
por: Sinha, Sidharth, et al.
Publicado: (2026)
Debugging the Debuggers: Failure-Anchored Structured Recovery for Software Engineering Agents
por: Zhao, Chenyu, et al.
Publicado: (2026)
por: Zhao, Chenyu, et al.
Publicado: (2026)
WebXSkill: Skill Learning for Autonomous Web Agents
por: Wang, Zhaoyang, et al.
Publicado: (2026)
por: Wang, Zhaoyang, et al.
Publicado: (2026)
Intelligent Router for LLM Workloads: Improving Performance Through Workload-Aware Load Balancing
por: Jain, Kunal, et al.
Publicado: (2024)
por: Jain, Kunal, et al.
Publicado: (2024)
Building AI Agents for Autonomous Clouds: Challenges and Design Principles
por: Shetty, Manish, et al.
Publicado: (2024)
por: Shetty, Manish, et al.
Publicado: (2024)
From Reasoning to Answer: Empirical, Attention-Based and Mechanistic Insights into Distilled DeepSeek R1 Models
por: Zhang, Jue, et al.
Publicado: (2025)
por: Zhang, Jue, et al.
Publicado: (2025)
Attention Enhanced Entity Recommendation for Intelligent Monitoring in Cloud Systems
por: Hussain, Fiza, et al.
Publicado: (2025)
por: Hussain, Fiza, et al.
Publicado: (2025)
ACON: Optimizing Context Compression for Long-horizon LLM Agents
por: Kang, Minki, et al.
Publicado: (2025)
por: Kang, Minki, et al.
Publicado: (2025)
COIN: Chance-Constrained Imitation Learning for Uncertainty-aware Adaptive Resource Oversubscription Policy
por: Wang, Lu, et al.
Publicado: (2024)
por: Wang, Lu, et al.
Publicado: (2024)
Exploring How LLMs Capture and Represent Domain-Specific Knowledge
por: Garcia, Mirian Hipolito, et al.
Publicado: (2025)
por: Garcia, Mirian Hipolito, et al.
Publicado: (2025)
Simulating Environments with Reasoning Models for Agent Training
por: Li, Yuetai, et al.
Publicado: (2025)
por: Li, Yuetai, et al.
Publicado: (2025)
Exploring LLM-based Agents for Root Cause Analysis
por: Roy, Devjeet, et al.
Publicado: (2024)
por: Roy, Devjeet, et al.
Publicado: (2024)
Clustering High-dimensional Data: Balancing Abstraction and Representation Tutorial at AAAI 2026
por: Plant, Claudia, et al.
Publicado: (2026)
por: Plant, Claudia, et al.
Publicado: (2026)
Provable and Practical In-Context Policy Optimization for Self-Improvement
por: Yu, Tianrun, et al.
Publicado: (2026)
por: Yu, Tianrun, et al.
Publicado: (2026)
SageServe: Optimizing LLM Serving on Cloud Data Centers with Forecast Aware Auto-Scaling
por: Jaiswal, Shashwat, et al.
Publicado: (2025)
por: Jaiswal, Shashwat, et al.
Publicado: (2025)
CARMO: Dynamic Criteria Generation for Context-Aware Reward Modelling
por: Gupta, Taneesh, et al.
Publicado: (2024)
por: Gupta, Taneesh, et al.
Publicado: (2024)
eARCO: Efficient Automated Root Cause Analysis with Prompt Optimization
por: Goel, Drishti, et al.
Publicado: (2025)
por: Goel, Drishti, et al.
Publicado: (2025)
Contrastive Attribution in the Wild: An Interpretability Analysis of LLM Failures on Realistic Benchmarks
por: Tan, Rongyuan, et al.
Publicado: (2026)
por: Tan, Rongyuan, et al.
Publicado: (2026)
EcoAct: Economic Agent Determines When to Register What Action
por: Zhang, Shaokun, et al.
Publicado: (2024)
por: Zhang, Shaokun, et al.
Publicado: (2024)
Neural Causal Abstractions
por: Xia, Kevin, et al.
Publicado: (2024)
por: Xia, Kevin, et al.
Publicado: (2024)
Ejemplares similares
-
Hybrid-RACA: Hybrid Retrieval-Augmented Composition Assistance for Real-time Text Prediction
por: Xia, Menglin, et al.
Publicado: (2023) -
AMPO: Active Multi-Preference Optimization for Self-play Preference Selection
por: Gupta, Taneesh, et al.
Publicado: (2025) -
REFA: Reference Free Alignment for multi-preference optimization
por: Gupta, Taneesh, et al.
Publicado: (2024) -
LEGOMem: Modular Procedural Memory for Multi-agent LLM Systems for Workflow Automation
por: Han, Dongge, et al.
Publicado: (2025) -
Multi-Preference Optimization: Generalizing DPO via Set-Level Contrasts
por: Gupta, Taneesh, et al.
Publicado: (2024)