Guardado en:
| Autores principales: | Zhang, Genghan, Liang, Weixin, Hsu, Olivia, Olukotun, Kunle |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2502.02534 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
AccelOpt: A Self-Improving LLM Agentic System for AI Accelerator Kernel Optimization
por: Zhang, Genghan, et al.
Publicado: (2025)
por: Zhang, Genghan, et al.
Publicado: (2025)
Agentic Plan Caching: Test-Time Memory for Fast and Cost-Efficient LLM Agents
por: Zhang, Qizheng, et al.
Publicado: (2025)
por: Zhang, Qizheng, et al.
Publicado: (2025)
Streaming Tensor Programs: A Streaming Abstraction for Dynamic Parallelism
por: Sohn, Gina, et al.
Publicado: (2025)
por: Sohn, Gina, et al.
Publicado: (2025)
LowRA: Accurate and Efficient LoRA Fine-Tuning of LLMs under 2 Bits
por: Zhou, Zikai, et al.
Publicado: (2025)
por: Zhou, Zikai, et al.
Publicado: (2025)
Compilation of Modular and General Sparse Workspaces
por: Zhang, Genghan, et al.
Publicado: (2024)
por: Zhang, Genghan, et al.
Publicado: (2024)
FuseFlow: A Fusion-Centric Compilation Framework for Sparse Deep Learning on Streaming Dataflow
por: Lacouture, Rubens, et al.
Publicado: (2025)
por: Lacouture, Rubens, et al.
Publicado: (2025)
Agentic Context Engineering: Evolving Contexts for Self-Improving Language Models
por: Zhang, Qizheng, et al.
Publicado: (2025)
por: Zhang, Qizheng, et al.
Publicado: (2025)
DFModel: Design Space Optimization of Large-Scale Systems Exploiting Dataflow Mappings
por: Ko, Sho, et al.
Publicado: (2024)
por: Ko, Sho, et al.
Publicado: (2024)
Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity
por: Liang, Weixin, et al.
Publicado: (2025)
por: Liang, Weixin, et al.
Publicado: (2025)
Cyclotron: Compilation of Recurrences to Distributed and Systolic Architectures
por: Sundram, Shiv, et al.
Publicado: (2025)
por: Sundram, Shiv, et al.
Publicado: (2025)
SSM-RDU: A Reconfigurable Dataflow Unit for Long-Sequence State-Space Models
por: Ko, Sho, et al.
Publicado: (2025)
por: Ko, Sho, et al.
Publicado: (2025)
Combee: Scaling Prompt Learning for Self-Improving Language Model Agents
por: Li, Hanchen, et al.
Publicado: (2026)
por: Li, Hanchen, et al.
Publicado: (2026)
Sieve: Dynamic Expert-Aware PIM Acceleration for Evolving Mixture-of-Experts Models
por: Kim, Jungwoo, et al.
Publicado: (2026)
por: Kim, Jungwoo, et al.
Publicado: (2026)
Implementing and Optimizing the Scaled Dot-Product Attention on Streaming Dataflow
por: Sohn, Gina, et al.
Publicado: (2024)
por: Sohn, Gina, et al.
Publicado: (2024)
Computational Approaches to Understanding Large Language Model Impact on Writing and Information Ecosystems
por: Liang, Weixin
Publicado: (2025)
por: Liang, Weixin
Publicado: (2025)
Canvas: End-to-End Kernel Architecture Search in Neural Networks
por: Zhao, Chenggang, et al.
Publicado: (2023)
por: Zhao, Chenggang, et al.
Publicado: (2023)
One-Eval: An Agentic System for Automated and Traceable LLM Evaluation
por: Shen, Chengyu, et al.
Publicado: (2026)
por: Shen, Chengyu, et al.
Publicado: (2026)
From Replication to Redesign: Exploring Pairwise Comparisons for LLM-Based Peer Review
por: Zhang, Yaohui, et al.
Publicado: (2025)
por: Zhang, Yaohui, et al.
Publicado: (2025)
CATS: Contextually-Aware Thresholding for Sparsity in Large Language Models
por: Lee, Donghyun, et al.
Publicado: (2024)
por: Lee, Donghyun, et al.
Publicado: (2024)
A-MEM: Agentic Memory for LLM Agents
por: Xu, Wujiang, et al.
Publicado: (2025)
por: Xu, Wujiang, et al.
Publicado: (2025)
I-MCTS: Enhancing Agentic AutoML via Introspective Monte Carlo Tree Search
por: Liang, Zujie, et al.
Publicado: (2025)
por: Liang, Zujie, et al.
Publicado: (2025)
GRATH: Gradual Self-Truthifying for Large Language Models
por: Chen, Weixin, et al.
Publicado: (2024)
por: Chen, Weixin, et al.
Publicado: (2024)
Self-rationalization improves LLM as a fine-grained judge
por: Trivedi, Prapti, et al.
Publicado: (2024)
por: Trivedi, Prapti, et al.
Publicado: (2024)
AgenTracer: Who Is Inducing Failure in the LLM Agentic Systems?
por: Zhang, Guibin, et al.
Publicado: (2025)
por: Zhang, Guibin, et al.
Publicado: (2025)
Agentic Operator Generation for ML ASICs
por: Hammond, Alec M., et al.
Publicado: (2025)
por: Hammond, Alec M., et al.
Publicado: (2025)
MobEvolve: An Agentic Self-Evolving Heuristic System for Interpretable Human Mobility Generation
por: He, Junlin, et al.
Publicado: (2026)
por: He, Junlin, et al.
Publicado: (2026)
Oblivion: Self-Adaptive Agentic Memory Control through Decay-Driven Activation
por: Rana, Ashish, et al.
Publicado: (2026)
por: Rana, Ashish, et al.
Publicado: (2026)
AgenticGEO: A Self-Evolving Agentic System for Generative Engine Optimization
por: Yuan, Jiaqi, et al.
Publicado: (2026)
por: Yuan, Jiaqi, et al.
Publicado: (2026)
SelfBudgeter: Adaptive Token Allocation for Efficient LLM Reasoning
por: Li, Zheng, et al.
Publicado: (2025)
por: Li, Zheng, et al.
Publicado: (2025)
The Widespread Adoption of Large Language Model-Assisted Writing Across Society
por: Liang, Weixin, et al.
Publicado: (2025)
por: Liang, Weixin, et al.
Publicado: (2025)
IsolateGPT: An Execution Isolation Architecture for LLM-Based Agentic Systems
por: Wu, Yuhao, et al.
Publicado: (2024)
por: Wu, Yuhao, et al.
Publicado: (2024)
Vectors Are Not Neutral: Sensitive-Information Inference from Exported LLM Representations in Summarization
por: Liu, Weixin, et al.
Publicado: (2026)
por: Liu, Weixin, et al.
Publicado: (2026)
Adaptive Monitoring and Real-World Evaluation of Agentic AI Systems
por: Shukla, Manish
Publicado: (2025)
por: Shukla, Manish
Publicado: (2025)
MARS: Memory-Enhanced Agents with Reflective Self-improvement
por: Liang, Xuechen, et al.
Publicado: (2025)
por: Liang, Xuechen, et al.
Publicado: (2025)
OpenNovelty: An LLM-powered Agentic System for Verifiable Scholarly Novelty Assessment
por: Zhang, Ming, et al.
Publicado: (2026)
por: Zhang, Ming, et al.
Publicado: (2026)
ALRM: Agentic LLM for Robotic Manipulation
por: Santos, Vitor Gaboardi dos, et al.
Publicado: (2026)
por: Santos, Vitor Gaboardi dos, et al.
Publicado: (2026)
A Comprehensive Survey on Benchmarks and Solutions in Software Engineering of LLM-Empowered Agentic System
por: Guo, Jiale, et al.
Publicado: (2025)
por: Guo, Jiale, et al.
Publicado: (2025)
PubMedCausal: A Span-Level Annotated Corpus for Causal Relation Extraction in Biomedical Text
por: Kunle-John, Ifeoluwa, et al.
Publicado: (2026)
por: Kunle-John, Ifeoluwa, et al.
Publicado: (2026)
Position: The Real Barrier to LLM Agent Usability is Agentic ROI
por: Liu, Weiwen, et al.
Publicado: (2025)
por: Liu, Weiwen, et al.
Publicado: (2025)
Beyond Text-to-SQL: An Agentic LLM System for Governed Enterprise Analytics APIs
por: Singh, Gundeep, et al.
Publicado: (2026)
por: Singh, Gundeep, et al.
Publicado: (2026)
Ejemplares similares
-
AccelOpt: A Self-Improving LLM Agentic System for AI Accelerator Kernel Optimization
por: Zhang, Genghan, et al.
Publicado: (2025) -
Agentic Plan Caching: Test-Time Memory for Fast and Cost-Efficient LLM Agents
por: Zhang, Qizheng, et al.
Publicado: (2025) -
Streaming Tensor Programs: A Streaming Abstraction for Dynamic Parallelism
por: Sohn, Gina, et al.
Publicado: (2025) -
LowRA: Accurate and Efficient LoRA Fine-Tuning of LLMs under 2 Bits
por: Zhou, Zikai, et al.
Publicado: (2025) -
Compilation of Modular and General Sparse Workspaces
por: Zhang, Genghan, et al.
Publicado: (2024)