Salvato in:
| Autori principali: | Zhang, Genghan, Liang, Weixin, Hsu, Olivia, Olukotun, Kunle |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2502.02534 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
AccelOpt: A Self-Improving LLM Agentic System for AI Accelerator Kernel Optimization
di: Zhang, Genghan, et al.
Pubblicazione: (2025)
di: Zhang, Genghan, et al.
Pubblicazione: (2025)
Agentic Plan Caching: Test-Time Memory for Fast and Cost-Efficient LLM Agents
di: Zhang, Qizheng, et al.
Pubblicazione: (2025)
di: Zhang, Qizheng, et al.
Pubblicazione: (2025)
Streaming Tensor Programs: A Streaming Abstraction for Dynamic Parallelism
di: Sohn, Gina, et al.
Pubblicazione: (2025)
di: Sohn, Gina, et al.
Pubblicazione: (2025)
LowRA: Accurate and Efficient LoRA Fine-Tuning of LLMs under 2 Bits
di: Zhou, Zikai, et al.
Pubblicazione: (2025)
di: Zhou, Zikai, et al.
Pubblicazione: (2025)
Compilation of Modular and General Sparse Workspaces
di: Zhang, Genghan, et al.
Pubblicazione: (2024)
di: Zhang, Genghan, et al.
Pubblicazione: (2024)
FuseFlow: A Fusion-Centric Compilation Framework for Sparse Deep Learning on Streaming Dataflow
di: Lacouture, Rubens, et al.
Pubblicazione: (2025)
di: Lacouture, Rubens, et al.
Pubblicazione: (2025)
Agentic Context Engineering: Evolving Contexts for Self-Improving Language Models
di: Zhang, Qizheng, et al.
Pubblicazione: (2025)
di: Zhang, Qizheng, et al.
Pubblicazione: (2025)
DFModel: Design Space Optimization of Large-Scale Systems Exploiting Dataflow Mappings
di: Ko, Sho, et al.
Pubblicazione: (2024)
di: Ko, Sho, et al.
Pubblicazione: (2024)
Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity
di: Liang, Weixin, et al.
Pubblicazione: (2025)
di: Liang, Weixin, et al.
Pubblicazione: (2025)
Cyclotron: Compilation of Recurrences to Distributed and Systolic Architectures
di: Sundram, Shiv, et al.
Pubblicazione: (2025)
di: Sundram, Shiv, et al.
Pubblicazione: (2025)
SSM-RDU: A Reconfigurable Dataflow Unit for Long-Sequence State-Space Models
di: Ko, Sho, et al.
Pubblicazione: (2025)
di: Ko, Sho, et al.
Pubblicazione: (2025)
Combee: Scaling Prompt Learning for Self-Improving Language Model Agents
di: Li, Hanchen, et al.
Pubblicazione: (2026)
di: Li, Hanchen, et al.
Pubblicazione: (2026)
Sieve: Dynamic Expert-Aware PIM Acceleration for Evolving Mixture-of-Experts Models
di: Kim, Jungwoo, et al.
Pubblicazione: (2026)
di: Kim, Jungwoo, et al.
Pubblicazione: (2026)
Implementing and Optimizing the Scaled Dot-Product Attention on Streaming Dataflow
di: Sohn, Gina, et al.
Pubblicazione: (2024)
di: Sohn, Gina, et al.
Pubblicazione: (2024)
Computational Approaches to Understanding Large Language Model Impact on Writing and Information Ecosystems
di: Liang, Weixin
Pubblicazione: (2025)
di: Liang, Weixin
Pubblicazione: (2025)
Canvas: End-to-End Kernel Architecture Search in Neural Networks
di: Zhao, Chenggang, et al.
Pubblicazione: (2023)
di: Zhao, Chenggang, et al.
Pubblicazione: (2023)
One-Eval: An Agentic System for Automated and Traceable LLM Evaluation
di: Shen, Chengyu, et al.
Pubblicazione: (2026)
di: Shen, Chengyu, et al.
Pubblicazione: (2026)
From Replication to Redesign: Exploring Pairwise Comparisons for LLM-Based Peer Review
di: Zhang, Yaohui, et al.
Pubblicazione: (2025)
di: Zhang, Yaohui, et al.
Pubblicazione: (2025)
CATS: Contextually-Aware Thresholding for Sparsity in Large Language Models
di: Lee, Donghyun, et al.
Pubblicazione: (2024)
di: Lee, Donghyun, et al.
Pubblicazione: (2024)
A-MEM: Agentic Memory for LLM Agents
di: Xu, Wujiang, et al.
Pubblicazione: (2025)
di: Xu, Wujiang, et al.
Pubblicazione: (2025)
I-MCTS: Enhancing Agentic AutoML via Introspective Monte Carlo Tree Search
di: Liang, Zujie, et al.
Pubblicazione: (2025)
di: Liang, Zujie, et al.
Pubblicazione: (2025)
GRATH: Gradual Self-Truthifying for Large Language Models
di: Chen, Weixin, et al.
Pubblicazione: (2024)
di: Chen, Weixin, et al.
Pubblicazione: (2024)
Self-rationalization improves LLM as a fine-grained judge
di: Trivedi, Prapti, et al.
Pubblicazione: (2024)
di: Trivedi, Prapti, et al.
Pubblicazione: (2024)
AgenTracer: Who Is Inducing Failure in the LLM Agentic Systems?
di: Zhang, Guibin, et al.
Pubblicazione: (2025)
di: Zhang, Guibin, et al.
Pubblicazione: (2025)
Agentic Operator Generation for ML ASICs
di: Hammond, Alec M., et al.
Pubblicazione: (2025)
di: Hammond, Alec M., et al.
Pubblicazione: (2025)
MobEvolve: An Agentic Self-Evolving Heuristic System for Interpretable Human Mobility Generation
di: He, Junlin, et al.
Pubblicazione: (2026)
di: He, Junlin, et al.
Pubblicazione: (2026)
Oblivion: Self-Adaptive Agentic Memory Control through Decay-Driven Activation
di: Rana, Ashish, et al.
Pubblicazione: (2026)
di: Rana, Ashish, et al.
Pubblicazione: (2026)
AgenticGEO: A Self-Evolving Agentic System for Generative Engine Optimization
di: Yuan, Jiaqi, et al.
Pubblicazione: (2026)
di: Yuan, Jiaqi, et al.
Pubblicazione: (2026)
SelfBudgeter: Adaptive Token Allocation for Efficient LLM Reasoning
di: Li, Zheng, et al.
Pubblicazione: (2025)
di: Li, Zheng, et al.
Pubblicazione: (2025)
The Widespread Adoption of Large Language Model-Assisted Writing Across Society
di: Liang, Weixin, et al.
Pubblicazione: (2025)
di: Liang, Weixin, et al.
Pubblicazione: (2025)
IsolateGPT: An Execution Isolation Architecture for LLM-Based Agentic Systems
di: Wu, Yuhao, et al.
Pubblicazione: (2024)
di: Wu, Yuhao, et al.
Pubblicazione: (2024)
Vectors Are Not Neutral: Sensitive-Information Inference from Exported LLM Representations in Summarization
di: Liu, Weixin, et al.
Pubblicazione: (2026)
di: Liu, Weixin, et al.
Pubblicazione: (2026)
Adaptive Monitoring and Real-World Evaluation of Agentic AI Systems
di: Shukla, Manish
Pubblicazione: (2025)
di: Shukla, Manish
Pubblicazione: (2025)
MARS: Memory-Enhanced Agents with Reflective Self-improvement
di: Liang, Xuechen, et al.
Pubblicazione: (2025)
di: Liang, Xuechen, et al.
Pubblicazione: (2025)
OpenNovelty: An LLM-powered Agentic System for Verifiable Scholarly Novelty Assessment
di: Zhang, Ming, et al.
Pubblicazione: (2026)
di: Zhang, Ming, et al.
Pubblicazione: (2026)
ALRM: Agentic LLM for Robotic Manipulation
di: Santos, Vitor Gaboardi dos, et al.
Pubblicazione: (2026)
di: Santos, Vitor Gaboardi dos, et al.
Pubblicazione: (2026)
A Comprehensive Survey on Benchmarks and Solutions in Software Engineering of LLM-Empowered Agentic System
di: Guo, Jiale, et al.
Pubblicazione: (2025)
di: Guo, Jiale, et al.
Pubblicazione: (2025)
PubMedCausal: A Span-Level Annotated Corpus for Causal Relation Extraction in Biomedical Text
di: Kunle-John, Ifeoluwa, et al.
Pubblicazione: (2026)
di: Kunle-John, Ifeoluwa, et al.
Pubblicazione: (2026)
Position: The Real Barrier to LLM Agent Usability is Agentic ROI
di: Liu, Weiwen, et al.
Pubblicazione: (2025)
di: Liu, Weiwen, et al.
Pubblicazione: (2025)
Beyond Text-to-SQL: An Agentic LLM System for Governed Enterprise Analytics APIs
di: Singh, Gundeep, et al.
Pubblicazione: (2026)
di: Singh, Gundeep, et al.
Pubblicazione: (2026)
Documenti analoghi
-
AccelOpt: A Self-Improving LLM Agentic System for AI Accelerator Kernel Optimization
di: Zhang, Genghan, et al.
Pubblicazione: (2025) -
Agentic Plan Caching: Test-Time Memory for Fast and Cost-Efficient LLM Agents
di: Zhang, Qizheng, et al.
Pubblicazione: (2025) -
Streaming Tensor Programs: A Streaming Abstraction for Dynamic Parallelism
di: Sohn, Gina, et al.
Pubblicazione: (2025) -
LowRA: Accurate and Efficient LoRA Fine-Tuning of LLMs under 2 Bits
di: Zhou, Zikai, et al.
Pubblicazione: (2025) -
Compilation of Modular and General Sparse Workspaces
di: Zhang, Genghan, et al.
Pubblicazione: (2024)