VeriLocc: End-to-End Cross-Architecture Register Allocation via LLM
Fuente:
arXiv
Saved in:
| Main Authors: | Jin, Lesheng, Ruan, Zhenyuan, Mai, Haohui, Shang, Jingbo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dissecting the Impact of Mobile DVFS Governors on LLM Inference Performance and Energy Efficiency
by: Zhang, Zongpu, et al.
Published: (2025)
by: Zhang, Zongpu, et al.
Published: (2025)
Horizon-LM: A RAM-Centric Architecture for LLM Training
by: Yuan, Zhengqing, et al.
Published: (2026)
by: Yuan, Zhengqing, et al.
Published: (2026)
AIOS: LLM Agent Operating System
by: Mei, Kai, et al.
Published: (2024)
by: Mei, Kai, et al.
Published: (2024)
Columbo: Low Level End-to-End System Traces through Modular Full-System Simulation
by: Görgen, Jakob, et al.
Published: (2024)
by: Görgen, Jakob, et al.
Published: (2024)
Performance Isolation and Semantic Determinism in Efficient GPU Spatial Sharing
by: Yang, Zhenyuan, et al.
Published: (2026)
by: Yang, Zhenyuan, et al.
Published: (2026)
Cerebrum (AIOS SDK): A Platform for Agent Development, Deployment, Distribution, and Discovery
by: Rama, Balaji, et al.
Published: (2025)
by: Rama, Balaji, et al.
Published: (2025)
RUISA Operational Ecosystem Architecture
by: AL Mohtar, Mouayad
Published: (2026)
by: AL Mohtar, Mouayad
Published: (2026)
A TRRIP Down Memory Lane: Temperature-Based Re-Reference Interval Prediction For Instruction Caching
by: Kao, Henry, et al.
Published: (2025)
by: Kao, Henry, et al.
Published: (2025)
MegaTrain: Full Precision Training of 100B+ Parameter Large Language Models on a Single GPU
by: Yuan, Zhengqing, et al.
Published: (2026)
by: Yuan, Zhengqing, et al.
Published: (2026)
Potential of WebAssembly for Embedded Systems
by: Wallentowitz, Stefan, et al.
Published: (2024)
by: Wallentowitz, Stefan, et al.
Published: (2024)
OBASE: Object-Based Address-Space Engineering to Improve Memory Tiering
by: Banakar, Vinay, et al.
Published: (2026)
by: Banakar, Vinay, et al.
Published: (2026)
Getting a Handle on Unmanaged Memory
by: Wanninger, Nick, et al.
Published: (2024)
by: Wanninger, Nick, et al.
Published: (2024)
ConsumerBench: Benchmarking Generative AI Applications on End-User Devices
by: Gu, Yile, et al.
Published: (2025)
by: Gu, Yile, et al.
Published: (2025)
Scalable and Accurate Application-Level Crash-Consistency Testing via Representative Testing
by: Gu, Yile, et al.
Published: (2025)
by: Gu, Yile, et al.
Published: (2025)
Talyxion: From Speculation to Optimization in Risk Managed Crypto Portfolio Allocation
by: Nguyen, Thanh
Published: (2025)
by: Nguyen, Thanh
Published: (2025)
SARA: A Stall-Aware Memory Allocation Strategy for Mixed-Criticality Systems
by: Lee, Meng-Chia, et al.
Published: (2025)
by: Lee, Meng-Chia, et al.
Published: (2025)
OSVBench: Benchmarking LLMs on Specification Generation Tasks for Operating System Verification
by: Li, Shangyu, et al.
Published: (2025)
by: Li, Shangyu, et al.
Published: (2025)
E-Mapper: Energy-Efficient Resource Allocation for Traditional Operating Systems on Heterogeneous Processors
by: Smejkal, Till, et al.
Published: (2024)
by: Smejkal, Till, et al.
Published: (2024)
Tidying Up the Address Space
by: Banakar, Vinay, et al.
Published: (2025)
by: Banakar, Vinay, et al.
Published: (2025)
Quine: Realizing LLM Agents as Native POSIX Processes
by: Ke, Hao
Published: (2026)
by: Ke, Hao
Published: (2026)
RTOS Architectures that Solve the Diminishing Bandwidth Problem
by: Arakji, Mazen
Published: (2025)
by: Arakji, Mazen
Published: (2025)
CvxCluster: Solving Large, Complex, Granular Resource Allocation Problems 100-1000x Faster
by: Nnorom Jr, Obi, et al.
Published: (2026)
by: Nnorom Jr, Obi, et al.
Published: (2026)
Exploiting Application-to-Architecture Dependencies for Designing Scalable OS
by: Xiao, Yao, et al.
Published: (2025)
by: Xiao, Yao, et al.
Published: (2025)
RTP-LLM: High-Performance Alibaba LLM Inference Engine
by: Tan, Boyu, et al.
Published: (2026)
by: Tan, Boyu, et al.
Published: (2026)
A Task Equalization Allocation Algorithm Incorporating Blocking Estimation and Resource Similarity Analysis for Vehicle Control Real-Time Systems
by: Duan, Qianlong, et al.
Published: (2025)
by: Duan, Qianlong, et al.
Published: (2025)
CORD: Co-design of Resource Allocation and Deadline Decomposition with Generative Profiling
by: Gifford, Robert, et al.
Published: (2025)
by: Gifford, Robert, et al.
Published: (2025)
Pomegranate: A Lightweight Compartmentalization Architecture using Virtualization Extensions
by: Raja, Shriram, et al.
Published: (2026)
by: Raja, Shriram, et al.
Published: (2026)
Reinforcement Learning for Dynamic Memory Allocation
by: Lim, Arisrei, et al.
Published: (2024)
by: Lim, Arisrei, et al.
Published: (2024)
Compiling Away the Overhead of Race Detection
by: Paznikov, Alexey, et al.
Published: (2025)
by: Paznikov, Alexey, et al.
Published: (2025)
Sockeye: a language for analyzing hardware documentation
by: Fiedler, Ben, et al.
Published: (2025)
by: Fiedler, Ben, et al.
Published: (2025)
Futureproof Static Memory Planning
by: Lamprakos, Christos, et al.
Published: (2025)
by: Lamprakos, Christos, et al.
Published: (2025)
vNV-Heap: An Ownership-Based Virtually Non-Volatile Heap for Embedded Systems
by: Gerber, Markus Elias, et al.
Published: (2025)
by: Gerber, Markus Elias, et al.
Published: (2025)
Scaling Inter-procedural Dataflow Analysis on the Cloud
by: Sun, Zewen, et al.
Published: (2024)
by: Sun, Zewen, et al.
Published: (2024)
LLM as a System Service on Mobile Devices
by: Yin, Wangsong, et al.
Published: (2024)
by: Yin, Wangsong, et al.
Published: (2024)
Analyzing Configuration Dependencies of File Systems
by: Mahmud, Tabassum, et al.
Published: (2025)
by: Mahmud, Tabassum, et al.
Published: (2025)
SSV: Sparse Speculative Verification for Efficient LLM Inference
by: Wang, Zhibin, et al.
Published: (2026)
by: Wang, Zhibin, et al.
Published: (2026)
Extending Data Spatial Semantics for Scale Agnostic Programming
by: Mars, Jason
Published: (2025)
by: Mars, Jason
Published: (2025)
Towards Efficient and Practical GPU Multitasking in the Era of LLM
by: Xing, Jiarong, et al.
Published: (2025)
by: Xing, Jiarong, et al.
Published: (2025)
Towards High-Goodput LLM Serving with Prefill-decode Multiplexing
by: Chen, Yukang, et al.
Published: (2025)
by: Chen, Yukang, et al.
Published: (2025)
Don't Let AI Agents YOLO Your Files: Shifting Information and Control to Filesystems for Agent Safety and Autonomy
by: Zhong, Shawn Wanxiang, et al.
Published: (2026)
by: Zhong, Shawn Wanxiang, et al.
Published: (2026)
Similar Items
-
Dissecting the Impact of Mobile DVFS Governors on LLM Inference Performance and Energy Efficiency
by: Zhang, Zongpu, et al.
Published: (2025) -
Horizon-LM: A RAM-Centric Architecture for LLM Training
by: Yuan, Zhengqing, et al.
Published: (2026) -
AIOS: LLM Agent Operating System
by: Mei, Kai, et al.
Published: (2024) -
Columbo: Low Level End-to-End System Traces through Modular Full-System Simulation
by: Görgen, Jakob, et al.
Published: (2024) -
Performance Isolation and Semantic Determinism in Efficient GPU Spatial Sharing
by: Yang, Zhenyuan, et al.
Published: (2026)