Composer: A Search Framework for Hybrid Neural Architecture Design
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Acun, Bilge, Sinha, Prasoon, Ardalani, Newsha, Bae, Sangmin, Golden, Alicia, Lin, Chien-Yu, Madhyastha, Meghana, Sun, Fei, Yadwadkar, Neeraja J., Wu, Carole-Jean |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
To 2:4 Sparsity and Beyond: Neuron-level Activation Function to Accelerate LLM Pre-Training
von: Madhyastha, Meghana, et al.
Veröffentlicht: (2026)
von: Madhyastha, Meghana, et al.
Veröffentlicht: (2026)
MAD Max Beyond Single-Node: Enabling Large Machine Learning Model Acceleration on Distributed Systems
von: Hsia, Samuel, et al.
Veröffentlicht: (2023)
von: Hsia, Samuel, et al.
Veröffentlicht: (2023)
Hybrid Architectures for Language Models: Systematic Analysis and Design Insights
von: Bae, Sangmin, et al.
Veröffentlicht: (2025)
von: Bae, Sangmin, et al.
Veröffentlicht: (2025)
Shabari: Delayed Decision-Making for Faster and Efficient Serverless Functions
von: Sinha, Prasoon, et al.
Veröffentlicht: (2024)
von: Sinha, Prasoon, et al.
Veröffentlicht: (2024)
Agentic Discovery of Neural Architectures: AIRA-Compose and AIRA-Design
von: Pepe, Alberto, et al.
Veröffentlicht: (2026)
von: Pepe, Alberto, et al.
Veröffentlicht: (2026)
iServe: An Intent-based Serving System for LLMs
von: Liakopoulos, Dimitrios, et al.
Veröffentlicht: (2025)
von: Liakopoulos, Dimitrios, et al.
Veröffentlicht: (2025)
CATransformers: Carbon Aware Transformers Through Joint Model-Hardware Optimization
von: Wang, Irene, et al.
Veröffentlicht: (2025)
von: Wang, Irene, et al.
Veröffentlicht: (2025)
LiVeAction: a Lightweight, Versatile, and Asymmetric Neural Codec Design for Real-time Operation
von: Jacobellis, Dan, et al.
Veröffentlicht: (2026)
von: Jacobellis, Dan, et al.
Veröffentlicht: (2026)
Beyond Efficiency: Scaling AI Sustainably
von: Wu, Carole-Jean, et al.
Veröffentlicht: (2024)
von: Wu, Carole-Jean, et al.
Veröffentlicht: (2024)
FRAPPE: Full Input, Residual Output Autoencoding with Projection Pursuit Encoder
von: Jacobellis, Dan, et al.
Veröffentlicht: (2026)
von: Jacobellis, Dan, et al.
Veröffentlicht: (2026)
Learned Compression for Compressed Learning
von: Jacobellis, Dan, et al.
Veröffentlicht: (2024)
von: Jacobellis, Dan, et al.
Veröffentlicht: (2024)
Oneiros: KV Cache Optimization through Parameter Remapping for Multi-tenant LLM Serving
von: Li, Ruihao, et al.
Veröffentlicht: (2025)
von: Li, Ruihao, et al.
Veröffentlicht: (2025)
Quagmires in SFT-RL Post-Training: When High SFT Scores Mislead and What to Use Instead
von: Kang, Feiyang, et al.
Veröffentlicht: (2025)
von: Kang, Feiyang, et al.
Veröffentlicht: (2025)
Unlocking the Potential of Renewable Energy Through Curtailment Prediction
von: Acun, Bilge, et al.
Veröffentlicht: (2024)
von: Acun, Bilge, et al.
Veröffentlicht: (2024)
Machine Perceptual Quality: Evaluating the Impact of Severe Lossy Compression on Audio and Image Models
von: Jacobellis, Dan, et al.
Veröffentlicht: (2024)
von: Jacobellis, Dan, et al.
Veröffentlicht: (2024)
Is Flash Attention Stable?
von: Golden, Alicia, et al.
Veröffentlicht: (2024)
von: Golden, Alicia, et al.
Veröffentlicht: (2024)
Generative AI Beyond LLMs: System Implications of Multi-Modal Generation
von: Golden, Alicia, et al.
Veröffentlicht: (2023)
von: Golden, Alicia, et al.
Veröffentlicht: (2023)
Masked Matrix Multiplication for Emergent Sparsity
von: Wheatman, Brian, et al.
Veröffentlicht: (2024)
von: Wheatman, Brian, et al.
Veröffentlicht: (2024)
Old is Gold: Optimizing Single-threaded Applications with Exgen-Malloc
von: Li, Ruihao, et al.
Veröffentlicht: (2025)
von: Li, Ruihao, et al.
Veröffentlicht: (2025)
Gecko: An Efficient Neural Architecture Inherently Processing Sequences with Arbitrary Lengths
von: Ma, Xuezhe, et al.
Veröffentlicht: (2026)
von: Ma, Xuezhe, et al.
Veröffentlicht: (2026)
SPEC CPU2026: Characterization, Representativeness, and Cross-Suite Comparison
von: Li, Ruihao, et al.
Veröffentlicht: (2026)
von: Li, Ruihao, et al.
Veröffentlicht: (2026)
A Cognitively Grounded Bayesian Framework for Misinformation Susceptibility
von: Madhyastha, Pranava
Veröffentlicht: (2026)
von: Madhyastha, Pranava
Veröffentlicht: (2026)
Demystifying Synthetic Data in LLM Pre-training: A Systematic Study of Scaling Laws, Benefits, and Pitfalls
von: Kang, Feiyang, et al.
Veröffentlicht: (2025)
von: Kang, Feiyang, et al.
Veröffentlicht: (2025)
Scalable LLM Reasoning Acceleration with Low-rank Distillation
von: Dong, Harry, et al.
Veröffentlicht: (2025)
von: Dong, Harry, et al.
Veröffentlicht: (2025)
Simulating Rumor Spreading in Social Networks using LLM Agents
von: Hu, Tianrui, et al.
Veröffentlicht: (2025)
von: Hu, Tianrui, et al.
Veröffentlicht: (2025)
DeDelayed: Deleting Remote Inference Delay via On-Device Correction
von: Jacobellis, Dan, et al.
Veröffentlicht: (2025)
von: Jacobellis, Dan, et al.
Veröffentlicht: (2025)
CHAI: Clustered Head Attention for Efficient LLM Inference
von: Agarwal, Saurabh, et al.
Veröffentlicht: (2024)
von: Agarwal, Saurabh, et al.
Veröffentlicht: (2024)
Annotating the Pangenome Reveals the Diversity in the Genetic Basis for Metabolic Enzymes
von: Ardalani, Omid
Veröffentlicht: (2025)
von: Ardalani, Omid
Veröffentlicht: (2025)
Accelerating Large Language Model Inference via Early-Exiting Algorithms
von: Bae, Sangmin
Veröffentlicht: (2025)
von: Bae, Sangmin
Veröffentlicht: (2025)
Sieve: Multimodal Dataset Pruning Using Image Captioning Models
von: Mahmoud, Anas, et al.
Veröffentlicht: (2023)
von: Mahmoud, Anas, et al.
Veröffentlicht: (2023)
The xPU-athalon: Quantifying the Competition of AI Acceleration
von: Golden, Alicia, et al.
Veröffentlicht: (2026)
von: Golden, Alicia, et al.
Veröffentlicht: (2026)
Towards Decentralized and Sustainable Foundation Model Training with the Edge
von: Xue, Leyang, et al.
Veröffentlicht: (2025)
von: Xue, Leyang, et al.
Veröffentlicht: (2025)
Spectroscopic Search for Topological Protection in Open Quantum Hardware: The Dissipative Mixed Hodge Module Approach
von: Saurabh, Prasoon
Veröffentlicht: (2025)
von: Saurabh, Prasoon
Veröffentlicht: (2025)
Case report of high origin of radial, ulnar, and profunda brachii arteries, its clinical implications and review of the literature
von: Sampath Madhyastha
Veröffentlicht: (2009)
von: Sampath Madhyastha
Veröffentlicht: (2009)
Enhancing Gluten‐Free Cake Quality With Germinated and Ungerminated Mung Bean Flours: A Comparative Analysis
von: Sultan Acun
Veröffentlicht: (2026)
von: Sultan Acun
Veröffentlicht: (2026)
SoK: A Systems Perspective on Compound AI Threats and Countermeasures
von: Banerjee, Sarbartha, et al.
Veröffentlicht: (2024)
von: Banerjee, Sarbartha, et al.
Veröffentlicht: (2024)
Access to Periodicals: Search Key versus Keyword.
von: Golden, Susan U., et al.
Veröffentlicht: (1983)
von: Golden, Susan U., et al.
Veröffentlicht: (1983)
Hybrid Quantum-Classical Neural Architecture Search
von: Marchisio, Alberto, et al.
Veröffentlicht: (2026)
von: Marchisio, Alberto, et al.
Veröffentlicht: (2026)
On Harnessing Idle Compute at the Edge for Foundation Model Training
von: Xue, Leyang, et al.
Veröffentlicht: (2025)
von: Xue, Leyang, et al.
Veröffentlicht: (2025)
Q-PhotoNAS: Hybrid Quantum Neural Architecture Search Framework on Photonic Devices
von: Elnakhal, Farah, et al.
Veröffentlicht: (2026)
von: Elnakhal, Farah, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
To 2:4 Sparsity and Beyond: Neuron-level Activation Function to Accelerate LLM Pre-Training
von: Madhyastha, Meghana, et al.
Veröffentlicht: (2026) -
MAD Max Beyond Single-Node: Enabling Large Machine Learning Model Acceleration on Distributed Systems
von: Hsia, Samuel, et al.
Veröffentlicht: (2023) -
Hybrid Architectures for Language Models: Systematic Analysis and Design Insights
von: Bae, Sangmin, et al.
Veröffentlicht: (2025) -
Shabari: Delayed Decision-Making for Faster and Efficient Serverless Functions
von: Sinha, Prasoon, et al.
Veröffentlicht: (2024) -
Agentic Discovery of Neural Architectures: AIRA-Compose and AIRA-Design
von: Pepe, Alberto, et al.
Veröffentlicht: (2026)