Convergence Theory for Iterative LLM-Based Neural Architecture Search: A Parametric Cross-Entropy Framework with Closed-Form Proxy Reliability
Fuente:
arXiv
Guardado en:
| Autores principales: | Adhikari, Santosh Premi, Timofte, Radu, Ignatov, Dmitry |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Delta-Based Neural Architecture Search: LLM Fine-Tuning via Code Diffs
por: Adhikari, Santosh Premi, et al.
Publicado: (2026)
por: Adhikari, Santosh Premi, et al.
Publicado: (2026)
Resource-Efficient Iterative LLM-Based NAS with Feedback Memory
por: Gu, Xiaojie, et al.
Publicado: (2026)
por: Gu, Xiaojie, et al.
Publicado: (2026)
From Memorization to Creativity: LLM as a Designer of Novel Neural Architectures
por: Khalid, Waleed, et al.
Publicado: (2026)
por: Khalid, Waleed, et al.
Publicado: (2026)
Preparation of Fractal-Inspired Computational Architectures for Advanced Large Language Model Analysis
por: Mittal, Yash, et al.
Publicado: (2025)
por: Mittal, Yash, et al.
Publicado: (2025)
LLM as a Neural Architect: Controlled Generation of Image Captioning Models Under Strict API Contracts
por: Jesani, Krunal, et al.
Publicado: (2025)
por: Jesani, Krunal, et al.
Publicado: (2025)
Virtually Enriched NYU Depth V2 Dataset for Monocular Depth Estimation: Do We Need Artificial Augmentation?
por: Ignatov, Dmitry, et al.
Publicado: (2024)
por: Ignatov, Dmitry, et al.
Publicado: (2024)
From Code to Prediction: Fine-Tuning LLMs for Neural Network Performance Classification in NNGPT
por: Hanouneh, Mahmoud, et al.
Publicado: (2026)
por: Hanouneh, Mahmoud, et al.
Publicado: (2026)
From Brute Force to Semantic Insight: Performance-Guided Data Transformation Design with LLMs
por: Shrestha, Usha, et al.
Publicado: (2026)
por: Shrestha, Usha, et al.
Publicado: (2026)
Factorized Latent Dynamics for Video JEPA: An Empirical Study of Auxiliary Objectives
por: Premi, Santosh
Publicado: (2026)
por: Premi, Santosh
Publicado: (2026)
Real Image Denoising with Knowledge Distillation for High-Performance Mobile NPUs
por: Kayani, Faraz, et al.
Publicado: (2026)
por: Kayani, Faraz, et al.
Publicado: (2026)
Closed-Loop LLM Discovery of Non-Standard Channel Priors in Vision Models
por: Uzun, Tolgay Atinc, et al.
Publicado: (2026)
por: Uzun, Tolgay Atinc, et al.
Publicado: (2026)
VIST-GPT: Ushering in the Era of Visual Storytelling with LLMs?
por: Gado, Mohamed, et al.
Publicado: (2025)
por: Gado, Mohamed, et al.
Publicado: (2025)
Enhancing LLM-Based Neural Network Generation: Few-Shot Prompting and Efficient Validation for Automated Architecture Design
por: Duvvuri, Raghuvir, et al.
Publicado: (2025)
por: Duvvuri, Raghuvir, et al.
Publicado: (2025)
Optuna vs Code Llama: Are LLMs a New Paradigm for Hyperparameter Tuning?
por: Kochnev, Roman, et al.
Publicado: (2025)
por: Kochnev, Roman, et al.
Publicado: (2025)
A Retrieval-Augmented Generation Approach to Extracting Algorithmic Logic from Neural Networks
por: Khalid, Waleed, et al.
Publicado: (2025)
por: Khalid, Waleed, et al.
Publicado: (2025)
OptiProxy-NAS: Optimization Proxy based End-to-End Neural Architecture Search
por: Lyu, Bo, et al.
Publicado: (2025)
por: Lyu, Bo, et al.
Publicado: (2025)
Closing the Gap: Achieving Global Convergence (Last Iterate) of Actor-Critic under Markovian Sampling with Neural Network Parametrization
por: Gaur, Mudit, et al.
Publicado: (2024)
por: Gaur, Mudit, et al.
Publicado: (2024)
AugmentGest: Can Random Data Cropping Augmentation Boost Gesture Recognition Performance?
por: Aboudeshish, Nada, et al.
Publicado: (2025)
por: Aboudeshish, Nada, et al.
Publicado: (2025)
On The Global Convergence Of Online RLHF With Neural Parametrization
por: Gaur, Mudit, et al.
Publicado: (2024)
por: Gaur, Mudit, et al.
Publicado: (2024)
Learning Transformer-based World Models with Contrastive Predictive Coding
por: Burchi, Maxime, et al.
Publicado: (2025)
por: Burchi, Maxime, et al.
Publicado: (2025)
Accurate and Efficient World Modeling with Masked Latent Transformers
por: Burchi, Maxime, et al.
Publicado: (2025)
por: Burchi, Maxime, et al.
Publicado: (2025)
LLM-NAS: LLM-driven Hardware-Aware Neural Architecture Search
por: Zhu, Hengyi, et al.
Publicado: (2025)
por: Zhu, Hengyi, et al.
Publicado: (2025)
CrossNAS: A Cross-Layer Neural Architecture Search Framework for PIM Systems
por: Amin, Md Hasibul, et al.
Publicado: (2025)
por: Amin, Md Hasibul, et al.
Publicado: (2025)
SeqNAS: Neural Architecture Search for Event Sequence Classification
por: Udovichenko, Igor, et al.
Publicado: (2024)
por: Udovichenko, Igor, et al.
Publicado: (2024)
Iterate to Accelerate: A Unified Framework for Iterative Reasoning and Feedback Convergence
por: Fein-Ashley, Jacob
Publicado: (2025)
por: Fein-Ashley, Jacob
Publicado: (2025)
ProxyKV: Cross-Model Proxy Pruning for Efficient Long-Context LLM Inference
por: Li, Junjie, et al.
Publicado: (2026)
por: Li, Junjie, et al.
Publicado: (2026)
TrashDet: Iterative Neural Architecture Search for Efficient Waste Detection
por: Tran, Tony, et al.
Publicado: (2025)
por: Tran, Tony, et al.
Publicado: (2025)
Zero-Shot Neural Architecture Search: Challenges, Solutions, and Opportunities
por: Li, Guihong, et al.
Publicado: (2023)
por: Li, Guihong, et al.
Publicado: (2023)
MobileAgeNet: Lightweight Facial Age Estimation for Mobile Deployment
por: Kumar, Arun, et al.
Publicado: (2026)
por: Kumar, Arun, et al.
Publicado: (2026)
Composer: A Search Framework for Hybrid Neural Architecture Design
por: Acun, Bilge, et al.
Publicado: (2025)
por: Acun, Bilge, et al.
Publicado: (2025)
TG-NAS: Generalizable Zero-Cost Proxies with Operator Description Embedding and Graph Learning for Efficient Neural Architecture Search
por: Qiao, Ye, et al.
Publicado: (2024)
por: Qiao, Ye, et al.
Publicado: (2024)
AZ-NAS: Assembling Zero-Cost Proxies for Network Architecture Search
por: Lee, Junghyup, et al.
Publicado: (2024)
por: Lee, Junghyup, et al.
Publicado: (2024)
Provably Reliable Classifier Guidance via Cross-Entropy Control
por: Sahu, Sharan, et al.
Publicado: (2026)
por: Sahu, Sharan, et al.
Publicado: (2026)
Exploring the Relationship between Brain Hemisphere States and Frequency Bands through Classical Machine Learning and Deep Learning Optimization Techniques with Neurofeedback
por: Islam, Robiul, et al.
Publicado: (2025)
por: Islam, Robiul, et al.
Publicado: (2025)
LEMUR Neural Network Dataset: Towards Seamless AutoML
por: Goodarzi, Arash Torabi, et al.
Publicado: (2025)
por: Goodarzi, Arash Torabi, et al.
Publicado: (2025)
Reliable Policy Iteration: Performance Robustness Across Architecture and Environment Perturbations
por: Eshwar, S. R., et al.
Publicado: (2025)
por: Eshwar, S. R., et al.
Publicado: (2025)
Hybrid Iterative Solvers with Geometry-Aware Neural Preconditioners for Parametric PDEs
por: Lee, Youngkyu, et al.
Publicado: (2025)
por: Lee, Youngkyu, et al.
Publicado: (2025)
MLIPilot: LLM-Driven Auto-Research for Machine-Learned Interatomic Potentials
por: Osaro, Etinosa, et al.
Publicado: (2026)
por: Osaro, Etinosa, et al.
Publicado: (2026)
Structured Progressive Knowledge Activation for LLM-Driven Neural Architecture Search
por: Liu, Zhen, et al.
Publicado: (2026)
por: Liu, Zhen, et al.
Publicado: (2026)
NNGPT: Rethinking AutoML with Large Language Models
por: Kochnev, Roman, et al.
Publicado: (2025)
por: Kochnev, Roman, et al.
Publicado: (2025)
Ejemplares similares
-
Delta-Based Neural Architecture Search: LLM Fine-Tuning via Code Diffs
por: Adhikari, Santosh Premi, et al.
Publicado: (2026) -
Resource-Efficient Iterative LLM-Based NAS with Feedback Memory
por: Gu, Xiaojie, et al.
Publicado: (2026) -
From Memorization to Creativity: LLM as a Designer of Novel Neural Architectures
por: Khalid, Waleed, et al.
Publicado: (2026) -
Preparation of Fractal-Inspired Computational Architectures for Advanced Large Language Model Analysis
por: Mittal, Yash, et al.
Publicado: (2025) -
LLM as a Neural Architect: Controlled Generation of Image Captioning Models Under Strict API Contracts
por: Jesani, Krunal, et al.
Publicado: (2025)