Enhancing LLM-Based Neural Network Generation: Few-Shot Prompting and Efficient Validation for Automated Architecture Design
Fuente:
arXiv
Salvato in:
| Autori principali: | Duvvuri, Raghuvir, Vysyaraju, Chandini, Goyal, Avi, Ignatov, Dmitry, Timofte, Radu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
From Memorization to Creativity: LLM as a Designer of Novel Neural Architectures
di: Khalid, Waleed, et al.
Pubblicazione: (2026)
di: Khalid, Waleed, et al.
Pubblicazione: (2026)
NNGPT: Rethinking AutoML with Large Language Models
di: Kochnev, Roman, et al.
Pubblicazione: (2025)
di: Kochnev, Roman, et al.
Pubblicazione: (2025)
Delta-Based Neural Architecture Search: LLM Fine-Tuning via Code Diffs
di: Adhikari, Santosh Premi, et al.
Pubblicazione: (2026)
di: Adhikari, Santosh Premi, et al.
Pubblicazione: (2026)
A Retrieval-Augmented Generation Approach to Extracting Algorithmic Logic from Neural Networks
di: Khalid, Waleed, et al.
Pubblicazione: (2025)
di: Khalid, Waleed, et al.
Pubblicazione: (2025)
LLM as a Neural Architect: Controlled Generation of Image Captioning Models Under Strict API Contracts
di: Jesani, Krunal, et al.
Pubblicazione: (2025)
di: Jesani, Krunal, et al.
Pubblicazione: (2025)
From Code to Prediction: Fine-Tuning LLMs for Neural Network Performance Classification in NNGPT
di: Hanouneh, Mahmoud, et al.
Pubblicazione: (2026)
di: Hanouneh, Mahmoud, et al.
Pubblicazione: (2026)
Preparation of Fractal-Inspired Computational Architectures for Advanced Large Language Model Analysis
di: Mittal, Yash, et al.
Pubblicazione: (2025)
di: Mittal, Yash, et al.
Pubblicazione: (2025)
Virtually Enriched NYU Depth V2 Dataset for Monocular Depth Estimation: Do We Need Artificial Augmentation?
di: Ignatov, Dmitry, et al.
Pubblicazione: (2024)
di: Ignatov, Dmitry, et al.
Pubblicazione: (2024)
From Brute Force to Semantic Insight: Performance-Guided Data Transformation Design with LLMs
di: Shrestha, Usha, et al.
Pubblicazione: (2026)
di: Shrestha, Usha, et al.
Pubblicazione: (2026)
Closed-Loop LLM Discovery of Non-Standard Channel Priors in Vision Models
di: Uzun, Tolgay Atinc, et al.
Pubblicazione: (2026)
di: Uzun, Tolgay Atinc, et al.
Pubblicazione: (2026)
Resource-Efficient Iterative LLM-Based NAS with Feedback Memory
di: Gu, Xiaojie, et al.
Pubblicazione: (2026)
di: Gu, Xiaojie, et al.
Pubblicazione: (2026)
AugmentGest: Can Random Data Cropping Augmentation Boost Gesture Recognition Performance?
di: Aboudeshish, Nada, et al.
Pubblicazione: (2025)
di: Aboudeshish, Nada, et al.
Pubblicazione: (2025)
Convergence Theory for Iterative LLM-Based Neural Architecture Search: A Parametric Cross-Entropy Framework with Closed-Form Proxy Reliability
di: Adhikari, Santosh Premi, et al.
Pubblicazione: (2026)
di: Adhikari, Santosh Premi, et al.
Pubblicazione: (2026)
MobileAgeNet: Lightweight Facial Age Estimation for Mobile Deployment
di: Kumar, Arun, et al.
Pubblicazione: (2026)
di: Kumar, Arun, et al.
Pubblicazione: (2026)
Real Image Denoising with Knowledge Distillation for High-Performance Mobile NPUs
di: Kayani, Faraz, et al.
Pubblicazione: (2026)
di: Kayani, Faraz, et al.
Pubblicazione: (2026)
VIST-GPT: Ushering in the Era of Visual Storytelling with LLMs?
di: Gado, Mohamed, et al.
Pubblicazione: (2025)
di: Gado, Mohamed, et al.
Pubblicazione: (2025)
Accurate and Efficient World Modeling with Masked Latent Transformers
di: Burchi, Maxime, et al.
Pubblicazione: (2025)
di: Burchi, Maxime, et al.
Pubblicazione: (2025)
DiTVR: Zero-Shot Diffusion Transformer for Video Restoration
di: Gao, Sicheng, et al.
Pubblicazione: (2025)
di: Gao, Sicheng, et al.
Pubblicazione: (2025)
Efficient Few-Shot Neural Architecture Search by Counting the Number of Nonlinear Functions
di: Oh, Youngmin, et al.
Pubblicazione: (2024)
di: Oh, Youngmin, et al.
Pubblicazione: (2024)
WaveHiT-SR: Hierarchical Wavelet Network for Efficient Image Super-Resolution
di: Ali, Fayaz, et al.
Pubblicazione: (2025)
di: Ali, Fayaz, et al.
Pubblicazione: (2025)
Automating MedSAM by Learning Prompts with Weak Few-Shot Supervision
di: Gaillochet, Mélanie, et al.
Pubblicazione: (2024)
di: Gaillochet, Mélanie, et al.
Pubblicazione: (2024)
Practical Manipulation Model for Robust Deepfake Detection
di: Hopf, Benedikt, et al.
Pubblicazione: (2025)
di: Hopf, Benedikt, et al.
Pubblicazione: (2025)
Streaming Neural Images
di: Conde, Marcos V., et al.
Pubblicazione: (2024)
di: Conde, Marcos V., et al.
Pubblicazione: (2024)
LEMUR Neural Network Dataset: Towards Seamless AutoML
di: Goodarzi, Arash Torabi, et al.
Pubblicazione: (2025)
di: Goodarzi, Arash Torabi, et al.
Pubblicazione: (2025)
Enhanced Super-Resolution Training via Mimicked Alignment for Real-World Scenes
di: Elezabi, Omar, et al.
Pubblicazione: (2024)
di: Elezabi, Omar, et al.
Pubblicazione: (2024)
AdaptSR: Low-Rank Adaptation for Efficient and Scalable Real-World Super-Resolution
di: Korkmaz, Cansu, et al.
Pubblicazione: (2025)
di: Korkmaz, Cansu, et al.
Pubblicazione: (2025)
MuDreamer: Learning Predictive World Models without Reconstruction
di: Burchi, Maxime, et al.
Pubblicazione: (2024)
di: Burchi, Maxime, et al.
Pubblicazione: (2024)
Learned Lightweight Smartphone ISP with Unpaired Data
di: Arhire, Andrei, et al.
Pubblicazione: (2025)
di: Arhire, Andrei, et al.
Pubblicazione: (2025)
Higher fidelity perceptual image and video compression with a latent conditioned residual denoising diffusion model
di: Brenig, Jonas, et al.
Pubblicazione: (2025)
di: Brenig, Jonas, et al.
Pubblicazione: (2025)
Toward Efficient Deep Blind RAW Image Restoration
di: Conde, Marcos V., et al.
Pubblicazione: (2024)
di: Conde, Marcos V., et al.
Pubblicazione: (2024)
Learning Transformer-based World Models with Contrastive Predictive Coding
di: Burchi, Maxime, et al.
Pubblicazione: (2025)
di: Burchi, Maxime, et al.
Pubblicazione: (2025)
LAFR: Efficient Diffusion-based Blind Face Restoration via Latent Codebook Alignment Adapter
di: Li, Runyi, et al.
Pubblicazione: (2025)
di: Li, Runyi, et al.
Pubblicazione: (2025)
mBLIP: Efficient Bootstrapping of Multilingual Vision-LLMs
di: Geigle, Gregor, et al.
Pubblicazione: (2023)
di: Geigle, Gregor, et al.
Pubblicazione: (2023)
Zero-Shot Neural Architecture Search: Challenges, Solutions, and Opportunities
di: Li, Guihong, et al.
Pubblicazione: (2023)
di: Li, Guihong, et al.
Pubblicazione: (2023)
Enabling Validation for Robust Few-Shot Recognition
di: Wang, Hanxin, et al.
Pubblicazione: (2025)
di: Wang, Hanxin, et al.
Pubblicazione: (2025)
Efficient Perceptual Image Super Resolution: AIM 2025 Study and Benchmark
di: Longarela, Bruno, et al.
Pubblicazione: (2025)
di: Longarela, Bruno, et al.
Pubblicazione: (2025)
APSeg: Auto-Prompt Network for Cross-Domain Few-Shot Semantic Segmentation
di: He, Weizhao, et al.
Pubblicazione: (2024)
di: He, Weizhao, et al.
Pubblicazione: (2024)
Cat: Post-Training Quantization Error Reduction via Cluster-based Affine Transformation
di: Zoljodi, Ali, et al.
Pubblicazione: (2025)
di: Zoljodi, Ali, et al.
Pubblicazione: (2025)
The Return of Structural Handwritten Mathematical Expression Recognition
di: Seitz, Jakob, et al.
Pubblicazione: (2025)
di: Seitz, Jakob, et al.
Pubblicazione: (2025)
The Regularizing Power of Language-Training Deepfake Detectors
di: Hopf, Benedikt, et al.
Pubblicazione: (2026)
di: Hopf, Benedikt, et al.
Pubblicazione: (2026)
Documenti analoghi
-
From Memorization to Creativity: LLM as a Designer of Novel Neural Architectures
di: Khalid, Waleed, et al.
Pubblicazione: (2026) -
NNGPT: Rethinking AutoML with Large Language Models
di: Kochnev, Roman, et al.
Pubblicazione: (2025) -
Delta-Based Neural Architecture Search: LLM Fine-Tuning via Code Diffs
di: Adhikari, Santosh Premi, et al.
Pubblicazione: (2026) -
A Retrieval-Augmented Generation Approach to Extracting Algorithmic Logic from Neural Networks
di: Khalid, Waleed, et al.
Pubblicazione: (2025) -
LLM as a Neural Architect: Controlled Generation of Image Captioning Models Under Strict API Contracts
di: Jesani, Krunal, et al.
Pubblicazione: (2025)