Automated Model Discovery via Multi-modal & Multi-step Pipeline
Fuente:
arXiv
Saved in:
| Main Authors: | Jung-Mok, Lee, Hyeon-Woo, Nam, Ye-Bin, Moon, Nam, Junhyun, Oh, Tae-Hyun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VLM's Eye Examination: Instruct and Inspect Visual Competency of Vision Language Models
by: Hyeon-Woo, Nam, et al.
Published: (2024)
by: Hyeon-Woo, Nam, et al.
Published: (2024)
SMILE-Next: Teaching Large Language Models to Detect, Classify, and Reason about Laughter
by: Jung-Mok, Lee, et al.
Published: (2026)
by: Jung-Mok, Lee, et al.
Published: (2026)
BEAF: Observing BEfore-AFter Changes to Evaluate Hallucination in Vision-language Models
by: Ye-Bin, Moon, et al.
Published: (2024)
by: Ye-Bin, Moon, et al.
Published: (2024)
A Language-Guided Bayesian Optimization for Efficient LoRA Hyperparameter Search
by: Seong-Eun, Baek, et al.
Published: (2026)
by: Seong-Eun, Baek, et al.
Published: (2026)
Scratching Visual Transformer's Back with Uniform Attention
by: Hyeon-Woo, Nam, et al.
Published: (2022)
by: Hyeon-Woo, Nam, et al.
Published: (2022)
EcoScaleNet: A Lightweight Multi Kernel Network for Long Sequence 12 lead ECG Classification
by: Kang, Dong-Hyeon, et al.
Published: (2025)
by: Kang, Dong-Hyeon, et al.
Published: (2025)
mEOL: Training-Free Instruction-Guided Multimodal Embedder for Vector Graphics and Image Retrieval
by: Kim, Kyeong Seon, et al.
Published: (2026)
by: Kim, Kyeong Seon, et al.
Published: (2026)
Early Failure Detection and Intervention in Video Diffusion Models
by: Byung-Ki, Kwon, et al.
Published: (2026)
by: Byung-Ki, Kwon, et al.
Published: (2026)
VizECGNet: Visual ECG Image Network for Cardiovascular Diseases Classification with Multi-Modal Training and Knowledge Distillation
by: Nam, Ju-Hyeon, et al.
Published: (2024)
by: Nam, Ju-Hyeon, et al.
Published: (2024)
Learning Correlation-aware Aleatoric Uncertainty for 3D Hand Pose Estimation
by: Chae-Yeon, Lee, et al.
Published: (2025)
by: Chae-Yeon, Lee, et al.
Published: (2025)
SMILE: Multimodal Dataset for Understanding Laughter in Video with Language Models
by: Hyun, Lee, et al.
Published: (2023)
by: Hyun, Lee, et al.
Published: (2023)
MC2SleepNet: Multi-modal Cross-masking with Contrastive Learning for Sleep Stage Classification
by: Na, Younghoon, et al.
Published: (2025)
by: Na, Younghoon, et al.
Published: (2025)
Biomarker Discovery with Quantum Neural Networks: A Case-study in CTLA4-Activation Pathways
by: Nguyen, Nam
Published: (2023)
by: Nguyen, Nam
Published: (2023)
MathReader : Text-to-Speech for Mathematical Documents
by: Hyeon, Sieun, et al.
Published: (2025)
by: Hyeon, Sieun, et al.
Published: (2025)
SYNAuG: Exploiting Synthetic Data for Data Imbalance Problems
by: Ye-Bin, Moon, et al.
Published: (2023)
by: Ye-Bin, Moon, et al.
Published: (2023)
Learning to Better Search with Language Models via Guided Reinforced Self-Training
by: Moon, Seungyong, et al.
Published: (2024)
by: Moon, Seungyong, et al.
Published: (2024)
Spotlight-TTS: Spotlighting the Style via Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech
by: Kim, Nam-Gyu, et al.
Published: (2025)
by: Kim, Nam-Gyu, et al.
Published: (2025)
Visual Delta Generator with Large Multi-modal Models for Semi-supervised Composed Image Retrieval
by: Jang, Young Kyun, et al.
Published: (2024)
by: Jang, Young Kyun, et al.
Published: (2024)
Illuminating Salient Contributions in Neuron Activation with Attribution Equilibrium
by: Nam, Woo-Jeoung, et al.
Published: (2022)
by: Nam, Woo-Jeoung, et al.
Published: (2022)
Towards Better Visualizing the Decision Basis of Networks via Unfold and Conquer Attribution Guidance
by: Hong, Jung-Ho, et al.
Published: (2023)
by: Hong, Jung-Ho, et al.
Published: (2023)
MathSpeech: Leveraging Small LMs for Accurate Conversion in Mathematical Speech-to-Formula
by: Hyeon, Sieun, et al.
Published: (2024)
by: Hyeon, Sieun, et al.
Published: (2024)
Understanding Physical Properties of Unseen Deformable Objects by Leveraging Large Language Models and Robot Actions
by: Park, Changmin, et al.
Published: (2025)
by: Park, Changmin, et al.
Published: (2025)
Multi-View Attention Multiple-Instance Learning Enhanced by LLM Reasoning for Cognitive Distortion Detection
by: Kim, Jun Seo, et al.
Published: (2025)
by: Kim, Jun Seo, et al.
Published: (2025)
LAMB: LLM-based Audio Captioning with Modality Gap Bridging via Cauchy-Schwarz Divergence
by: Lee, Hyeongkeun, et al.
Published: (2026)
by: Lee, Hyeongkeun, et al.
Published: (2026)
M2SFormer: Multi-Spectral and Multi-Scale Attention with Edge-Aware Difficulty Guidance for Image Forgery Localization
by: Nam, Ju-Hyeon, et al.
Published: (2025)
by: Nam, Ju-Hyeon, et al.
Published: (2025)
Multi-modal Multi-kernel Graph Learning for Autism Prediction and Biomarker Discovery
by: Liu, Jin, et al.
Published: (2023)
by: Liu, Jin, et al.
Published: (2023)
QuantEvolve: Automating Quantitative Strategy Discovery through Multi-Agent Evolutionary Framework
by: Yun, Junhyeog, et al.
Published: (2025)
by: Yun, Junhyeog, et al.
Published: (2025)
FacEDiT: Unified Talking Face Editing and Generation via Facial Motion Infilling
by: Sung-Bin, Kim, et al.
Published: (2025)
by: Sung-Bin, Kim, et al.
Published: (2025)
Learning Context-Conditioned Predicate Semantics via Prototype Feedback
by: Jung, NamGyu, et al.
Published: (2026)
by: Jung, NamGyu, et al.
Published: (2026)
GeoEvolve: Automating Geospatial Model Discovery via Multi-Agent Large Language Models
by: Luo, Peng, et al.
Published: (2025)
by: Luo, Peng, et al.
Published: (2025)
Training Greedy Policy for Proposal Batch Selection in Expensive Multi-Objective Combinatorial Optimization
by: Lee, Deokjae, et al.
Published: (2024)
by: Lee, Deokjae, et al.
Published: (2024)
Mitigating Attention Localization in Small Scale: Self-Attention Refinement via One-step Belief Propagation
by: Lee, Nakyung, et al.
Published: (2025)
by: Lee, Nakyung, et al.
Published: (2025)
TransGUNet: Transformer Meets Graph-based Skip Connection for Medical Image Segmentation
by: Nam, Ju-Hyeon, et al.
Published: (2025)
by: Nam, Ju-Hyeon, et al.
Published: (2025)
KPC-cF: Aspect-Based Sentiment Analysis via Implicit-Feature Alignment with Corpus Filtering
by: Nam, Kibeom
Published: (2024)
by: Nam, Kibeom
Published: (2024)
An Automated Multi-modal Evaluation Framework for Mobile Intelligent Assistants Based on Large Language Models and Multi-Agent Collaboration
by: Wang, Meiping, et al.
Published: (2025)
by: Wang, Meiping, et al.
Published: (2025)
Designing and Evaluating Multi-Chatbot Interface for Human-AI Communication: Preliminary Findings from a Persuasion Task
by: Yoon, Sion, et al.
Published: (2024)
by: Yoon, Sion, et al.
Published: (2024)
Multi-modal Generative AI: Multi-modal LLMs, Diffusions, and the Unification
by: Wang, Xin, et al.
Published: (2024)
by: Wang, Xin, et al.
Published: (2024)
HLER: Human-in-the-Loop Economic Research via Multi-Agent Pipelines for Empirical Discovery
by: Zhu, Chen, et al.
Published: (2026)
by: Zhu, Chen, et al.
Published: (2026)
MathBridge: A Large Corpus Dataset for Translating Spoken Mathematical Expressions into $LaTeX$ Formulas for Improved Readability
by: Jung, Kyudan, et al.
Published: (2024)
by: Jung, Kyudan, et al.
Published: (2024)
Federated Learning and RAG Integration: A Scalable Approach for Medical Large Language Models
by: Jung, Jincheol, et al.
Published: (2024)
by: Jung, Jincheol, et al.
Published: (2024)
Similar Items
-
VLM's Eye Examination: Instruct and Inspect Visual Competency of Vision Language Models
by: Hyeon-Woo, Nam, et al.
Published: (2024) -
SMILE-Next: Teaching Large Language Models to Detect, Classify, and Reason about Laughter
by: Jung-Mok, Lee, et al.
Published: (2026) -
BEAF: Observing BEfore-AFter Changes to Evaluate Hallucination in Vision-language Models
by: Ye-Bin, Moon, et al.
Published: (2024) -
A Language-Guided Bayesian Optimization for Efficient LoRA Hyperparameter Search
by: Seong-Eun, Baek, et al.
Published: (2026) -
Scratching Visual Transformer's Back with Uniform Attention
by: Hyeon-Woo, Nam, et al.
Published: (2022)