Latency Analysis and Optimization of Alpamayo 1 via Efficient Trajectory Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jeon, Yunseong, Lee, Namcheol, Lee, Yoonsu, Park, Jangwoon, Ahn, Sol, Kim, Jong-Chan, Hong, Seongsoo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
OOM-Free Alpamayo via CPU-GPU Memory Swapping for Vision-Language-Action Models
von: Roh, Seungwoo, et al.
Veröffentlicht: (2026)
von: Roh, Seungwoo, et al.
Veröffentlicht: (2026)
Integrated Framework for LLM Evaluation with Answer Generation
von: Lee, Sujeong, et al.
Veröffentlicht: (2025)
von: Lee, Sujeong, et al.
Veröffentlicht: (2025)
Melon Fruit Detection and Quality Assessment Using Generative AI-Based Image Data Augmentation
von: Yoon, Seungri, et al.
Veröffentlicht: (2024)
von: Yoon, Seungri, et al.
Veröffentlicht: (2024)
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses
von: Lee, Sujeong, et al.
Veröffentlicht: (2025)
von: Lee, Sujeong, et al.
Veröffentlicht: (2025)
PaCA: Partial Connection Adaptation for Efficient Fine-Tuning
von: Woo, Sunghyeon, et al.
Veröffentlicht: (2025)
von: Woo, Sunghyeon, et al.
Veröffentlicht: (2025)
Optimized Layerwise Approximation for Efficient Private Inference on Fully Homomorphic Encryption
von: Lee, Junghyun, et al.
Veröffentlicht: (2023)
von: Lee, Junghyun, et al.
Veröffentlicht: (2023)
Motion Manifold Flow Primitives for Task-Conditioned Trajectory Generation under Complex Task-Motion Dependencies
von: Lee, Yonghyeon, et al.
Veröffentlicht: (2024)
von: Lee, Yonghyeon, et al.
Veröffentlicht: (2024)
Efficient LLM Collaboration via Planning
von: Lee, Byeongchan, et al.
Veröffentlicht: (2025)
von: Lee, Byeongchan, et al.
Veröffentlicht: (2025)
TT-SEAL: TTD-Aware Selective Encryption for Adversarially-Robust and Low-Latency Edge AI
von: Min, Kyeongpil, et al.
Veröffentlicht: (2026)
von: Min, Kyeongpil, et al.
Veröffentlicht: (2026)
Alpamayo-R1: Bridging Reasoning and Action Prediction for Generalizable Autonomous Driving in the Long Tail
von: NVIDIA, et al.
Veröffentlicht: (2025)
von: NVIDIA, et al.
Veröffentlicht: (2025)
Guided Trajectory Generation with Diffusion Models for Offline Model-based Optimization
von: Yun, Taeyoung, et al.
Veröffentlicht: (2024)
von: Yun, Taeyoung, et al.
Veröffentlicht: (2024)
DyLLM: Efficient Diffusion LLM Inference via Saliency-based Token Selection and Partial Attention
von: Lee, Younjoo, et al.
Veröffentlicht: (2026)
von: Lee, Younjoo, et al.
Veröffentlicht: (2026)
ASKD-Whisper: Adaptive Self-knowledge Distillation for Efficient and Low-Latency Automatic Speech Recognition
von: Lee, Junseok, et al.
Veröffentlicht: (2026)
von: Lee, Junseok, et al.
Veröffentlicht: (2026)
Improving SMOTE via Fusing Conditional VAE for Data-adaptive Noise Filtering
von: Hong, Sungchul, et al.
Veröffentlicht: (2024)
von: Hong, Sungchul, et al.
Veröffentlicht: (2024)
Scene Graph Generation Strategy with Co-occurrence Knowledge and Learnable Term Frequency
von: Kim, Hyeongjin, et al.
Veröffentlicht: (2024)
von: Kim, Hyeongjin, et al.
Veröffentlicht: (2024)
Multidimensional Adaptive Coefficient for Inference Trajectory Optimization in Flow and Diffusion
von: Lee, Dohoon, et al.
Veröffentlicht: (2024)
von: Lee, Dohoon, et al.
Veröffentlicht: (2024)
Cold-start Bundle Recommendation via Popularity-based Coalescence and Curriculum Heating
von: Jeon, Hyunsik, et al.
Veröffentlicht: (2023)
von: Jeon, Hyunsik, et al.
Veröffentlicht: (2023)
LLM-Driven Learning Analytics Dashboard for Teachers in EFL Writing Education
von: Kim, Minsun, et al.
Veröffentlicht: (2024)
von: Kim, Minsun, et al.
Veröffentlicht: (2024)
See and Fix the Flaws: Enabling VLMs and Diffusion Models to Comprehend Visual Artifacts via Agentic Data Synthesis
von: Park, Jaehyun, et al.
Veröffentlicht: (2026)
von: Park, Jaehyun, et al.
Veröffentlicht: (2026)
Investigating Long-term Training for Remote Sensing Object Detection
von: Park, JongHyun, et al.
Veröffentlicht: (2024)
von: Park, JongHyun, et al.
Veröffentlicht: (2024)
FlowAlign: Trajectory-Regularized, Inversion-Free Flow-based Image Editing
von: Kim, Jeongsol, et al.
Veröffentlicht: (2025)
von: Kim, Jeongsol, et al.
Veröffentlicht: (2025)
Beyond the Final Answer: Evaluating the Reasoning Trajectories of Tool-Augmented Agents
von: Kim, Wonjoong, et al.
Veröffentlicht: (2025)
von: Kim, Wonjoong, et al.
Veröffentlicht: (2025)
LookaheadKV: Fast and Accurate KV Cache Eviction by Glimpsing into the Future without Generation
von: Ahn, Jinwoo, et al.
Veröffentlicht: (2026)
von: Ahn, Jinwoo, et al.
Veröffentlicht: (2026)
Patch Rebirth: Toward Fast and Transferable Model Inversion of Vision Transformers
von: Heo, Seongsoo, et al.
Veröffentlicht: (2025)
von: Heo, Seongsoo, et al.
Veröffentlicht: (2025)
SMMF: Square-Matricized Momentum Factorization for Memory-Efficient Optimization
von: Park, Kwangryeol, et al.
Veröffentlicht: (2024)
von: Park, Kwangryeol, et al.
Veröffentlicht: (2024)
Unicorn: U-Net for Sea Ice Forecasting with Convolutional Neural Ordinary Differential Equations
von: Park, Jaesung, et al.
Veröffentlicht: (2024)
von: Park, Jaesung, et al.
Veröffentlicht: (2024)
Unlocking Robust Semantic Segmentation Performance via Label-only Elastic Deformations against Implicit Label Noise
von: Kim, Yechan, et al.
Veröffentlicht: (2025)
von: Kim, Yechan, et al.
Veröffentlicht: (2025)
HH-PIM: Dynamic Optimization of Power and Performance with Heterogeneous-Hybrid PIM for Edge AI Devices
von: Jeon, Sangmin, et al.
Veröffentlicht: (2025)
von: Jeon, Sangmin, et al.
Veröffentlicht: (2025)
Two-Stage Grid Optimization for Group-wise Quantization of LLMs
von: Kim, Junhan, et al.
Veröffentlicht: (2026)
von: Kim, Junhan, et al.
Veröffentlicht: (2026)
GTA: Generative Trajectory Augmentation with Guidance for Offline Reinforcement Learning
von: Lee, Jaewoo, et al.
Veröffentlicht: (2024)
von: Lee, Jaewoo, et al.
Veröffentlicht: (2024)
DSG-KD: Knowledge Distillation from Domain-Specific to General Language Models
von: Cho, Sangyeon, et al.
Veröffentlicht: (2024)
von: Cho, Sangyeon, et al.
Veröffentlicht: (2024)
SLM-Based Agentic AI with P-C-G: Optimized for Korean Tool Use
von: Jeon, Changhyun, et al.
Veröffentlicht: (2025)
von: Jeon, Changhyun, et al.
Veröffentlicht: (2025)
Beyond Line-Level Filtering for the Pretraining Corpora of LLMs
von: Park, Chanwoo, et al.
Veröffentlicht: (2025)
von: Park, Chanwoo, et al.
Veröffentlicht: (2025)
Generalized Consistency Trajectory Models for Image Manipulation
von: Kim, Beomsu, et al.
Veröffentlicht: (2024)
von: Kim, Beomsu, et al.
Veröffentlicht: (2024)
On Predicting Post-Click Conversion Rate via Counterfactual Inference
von: Ahn, Junhyung, et al.
Veröffentlicht: (2025)
von: Ahn, Junhyung, et al.
Veröffentlicht: (2025)
PCPO: Proportionate Credit Policy Optimization for Aligning Image Generation Models
von: Lee, Jeongjae, et al.
Veröffentlicht: (2025)
von: Lee, Jeongjae, et al.
Veröffentlicht: (2025)
"When to Hand Off, When to Work Together": Expanding Human-Agent Co-Creative Collaboration through Concurrent Interaction
von: Son, Kihoon, et al.
Veröffentlicht: (2026)
von: Son, Kihoon, et al.
Veröffentlicht: (2026)
A Two-Step Approach for Data-Efficient French Pronunciation Learning
von: Lee, Hoyeon, et al.
Veröffentlicht: (2024)
von: Lee, Hoyeon, et al.
Veröffentlicht: (2024)
EPIC: Graph Augmentation with Edit Path Interpolation via Learnable Cost
von: Heo, Jaeseung, et al.
Veröffentlicht: (2023)
von: Heo, Jaeseung, et al.
Veröffentlicht: (2023)
PAIR: Prefix-Aware Internal Reward Model for Multi-Turn Agent Optimization
von: Kim, Wonjoong, et al.
Veröffentlicht: (2026)
von: Kim, Wonjoong, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
OOM-Free Alpamayo via CPU-GPU Memory Swapping for Vision-Language-Action Models
von: Roh, Seungwoo, et al.
Veröffentlicht: (2026) -
Integrated Framework for LLM Evaluation with Answer Generation
von: Lee, Sujeong, et al.
Veröffentlicht: (2025) -
Melon Fruit Detection and Quality Assessment Using Generative AI-Based Image Data Augmentation
von: Yoon, Seungri, et al.
Veröffentlicht: (2024) -
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses
von: Lee, Sujeong, et al.
Veröffentlicht: (2025) -
PaCA: Partial Connection Adaptation for Efficient Fine-Tuning
von: Woo, Sunghyeon, et al.
Veröffentlicht: (2025)