Salvato in:
| Autori principali: | Chung, Dahyun, Shin, Donghyun, Sung, Yujin, Moon, Seunggi, Jeon, Jinwoo, Lee, Byung-Jun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2511.13036 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Vision-aligned Latent Reasoning for Multi-modal Large Language Model
di: Jeon, Byungwoo, et al.
Pubblicazione: (2026)
di: Jeon, Byungwoo, et al.
Pubblicazione: (2026)
Iterative Prompt Refinement for Safer Text-to-Image Generation
di: Jeon, Jinwoo, et al.
Pubblicazione: (2025)
di: Jeon, Jinwoo, et al.
Pubblicazione: (2025)
Self-Refining Language Model Anonymizers via Adversarial Distillation
di: Kim, Kyuyoung, et al.
Pubblicazione: (2025)
di: Kim, Kyuyoung, et al.
Pubblicazione: (2025)
AMRG: Extend Vision Language Models for Automatic Mammography Report Generation
di: Sung, Nak-Jun, et al.
Pubblicazione: (2025)
di: Sung, Nak-Jun, et al.
Pubblicazione: (2025)
Co‐Stimuli‐Driven 2D WSe 2 Optoelectronic Synapses for Neuromorphic Computing
di: Junho Sung, et al.
Pubblicazione: (2025)
di: Junho Sung, et al.
Pubblicazione: (2025)
CLIP Meets Diffusion: A Synergistic Approach to Anomaly Detection
di: Lee, Byeongchan, et al.
Pubblicazione: (2025)
di: Lee, Byeongchan, et al.
Pubblicazione: (2025)
Cog3DMap: Multi-View Vision-Language Reasoning with 3D Cognitive Maps
di: Gwak, Chanyoung, et al.
Pubblicazione: (2026)
di: Gwak, Chanyoung, et al.
Pubblicazione: (2026)
ContextVLA: Vision-Language-Action Model with Amortized Multi-Frame Context
di: Jang, Huiwon, et al.
Pubblicazione: (2025)
di: Jang, Huiwon, et al.
Pubblicazione: (2025)
A Parameter-efficient Language Extension Framework for Multilingual ASR
di: Liu, Wei, et al.
Pubblicazione: (2024)
di: Liu, Wei, et al.
Pubblicazione: (2024)
EZ-Sort: Efficient Pairwise Comparison via Zero-Shot CLIP-Based Pre-Ordering and Human-in-the-Loop Sorting
di: Park, Yujin, et al.
Pubblicazione: (2025)
di: Park, Yujin, et al.
Pubblicazione: (2025)
FLEX: A Benchmark for Evaluating Robustness of Fairness in Large Language Models
di: Jung, Dahyun, et al.
Pubblicazione: (2025)
di: Jung, Dahyun, et al.
Pubblicazione: (2025)
Hierarchical Vision Language Action Model Using Success and Failure Demonstrations
di: Park, Jeongeun, et al.
Pubblicazione: (2025)
di: Park, Jeongeun, et al.
Pubblicazione: (2025)
Dual-Stream Diffusion for World-Model Augmented Vision-Language-Action Model
di: Won, John, et al.
Pubblicazione: (2025)
di: Won, John, et al.
Pubblicazione: (2025)
SPARK: Multi-Vision Sensor Perception and Reasoning Benchmark for Large-scale Vision-Language Models
di: Yu, Youngjoon, et al.
Pubblicazione: (2024)
di: Yu, Youngjoon, et al.
Pubblicazione: (2024)
Modular Sensory Stream for Integrating Physical Feedback in Vision-Language-Action Models
di: Lee, Jimin, et al.
Pubblicazione: (2026)
di: Lee, Jimin, et al.
Pubblicazione: (2026)
Goal-driven Bayesian Optimal Experimental Design for Robust Decision-Making Under Model Uncertainty
di: Go, Jinwoo, et al.
Pubblicazione: (2026)
di: Go, Jinwoo, et al.
Pubblicazione: (2026)
RedacBench: Can AI Erase Your Secrets?
di: Jeon, Hyunjun, et al.
Pubblicazione: (2026)
di: Jeon, Hyunjun, et al.
Pubblicazione: (2026)
LLM4SGG: Large Language Models for Weakly Supervised Scene Graph Generation
di: Kim, Kibum, et al.
Pubblicazione: (2023)
di: Kim, Kibum, et al.
Pubblicazione: (2023)
Phantom of Latent for Large Language and Vision Models
di: Lee, Byung-Kwan, et al.
Pubblicazione: (2024)
di: Lee, Byung-Kwan, et al.
Pubblicazione: (2024)
Lightweight Unpaired Smartphone ISP Transfer with Semantic Pseudo-Pairing
di: Cho, Yujin, et al.
Pubblicazione: (2026)
di: Cho, Yujin, et al.
Pubblicazione: (2026)
MIMO Detection under Hardware Impairments: Data Augmentation With Boosting
di: Kang, Yujin, et al.
Pubblicazione: (2024)
di: Kang, Yujin, et al.
Pubblicazione: (2024)
TroL: Traversal of Layers for Large Language and Vision Models
di: Lee, Byung-Kwan, et al.
Pubblicazione: (2024)
di: Lee, Byung-Kwan, et al.
Pubblicazione: (2024)
Hybrid tracker based optimal path tracking system for complex road environments for autonomous driving
di: Seo, Eunbin, et al.
Pubblicazione: (2021)
di: Seo, Eunbin, et al.
Pubblicazione: (2021)
A Survey on Inference Engines for Large Language Models: Perspectives on Optimization and Efficiency
di: Park, Sihyeong, et al.
Pubblicazione: (2025)
di: Park, Sihyeong, et al.
Pubblicazione: (2025)
Space‐Efficient Logical Qubit Architecture with a Bus for Magic State Consumption
di: Yujin Kang, et al.
Pubblicazione: (2025)
di: Yujin Kang, et al.
Pubblicazione: (2025)
A Review on Proprietary Accelerators for Large Language Models
di: Park, Sihyeong, et al.
Pubblicazione: (2025)
di: Park, Sihyeong, et al.
Pubblicazione: (2025)
SpatialBoost: Enhancing Visual Representation through Language-Guided Reasoning
di: Jeon, Byungwoo, et al.
Pubblicazione: (2026)
di: Jeon, Byungwoo, et al.
Pubblicazione: (2026)
Verifier-free Test-Time Sampling for Vision Language Action Models
di: Jang, Suhyeok, et al.
Pubblicazione: (2025)
di: Jang, Suhyeok, et al.
Pubblicazione: (2025)
Learning Unified Distance Metric Across Diverse Data Distributions with Parameter-Efficient Transfer Learning
di: Kim, Sungyeon, et al.
Pubblicazione: (2023)
di: Kim, Sungyeon, et al.
Pubblicazione: (2023)
Operon: Incremental Construction of Ragged Data via Named Dimensions
di: Moon, Sungbin, et al.
Pubblicazione: (2025)
di: Moon, Sungbin, et al.
Pubblicazione: (2025)
TD3Net: A temporal densely connected multi-dilated convolutional network for lipreading
di: Lee, Byung Hoon, et al.
Pubblicazione: (2025)
di: Lee, Byung Hoon, et al.
Pubblicazione: (2025)
FALCON: False-Negative Aware Learning of Contrastive Negatives in Vision-Language Alignment
di: Kim, Myunsoo, et al.
Pubblicazione: (2025)
di: Kim, Myunsoo, et al.
Pubblicazione: (2025)
AHS: Adaptive Head Synthesis via Synthetic Data Augmentations
di: Kang, Taewoong, et al.
Pubblicazione: (2026)
di: Kang, Taewoong, et al.
Pubblicazione: (2026)
CLIP-SLA: Parameter-Efficient CLIP Adaptation for Continuous Sign Language Recognition
di: Alyami, Sarah, et al.
Pubblicazione: (2025)
di: Alyami, Sarah, et al.
Pubblicazione: (2025)
Weakly Supervised Video Scene Graph Generation via Natural Language Supervision
di: Kim, Kibum, et al.
Pubblicazione: (2025)
di: Kim, Kibum, et al.
Pubblicazione: (2025)
ViewSplat: View-Adaptive Dynamic Gaussian Splatting for Feed-Forward Synthesis
di: Jeong, Moonyeon, et al.
Pubblicazione: (2026)
di: Jeong, Moonyeon, et al.
Pubblicazione: (2026)
Data-Efficient Molecular Generation with Hierarchical Textual Inversion
di: Kim, Seojin, et al.
Pubblicazione: (2024)
di: Kim, Seojin, et al.
Pubblicazione: (2024)
Vector Quantization for Deep-Learning-Based CSI Feedback in Massive MIMO Systems
di: Shin, Junyong, et al.
Pubblicazione: (2024)
di: Shin, Junyong, et al.
Pubblicazione: (2024)
A deterministic proof of Loewner energy reversibility via local reversals
di: Sung, Jinwoo
Pubblicazione: (2024)
di: Sung, Jinwoo
Pubblicazione: (2024)
HAMLET: Switch your Vision-Language-Action Model into a History-Aware Policy
di: Koo, Myungkyu, et al.
Pubblicazione: (2025)
di: Koo, Myungkyu, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Vision-aligned Latent Reasoning for Multi-modal Large Language Model
di: Jeon, Byungwoo, et al.
Pubblicazione: (2026) -
Iterative Prompt Refinement for Safer Text-to-Image Generation
di: Jeon, Jinwoo, et al.
Pubblicazione: (2025) -
Self-Refining Language Model Anonymizers via Adversarial Distillation
di: Kim, Kyuyoung, et al.
Pubblicazione: (2025) -
AMRG: Extend Vision Language Models for Automatic Mammography Report Generation
di: Sung, Nak-Jun, et al.
Pubblicazione: (2025) -
Co‐Stimuli‐Driven 2D WSe 2 Optoelectronic Synapses for Neuromorphic Computing
di: Junho Sung, et al.
Pubblicazione: (2025)