OmniACBench: A Benchmark for Evaluating Context-Grounded Acoustic Control in Omni-Modal Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Seunghee, Park, Bumkyu, Jung, Kyudan, Lee, Joosung, Kim, Soyoon, Kim, Jeonghoon, Kim, Taeuk, Jo, Hwiyeol |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
OMHBench: Benchmarking Balanced and Grounded Omni-Modal Multi-Hop Reasoning
von: Kim, Seunghee, et al.
Veröffentlicht: (2025)
von: Kim, Seunghee, et al.
Veröffentlicht: (2025)
Enhancing Hallucination Detection via Future Context
von: Lee, Joosung, et al.
Veröffentlicht: (2025)
von: Lee, Joosung, et al.
Veröffentlicht: (2025)
What Models Know, How Well They Know It: Knowledge-Weighted Fine-Tuning for Learning When to Say "I Don't Know"
von: Lee, Joosung, et al.
Veröffentlicht: (2026)
von: Lee, Joosung, et al.
Veröffentlicht: (2026)
FCMR: Robust Evaluation of Financial Cross-Modal Multi-Hop Reasoning
von: Kim, Seunghee, et al.
Veröffentlicht: (2024)
von: Kim, Seunghee, et al.
Veröffentlicht: (2024)
Sommelier: Scalable Open Multi-turn Audio Pre-processing for Full-duplex Speech Language Models
von: Jung, Kyudan, et al.
Veröffentlicht: (2026)
von: Jung, Kyudan, et al.
Veröffentlicht: (2026)
SNAP: Speaker Nulling for Artifact Projection in Speech Deepfake Detection
von: Jung, Kyudan, et al.
Veröffentlicht: (2026)
von: Jung, Kyudan, et al.
Veröffentlicht: (2026)
Finding Answers in Thought Matters: Revisiting Evaluation on Large Language Models with Reasoning
von: Jo, Hwiyeol, et al.
Veröffentlicht: (2025)
von: Jo, Hwiyeol, et al.
Veröffentlicht: (2025)
Investigating the Influence of Prompt-Specific Shortcuts in AI Generated Text Detection
von: Park, Choonghyun, et al.
Veröffentlicht: (2024)
von: Park, Choonghyun, et al.
Veröffentlicht: (2024)
MAGIC: A Multi-Hop and Graph-Based Benchmark for Inter-Context Conflicts in Retrieval-Augmented Generation
von: Lee, Jungyeon, et al.
Veröffentlicht: (2025)
von: Lee, Jungyeon, et al.
Veröffentlicht: (2025)
Cross-lingual Collapse: How Language-Centric Foundation Models Shape Reasoning in Large Language Models
von: Park, Cheonbok, et al.
Veröffentlicht: (2025)
von: Park, Cheonbok, et al.
Veröffentlicht: (2025)
Enhanced Facet Generation with LLM Editing
von: Lee, Joosung, et al.
Veröffentlicht: (2024)
von: Lee, Joosung, et al.
Veröffentlicht: (2024)
Dynin-Omni: Omnimodal Unified Large Diffusion Language Model
von: Kim, Jaeik, et al.
Veröffentlicht: (2026)
von: Kim, Jaeik, et al.
Veröffentlicht: (2026)
Improving Multi-hop Logical Reasoning in Knowledge Graphs with Context-Aware Query Representation Learning
von: Kim, Jeonghoon, et al.
Veröffentlicht: (2024)
von: Kim, Jeonghoon, et al.
Veröffentlicht: (2024)
Benchmarks Are Not That Out of Distribution: Word Overlap Predicts Performance
von: Chung, Woojin, et al.
Veröffentlicht: (2026)
von: Chung, Woojin, et al.
Veröffentlicht: (2026)
FutureOmni: Evaluating Future Forecasting from Omni-Modal Context for Multimodal LLMs
von: Chen, Qian, et al.
Veröffentlicht: (2026)
von: Chen, Qian, et al.
Veröffentlicht: (2026)
Maternal wellbeing amidst English fever: An integrative framework of vicarious pride, empathy, agency and hope (M–VEAH)
von: Yeji Han, et al.
Veröffentlicht: (2026)
von: Yeji Han, et al.
Veröffentlicht: (2026)
Latent Preference Modeling for Cross-Session Personalized Tool Calling
von: Yoon, Yejin, et al.
Veröffentlicht: (2026)
von: Yoon, Yejin, et al.
Veröffentlicht: (2026)
Adaptive Contrastive Decoding in Retrieval-Augmented Generation for Handling Noisy Contexts
von: Kim, Youna, et al.
Veröffentlicht: (2024)
von: Kim, Youna, et al.
Veröffentlicht: (2024)
BlendX: Complex Multi-Intent Detection with Blended Patterns
von: Yoon, Yejin, et al.
Veröffentlicht: (2024)
von: Yoon, Yejin, et al.
Veröffentlicht: (2024)
HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context
von: Yang, Qize, et al.
Veröffentlicht: (2025)
von: Yang, Qize, et al.
Veröffentlicht: (2025)
When to Speak, When to Abstain: Contrastive Decoding with Abstention
von: Kim, Hyuhng Joon, et al.
Veröffentlicht: (2024)
von: Kim, Hyuhng Joon, et al.
Veröffentlicht: (2024)
Aligning Language Models to Explicitly Handle Ambiguity
von: Kim, Hyuhng Joon, et al.
Veröffentlicht: (2024)
von: Kim, Hyuhng Joon, et al.
Veröffentlicht: (2024)
K-BrowseComp: A Web Browsing Agent Benchmark Grounded in Korean Contexts
von: Lee, Nahyun, et al.
Veröffentlicht: (2026)
von: Lee, Nahyun, et al.
Veröffentlicht: (2026)
Revisiting the Impact of Pursuing Modularity for Code Generation
von: Kang, Deokyeong, et al.
Veröffentlicht: (2024)
von: Kang, Deokyeong, et al.
Veröffentlicht: (2024)
ADVICE: Answer-Dependent Verbalized Confidence Estimation
von: Seo, Ki Jung, et al.
Veröffentlicht: (2025)
von: Seo, Ki Jung, et al.
Veröffentlicht: (2025)
Subgraph-Aware Training of Language Models for Knowledge Graph Completion Using Structure-Aware Contrastive Learning
von: Ko, Youmin, et al.
Veröffentlicht: (2024)
von: Ko, Youmin, et al.
Veröffentlicht: (2024)
ReGUIDE: Data Efficient GUI Grounding via Spatial Reasoning and Search
von: Lee, Hyunseok, et al.
Veröffentlicht: (2025)
von: Lee, Hyunseok, et al.
Veröffentlicht: (2025)
Exploiting Vocabulary Frequency Imbalance in Language Model Pre-training
von: Chung, Woojin, et al.
Veröffentlicht: (2025)
von: Chung, Woojin, et al.
Veröffentlicht: (2025)
Lightweight Audio Segmentation for Long-form Speech Translation
von: Lee, Jaesong, et al.
Veröffentlicht: (2024)
von: Lee, Jaesong, et al.
Veröffentlicht: (2024)
KGMEL: Knowledge Graph-Enhanced Multimodal Entity Linking
von: Kim, Juyeon, et al.
Veröffentlicht: (2025)
von: Kim, Juyeon, et al.
Veröffentlicht: (2025)
OmniSIFT: Modality-Asymmetric Token Compression for Efficient Omni-modal Large Language Models
von: Ding, Yue, et al.
Veröffentlicht: (2026)
von: Ding, Yue, et al.
Veröffentlicht: (2026)
Omni-RRM: Advancing Omni Reward Modeling via Automatic Rubric-Grounded Preference Synthesis
von: Kong, Zicheng, et al.
Veröffentlicht: (2026)
von: Kong, Zicheng, et al.
Veröffentlicht: (2026)
TeXBLEU: Automatic Metric for Evaluate LaTeX Format
von: Jung, Kyudan, et al.
Veröffentlicht: (2024)
von: Jung, Kyudan, et al.
Veröffentlicht: (2024)
PK-ICR: Persona-Knowledge Interactive Context Retrieval for Grounded Dialogue
von: Oh, Minsik, et al.
Veröffentlicht: (2023)
von: Oh, Minsik, et al.
Veröffentlicht: (2023)
ZeroDL: Zero-shot Distribution Learning for Text Clustering via Large Language Models
von: Jo, Hwiyeol, et al.
Veröffentlicht: (2024)
von: Jo, Hwiyeol, et al.
Veröffentlicht: (2024)
LUT-GEMM: Quantized Matrix Multiplication based on LUTs for Efficient Inference in Large-Scale Generative Language Models
von: Park, Gunho, et al.
Veröffentlicht: (2022)
von: Park, Gunho, et al.
Veröffentlicht: (2022)
Open Ko-LLM Leaderboard: Evaluating Large Language Models in Korean with Ko-H5 Benchmark
von: Park, Chanjun, et al.
Veröffentlicht: (2024)
von: Park, Chanjun, et al.
Veröffentlicht: (2024)
OmniPlay: Benchmarking Omni-Modal Models on Omni-Modal Game Playing
von: Bie, Fuqing, et al.
Veröffentlicht: (2025)
von: Bie, Fuqing, et al.
Veröffentlicht: (2025)
Omni-Captioner: Data Pipeline, Models, and Benchmark for Omni Detailed Perception
von: Ma, Ziyang, et al.
Veröffentlicht: (2025)
von: Ma, Ziyang, et al.
Veröffentlicht: (2025)
Omni-Persona: Systematic Benchmarking and Improving Omnimodal Personalization
von: Oh, Yeongtak, et al.
Veröffentlicht: (2026)
von: Oh, Yeongtak, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
OMHBench: Benchmarking Balanced and Grounded Omni-Modal Multi-Hop Reasoning
von: Kim, Seunghee, et al.
Veröffentlicht: (2025) -
Enhancing Hallucination Detection via Future Context
von: Lee, Joosung, et al.
Veröffentlicht: (2025) -
What Models Know, How Well They Know It: Knowledge-Weighted Fine-Tuning for Learning When to Say "I Don't Know"
von: Lee, Joosung, et al.
Veröffentlicht: (2026) -
FCMR: Robust Evaluation of Financial Cross-Modal Multi-Hop Reasoning
von: Kim, Seunghee, et al.
Veröffentlicht: (2024) -
Sommelier: Scalable Open Multi-turn Audio Pre-processing for Full-duplex Speech Language Models
von: Jung, Kyudan, et al.
Veröffentlicht: (2026)