Unlocking Cognitive Capabilities and Analyzing the Perception-Logic Trade-off
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Longyin, Sun, Shuo, He, Yingxu, Lewis, Won Cheng Yi, Shahrin, Muhammad Huzaifah Bin Md, Sailor, Hardik Bhupendra, Wong, Heng Meng Jeremy, Vangani, Tarun Kumar, Ma, Yi, Wang, Qiongqiong, Pham, Minh Duc, Jiang, Ridong, Li, Jingtao, Liao, Jingyi, Liu, Zhuohan, Lu, Yanfeng, Gupta, Manas, Aw, Ai Ti |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Contextual Paralinguistic Data Creation for Multi-Modal Speech-LLM: Data Condensation and Spoken QA Generation
by: Wang, Qiongqiong, et al.
Published: (2025)
by: Wang, Qiongqiong, et al.
Published: (2025)
Benchmarking Contextual and Paralinguistic Reasoning in Speech-LLMs: A Case Study with In-the-Wild Data
by: Wang, Qiongqiong, et al.
Published: (2025)
by: Wang, Qiongqiong, et al.
Published: (2025)
MERaLiON-SpeechEncoder: Towards a Speech Foundation Model for Singapore and Beyond
by: Huzaifah, Muhammad, et al.
Published: (2024)
by: Huzaifah, Muhammad, et al.
Published: (2024)
MERaLiON-SER: Robust Speech Emotion Recognition Model for English and SEA Languages
by: Sailor, Hardik B., et al.
Published: (2025)
by: Sailor, Hardik B., et al.
Published: (2025)
Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models
by: Wang, Qiongqiong, et al.
Published: (2025)
by: Wang, Qiongqiong, et al.
Published: (2025)
MERaLiON-TextLLM: Cross-Lingual Understanding of Large Language Models in Chinese, Indonesian, Malay, and Singlish
by: Huang, Xin, et al.
Published: (2024)
by: Huang, Xin, et al.
Published: (2024)
Improving Multilingual Social Media Insights: Aspect-based Comment Analysis
by: Zhang, Longyin, et al.
Published: (2025)
by: Zhang, Longyin, et al.
Published: (2025)
Beyond Classification: Towards Speech Emotion Reasoning with Multitask AudioLLMs
by: Zhang, Wenyu, et al.
Published: (2025)
by: Zhang, Wenyu, et al.
Published: (2025)
MoLEx: Mixture of LoRA Experts in Speech Self-Supervised Models for Audio Deepfake Detection
by: Pan, Zihan, et al.
Published: (2025)
by: Pan, Zihan, et al.
Published: (2025)
Direct Preference Optimization for English-Mandarin Code-Switching Speech Recognition in Audio LLMs
by: Quang, Trung Nguyen, et al.
Published: (2026)
by: Quang, Trung Nguyen, et al.
Published: (2026)
Attentive Merging of Hidden Embeddings from Pre-trained Speech Model for Anti-spoofing Detection
by: Pan, Zihan, et al.
Published: (2024)
by: Pan, Zihan, et al.
Published: (2024)
MERaLiON-AudioLLM: Bridging Audio and Language with Large Language Models
by: He, Yingxu, et al.
Published: (2024)
by: He, Yingxu, et al.
Published: (2024)
Speech Foundation Model Ensembles for the Controlled Singing Voice Deepfake Detection (CtrSVDD) Challenge 2024
by: Guragain, Anmol, et al.
Published: (2024)
by: Guragain, Anmol, et al.
Published: (2024)
Quantizer-Aware Hierarchical Neural Codec Modeling for Speech Deepfake Detection
by: Wu, Jinyang, et al.
Published: (2026)
by: Wu, Jinyang, et al.
Published: (2026)
MoWE-Audio: Multitask AudioLLMs with Mixture of Weak Encoders
by: Zhang, Wenyu, et al.
Published: (2024)
by: Zhang, Wenyu, et al.
Published: (2024)
Advancing Singlish Understanding: Bridging the Gap with Datasets and Multimodal Models
by: Wang, Bin, et al.
Published: (2025)
by: Wang, Bin, et al.
Published: (2025)
IFEval-Audio: Benchmarking Instruction-Following Capability in Audio-based Large Language Models
by: Gao, Yiming, et al.
Published: (2025)
by: Gao, Yiming, et al.
Published: (2025)
SEA-Spoof: Bridging The Gap in Multilingual Audio Deepfake Detection for South-East Asian
by: Wu, Jinyang, et al.
Published: (2025)
by: Wu, Jinyang, et al.
Published: (2025)
Towards Quantifying and Reducing Language Mismatch Effects in Cross-Lingual Speech Anti-Spoofing
by: Liu, Tianchi, et al.
Published: (2024)
by: Liu, Tianchi, et al.
Published: (2024)
Bridging Linguistic Gaps: Cross-Lingual Mapping in Pre-Training and Dataset for Enhanced Multilingual LLM Performance
by: Zheng, Weihua, et al.
Published: (2026)
by: Zheng, Weihua, et al.
Published: (2026)
AdaMCoT: Rethinking Cross-Lingual Factual Reasoning through Adaptive Multilingual Chain-of-Thought
by: Zheng, Weihua, et al.
Published: (2025)
by: Zheng, Weihua, et al.
Published: (2025)
Interpolating Speaker Identities in Embedding Space for Data Expansion
by: Liu, Tianchi, et al.
Published: (2025)
by: Liu, Tianchi, et al.
Published: (2025)
Optimizing CNN Using HPC Tools
by: Rahman, Shahrin
Published: (2024)
by: Rahman, Shahrin
Published: (2024)
AudioBench: A Universal Benchmark for Audio Large Language Models
by: Wang, Bin, et al.
Published: (2024)
by: Wang, Bin, et al.
Published: (2024)
Adversarial synthesis based data-augmentation for code-switched spoken language identification
by: Shastri, Parth, et al.
Published: (2022)
by: Shastri, Parth, et al.
Published: (2022)
California Research Institute on the Integration of Students with Severe Disabilities. Final Report, Years 1987-1992.
by: Sailor, Wayne, et al.
Published: (1993)
by: Sailor, Wayne, et al.
Published: (1993)
Capability-aware Prompt Reformulation Learning for Text-to-Image Generation
by: Zhan, Jingtao, et al.
Published: (2024)
by: Zhan, Jingtao, et al.
Published: (2024)
CCL-XCoT: An Efficient Cross-Lingual Knowledge Transfer Method for Mitigating Hallucination Generation
by: Zheng, Weihua, et al.
Published: (2025)
by: Zheng, Weihua, et al.
Published: (2025)
Staphylococcus aureus subcapsular splenic abscess and associated empyema in the setting of tocilizumab therapy: A case report
by: Audrey Lee, et al.
Published: (2024)
by: Audrey Lee, et al.
Published: (2024)
Non‐Reciprocal Terahertz Topological Sensor on a Silicon Chip
by: Ridong Jia, et al.
Published: (2025)
by: Ridong Jia, et al.
Published: (2025)
SeaEval for Multilingual Foundation Models: From Cross-Lingual Alignment to Cultural Reasoning
by: Wang, Bin, et al.
Published: (2023)
by: Wang, Bin, et al.
Published: (2023)
A scalable narrow linewidth high power laser for barium ion optical qubit
by: Ahmadi, Morteza, et al.
Published: (2023)
by: Ahmadi, Morteza, et al.
Published: (2023)
A semiempirical downscaling approach for predicting regional temperature impacts associated with climate change
by: David J. Sailor and Xiangshang Li
Published: (1999)
by: David J. Sailor and Xiangshang Li
Published: (1999)
Photonic Supercoupling in Silicon Topological Waveguides
by: Jia, Ridong, et al.
Published: (2024)
by: Jia, Ridong, et al.
Published: (2024)
Photonic Supercoupling in Silicon Topological Waveguides
by: Ridong Jia, et al.
Published: (2024)
by: Ridong Jia, et al.
Published: (2024)
Biomass‐Modified Carbon Composites: Unlocking New Horizons in Electromagnetic Absorption
by: Mengyao Tang, et al.
Published: (2025)
by: Mengyao Tang, et al.
Published: (2025)
DRESS syndrome with multiorgan involvement and HHV‐6 reactivation in the absence of a drug trigger
by: Yi Tong Vincent Aw, et al.
Published: (2024)
by: Yi Tong Vincent Aw, et al.
Published: (2024)
Using Artificial Intelligence to Unlock Crowdfunding Success for Small Businesses
by: Ye, Teng, et al.
Published: (2024)
by: Ye, Teng, et al.
Published: (2024)
A Study in Markov Chains, Loop-Erased Random Walk and Loop Soups
by: Gu, Zhuohan
Published: (2024)
by: Gu, Zhuohan
Published: (2024)
Child Marriage in Ethnic Minority Communities in Vietnam: Current Situation and Solutions
by: Duc Tran Minh
Published: (2025)
by: Duc Tran Minh
Published: (2025)
Similar Items
-
Contextual Paralinguistic Data Creation for Multi-Modal Speech-LLM: Data Condensation and Spoken QA Generation
by: Wang, Qiongqiong, et al.
Published: (2025) -
Benchmarking Contextual and Paralinguistic Reasoning in Speech-LLMs: A Case Study with In-the-Wild Data
by: Wang, Qiongqiong, et al.
Published: (2025) -
MERaLiON-SpeechEncoder: Towards a Speech Foundation Model for Singapore and Beyond
by: Huzaifah, Muhammad, et al.
Published: (2024) -
MERaLiON-SER: Robust Speech Emotion Recognition Model for English and SEA Languages
by: Sailor, Hardik B., et al.
Published: (2025) -
Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models
by: Wang, Qiongqiong, et al.
Published: (2025)