Incentivizing Consistent, Effective and Scalable Reasoning Capability in Audio LLMs via Reasoning Process Rewards
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fan, Jiajun, Ren, Roger, Li, Jingyuan, Pandey, Rahul, Shivakumar, Prashanth Gurunath, Bulyko, Ivan, Gandhe, Ankur, Liu, Ge, Gu, Yile |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Group Relative Policy Optimization for Speech Recognition
von: Shivakumar, Prashanth Gurunath, et al.
Veröffentlicht: (2025)
von: Shivakumar, Prashanth Gurunath, et al.
Veröffentlicht: (2025)
Align-SLM: Textless Spoken Language Models with Reinforcement Learning from AI Feedback
von: Lin, Guan-Ting, et al.
Veröffentlicht: (2024)
von: Lin, Guan-Ting, et al.
Veröffentlicht: (2024)
Multi-Modal Retrieval For Large Language Model Based Speech Recognition
von: Kolehmainen, Jari, et al.
Veröffentlicht: (2024)
von: Kolehmainen, Jari, et al.
Veröffentlicht: (2024)
Speech Recognition Rescoring with Large Speech-Text Foundation Models
von: Shivakumar, Prashanth Gurunath, et al.
Veröffentlicht: (2024)
von: Shivakumar, Prashanth Gurunath, et al.
Veröffentlicht: (2024)
Paralinguistics-Enhanced Large Language Modeling of Spoken Dialogue
von: Lin, Guan-Ting, et al.
Veröffentlicht: (2023)
von: Lin, Guan-Ting, et al.
Veröffentlicht: (2023)
Towards ASR Robust Spoken Language Understanding Through In-Context Learning With Word Confusion Networks
von: Everson, Kevin, et al.
Veröffentlicht: (2024)
von: Everson, Kevin, et al.
Veröffentlicht: (2024)
PROCTER: PROnunciation-aware ConTextual adaptER for personalized speech recognition in neural transducers
von: Pandey, Rahul, et al.
Veröffentlicht: (2023)
von: Pandey, Rahul, et al.
Veröffentlicht: (2023)
Investigating Training Strategies and Model Robustness of Low-Rank Adaptation for Language Modeling in Speech Recognition
von: Yu, Yu, et al.
Veröffentlicht: (2024)
von: Yu, Yu, et al.
Veröffentlicht: (2024)
Phone Duration Modeling for Speaker Age Estimation in Children
von: Shivakumar, Prashanth Gurunath, et al.
Veröffentlicht: (2021)
von: Shivakumar, Prashanth Gurunath, et al.
Veröffentlicht: (2021)
Low-rank Adaptation of Large Language Model Rescoring for Parameter-Efficient Speech Recognition
von: Yu, Yu, et al.
Veröffentlicht: (2023)
von: Yu, Yu, et al.
Veröffentlicht: (2023)
RAG-R1: Incentivizing the Search and Reasoning Capabilities of LLMs through Multi-query Parallelism
von: Tan, Zhiwen, et al.
Veröffentlicht: (2025)
von: Tan, Zhiwen, et al.
Veröffentlicht: (2025)
DINO-R1: Incentivizing Reasoning Capability in Vision Foundation Models
von: Pan, Chenbin, et al.
Veröffentlicht: (2025)
von: Pan, Chenbin, et al.
Veröffentlicht: (2025)
Reinforcement Learning with Verifiable Rewards Implicitly Incentivizes Correct Reasoning in Base LLMs
von: Wen, Xumeng, et al.
Veröffentlicht: (2025)
von: Wen, Xumeng, et al.
Veröffentlicht: (2025)
rSIM: Incentivizing Reasoning Capabilities of LLMs via Reinforced Strategy Injection
von: Chen, Sijia, et al.
Veröffentlicht: (2025)
von: Chen, Sijia, et al.
Veröffentlicht: (2025)
DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
von: DeepSeek-AI, et al.
Veröffentlicht: (2025)
von: DeepSeek-AI, et al.
Veröffentlicht: (2025)
R1-T1: Fully Incentivizing Translation Capability in LLMs via Reasoning Learning
von: He, Minggui, et al.
Veröffentlicht: (2025)
von: He, Minggui, et al.
Veröffentlicht: (2025)
Streaming Speech-to-Confusion Network Speech Recognition
von: Filimonov, Denis, et al.
Veröffentlicht: (2023)
von: Filimonov, Denis, et al.
Veröffentlicht: (2023)
Audio-Reasoner: Improving Reasoning Capability in Large Audio Language Models
von: Xie, Zhifei, et al.
Veröffentlicht: (2025)
von: Xie, Zhifei, et al.
Veröffentlicht: (2025)
R-Log: Incentivizing Log Analysis Capability in LLMs via Reasoning-based Reinforcement Learning
von: Liu, Yilun, et al.
Veröffentlicht: (2025)
von: Liu, Yilun, et al.
Veröffentlicht: (2025)
Generative Speech Recognition Error Correction with Large Language Models and Task-Activating Prompting
von: Yang, Chao-Han Huck, et al.
Veröffentlicht: (2023)
von: Yang, Chao-Han Huck, et al.
Veröffentlicht: (2023)
DeduCE: Deductive Consistency as a Framework to Evaluate LLM Reasoning
von: Pandey, Atharva, et al.
Veröffentlicht: (2025)
von: Pandey, Atharva, et al.
Veröffentlicht: (2025)
Graph-R1: Incentivizing the Zero-Shot Graph Learning Capability in LLMs via Explicit Reasoning
von: Wu, Yicong, et al.
Veröffentlicht: (2025)
von: Wu, Yicong, et al.
Veröffentlicht: (2025)
Enhancing the Code Reasoning Capabilities of LLMs via Consistency-based Reinforcement Learning
von: Qin, Zhanyue, et al.
Veröffentlicht: (2026)
von: Qin, Zhanyue, et al.
Veröffentlicht: (2026)
Chart-based Reasoning: Transferring Capabilities from LLMs to VLMs
von: Carbune, Victor, et al.
Veröffentlicht: (2024)
von: Carbune, Victor, et al.
Veröffentlicht: (2024)
Context Reasoner: Incentivizing Reasoning Capability for Contextualized Privacy and Safety Compliance via Reinforcement Learning
von: Hu, Wenbin, et al.
Veröffentlicht: (2025)
von: Hu, Wenbin, et al.
Veröffentlicht: (2025)
HiMed: Incentivizing Hindi Reasoning in Medical LLMs
von: Jiang, Dingfeng, et al.
Veröffentlicht: (2026)
von: Jiang, Dingfeng, et al.
Veröffentlicht: (2026)
SoundMind: RL-Incentivized Logic Reasoning for Audio-Language Models
von: Diao, Xingjian, et al.
Veröffentlicht: (2025)
von: Diao, Xingjian, et al.
Veröffentlicht: (2025)
Are Your LLMs Capable of Stable Reasoning?
von: Liu, Junnan, et al.
Veröffentlicht: (2024)
von: Liu, Junnan, et al.
Veröffentlicht: (2024)
MolReasoner: Toward Effective and Interpretable Reasoning for Molecular LLMs
von: Zhao, Guojiang, et al.
Veröffentlicht: (2025)
von: Zhao, Guojiang, et al.
Veröffentlicht: (2025)
Reasoning Under Uncertainty: Exploring Probabilistic Reasoning Capabilities of LLMs
von: Pournemat, Mobina, et al.
Veröffentlicht: (2025)
von: Pournemat, Mobina, et al.
Veröffentlicht: (2025)
Benchmarking Spatiotemporal Reasoning in LLMs and Reasoning Models: Capabilities and Challenges
von: Quan, Pengrui, et al.
Veröffentlicht: (2025)
von: Quan, Pengrui, et al.
Veröffentlicht: (2025)
LogicReward: Incentivizing LLM Reasoning via Step-Wise Logical Supervision
von: Xu, Jundong, et al.
Veröffentlicht: (2025)
von: Xu, Jundong, et al.
Veröffentlicht: (2025)
Prioritizing the Best: Incentivizing Reliable Multimodal Reasoning by Rewarding Beyond Answer Correctness
von: Jia, Mengzhao, et al.
Veröffentlicht: (2026)
von: Jia, Mengzhao, et al.
Veröffentlicht: (2026)
Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models
von: Huang, Wenxuan, et al.
Veröffentlicht: (2025)
von: Huang, Wenxuan, et al.
Veröffentlicht: (2025)
Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding
von: Chen, Haolin, et al.
Veröffentlicht: (2024)
von: Chen, Haolin, et al.
Veröffentlicht: (2024)
How Numerical Precision Affects Arithmetical Reasoning Capabilities of LLMs
von: Feng, Guhao, et al.
Veröffentlicht: (2024)
von: Feng, Guhao, et al.
Veröffentlicht: (2024)
Pixel Reasoner: Incentivizing Pixel-Space Reasoning with Curiosity-Driven Reinforcement Learning
von: Wang, Haozhe, et al.
Veröffentlicht: (2025)
von: Wang, Haozhe, et al.
Veröffentlicht: (2025)
Evaluating Consistency and Reasoning Capabilities of Large Language Models
von: Saxena, Yash, et al.
Veröffentlicht: (2024)
von: Saxena, Yash, et al.
Veröffentlicht: (2024)
Enhancing Reasoning Capabilities in SLMs with Reward Guided Dataset Distillation
von: Padarha, Shreyansh
Veröffentlicht: (2025)
von: Padarha, Shreyansh
Veröffentlicht: (2025)
InfiAlign: A Scalable and Sample-Efficient Framework for Aligning LLMs to Enhance Reasoning Capabilities
von: Cai, Shuo, et al.
Veröffentlicht: (2025)
von: Cai, Shuo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Group Relative Policy Optimization for Speech Recognition
von: Shivakumar, Prashanth Gurunath, et al.
Veröffentlicht: (2025) -
Align-SLM: Textless Spoken Language Models with Reinforcement Learning from AI Feedback
von: Lin, Guan-Ting, et al.
Veröffentlicht: (2024) -
Multi-Modal Retrieval For Large Language Model Based Speech Recognition
von: Kolehmainen, Jari, et al.
Veröffentlicht: (2024) -
Speech Recognition Rescoring with Large Speech-Text Foundation Models
von: Shivakumar, Prashanth Gurunath, et al.
Veröffentlicht: (2024) -
Paralinguistics-Enhanced Large Language Modeling of Spoken Dialogue
von: Lin, Guan-Ting, et al.
Veröffentlicht: (2023)