Towards Analyzing and Understanding the Limitations of VAPO: A Theoretical Perspective
Fuente:
arXiv
Salvato in:
| Autori principali: | Shao, Jintian, Cheng, Yiming |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Towards Analyzing and Understanding the Limitations of VAPO: A Theoretical Perspective
di: Shao, Jintian, et al.
Pubblicazione: (2025)
di: Shao, Jintian, et al.
Pubblicazione: (2025)
CoT is Not True Reasoning, It Is Just a Tight Constraint to Imitate: A Theory Perspective
di: Shao, Jintian, et al.
Pubblicazione: (2025)
di: Shao, Jintian, et al.
Pubblicazione: (2025)
Towards Analyzing and Understanding the Limitations of DPO: A Theoretical Perspective
di: Feng, Duanyu, et al.
Pubblicazione: (2024)
di: Feng, Duanyu, et al.
Pubblicazione: (2024)
Power-Law Decay Loss for Large Language Model Finetuning: A Theory Perspective
di: Shao, Jintian
Pubblicazione: (2025)
di: Shao, Jintian
Pubblicazione: (2025)
Analyzing Consumer Reviews for Understanding Drivers of Hotels Ratings: An Indian Perspective
di: Dasgupta, Subhasis, et al.
Pubblicazione: (2024)
di: Dasgupta, Subhasis, et al.
Pubblicazione: (2024)
Towards a Theoretical Understanding of Synthetic Data in LLM Post-Training: A Reverse-Bottleneck Perspective
di: Gan, Zeyu, et al.
Pubblicazione: (2024)
di: Gan, Zeyu, et al.
Pubblicazione: (2024)
Towards Distillation-Resistant Large Language Models: An Information-Theoretic Perspective
di: Fang, Hao, et al.
Pubblicazione: (2026)
di: Fang, Hao, et al.
Pubblicazione: (2026)
Why Code, Why Now: An Information-Theoretic Perspective on the Limits of Machine Learning
di: Zhao, Zhimin
Pubblicazione: (2026)
di: Zhao, Zhimin
Pubblicazione: (2026)
Analyzing Wrap-Up Effects through an Information-Theoretic Lens
di: Meister, Clara, et al.
Pubblicazione: (2022)
di: Meister, Clara, et al.
Pubblicazione: (2022)
Understanding and Addressing the Under-Translation Problem from the Perspective of Decoding Objective
di: Shao, Chenze, et al.
Pubblicazione: (2024)
di: Shao, Chenze, et al.
Pubblicazione: (2024)
Feature Resemblance: Towards a Theoretical Understanding of Analogical Reasoning in Transformers
di: Xu, Ruichen, et al.
Pubblicazione: (2026)
di: Xu, Ruichen, et al.
Pubblicazione: (2026)
Towards a Theoretical Understanding of the 'Reversal Curse' via Training Dynamics
di: Zhu, Hanlin, et al.
Pubblicazione: (2024)
di: Zhu, Hanlin, et al.
Pubblicazione: (2024)
Analyzing Diversity in Healthcare LLM Research: A Scientometric Perspective
di: Restrepo, David, et al.
Pubblicazione: (2024)
di: Restrepo, David, et al.
Pubblicazione: (2024)
VQ-Logits: Compressing the Output Bottleneck of Large Language Models via Vector Quantized Logits
di: Shao, Jintian, et al.
Pubblicazione: (2025)
di: Shao, Jintian, et al.
Pubblicazione: (2025)
Toward Understanding Unlearning Difficulty: A Mechanistic Perspective and Circuit-Guided Difficulty Metric
di: Cheng, Jiali, et al.
Pubblicazione: (2026)
di: Cheng, Jiali, et al.
Pubblicazione: (2026)
Theoretical Understanding of In-Context Learning in Shallow Transformers with Unstructured Data
di: Xing, Yue, et al.
Pubblicazione: (2024)
di: Xing, Yue, et al.
Pubblicazione: (2024)
Measuring and Analyzing Subjective Uncertainty in Scientific Communications
di: Sourati, Jamshid, et al.
Pubblicazione: (2025)
di: Sourati, Jamshid, et al.
Pubblicazione: (2025)
Benchmarking Abstract and Reasoning Abilities Through A Theoretical Perspective
di: Ma, Qingchuan, et al.
Pubblicazione: (2025)
di: Ma, Qingchuan, et al.
Pubblicazione: (2025)
Analyzing and Mitigating Object Hallucination: A Training Bias Perspective
di: Li, Yifan, et al.
Pubblicazione: (2025)
di: Li, Yifan, et al.
Pubblicazione: (2025)
Med-HEAL: Analyzing and Mitigating Hallucinations in Medical LLMs with Hallucination-Aware In-Context Learning
di: Liao, Yiming, et al.
Pubblicazione: (2026)
di: Liao, Yiming, et al.
Pubblicazione: (2026)
Towards Cross-lingual Values Judgment: A Consensus-Pluralism Perspective
di: Chen, Yukun, et al.
Pubblicazione: (2026)
di: Chen, Yukun, et al.
Pubblicazione: (2026)
An Information-Theoretic Approach to Analyze NLP Classification Tasks
di: Wang, Luran, et al.
Pubblicazione: (2024)
di: Wang, Luran, et al.
Pubblicazione: (2024)
Understanding and Analyzing Inappropriately Targeting Language in Online Discourse: A Comparative Annotation Study
di: Barbarestani, Baran, et al.
Pubblicazione: (2025)
di: Barbarestani, Baran, et al.
Pubblicazione: (2025)
Unveiling Cultural Blind Spots: Analyzing the Limitations of mLLMs in Procedural Text Comprehension
di: Yari, Amir Hossein, et al.
Pubblicazione: (2025)
di: Yari, Amir Hossein, et al.
Pubblicazione: (2025)
Theoretical Limits of Language Model Alignment
di: Paes, Lucas Monteiro, et al.
Pubblicazione: (2026)
di: Paes, Lucas Monteiro, et al.
Pubblicazione: (2026)
Do LVLMs Understand Charts? Analyzing and Correcting Factual Errors in Chart Captioning
di: Huang, Kung-Hsiang, et al.
Pubblicazione: (2023)
di: Huang, Kung-Hsiang, et al.
Pubblicazione: (2023)
LinguaGame: A Linguistically Grounded Game-Theoretic Paradigm for Multi-Agent Dialogue Generation
di: Ye, Yuxiao, et al.
Pubblicazione: (2026)
di: Ye, Yuxiao, et al.
Pubblicazione: (2026)
Towards Understanding Safety Alignment: A Mechanistic Perspective from Safety Neurons
di: Chen, Jianhui, et al.
Pubblicazione: (2024)
di: Chen, Jianhui, et al.
Pubblicazione: (2024)
On the Theoretical Limitations of Embedding-Based Retrieval
di: Weller, Orion, et al.
Pubblicazione: (2025)
di: Weller, Orion, et al.
Pubblicazione: (2025)
A Theoretical Understanding of Self-Correction through In-context Alignment
di: Wang, Yifei, et al.
Pubblicazione: (2024)
di: Wang, Yifei, et al.
Pubblicazione: (2024)
Reranking Laws for Language Generation: A Communication-Theoretic Perspective
di: Farinhas, António, et al.
Pubblicazione: (2024)
di: Farinhas, António, et al.
Pubblicazione: (2024)
Do Large Language Models Truly Understand Geometric Structures?
di: Wang, Xiaofeng, et al.
Pubblicazione: (2025)
di: Wang, Xiaofeng, et al.
Pubblicazione: (2025)
LIDAO: Towards Limited Interventions for Debiasing (Large) Language Models
di: Liu, Tianci, et al.
Pubblicazione: (2024)
di: Liu, Tianci, et al.
Pubblicazione: (2024)
Visual Thoughts: A Unified Perspective of Understanding Multimodal Chain-of-Thought
di: Cheng, Zihui, et al.
Pubblicazione: (2025)
di: Cheng, Zihui, et al.
Pubblicazione: (2025)
A Theoretical Perspective for Speculative Decoding Algorithm
di: Yin, Ming, et al.
Pubblicazione: (2024)
di: Yin, Ming, et al.
Pubblicazione: (2024)
Analyzing Temporal Complex Events with Large Language Models? A Benchmark towards Temporal, Long Context Understanding
di: Zhang, Zhihan, et al.
Pubblicazione: (2024)
di: Zhang, Zhihan, et al.
Pubblicazione: (2024)
Towards Spoken Language Understanding via Multi-level Multi-grained Contrastive Learning
di: Cheng, Xuxin, et al.
Pubblicazione: (2024)
di: Cheng, Xuxin, et al.
Pubblicazione: (2024)
Understanding the Capabilities and Limitations of Large Language Models for Cultural Commonsense
di: Shen, Siqi, et al.
Pubblicazione: (2024)
di: Shen, Siqi, et al.
Pubblicazione: (2024)
Theoretical Benefit and Limitation of Diffusion Language Model
di: Feng, Guhao, et al.
Pubblicazione: (2025)
di: Feng, Guhao, et al.
Pubblicazione: (2025)
ComplexFormer: Disruptively Advancing Transformer Inference Ability via Head-Specific Complex Vector Attention
di: Shao, Jintian, et al.
Pubblicazione: (2025)
di: Shao, Jintian, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Towards Analyzing and Understanding the Limitations of VAPO: A Theoretical Perspective
di: Shao, Jintian, et al.
Pubblicazione: (2025) -
CoT is Not True Reasoning, It Is Just a Tight Constraint to Imitate: A Theory Perspective
di: Shao, Jintian, et al.
Pubblicazione: (2025) -
Towards Analyzing and Understanding the Limitations of DPO: A Theoretical Perspective
di: Feng, Duanyu, et al.
Pubblicazione: (2024) -
Power-Law Decay Loss for Large Language Model Finetuning: A Theory Perspective
di: Shao, Jintian
Pubblicazione: (2025) -
Analyzing Consumer Reviews for Understanding Drivers of Hotels Ratings: An Indian Perspective
di: Dasgupta, Subhasis, et al.
Pubblicazione: (2024)