LambdaPO: A Lambda Style Policy Optimization for Reasoning Language Models
Fuente:
arXiv
Saved in:
| Main Author: | arXiv, Redacted by |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LamPO: A Lambda Style Policy Optimization for Reasoning Language Models
by: arXiv, Redacted by
Published: (2026)
by: arXiv, Redacted by
Published: (2026)
The Llama 4 Herd: Architecture, Training, Evaluation, and Deployment Notes
by: arXiv, Redacted by
Published: (2026)
by: arXiv, Redacted by
Published: (2026)
Removed by arXiv
by: arXiv, Removed by
Published: (2023)
by: arXiv, Removed by
Published: (2023)
Lambdas at the Far Edge: a Tale of Flying Lambdas and Lambdas on Wheels
by: Audrito, Giorgio, et al.
Published: (2026)
by: Audrito, Giorgio, et al.
Published: (2026)
Lax Modal Lambda Calculi
by: Valliappan, Nachiappan
Published: (2025)
by: Valliappan, Nachiappan
Published: (2025)
The Lambda Calculus is Quantifiable
by: Maestracci, Valentin, et al.
Published: (2024)
by: Maestracci, Valentin, et al.
Published: (2024)
On Decidable and Undecidable Extensions of Simply Typed Lambda Calculus
by: Kobayashi, Naoki
Published: (2024)
by: Kobayashi, Naoki
Published: (2024)
Opportunistically Parallel Lambda Calculus
by: Mell, Stephen, et al.
Published: (2024)
by: Mell, Stephen, et al.
Published: (2024)
A Gradual Probabilistic Lambda Calculus
by: Ye, Wenjia, et al.
Published: (2026)
by: Ye, Wenjia, et al.
Published: (2026)
HiPO: Hybrid Policy Optimization for Dynamic Reasoning in LLMs
by: Deng, Ken, et al.
Published: (2025)
by: Deng, Ken, et al.
Published: (2025)
Intersection Types for a Computational Lambda-Calculus with Global State
by: de'Liguoro, Ugo, et al.
Published: (2021)
by: de'Liguoro, Ugo, et al.
Published: (2021)
Lambda: Learning Matchable Prior For Entity Alignment with Unlabeled Dangling Cases
by: Yin, Hang, et al.
Published: (2024)
by: Yin, Hang, et al.
Published: (2024)
African Lambdas II: Formal Semantics of African Languages—The Verbal and Clausal Domain
by: Malte Zimmermann
Published: (2025)
by: Malte Zimmermann
Published: (2025)
Bidirectional Interpolation for the Lambda-Calculus -- Revisiting and Formalising Craig-Čubrić Interpolation
by: Bertrand, Meven Lennon, et al.
Published: (2026)
by: Bertrand, Meven Lennon, et al.
Published: (2026)
On Representability of Multiple-Valued Functions by Linear Lambda Terms Typed with Second-order Polymorphic Type System
by: Matsuoka, Satoshi
Published: (2026)
by: Matsuoka, Satoshi
Published: (2026)
Analyzing the Resource Utilization of Lambda Functions on Mobile Devices: Case Studies on Kotlin and Swift
by: Ejimuda, Chibundom U., et al.
Published: (2025)
by: Ejimuda, Chibundom U., et al.
Published: (2025)
Nominal Algebraic-Coalgebraic Data Types, with Applications to Infinitary Lambda-Calculi
by: Cerda, Rémy
Published: (2025)
by: Cerda, Rémy
Published: (2025)
Authorship Style Transfer with Policy Optimization
by: Liu, Shuai, et al.
Published: (2024)
by: Liu, Shuai, et al.
Published: (2024)
Adding Negation to Lambda Mu
by: van Bakel, Steffen
Published: (2021)
by: van Bakel, Steffen
Published: (2021)
From Compactifying Lambda-Letrec Terms to Recognizing Regular-Expression Processes
by: Grabmayer, Clemens
Published: (2024)
by: Grabmayer, Clemens
Published: (2024)
$λ_A$: A Typed Lambda Calculus for LLM Agent Composition
by: Liu, Qin
Published: (2026)
by: Liu, Qin
Published: (2026)
Role of Lambda N and Lambda NN interaction parameters on binding energy of Lambda H4 and Lambda H4*
by: Sharma, Bhupali
Published: (2024)
by: Sharma, Bhupali
Published: (2024)
PrAg-PO: Prompt Augmented Policy Optimization for Robust and Diverse Mathematical Reasoning
by: Lu, Wenquan, et al.
Published: (2026)
by: Lu, Wenquan, et al.
Published: (2026)
Reduction Strategies in the Lambda Calculus and Their Implementation through Derivable Abstract Machines: Introduction
by: Drab, Tomasz
Published: (2024)
by: Drab, Tomasz
Published: (2024)
CiPO: Counterfactual Unlearning for Large Reasoning Models through Iterative Preference Optimization
by: Li, Junyi, et al.
Published: (2026)
by: Li, Junyi, et al.
Published: (2026)
Groups and Inverse Semigroups in Lambda Calculus
by: Bucciarelli, Antonio, et al.
Published: (2026)
by: Bucciarelli, Antonio, et al.
Published: (2026)
StepPO: Step-Aligned Policy Optimization for Agentic Reinforcement Learning
by: Wang, Daoyu, et al.
Published: (2026)
by: Wang, Daoyu, et al.
Published: (2026)
Combining On-Policy Optimization and Distillation for Long-Context Reasoning in Large Language Models
by: Ramos, Miguel Moura, et al.
Published: (2026)
by: Ramos, Miguel Moura, et al.
Published: (2026)
RePO: Replay-Enhanced Policy Optimization
by: Li, Siheng, et al.
Published: (2025)
by: Li, Siheng, et al.
Published: (2025)
A Rewriting Theory for Quantum Lambda-Calculus
by: Faggian, Claudia, et al.
Published: (2024)
by: Faggian, Claudia, et al.
Published: (2024)
AMaPO: Adaptive Margin-attached Preference Optimization for Language Model Alignment
by: Deng, Ruibo, et al.
Published: (2025)
by: Deng, Ruibo, et al.
Published: (2025)
Towards a Neural Lambda Calculus: Neurosymbolic AI Applied to the Foundations of Functional Programming
by: Flach, João, et al.
Published: (2023)
by: Flach, João, et al.
Published: (2023)
SeaPO: Strategic Error Amplification for Robust Preference Optimization of Large Language Models
by: Rao, Jun, et al.
Published: (2025)
by: Rao, Jun, et al.
Published: (2025)
Term Orders for Optimistic Lambda-Superposition
by: Bentkamp, Alexander, et al.
Published: (2025)
by: Bentkamp, Alexander, et al.
Published: (2025)
Lambda Nordica
Published: (2019)
Published: (2019)
Comprehensive Review of Performance Optimization Strategies for Serverless Applications on AWS Lambda
by: Bechir, Mohamed Lemine El, et al.
Published: (2024)
by: Bechir, Mohamed Lemine El, et al.
Published: (2024)
AT$^2$PO: Agentic Turn-based Policy Optimization via Tree Search
by: Zong, Zefang, et al.
Published: (2026)
by: Zong, Zefang, et al.
Published: (2026)
TreePO: Bridging the Gap of Policy Optimization and Efficacy and Inference Efficiency with Heuristic Tree-based Modeling
by: Li, Yizhi, et al.
Published: (2025)
by: Li, Yizhi, et al.
Published: (2025)
ConfPO: Exploiting Policy Model Confidence for Critical Token Selection in Preference Optimization
by: Yoon, Hee Suk, et al.
Published: (2025)
by: Yoon, Hee Suk, et al.
Published: (2025)
DiffPO: Diffusion-styled Preference Optimization for Efficient Inference-Time Alignment of Large Language Models
by: Chen, Ruizhe, et al.
Published: (2025)
by: Chen, Ruizhe, et al.
Published: (2025)
Similar Items
-
LamPO: A Lambda Style Policy Optimization for Reasoning Language Models
by: arXiv, Redacted by
Published: (2026) -
The Llama 4 Herd: Architecture, Training, Evaluation, and Deployment Notes
by: arXiv, Redacted by
Published: (2026) -
Removed by arXiv
by: arXiv, Removed by
Published: (2023) -
Lambdas at the Far Edge: a Tale of Flying Lambdas and Lambdas on Wheels
by: Audrito, Giorgio, et al.
Published: (2026) -
Lax Modal Lambda Calculi
by: Valliappan, Nachiappan
Published: (2025)