In harmony with gpt-oss
Fuente:
arXiv
Saved in:
| Main Author: | Mavrin, Borislav |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
gpt-oss-120b & gpt-oss-20b Model Card
by: OpenAI, et al.
Published: (2025)
by: OpenAI, et al.
Published: (2025)
In AI Sweet Harmony: Sociopragmatic Guardrail Bypasses and Evaluation-Awareness in OpenAI gpt-oss-20b
by: Durner, Nils
Published: (2025)
by: Durner, Nils
Published: (2025)
Standardization of Psychiatric Diagnoses -- Role of Fine-tuned LLM Consortium and OpenAI-gpt-oss Reasoning LLM Enabled Decision Support System
by: Bandara, Eranga, et al.
Published: (2025)
by: Bandara, Eranga, et al.
Published: (2025)
Library learning with e-graphs on jazz harmony
by: Ren, Zeng, et al.
Published: (2026)
by: Ren, Zeng, et al.
Published: (2026)
Standardization of Neuromuscular Reflex Analysis -- Role of Fine-Tuned Vision-Language Model Consortium and OpenAI gpt-oss Reasoning LLM Enabled Decision Support System
by: Bandara, Eranga, et al.
Published: (2025)
by: Bandara, Eranga, et al.
Published: (2025)
Fine-Tuning Transformers: Vocabulary Transfer
by: Mosin, Vladislav, et al.
Published: (2021)
by: Mosin, Vladislav, et al.
Published: (2021)
The Path to Open Innovation: Peer-Review Under Fire (PRUF)
by: Billions, Ava, et al.
Published: (2025)
by: Billions, Ava, et al.
Published: (2025)
AlphaEvolve: A coding agent for scientific and algorithmic discovery
by: Novikov, Alexander, et al.
Published: (2025)
by: Novikov, Alexander, et al.
Published: (2025)
Interpretability-by-Design with Accurate Locally Additive Models and Conditional Feature Effects
by: Gkolemis, Vasilis, et al.
Published: (2026)
by: Gkolemis, Vasilis, et al.
Published: (2026)
HiVAE: Hierarchical Latent Variables for Scalable Theory of Mind
by: Doering, Nigel, et al.
Published: (2026)
by: Doering, Nigel, et al.
Published: (2026)
VAM: Verbalized Action Masking for Controllable Exploration in RL Post-Training -- A Chess Case Study
by: Zhang, Zhicheng, et al.
Published: (2026)
by: Zhang, Zhicheng, et al.
Published: (2026)
In-Context Learning in Linear vs. Quadratic Attention Models: An Empirical Study on Regression Tasks
by: Goel, Ayush, et al.
Published: (2026)
by: Goel, Ayush, et al.
Published: (2026)
A feature-stable and explainable machine learning framework for trustworthy decision-making under incomplete clinical data
by: Andrys-Olek, Justyna, et al.
Published: (2026)
by: Andrys-Olek, Justyna, et al.
Published: (2026)
AsynDBT: Asynchronous Distributed Bilevel Tuning for efficient In-Context Learning with Large Language Models
by: Ma, Hui, et al.
Published: (2026)
by: Ma, Hui, et al.
Published: (2026)
Investigating Target Class Influence on Neural Network Compressibility for Energy-Autonomous Avian Monitoring
by: Brolich, Nina, et al.
Published: (2026)
by: Brolich, Nina, et al.
Published: (2026)
Causal Neighbourhood Learning for Invariant Graph Representations
by: Job, Simi, et al.
Published: (2026)
by: Job, Simi, et al.
Published: (2026)
OffSeeker: Online Reinforcement Learning Is Not All You Need for Deep Research Agents
by: Zhou, Yuhang, et al.
Published: (2026)
by: Zhou, Yuhang, et al.
Published: (2026)
HAWX: A Hardware-Aware FrameWork for Fast and Scalable ApproXimation of DNNs
by: Nazari, Samira, et al.
Published: (2026)
by: Nazari, Samira, et al.
Published: (2026)
Global Low-Rank, Local Full-Rank: The Holographic Encoding of Learned Algorithms
by: Xu, Yongzhong
Published: (2026)
by: Xu, Yongzhong
Published: (2026)
Robust Exploration in Directed Controller Synthesis via Reinforcement Learning with Soft Mixture-of-Experts
by: Ubukata, Toshihide, et al.
Published: (2026)
by: Ubukata, Toshihide, et al.
Published: (2026)
Taming Preconditioner Drift: Unlocking the Potential of Second-Order Optimizers for Federated Learning on Non-IID Data
by: Liu, Junkang, et al.
Published: (2026)
by: Liu, Junkang, et al.
Published: (2026)
IR$^3$: Contrastive Inverse Reinforcement Learning for Interpretable Detection and Mitigation of Reward Hacking
by: Beigi, Mohammad, et al.
Published: (2026)
by: Beigi, Mohammad, et al.
Published: (2026)
Compositional Planning with Jumpy World Models
by: Farebrother, Jesse, et al.
Published: (2026)
by: Farebrother, Jesse, et al.
Published: (2026)
Do We Need Adam? Surprisingly Strong and Sparse Reinforcement Learning with SGD in LLMs
by: Mukherjee, Sagnik, et al.
Published: (2026)
by: Mukherjee, Sagnik, et al.
Published: (2026)
KBVQ-MoE: KLT-guided SVD with Bias-Corrected Vector Quantization for MoE Large Language Models
by: Xu, Zukang, et al.
Published: (2026)
by: Xu, Zukang, et al.
Published: (2026)
Imputation of Unknown Missingness in Sparse Electronic Health Records
by: Han, Jun, et al.
Published: (2026)
by: Han, Jun, et al.
Published: (2026)
Elimination-compensation pruning for fully-connected neural networks
by: Ballini, Enrico, et al.
Published: (2026)
by: Ballini, Enrico, et al.
Published: (2026)
ActionEngine: From Reactive to Programmatic GUI Agents via State Machine Memory
by: Zhong, Hongbin, et al.
Published: (2026)
by: Zhong, Hongbin, et al.
Published: (2026)
Regret-Guided Search Control for Efficient Learning in AlphaZero
by: Tsai, Yun-Jui, et al.
Published: (2026)
by: Tsai, Yun-Jui, et al.
Published: (2026)
Orthogonalized Policy Optimization:Policy Optimization as Orthogonal Projection in Hilbert Space
by: Zixian, Wang
Published: (2026)
by: Zixian, Wang
Published: (2026)
Scaling Laws for Precision in High-Dimensional Linear Regression
by: Zhang, Dechen, et al.
Published: (2026)
by: Zhang, Dechen, et al.
Published: (2026)
Soft Sequence Policy Optimization
by: Glazyrina, Svetlana, et al.
Published: (2026)
by: Glazyrina, Svetlana, et al.
Published: (2026)
Positional-aware Spatio-Temporal Network for Large-Scale Traffic Prediction
by: Chen, Runfei
Published: (2026)
by: Chen, Runfei
Published: (2026)
AviaSafe: A Physics-Informed Data-Driven Model for Aviation Safety-Critical Cloud Forecasts
by: Zhu, Zijian, et al.
Published: (2026)
by: Zhu, Zijian, et al.
Published: (2026)
A 1/R Law for Kurtosis Contrast in Balanced Mixtures
by: Bi, Yuda, et al.
Published: (2026)
by: Bi, Yuda, et al.
Published: (2026)
Structure and Redundancy in Large Language Models: A Spectral Study via Random Matrix Theory
by: Ettori, Davide
Published: (2026)
by: Ettori, Davide
Published: (2026)
Revisiting Chebyshev Polynomial and Anisotropic RBF Models for Tabular Regression
by: Gerber, Luciano, et al.
Published: (2026)
by: Gerber, Luciano, et al.
Published: (2026)
Operationalizing Fairness: Post-Hoc Threshold Optimization Under Hard Resource Limits
by: Singh, Moirangthem Tiken, et al.
Published: (2026)
by: Singh, Moirangthem Tiken, et al.
Published: (2026)
Correcting Human Labels for Rater Effects in AI Evaluation: An Item Response Theory Approach
by: Casabianca, Jodi M., et al.
Published: (2026)
by: Casabianca, Jodi M., et al.
Published: (2026)
SOTAlign: Semi-Supervised Alignment of Unimodal Vision and Language Models via Optimal Transport
by: Roschmann, Simon, et al.
Published: (2026)
by: Roschmann, Simon, et al.
Published: (2026)
Similar Items
-
gpt-oss-120b & gpt-oss-20b Model Card
by: OpenAI, et al.
Published: (2025) -
In AI Sweet Harmony: Sociopragmatic Guardrail Bypasses and Evaluation-Awareness in OpenAI gpt-oss-20b
by: Durner, Nils
Published: (2025) -
Standardization of Psychiatric Diagnoses -- Role of Fine-tuned LLM Consortium and OpenAI-gpt-oss Reasoning LLM Enabled Decision Support System
by: Bandara, Eranga, et al.
Published: (2025) -
Library learning with e-graphs on jazz harmony
by: Ren, Zeng, et al.
Published: (2026) -
Standardization of Neuromuscular Reflex Analysis -- Role of Fine-Tuned Vision-Language Model Consortium and OpenAI gpt-oss Reasoning LLM Enabled Decision Support System
by: Bandara, Eranga, et al.
Published: (2025)