Gespeichert in:
| Hauptverfasser: | Shorewala, Shivam, Yang, Zihao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2506.00348 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Anomaly Detection and Improvement of Clusters using Enhanced K-Means Algorithm
von: Shorewala, Vardhan, et al.
Veröffentlicht: (2025)
von: Shorewala, Vardhan, et al.
Veröffentlicht: (2025)
Skill Issues: An Analysis of CS:GO Skill Rating Systems
von: Bober-Irizar, Mikel, et al.
Veröffentlicht: (2024)
von: Bober-Irizar, Mikel, et al.
Veröffentlicht: (2024)
Safe RLHF Beyond Expectation: Stochastic Dominance for Universal Spectral Risk Control
von: Chittepu, Yaswanth, et al.
Veröffentlicht: (2026)
von: Chittepu, Yaswanth, et al.
Veröffentlicht: (2026)
Find A Winning Sign: Sign Is All We Need to Win the Lottery
von: Oh, Junghun, et al.
Veröffentlicht: (2025)
von: Oh, Junghun, et al.
Veröffentlicht: (2025)
Beyond Expectations: Learning with Stochastic Dominance Made Practical
von: Cen, Shicong, et al.
Veröffentlicht: (2024)
von: Cen, Shicong, et al.
Veröffentlicht: (2024)
Expectation Alignment: Handling Reward Misspecification in the Presence of Expectation Mismatch
von: Mechergui, Malek, et al.
Veröffentlicht: (2024)
von: Mechergui, Malek, et al.
Veröffentlicht: (2024)
MetaGen Blended RAG: Unlocking Zero-Shot Precision for Specialized Domain Question-Answering
von: Sawarkar, Kunal, et al.
Veröffentlicht: (2025)
von: Sawarkar, Kunal, et al.
Veröffentlicht: (2025)
Mini-Giants: "Small" Language Models and Open Source Win-Win
von: Zhou, Zhengping, et al.
Veröffentlicht: (2023)
von: Zhou, Zhengping, et al.
Veröffentlicht: (2023)
Cheating Automatic LLM Benchmarks: Null Models Achieve High Win Rates
von: Zheng, Xiaosen, et al.
Veröffentlicht: (2024)
von: Zheng, Xiaosen, et al.
Veröffentlicht: (2024)
Neural Expectation Operators
von: Qi, Qian
Veröffentlicht: (2025)
von: Qi, Qian
Veröffentlicht: (2025)
Mesh-based Super-resolution of Detonation Flows with Multiscale Graph Transformers
von: Barwey, Shivam, et al.
Veröffentlicht: (2025)
von: Barwey, Shivam, et al.
Veröffentlicht: (2025)
Sustained Gradient Alignment Mediates Subliminal Learning in a Multi-Step Setting: Evidence from MNIST Auxiliary Logit Distillation Experiment
von: Kitkana, Chayanon, et al.
Veröffentlicht: (2026)
von: Kitkana, Chayanon, et al.
Veröffentlicht: (2026)
Beyond Token-level Supervision: Unlocking the Potential of Decoding-based Regression via Reinforcement Learning
von: Chen, Ming, et al.
Veröffentlicht: (2025)
von: Chen, Ming, et al.
Veröffentlicht: (2025)
Unlocking In-Context Learning for Natural Datasets Beyond Language Modelling
von: Bratulić, Jelena, et al.
Veröffentlicht: (2025)
von: Bratulić, Jelena, et al.
Veröffentlicht: (2025)
UpSkill: Mutual Information Skill Learning for Structured Response Diversity in LLMs
von: Shah, Devan, et al.
Veröffentlicht: (2026)
von: Shah, Devan, et al.
Veröffentlicht: (2026)
eMargin: Revisiting Contrastive Learning with Margin-Based Separation
von: Shamba, Abdul-Kazeem, et al.
Veröffentlicht: (2025)
von: Shamba, Abdul-Kazeem, et al.
Veröffentlicht: (2025)
Discovering Data Structures: Nearest Neighbor Search and Beyond
von: Salemohamed, Omar, et al.
Veröffentlicht: (2024)
von: Salemohamed, Omar, et al.
Veröffentlicht: (2024)
Winning Amazon KDD Cup'24
von: Deotte, Chris, et al.
Veröffentlicht: (2024)
von: Deotte, Chris, et al.
Veröffentlicht: (2024)
Slow and Steady Wins the Race: Maintaining Plasticity with Hare and Tortoise Networks
von: Lee, Hojoon, et al.
Veröffentlicht: (2024)
von: Lee, Hojoon, et al.
Veröffentlicht: (2024)
Uncovering a Winning Lottery Ticket with Continuously Relaxed Bernoulli Gates
von: Tsayag, Itamar, et al.
Veröffentlicht: (2026)
von: Tsayag, Itamar, et al.
Veröffentlicht: (2026)
A Competition Winning Deep Reinforcement Learning Agent in microRTS
von: Goodfriend, Scott
Veröffentlicht: (2024)
von: Goodfriend, Scott
Veröffentlicht: (2024)
Marginals Before Conditionals
von: Sahasrabudhe, Mihir
Veröffentlicht: (2026)
von: Sahasrabudhe, Mihir
Veröffentlicht: (2026)
Generative Marginalization Models
von: Liu, Sulin, et al.
Veröffentlicht: (2023)
von: Liu, Sulin, et al.
Veröffentlicht: (2023)
A Unified Theory of $θ$-Expectations
von: Qi, Qian
Veröffentlicht: (2025)
von: Qi, Qian
Veröffentlicht: (2025)
Towards a Law of Iterated Expectations for Heuristic Estimators
von: Christiano, Paul, et al.
Veröffentlicht: (2024)
von: Christiano, Paul, et al.
Veröffentlicht: (2024)
GEPO: Group Expectation Policy Optimization for Stable Heterogeneous Reinforcement Learning
von: Zhang, Han, et al.
Veröffentlicht: (2025)
von: Zhang, Han, et al.
Veröffentlicht: (2025)
Hybrid Fourier Neural Operator-Plasma Fluid Model for Fast and Accurate Multiscale Simulations of High Power Microwave Breakdown
von: Pandya, Kalp, et al.
Veröffentlicht: (2025)
von: Pandya, Kalp, et al.
Veröffentlicht: (2025)
ReLU$^2$ Wins: Discovering Efficient Activation Functions for Sparse LLMs
von: Zhang, Zhengyan, et al.
Veröffentlicht: (2024)
von: Zhang, Zhengyan, et al.
Veröffentlicht: (2024)
LOCUS: Low-Dimensional Model Embeddings for Efficient Model Exploration, Comparison, and Selection
von: Patel, Shivam, et al.
Veröffentlicht: (2026)
von: Patel, Shivam, et al.
Veröffentlicht: (2026)
Segment, Shuffle, and Stitch: A Simple Layer for Improving Time-Series Representations
von: Grover, Shivam, et al.
Veröffentlicht: (2024)
von: Grover, Shivam, et al.
Veröffentlicht: (2024)
Correlated Proxies: A New Definition and Improved Mitigation for Reward Hacking
von: Laidlaw, Cassidy, et al.
Veröffentlicht: (2024)
von: Laidlaw, Cassidy, et al.
Veröffentlicht: (2024)
An Expectation-Maximization Algorithm for Domain Adaptation in Gaussian Causal Models
von: Javidian, Mohammad Ali
Veröffentlicht: (2026)
von: Javidian, Mohammad Ali
Veröffentlicht: (2026)
A Mean-Field Theory of $Θ$-Expectations
von: Qi, Qian
Veröffentlicht: (2025)
von: Qi, Qian
Veröffentlicht: (2025)
Recursive Inference Scaling: A Winning Path to Scalable Inference in Language and Multimodal Systems
von: Alabdulmohsin, Ibrahim, et al.
Veröffentlicht: (2025)
von: Alabdulmohsin, Ibrahim, et al.
Veröffentlicht: (2025)
ToolACE: Winning the Points of LLM Function Calling
von: Liu, Weiwen, et al.
Veröffentlicht: (2024)
von: Liu, Weiwen, et al.
Veröffentlicht: (2024)
Global Sensitivity Analysis for Engineering Design Based on Individual Conditional Expectations
von: Palar, Pramudita Satria, et al.
Veröffentlicht: (2025)
von: Palar, Pramudita Satria, et al.
Veröffentlicht: (2025)
Explaining Machine Learning Predictive Models through Conditional Expectation Methods
von: Ruiz-España, Silvia, et al.
Veröffentlicht: (2026)
von: Ruiz-España, Silvia, et al.
Veröffentlicht: (2026)
OptSkills: Learning Generalizable Optimization Skills from Problem Archetypes via Cluster-Based Distillation
von: Yang, Haochen, et al.
Veröffentlicht: (2026)
von: Yang, Haochen, et al.
Veröffentlicht: (2026)
Regulating Model Reliance on Non-Robust Features by Smoothing Input Marginal Density
von: Yang, Peiyu, et al.
Veröffentlicht: (2024)
von: Yang, Peiyu, et al.
Veröffentlicht: (2024)
Reinforcement Learning with Conditional Expectation Reward
von: Xiao, Changyi, et al.
Veröffentlicht: (2026)
von: Xiao, Changyi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Anomaly Detection and Improvement of Clusters using Enhanced K-Means Algorithm
von: Shorewala, Vardhan, et al.
Veröffentlicht: (2025) -
Skill Issues: An Analysis of CS:GO Skill Rating Systems
von: Bober-Irizar, Mikel, et al.
Veröffentlicht: (2024) -
Safe RLHF Beyond Expectation: Stochastic Dominance for Universal Spectral Risk Control
von: Chittepu, Yaswanth, et al.
Veröffentlicht: (2026) -
Find A Winning Sign: Sign Is All We Need to Win the Lottery
von: Oh, Junghun, et al.
Veröffentlicht: (2025) -
Beyond Expectations: Learning with Stochastic Dominance Made Practical
von: Cen, Shicong, et al.
Veröffentlicht: (2024)