Risultati della ricerca - (decoding OR coding) (algorithmsssicics OR (algorithm OR (algorithm OR algorithm)))
Cerca anche:
- decoding »
- coding »
- algorithmsssicics »
- algorithm »
-
8241
Multistability and intermingledness in complex high-dimensional data
Pubblicazione 2026Sommario: “... that analyses potentially multistable simulation data and decides algorithmically what are the alternative...”
Testo
Preprint -
8242
Robust Optimization for Mitigating Reward Hacking with Correlated Proxies
Pubblicazione 2026Sommario: “... environments show that our algorithms consistently outperform ORPO in worst-case returns, and offer improved...”
Testo
Preprint -
8243
Wave-Based Dispatch for Circuit Cutting in Hybrid HPC--Quantum Systems
Pubblicazione 2026Sommario: “... layers to parse quantum code, a wave-based coordinator that achieves pipeline concurrency via non...”
Testo
Preprint -
8244
CAR-EnKF: A Covariance-Adaptive and Recalibrated Ensemble Kalman Filter Framework
Pubblicazione 2026Sommario: “... online. The framework is algorithmically general and is specialized here to the stochastic EnKF...”
Testo
Preprint -
8245
NLPOpt-Net: A Learning Method for Nonlinear Optimization with Feasibility Guarantees
Pubblicazione 2026Sommario: “... the NN predictions further. NLPOpt-Net deploys an inversion-free, modified Chambolle-Pock algorithm...”
Testo
Preprint -
8246
Hitting Time Isomorphism for Multi-Stage Planning with Foundation Policies
Pubblicazione 2026Sommario: “... Isomorphic Embedding Learning (IEL) as a new goal-agnostic foundation policy learning algorithm that anchors...”
Testo
Preprint -
8247
WiCER: Wiki-memory Compile, Evaluate, Refine Iterative Knowledge Compilation for LLM Wiki Systems
Pubblicazione 2026Sommario: “... algorithm inspired by counterexample-guided abstraction refinement (CEGAR) that closes this gap. WiCER...”
Testo
Preprint -
8248
ReasonEdit: Towards Interpretable Image Editing Evaluation via Reinforcement Learning
Pubblicazione 2026Sommario: “... signals derived from RE-Reward and the Group Relative Policy Optimization (GRPO) algorithm to learn...”
Testo
Preprint -
8249
Self-Play Enhancement via Advantage-Weighted Refinement in Online Federated LLM Fine-Tuning with Real-Time Feedback
Pubblicazione 2026Sommario: “... Advantage-Weighted Refinement), an efficient online learning algorithm for federated LLM fine-tuning. SPEAR...”
Testo
Preprint -
8250
Learning CLI Agents with Structured Action Credit under Selective Observation
Pubblicazione 2026Sommario: “.... Beyond this underused action structure, CLI learning also couples two bottlenecks for coding agents...”
Testo
Preprint -
8251
Swarm Skills: A Portable, Self-Evolving Multi-Agent System Specification for Coordination Engineering
Pubblicazione 2026Sommario: “... framework-internal code or static configurations, preventing them from being shared across systems...”
Testo
Preprint -
8252
A categorical error sensitivity index (ISEC): A preventive ordinal decision-support measure for irrecoverable errors in manual data entry systems
Pubblicazione 2026Sommario: “... morphological transformation costs (through an adapted Damerau Levenshtein algorithm), and empirical frequency...”
Testo
Preprint -
8253
High-Rate Quantized Matrix Multiplication II
Pubblicazione 2026Sommario: “... to the problem of weighted mean squared error (WMSE) source coding, whose classical (reverse) waterfilling...”
Testo
Preprint -
8254
Linear-Time T-Gate Optimization via Random Abstraction
Pubblicazione 2026Sommario: “... must be encoded redundantly across many physical ones using quantum error-correcting codes. In most...”
Testo
Preprint -
8255
Adaptive Generate-Rank-Verify: Inference-Time Search with Costly Verification
Pubblicazione 2026Sommario: “... verifier, such as exact answer checking in mathematical reasoning or hidden-test execution in code...”
Testo
Preprint -
8256
Weasel: Out-of-Domain Generalization for Web Agents via Importance-Diversity Data Selection
Pubblicazione 2026Sommario: “... diversity over states, websites, and interaction patterns, solving efficiently with a greedy algorithm. We...”
Testo
Preprint -
8257
Vector Policy Optimization: Training for Diversity Improves Test-Time Search
Pubblicazione 2026Sommario: “... Optimization (VPO), an RL algorithm that explicitly trains policies to anticipate diverse downstream reward...”
Testo
Preprint -
8258
ThriftAttention: Selective Mixed Precision for Long-Context FP4 Attention
Pubblicazione 2026Sommario: “...Efficient attention algorithms are critical to mitigate the quadratic cost of attention in long...”
Testo
Preprint -
8259
Self-Refining Topology Optimization via an LLM-Based Multi-Agent Framework
Pubblicazione 2026Sommario: “... for prescribed objectives and constraints through well-established numerical algorithms. Throughout the workflow...”
Testo
Preprint -
8260
Polar: Agentic RL on Any Harness at Scale
Pubblicazione 2026Sommario: “.... This decoupled design makes Polar agnostic to agent harnesses, training infrastructure, and RL algorithms while...”
Testo
Preprint