Risultati della ricerca - (decoding OR coding) (algorithmsssicics OR (algorithm OR (algorithm OR algorithm)))

  1. 8241

    Multistability and intermingledness in complex high-dimensional data di Datseris, George, Lohmann, Johannes, Hamilton, Oisín, Haqq-Misra, Jacob

    Pubblicazione 2026
    Sommario: “... that analyses potentially multistable simulation data and decides algorithmically what are the alternative...”
    Testo
    Preprint
  2. 8242

    Robust Optimization for Mitigating Reward Hacking with Correlated Proxies di Liu, Zixuan, Sun, Xiaolin, Zheng, Zizhan

    Pubblicazione 2026
    Sommario: “... environments show that our algorithms consistently outperform ORPO in worst-case returns, and offer improved...”
    Testo
    Preprint
  3. 8243

    Wave-Based Dispatch for Circuit Cutting in Hybrid HPC--Quantum Systems di García-Raigada, Ricard S., Jorba, Josep, Iserte, Sergio

    Pubblicazione 2026
    Sommario: “... layers to parse quantum code, a wave-based coordinator that achieves pipeline concurrency via non...”
    Testo
    Preprint
  4. 8244

    CAR-EnKF: A Covariance-Adaptive and Recalibrated Ensemble Kalman Filter Framework di Jiang, Shida, Tao, Shengyu, Liu, Zihe, Moura, Scott

    Pubblicazione 2026
    Sommario: “... online. The framework is algorithmically general and is specialized here to the stochastic EnKF...”
    Testo
    Preprint
  5. 8245

    NLPOpt-Net: A Learning Method for Nonlinear Optimization with Feasibility Guarantees di Roy, Bimol Nath, Golder, Rahul, Hasan, MM Faruque

    Pubblicazione 2026
    Sommario: “... the NN predictions further. NLPOpt-Net deploys an inversion-free, modified Chambolle-Pock algorithm...”
    Testo
    Preprint
  6. 8246

    Hitting Time Isomorphism for Multi-Stage Planning with Foundation Policies di Boock, Magnus Victor, Akgül, Abdullah, Çelikok, Mustafa Mert, Kandemir, Melih

    Pubblicazione 2026
    Sommario: “... Isomorphic Embedding Learning (IEL) as a new goal-agnostic foundation policy learning algorithm that anchors...”
    Testo
    Preprint
  7. 8247

    WiCER: Wiki-memory Compile, Evaluate, Refine Iterative Knowledge Compilation for LLM Wiki Systems di Huerta, Juan M.

    Pubblicazione 2026
    Sommario: “... algorithm inspired by counterexample-guided abstraction refinement (CEGAR) that closes this gap. WiCER...”
    Testo
    Preprint
  8. 8248

    ReasonEdit: Towards Interpretable Image Editing Evaluation via Reinforcement Learning di Chen, Honghua, Xu, Zitong, Duan, Huiyu, Zhang, Xinyun, Min, Xiongkuo, Zhai, Guangtao

    Pubblicazione 2026
    Sommario: “... signals derived from RE-Reward and the Group Relative Policy Optimization (GRPO) algorithm to learn...”
    Testo
    Preprint
  9. 8249

    Self-Play Enhancement via Advantage-Weighted Refinement in Online Federated LLM Fine-Tuning with Real-Time Feedback di Lee, Seohyun, Fang, Wenzhi, Han, Dong-Jun, Hosseinalipour, Seyyedali, Brinton, Christopher G.

    Pubblicazione 2026
    Sommario: “... Advantage-Weighted Refinement), an efficient online learning algorithm for federated LLM fine-tuning. SPEAR...”
    Testo
    Preprint
  10. 8250

    Learning CLI Agents with Structured Action Credit under Selective Observation di Su, Haoyang, Wen, Ying

    Pubblicazione 2026
    Sommario: “.... Beyond this underused action structure, CLI learning also couples two bottlenecks for coding agents...”
    Testo
    Preprint
  11. 8251
  12. 8252
  13. 8253

    High-Rate Quantized Matrix Multiplication II di Ordentlich, Or, Polyanskiy, Yury

    Pubblicazione 2026
    Sommario: “... to the problem of weighted mean squared error (WMSE) source coding, whose classical (reverse) waterfilling...”
    Testo
    Preprint
  14. 8254

    Linear-Time T-Gate Optimization via Random Abstraction di Albarghouthi, Aws

    Pubblicazione 2026
    Sommario: “... must be encoded redundantly across many physical ones using quantum error-correcting codes. In most...”
    Testo
    Preprint
  15. 8255

    Adaptive Generate-Rank-Verify: Inference-Time Search with Costly Verification di Dughmi, Shaddin, Haghifam, Mahdi, Kalayci, Yusuf Hakan

    Pubblicazione 2026
    Sommario: “... verifier, such as exact answer checking in mathematical reasoning or hidden-test execution in code...”
    Testo
    Preprint
  16. 8256

    Weasel: Out-of-Domain Generalization for Web Agents via Importance-Diversity Data Selection di Zadeh, Fatemeh Pesaran, Choi, Seyeon, Lù, Xing Han, Reddy, Siva, Kim, Gunhee

    Pubblicazione 2026
    Sommario: “... diversity over states, websites, and interaction patterns, solving efficiently with a greedy algorithm. We...”
    Testo
    Preprint
  17. 8257

    Vector Policy Optimization: Training for Diversity Improves Test-Time Search di Bahlous-Boldi, Ryan, Puri, Isha, Shenfeld, Idan, Kumar, Akarsh, Damani, Mehul, Risi, Sebastian, Khattab, Omar, Hong, Zhang-Wei, Agrawal, Pulkit

    Pubblicazione 2026
    Sommario: “... Optimization (VPO), an RL algorithm that explicitly trains policies to anticipate diverse downstream reward...”
    Testo
    Preprint
  18. 8258

    ThriftAttention: Selective Mixed Precision for Long-Context FP4 Attention di Sharratt, Joe

    Pubblicazione 2026
    Sommario: “...Efficient attention algorithms are critical to mitigate the quadratic cost of attention in long...”
    Testo
    Preprint
  19. 8259

    Self-Refining Topology Optimization via an LLM-Based Multi-Agent Framework di Park, Hyunjee, Chung, Hayoung

    Pubblicazione 2026
    Sommario: “... for prescribed objectives and constraints through well-established numerical algorithms. Throughout the workflow...”
    Testo
    Preprint
  20. 8260

    Polar: Agentic RL on Any Harness at Scale di Xu, Binfeng, Zhang, Hao, Zhang, Shaokun, Han, Songyang, Liu, Mingjie, Hu, Jian, Diao, Shizhe, Jin, Zhenghui, Zou, Yunheng, Demoret, Michael, Kautz, Jan, Dong, Yi

    Pubblicazione 2026
    Sommario: “.... This decoupled design makes Polar agnostic to agent harnesses, training infrastructure, and RL algorithms while...”
    Testo
    Preprint