Risultati della ricerca - (coding OR Decoding) algorithm

  1. 8161

    Language Models that Think, Chat Better di Bhaskar, Adithya, Ye, Xi, Chen, Danqi

    Pubblicazione 2025
    Sommario: “... rule-based rewards in verifiable domains such as mathematics and code. However, RLVR leads to limited...”
    Testo
    Preprint
  2. 8162

    Active Attacks: Red-teaming LLMs via Adaptive Environments di Yun, Taeyoung, St-Charles, Pierre-Luc, Park, Jinkyoo, Bengio, Yoshua, Kim, Minsu

    Pubblicazione 2025
    Sommario: “... exploration, we introduce \textit{Active Attacks}, a novel RL-based red-teaming algorithm that adapts its...”
    Testo
    Preprint
  3. 8163

    From Static to Dynamic: Adaptive Monte Carlo Search for Mathematical Process Supervision di Ma, Jie, Qi, Shihao, Xing, Rui, Yin, Ziang, Wei, Bifan, Liu, Jun, Liu, Tongliang

    Pubblicazione 2025
    Sommario: “..., it enhances the path expansion through a Monte Carlo algorithm with a temporally adaptive policy that begins...”
    Testo
    Preprint
  4. 8164

    Vision-Zero: Scalable VLM Self-Improvement via Strategic Gamified Self-Play di Wang, Qinsi, Liu, Bo, Zhou, Tianyi, Shi, Jing, Lin, Yueqian, Chen, Yiran, Li, Hai Helen, Wan, Kun, Zhao, Wentian

    Pubblicazione 2025
    Sommario: “...-SPO), a novel training algorithm that alternates between Self-Play and reinforcement learning...”
    Testo
    Preprint
  5. 8165

    Who Should Have a Place on the Ark? Parameterized Algorithms for the Maximization of Phylogenetic Diversity di Schestag, Jannik

    Pubblicazione 2025
    Sommario: “... after which they can no longer be saved.For Time-PD,we present color-coding algorithms that prove...”
    Testo
    Preprint
  6. 8166

    FOLK: Fast Open-Vocabulary 3D Instance Segmentation via Label-guided Knowledge Distillation di Wu, Hongrui, Gao, Zhicheng, Cao, Jin, Yao, Kelu, Shen, Wen, Wei, Zhihua

    Pubblicazione 2025
    Sommario: “... propose a label-guided distillation algorithm to distill open-vocabulary knowledge from label-consistent 2...”
    Testo
    Preprint
  7. 8167

    Probabilistic Hyper-Graphs using Multiple Randomly Masked Autoencoders for Semi-supervised Multi-modal Multi-task Learning di Mihai-Cristian, Pîrvu, Leordeanu, Marius

    Pubblicazione 2025
    Sommario: “... the distribution of hyper-edges on each forward pass. Additionally, the model adapts the standard MAE algorithm...”
    Testo
    Preprint
  8. 8168

    FedAdamW: A Communication-Efficient Optimizer with Convergence and Generalization Guarantees for Federated Large Models di Liu, Junkang, Shang, Fanhua, Liu, Hongying, Tian, Yuxuan, Liu, Yuanyuan, Liu, Jin, Zhu, Kewen, Lin, Zhouchen

    Pubblicazione 2025
    Sommario: “... \underline{AdamW} algorithm, called \texttt{FedAdamW}, for training and fine-tuning various large models...”
    Testo
    Preprint
  9. 8169

    CORGI: Efficient Pattern Matching With Quadratic Guarantees di Weitekamp, Daniel

    Pubblicazione 2025
    Sommario: “... automatically generate rules from example-driven induction or code synthesis, they can easily produce worst-case...”
    Testo
    Preprint
  10. 8170

    MoDES: Accelerating Mixture-of-Experts Multimodal Large Language Models via Dynamic Expert Skipping di Huang, Yushi, Wang, Zining, Yuan, Zhihang, Ding, Yifu, Gong, Ruihao, Guo, Jinyang, Liu, Xianglong, Zhang, Jun

    Pubblicazione 2025
    Sommario: “... the optimal thresholds, we introduce a frontier search algorithm that exploits monotonicity properties...”
    Testo
    Preprint
  11. 8171

    Solving Spatial Supersensing Without Spatial Supersensing di Udandarao, Vishaal, Karthik, Shyamgopal, Nath, Surabhi S., Hochlehnert, Andreas, Bethge, Matthias, Prabhu, Ameya

    Pubblicazione 2025
    Sommario: “... object-count predictions unchanged; instead, Cambrian-S inference algorithm relies largely on a shortcut...”
    Testo
    Preprint
  12. 8172

    Spira: Exploiting Voxel Data Structural Properties for Efficient Sparse Convolution in Point Cloud Networks di Adamopoulos, Dionysios, Poulopoulou, Anastasia, Goumas, Georgios, Giannoula, Christina

    Pubblicazione 2025
    Sommario: “... for GPUs. Spira proposes (i) a high-performance one-shot search algorithm that builds the kernel map...”
    Testo
    Preprint
  13. 8173
  14. 8174

    Semantic Soft Bootstrapping: Long Context Reasoning in LLMs without Reinforcement Learning di Mitra, Purbesh, Ulukus, Sennur

    Pubblicazione 2025
    Sommario: “... optimization (GRPO), which is a commonly used RLVR algorithm. Our code is available at https://github.com...”
    Testo
    Preprint
  15. 8175

    How Far are Modern Trackers from UAV-Anti-UAV? A Million-Scale Benchmark and New Baseline di Zhang, Chunhui, Liu, Li, Zhang, Zhipeng, Wang, Yong, Wen, Hao, Zhou, Xi, Ge, Shiming, Wang, Yanfeng

    Pubblicazione 2025
    Sommario: “... of 50 modern deep tracking algorithms demonstrates that there is still significant room for improvement...”
    Testo
    Preprint
  16. 8176

    TraPO: A Semi-Supervised Reinforcement Learning Framework for Boosting LLM Reasoning di Yang, Shenzhi, Zhu, Guangcheng, Zheng, Xing, MA, Yingfan, Chen, Zhongqi, Song, Bowen, Wang, Weiqiang, Zhao, Junbo, Chen, Gang, Wang, Haobo

    Pubblicazione 2025
    Sommario: “... an effective policy optimization algorithm, TraPO, that identifies reliable unlabeled samples by matching...”
    Testo
    Preprint
  17. 8177

    HERO-Sign: Hierarchical Tuning and Efficient Compiler-Time GPU Optimizations for SPHINCS+ Signature Generation di Zhou, Yaoyun, Wang, Qian

    Pubblicazione 2025
    Sommario: “... algorithm that adapts fusion schemes to different GPU architectures. To further improve performance, HERO...”
    Testo
    Preprint
  18. 8178

    SP-Rank: A Dataset for Ranked Preferences with Secondary Information di Hosseini, Hadi, Mandal, Debmalya, Puhan, Amrit

    Pubblicazione 2025
    Sommario: “... algorithms that leverage both first-order preferences and second-order predictions in ranking tasks. Each...”
    Testo
    Preprint
  19. 8179

    TowerMind: A Tower Defence Game Learning Environment and Benchmark for LLM as Agents di Wang, Dawei, Zhou, Chengming, Zhao, Di, Liu, Xinyuan, Ma, Marci Chi, Ushaw, Gary, Davison, Richard

    Pubblicazione 2026
    Sommario: “... evaluate two classic reinforcement learning algorithms: Ape-X DQN and PPO. By offering a lightweight...”
    Testo
    Preprint
  20. 8180

    Streaming Operator Inference for Model Reduction of Large-Scale Dynamical Systems di Koike, Tomoki, Mohan, Prakash, de Frahan, Marc T. Henry, Bessac, Julie, Qian, Elizabeth

    Pubblicazione 2026
    Sommario: “... access to the underlying code or full model operators, i.e., non-intrusively. Traditional OpInf operates...”
    Testo
    Preprint