Suchergebnisse - (coding OR Decoding) (algorithm OR algorithmsssic)

  1. 8161

    Language Models that Think, Chat Better von Bhaskar, Adithya, Ye, Xi, Chen, Danqi

    Veröffentlicht 2025
    Inhaltsangabe: “… rule-based rewards in verifiable domains such as mathematics and code. However, RLVR leads to limited …”
    Volltext
    Preprint
  2. 8162

    Active Attacks: Red-teaming LLMs via Adaptive Environments von Yun, Taeyoung, St-Charles, Pierre-Luc, Park, Jinkyoo, Bengio, Yoshua, Kim, Minsu

    Veröffentlicht 2025
    Inhaltsangabe: “… exploration, we introduce \textit{Active Attacks}, a novel RL-based red-teaming algorithm that adapts its …”
    Volltext
    Preprint
  3. 8163

    From Static to Dynamic: Adaptive Monte Carlo Search for Mathematical Process Supervision von Ma, Jie, Qi, Shihao, Xing, Rui, Yin, Ziang, Wei, Bifan, Liu, Jun, Liu, Tongliang

    Veröffentlicht 2025
    Inhaltsangabe: “… , it enhances the path expansion through a Monte Carlo algorithm with a temporally adaptive policy that begins …”
    Volltext
    Preprint
  4. 8164

    Vision-Zero: Scalable VLM Self-Improvement via Strategic Gamified Self-Play von Wang, Qinsi, Liu, Bo, Zhou, Tianyi, Shi, Jing, Lin, Yueqian, Chen, Yiran, Li, Hai Helen, Wan, Kun, Zhao, Wentian

    Veröffentlicht 2025
    Inhaltsangabe: “… -SPO), a novel training algorithm that alternates between Self-Play and reinforcement learning …”
    Volltext
    Preprint
  5. 8165

    Who Should Have a Place on the Ark? Parameterized Algorithms for the Maximization of Phylogenetic Diversity von Schestag, Jannik

    Veröffentlicht 2025
    Inhaltsangabe: “… after which they can no longer be saved.For Time-PD,we present color-coding algorithms that prove …”
    Volltext
    Preprint
  6. 8166

    FOLK: Fast Open-Vocabulary 3D Instance Segmentation via Label-guided Knowledge Distillation von Wu, Hongrui, Gao, Zhicheng, Cao, Jin, Yao, Kelu, Shen, Wen, Wei, Zhihua

    Veröffentlicht 2025
    Inhaltsangabe: “… propose a label-guided distillation algorithm to distill open-vocabulary knowledge from label-consistent 2 …”
    Volltext
    Preprint
  7. 8167

    Probabilistic Hyper-Graphs using Multiple Randomly Masked Autoencoders for Semi-supervised Multi-modal Multi-task Learning von Mihai-Cristian, Pîrvu, Leordeanu, Marius

    Veröffentlicht 2025
    Inhaltsangabe: “… the distribution of hyper-edges on each forward pass. Additionally, the model adapts the standard MAE algorithm …”
    Volltext
    Preprint
  8. 8168

    FedAdamW: A Communication-Efficient Optimizer with Convergence and Generalization Guarantees for Federated Large Models von Liu, Junkang, Shang, Fanhua, Liu, Hongying, Tian, Yuxuan, Liu, Yuanyuan, Liu, Jin, Zhu, Kewen, Lin, Zhouchen

    Veröffentlicht 2025
    Inhaltsangabe: “… \underline{AdamW} algorithm, called \texttt{FedAdamW}, for training and fine-tuning various large models …”
    Volltext
    Preprint
  9. 8169

    CORGI: Efficient Pattern Matching With Quadratic Guarantees von Weitekamp, Daniel

    Veröffentlicht 2025
    Inhaltsangabe: “… automatically generate rules from example-driven induction or code synthesis, they can easily produce worst-case …”
    Volltext
    Preprint
  10. 8170

    MoDES: Accelerating Mixture-of-Experts Multimodal Large Language Models via Dynamic Expert Skipping von Huang, Yushi, Wang, Zining, Yuan, Zhihang, Ding, Yifu, Gong, Ruihao, Guo, Jinyang, Liu, Xianglong, Zhang, Jun

    Veröffentlicht 2025
    Inhaltsangabe: “… the optimal thresholds, we introduce a frontier search algorithm that exploits monotonicity properties …”
    Volltext
    Preprint
  11. 8171

    Solving Spatial Supersensing Without Spatial Supersensing von Udandarao, Vishaal, Karthik, Shyamgopal, Nath, Surabhi S., Hochlehnert, Andreas, Bethge, Matthias, Prabhu, Ameya

    Veröffentlicht 2025
    Inhaltsangabe: “… object-count predictions unchanged; instead, Cambrian-S inference algorithm relies largely on a shortcut …”
    Volltext
    Preprint
  12. 8172

    Spira: Exploiting Voxel Data Structural Properties for Efficient Sparse Convolution in Point Cloud Networks von Adamopoulos, Dionysios, Poulopoulou, Anastasia, Goumas, Georgios, Giannoula, Christina

    Veröffentlicht 2025
    Inhaltsangabe: “… for GPUs. Spira proposes (i) a high-performance one-shot search algorithm that builds the kernel map …”
    Volltext
    Preprint
  13. 8173
  14. 8174

    Semantic Soft Bootstrapping: Long Context Reasoning in LLMs without Reinforcement Learning von Mitra, Purbesh, Ulukus, Sennur

    Veröffentlicht 2025
    Inhaltsangabe: “… optimization (GRPO), which is a commonly used RLVR algorithm. Our code is available at https://github.com …”
    Volltext
    Preprint
  15. 8175

    How Far are Modern Trackers from UAV-Anti-UAV? A Million-Scale Benchmark and New Baseline von Zhang, Chunhui, Liu, Li, Zhang, Zhipeng, Wang, Yong, Wen, Hao, Zhou, Xi, Ge, Shiming, Wang, Yanfeng

    Veröffentlicht 2025
    Inhaltsangabe: “… of 50 modern deep tracking algorithms demonstrates that there is still significant room for improvement …”
    Volltext
    Preprint
  16. 8176

    TraPO: A Semi-Supervised Reinforcement Learning Framework for Boosting LLM Reasoning von Yang, Shenzhi, Zhu, Guangcheng, Zheng, Xing, MA, Yingfan, Chen, Zhongqi, Song, Bowen, Wang, Weiqiang, Zhao, Junbo, Chen, Gang, Wang, Haobo

    Veröffentlicht 2025
    Inhaltsangabe: “… an effective policy optimization algorithm, TraPO, that identifies reliable unlabeled samples by matching …”
    Volltext
    Preprint
  17. 8177

    HERO-Sign: Hierarchical Tuning and Efficient Compiler-Time GPU Optimizations for SPHINCS+ Signature Generation von Zhou, Yaoyun, Wang, Qian

    Veröffentlicht 2025
    Inhaltsangabe: “… algorithm that adapts fusion schemes to different GPU architectures. To further improve performance, HERO …”
    Volltext
    Preprint
  18. 8178

    SP-Rank: A Dataset for Ranked Preferences with Secondary Information von Hosseini, Hadi, Mandal, Debmalya, Puhan, Amrit

    Veröffentlicht 2025
    Inhaltsangabe: “… algorithms that leverage both first-order preferences and second-order predictions in ranking tasks. Each …”
    Volltext
    Preprint
  19. 8179

    TowerMind: A Tower Defence Game Learning Environment and Benchmark for LLM as Agents von Wang, Dawei, Zhou, Chengming, Zhao, Di, Liu, Xinyuan, Ma, Marci Chi, Ushaw, Gary, Davison, Richard

    Veröffentlicht 2026
    Inhaltsangabe: “… evaluate two classic reinforcement learning algorithms: Ape-X DQN and PPO. By offering a lightweight …”
    Volltext
    Preprint
  20. 8180

    Streaming Operator Inference for Model Reduction of Large-Scale Dynamical Systems von Koike, Tomoki, Mohan, Prakash, de Frahan, Marc T. Henry, Bessac, Julie, Qian, Elizabeth

    Veröffentlicht 2026
    Inhaltsangabe: “… access to the underlying code or full model operators, i.e., non-intrusively. Traditional OpInf operates …”
    Volltext
    Preprint