Risultati della ricerca - (decoding OR coding) (algorithm OR (algorithmics OR algorithmics))

  1. 9121

    ST-Raptor: LLM-Powered Semi-Structured Table Question Answering di Tang, Zirui, Niu, Boyu, Zhou, Xuanhe, Li, Boxiu, Zhou, Wei, Wang, Jiannan, Li, Guoliang, Zhang, Xinyi, Wu, Fan

    Pubblicazione 2025
    Sommario: “... with an effective algorithm for constructing the tree. Second, we define a set of basic tree operations to guide...”
    Testo
    Preprint
  2. 9122
  3. 9123

    Language Models that Think, Chat Better di Bhaskar, Adithya, Ye, Xi, Chen, Danqi

    Pubblicazione 2025
    Sommario: “... rule-based rewards in verifiable domains such as mathematics and code. However, RLVR leads to limited...”
    Testo
    Preprint
  4. 9124

    Active Attacks: Red-teaming LLMs via Adaptive Environments di Yun, Taeyoung, St-Charles, Pierre-Luc, Park, Jinkyoo, Bengio, Yoshua, Kim, Minsu

    Pubblicazione 2025
    Sommario: “... exploration, we introduce \textit{Active Attacks}, a novel RL-based red-teaming algorithm that adapts its...”
    Testo
    Preprint
  5. 9125

    From Static to Dynamic: Adaptive Monte Carlo Search for Mathematical Process Supervision di Ma, Jie, Qi, Shihao, Xing, Rui, Yin, Ziang, Wei, Bifan, Liu, Jun, Liu, Tongliang

    Pubblicazione 2025
    Sommario: “..., it enhances the path expansion through a Monte Carlo algorithm with a temporally adaptive policy that begins...”
    Testo
    Preprint
  6. 9126

    Vision-Zero: Scalable VLM Self-Improvement via Strategic Gamified Self-Play di Wang, Qinsi, Liu, Bo, Zhou, Tianyi, Shi, Jing, Lin, Yueqian, Chen, Yiran, Li, Hai Helen, Wan, Kun, Zhao, Wentian

    Pubblicazione 2025
    Sommario: “...-SPO), a novel training algorithm that alternates between Self-Play and reinforcement learning...”
    Testo
    Preprint
  7. 9127

    Who Should Have a Place on the Ark? Parameterized Algorithms for the Maximization of Phylogenetic Diversity di Schestag, Jannik

    Pubblicazione 2025
    Sommario: “... after which they can no longer be saved.For Time-PD,we present color-coding algorithms that prove...”
    Testo
    Preprint
  8. 9128

    FOLK: Fast Open-Vocabulary 3D Instance Segmentation via Label-guided Knowledge Distillation di Wu, Hongrui, Gao, Zhicheng, Cao, Jin, Yao, Kelu, Shen, Wen, Wei, Zhihua

    Pubblicazione 2025
    Sommario: “... propose a label-guided distillation algorithm to distill open-vocabulary knowledge from label-consistent 2...”
    Testo
    Preprint
  9. 9129

    Probabilistic Hyper-Graphs using Multiple Randomly Masked Autoencoders for Semi-supervised Multi-modal Multi-task Learning di Mihai-Cristian, Pîrvu, Leordeanu, Marius

    Pubblicazione 2025
    Sommario: “... the distribution of hyper-edges on each forward pass. Additionally, the model adapts the standard MAE algorithm...”
    Testo
    Preprint
  10. 9130

    FedAdamW: A Communication-Efficient Optimizer with Convergence and Generalization Guarantees for Federated Large Models di Liu, Junkang, Shang, Fanhua, Liu, Hongying, Tian, Yuxuan, Liu, Yuanyuan, Liu, Jin, Zhu, Kewen, Lin, Zhouchen

    Pubblicazione 2025
    Sommario: “... \underline{AdamW} algorithm, called \texttt{FedAdamW}, for training and fine-tuning various large models...”
    Testo
    Preprint
  11. 9131

    CORGI: Efficient Pattern Matching With Quadratic Guarantees di Weitekamp, Daniel

    Pubblicazione 2025
    Sommario: “... automatically generate rules from example-driven induction or code synthesis, they can easily produce worst-case...”
    Testo
    Preprint
  12. 9132

    Solving Spatial Supersensing Without Spatial Supersensing di Udandarao, Vishaal, Karthik, Shyamgopal, Nath, Surabhi S., Hochlehnert, Andreas, Bethge, Matthias, Prabhu, Ameya

    Pubblicazione 2025
    Sommario: “... object-count predictions unchanged; instead, Cambrian-S inference algorithm relies largely on a shortcut...”
    Testo
    Preprint
  13. 9133

    Spira: Exploiting Voxel Data Structural Properties for Efficient Sparse Convolution in Point Cloud Networks di Adamopoulos, Dionysios, Poulopoulou, Anastasia, Goumas, Georgios, Giannoula, Christina

    Pubblicazione 2025
    Sommario: “... for GPUs. Spira proposes (i) a high-performance one-shot search algorithm that builds the kernel map...”
    Testo
    Preprint
  14. 9134
  15. 9135

    Semantic Soft Bootstrapping: Long Context Reasoning in LLMs without Reinforcement Learning di Mitra, Purbesh, Ulukus, Sennur

    Pubblicazione 2025
    Sommario: “... optimization (GRPO), which is a commonly used RLVR algorithm. Our code is available at https://github.com...”
    Testo
    Preprint
  16. 9136

    How Far are Modern Trackers from UAV-Anti-UAV? A Million-Scale Benchmark and New Baseline di Zhang, Chunhui, Liu, Li, Zhang, Zhipeng, Wang, Yong, Wen, Hao, Zhou, Xi, Ge, Shiming, Wang, Yanfeng

    Pubblicazione 2025
    Sommario: “... of 50 modern deep tracking algorithms demonstrates that there is still significant room for improvement...”
    Testo
    Preprint
  17. 9137

    TraPO: A Semi-Supervised Reinforcement Learning Framework for Boosting LLM Reasoning di Yang, Shenzhi, Zhu, Guangcheng, Zheng, Xing, MA, Yingfan, Chen, Zhongqi, Song, Bowen, Wang, Weiqiang, Zhao, Junbo, Chen, Gang, Wang, Haobo

    Pubblicazione 2025
    Sommario: “... an effective policy optimization algorithm, TraPO, that identifies reliable unlabeled samples by matching...”
    Testo
    Preprint
  18. 9138

    HERO-Sign: Hierarchical Tuning and Efficient Compiler-Time GPU Optimizations for SPHINCS+ Signature Generation di Zhou, Yaoyun, Wang, Qian

    Pubblicazione 2025
    Sommario: “... algorithm that adapts fusion schemes to different GPU architectures. To further improve performance, HERO...”
    Testo
    Preprint
  19. 9139

    SP-Rank: A Dataset for Ranked Preferences with Secondary Information di Hosseini, Hadi, Mandal, Debmalya, Puhan, Amrit

    Pubblicazione 2025
    Sommario: “... algorithms that leverage both first-order preferences and second-order predictions in ranking tasks. Each...”
    Testo
    Preprint
  20. 9140

    TowerMind: A Tower Defence Game Learning Environment and Benchmark for LLM as Agents di Wang, Dawei, Zhou, Chengming, Zhao, Di, Liu, Xinyuan, Ma, Marci Chi, Ushaw, Gary, Davison, Richard

    Pubblicazione 2026
    Sommario: “... evaluate two classic reinforcement learning algorithms: Ape-X DQN and PPO. By offering a lightweight...”
    Testo
    Preprint