Résultats de la recherche - (coding OR Decoding) (algorithm OR (algorithmsssic OR algorithmsssic))
Recherches alternatives:
- Decoding »
- algorithm »
- algorithmsssic »
-
8161
Language Models that Think, Chat Better
Publié 2025Table des matières: “… rule-based rewards in verifiable domains such as mathematics and code. However, RLVR leads to limited…”
Accéder au texte intégral
Preprint -
8162
Active Attacks: Red-teaming LLMs via Adaptive Environments
Publié 2025Table des matières: “… exploration, we introduce \textit{Active Attacks}, a novel RL-based red-teaming algorithm that adapts its…”
Accéder au texte intégral
Preprint -
8163
From Static to Dynamic: Adaptive Monte Carlo Search for Mathematical Process Supervision
Publié 2025Table des matières: “…, it enhances the path expansion through a Monte Carlo algorithm with a temporally adaptive policy that begins…”
Accéder au texte intégral
Preprint -
8164
Vision-Zero: Scalable VLM Self-Improvement via Strategic Gamified Self-Play
Publié 2025Table des matières: “…-SPO), a novel training algorithm that alternates between Self-Play and reinforcement learning…”
Accéder au texte intégral
Preprint -
8165
Who Should Have a Place on the Ark? Parameterized Algorithms for the Maximization of Phylogenetic Diversity
Publié 2025Table des matières: “… after which they can no longer be saved.For Time-PD,we present color-coding algorithms that prove…”
Accéder au texte intégral
Preprint -
8166
FOLK: Fast Open-Vocabulary 3D Instance Segmentation via Label-guided Knowledge Distillation
Publié 2025Table des matières: “… propose a label-guided distillation algorithm to distill open-vocabulary knowledge from label-consistent 2…”
Accéder au texte intégral
Preprint -
8167
Probabilistic Hyper-Graphs using Multiple Randomly Masked Autoencoders for Semi-supervised Multi-modal Multi-task Learning
Publié 2025Table des matières: “… the distribution of hyper-edges on each forward pass. Additionally, the model adapts the standard MAE algorithm…”
Accéder au texte intégral
Preprint -
8168
FedAdamW: A Communication-Efficient Optimizer with Convergence and Generalization Guarantees for Federated Large Models
Publié 2025Table des matières: “… \underline{AdamW} algorithm, called \texttt{FedAdamW}, for training and fine-tuning various large models…”
Accéder au texte intégral
Preprint -
8169
CORGI: Efficient Pattern Matching With Quadratic Guarantees
Publié 2025Table des matières: “… automatically generate rules from example-driven induction or code synthesis, they can easily produce worst-case…”
Accéder au texte intégral
Preprint -
8170
MoDES: Accelerating Mixture-of-Experts Multimodal Large Language Models via Dynamic Expert Skipping
Publié 2025Table des matières: “… the optimal thresholds, we introduce a frontier search algorithm that exploits monotonicity properties…”
Accéder au texte intégral
Preprint -
8171
Solving Spatial Supersensing Without Spatial Supersensing
Publié 2025Table des matières: “… object-count predictions unchanged; instead, Cambrian-S inference algorithm relies largely on a shortcut…”
Accéder au texte intégral
Preprint -
8172
Spira: Exploiting Voxel Data Structural Properties for Efficient Sparse Convolution in Point Cloud Networks
Publié 2025Table des matières: “… for GPUs. Spira proposes (i) a high-performance one-shot search algorithm that builds the kernel map…”
Accéder au texte intégral
Preprint -
8173
MICCAI STS 2024 Challenge: Semi-Supervised Instance-Level Tooth Segmentation in Panoramic X-ray and CBCT Images
Publié 2025Table des matières: “…) registered teams. To ensure algorithmic excellence and full transparency, we rigorously evaluated the valid…”
Accéder au texte intégral
Preprint -
8174
Semantic Soft Bootstrapping: Long Context Reasoning in LLMs without Reinforcement Learning
Publié 2025Table des matières: “… optimization (GRPO), which is a commonly used RLVR algorithm. Our code is available at https://github.com…”
Accéder au texte intégral
Preprint -
8175
How Far are Modern Trackers from UAV-Anti-UAV? A Million-Scale Benchmark and New Baseline
Publié 2025Table des matières: “… of 50 modern deep tracking algorithms demonstrates that there is still significant room for improvement…”
Accéder au texte intégral
Preprint -
8176
TraPO: A Semi-Supervised Reinforcement Learning Framework for Boosting LLM Reasoning
Publié 2025Table des matières: “… an effective policy optimization algorithm, TraPO, that identifies reliable unlabeled samples by matching…”
Accéder au texte intégral
Preprint -
8177
HERO-Sign: Hierarchical Tuning and Efficient Compiler-Time GPU Optimizations for SPHINCS+ Signature Generation
Publié 2025Table des matières: “… algorithm that adapts fusion schemes to different GPU architectures. To further improve performance, HERO…”
Accéder au texte intégral
Preprint -
8178
SP-Rank: A Dataset for Ranked Preferences with Secondary Information
Publié 2025Table des matières: “… algorithms that leverage both first-order preferences and second-order predictions in ranking tasks. Each…”
Accéder au texte intégral
Preprint -
8179
TowerMind: A Tower Defence Game Learning Environment and Benchmark for LLM as Agents
Publié 2026Table des matières: “… evaluate two classic reinforcement learning algorithms: Ape-X DQN and PPO. By offering a lightweight…”
Accéder au texte intégral
Preprint -
8180
Streaming Operator Inference for Model Reduction of Large-Scale Dynamical Systems
Publié 2026Table des matières: “… access to the underlying code or full model operators, i.e., non-intrusively. Traditional OpInf operates…”
Accéder au texte intégral
Preprint