Risultati della ricerca - (coding OR Decoding) algorithm
Cerca anche:
-
8161
Language Models that Think, Chat Better
Pubblicazione 2025Sommario: “... rule-based rewards in verifiable domains such as mathematics and code. However, RLVR leads to limited...”
Testo
Preprint -
8162
Active Attacks: Red-teaming LLMs via Adaptive Environments
Pubblicazione 2025Sommario: “... exploration, we introduce \textit{Active Attacks}, a novel RL-based red-teaming algorithm that adapts its...”
Testo
Preprint -
8163
From Static to Dynamic: Adaptive Monte Carlo Search for Mathematical Process Supervision
Pubblicazione 2025Sommario: “..., it enhances the path expansion through a Monte Carlo algorithm with a temporally adaptive policy that begins...”
Testo
Preprint -
8164
Vision-Zero: Scalable VLM Self-Improvement via Strategic Gamified Self-Play
Pubblicazione 2025Sommario: “...-SPO), a novel training algorithm that alternates between Self-Play and reinforcement learning...”
Testo
Preprint -
8165
Who Should Have a Place on the Ark? Parameterized Algorithms for the Maximization of Phylogenetic Diversity
Pubblicazione 2025Sommario: “... after which they can no longer be saved.For Time-PD,we present color-coding algorithms that prove...”
Testo
Preprint -
8166
FOLK: Fast Open-Vocabulary 3D Instance Segmentation via Label-guided Knowledge Distillation
Pubblicazione 2025Sommario: “... propose a label-guided distillation algorithm to distill open-vocabulary knowledge from label-consistent 2...”
Testo
Preprint -
8167
Probabilistic Hyper-Graphs using Multiple Randomly Masked Autoencoders for Semi-supervised Multi-modal Multi-task Learning
Pubblicazione 2025Sommario: “... the distribution of hyper-edges on each forward pass. Additionally, the model adapts the standard MAE algorithm...”
Testo
Preprint -
8168
FedAdamW: A Communication-Efficient Optimizer with Convergence and Generalization Guarantees for Federated Large Models
Pubblicazione 2025Sommario: “... \underline{AdamW} algorithm, called \texttt{FedAdamW}, for training and fine-tuning various large models...”
Testo
Preprint -
8169
CORGI: Efficient Pattern Matching With Quadratic Guarantees
Pubblicazione 2025Sommario: “... automatically generate rules from example-driven induction or code synthesis, they can easily produce worst-case...”
Testo
Preprint -
8170
MoDES: Accelerating Mixture-of-Experts Multimodal Large Language Models via Dynamic Expert Skipping
Pubblicazione 2025Sommario: “... the optimal thresholds, we introduce a frontier search algorithm that exploits monotonicity properties...”
Testo
Preprint -
8171
Solving Spatial Supersensing Without Spatial Supersensing
Pubblicazione 2025Sommario: “... object-count predictions unchanged; instead, Cambrian-S inference algorithm relies largely on a shortcut...”
Testo
Preprint -
8172
Spira: Exploiting Voxel Data Structural Properties for Efficient Sparse Convolution in Point Cloud Networks
Pubblicazione 2025Sommario: “... for GPUs. Spira proposes (i) a high-performance one-shot search algorithm that builds the kernel map...”
Testo
Preprint -
8173
MICCAI STS 2024 Challenge: Semi-Supervised Instance-Level Tooth Segmentation in Panoramic X-ray and CBCT Images
Pubblicazione 2025Sommario: “...) registered teams. To ensure algorithmic excellence and full transparency, we rigorously evaluated the valid...”
Testo
Preprint -
8174
Semantic Soft Bootstrapping: Long Context Reasoning in LLMs without Reinforcement Learning
Pubblicazione 2025Sommario: “... optimization (GRPO), which is a commonly used RLVR algorithm. Our code is available at https://github.com...”
Testo
Preprint -
8175
How Far are Modern Trackers from UAV-Anti-UAV? A Million-Scale Benchmark and New Baseline
Pubblicazione 2025Sommario: “... of 50 modern deep tracking algorithms demonstrates that there is still significant room for improvement...”
Testo
Preprint -
8176
TraPO: A Semi-Supervised Reinforcement Learning Framework for Boosting LLM Reasoning
Pubblicazione 2025Sommario: “... an effective policy optimization algorithm, TraPO, that identifies reliable unlabeled samples by matching...”
Testo
Preprint -
8177
HERO-Sign: Hierarchical Tuning and Efficient Compiler-Time GPU Optimizations for SPHINCS+ Signature Generation
Pubblicazione 2025Sommario: “... algorithm that adapts fusion schemes to different GPU architectures. To further improve performance, HERO...”
Testo
Preprint -
8178
SP-Rank: A Dataset for Ranked Preferences with Secondary Information
Pubblicazione 2025Sommario: “... algorithms that leverage both first-order preferences and second-order predictions in ranking tasks. Each...”
Testo
Preprint -
8179
TowerMind: A Tower Defence Game Learning Environment and Benchmark for LLM as Agents
Pubblicazione 2026Sommario: “... evaluate two classic reinforcement learning algorithms: Ape-X DQN and PPO. By offering a lightweight...”
Testo
Preprint -
8180
Streaming Operator Inference for Model Reduction of Large-Scale Dynamical Systems
Pubblicazione 2026Sommario: “... access to the underlying code or full model operators, i.e., non-intrusively. Traditional OpInf operates...”
Testo
Preprint