Risultati della ricerca - decoding (algorithm OR (algorithmics OR Algorithmssics))

  1. 1261

    An Exploratory Framework for Future SETI Applications: Detecting Generative Reactivity via Language Models di Yu, Po-Chieh

    Pubblicazione 2025
    Sommario: “... responses in language models. Instead of assuming that extraterrestrial signals must be decoded, we evaluate...”
    Testo
    Preprint
  2. 1262

    Towards Bridging the Reward-Generation Gap in Direct Alignment Algorithms di Xiao, Zeguan, Chen, Yun, Chen, Guanhua, Tang, Ke

    Pubblicazione 2025
    Sommario: “... Feedback (RLHF) algorithms for aligning large language models (LLMs) with human preferences. However, DAAs...”
    Testo
    Preprint
  3. 1263

    ABS: Enforcing Constraint Satisfaction On Generated Sequences Via Automata-Guided Beam Search di Collura, Vincenzo, Tit, Karim, Bussi, Laura, Giunchiglia, Eleonora, Cordy, Maxime

    Pubblicazione 2025
    Sommario: “... ones, and beam search is commonly used to balance exploration and fluency during decoding. While deep...”
    Testo
    Preprint
  4. 1264

    A Study of Large Language Models for Patient Information Extraction: Model Architecture, Fine-Tuning Strategy, and Multi-task Instruction Tuning di Peng, Cheng, Dong, Xinyu, Lyu, Mengxian, Paredes, Daniel, Zhang, Yaoyun, Wu, Yonghui

    Pubblicazione 2025
    Sommario: “..., including: (1) encoder-only or decoder-only LLMs, (2) prompt-based parameter-efficient fine-tuning (PEFT...”
    Testo
    Preprint
  5. 1265

    Reading Images Like Texts: Sequential Image Understanding in Vision-Language Models di Li, Yueyan, Zhao, Chenggong, Zang, Zeyuan, Yuan, Caixia, Wang, Xiaojie

    Pubblicazione 2025
    Sommario: “... an instruction-agnostic token compression algorithm based on a plug-and-play visual decoder to improve decoding...”
    Testo
    Preprint
  6. 1266

    STARS: Synchronous Token Alignment for Robust Supervision in Large Language Models di Quamar, Mohammad Atif, Areeb, Mohammad, Kuznetsov, Mikhail, Ozmen, Muslum Ozgur, Celik, Z. Berkay

    Pubblicazione 2025
    Sommario: “..., we introduce STARS, Synchronous Token Alignment for Robust Supervision, a decoding-time algorithm...”
    Testo
    Preprint
  7. 1267

    An Information-Minimal Geometry for Qubit-Efficient Optimization di Ma, Gordon, Angelakis, Dimitris G.

    Pubblicazione 2025
    Sommario: “... algorithms explore exponentially large Hilbert spaces. We recast qubit-efficient optimization as a geometric...”
    Testo
    Preprint
  8. 1268

    Evolving Diffusion and Flow Matching Policies for Online Reinforcement Learning di Zhang, Chubin, Wan, Zhenglin, Chen, Feng, Yang, Fuchao, Feng, Lang, Zhou, Yaxin, Yu, Xingrui, You, Yang, Tsang, Ivor, An, Bo

    Pubblicazione 2025
    Sommario: “... generation. Building on this, we introduce GoRL (Generative Online Reinforcement Learning), an algorithm...”
    Testo
    Preprint
  9. 1269

    Short-Context Dominance: How Much Local Context Natural Language Actually Needs? di Vakilian, Vala, Wang, Zimeng, Rawat, Ankit Singh, Thrampoulidis, Christos

    Pubblicazione 2025
    Sommario: “... with sampling strategies beyond greedy decoding. Our experiments validate that simple thresholding of the metric...”
    Testo
    Preprint
  10. 1270

    Hold Onto That Thought: Assessing KV Cache Compression On Reasoning di Liu, Minghui, Palnitkar, Aadi, Rabbani, Tahseen, Jae, Hyunwoo, Sang, Kyle Rui, Yao, Dixi, Shabihi, Shayan, Zhao, Fuheng, Li, Tian, Zhang, Ce, Huang, Furong, Zhang, Kunpeng

    Pubblicazione 2025
    Sommario: “... attention computations, grows linearly with context length. A suite of compression algorithms has been...”
    Testo
    Preprint
  11. 1271

    LoPA: Scaling dLLM Inference via Lookahead Parallel Decoding di Xu, Chenkai, Jin, Yijie, Li, Jiajun, Tu, Yi, Long, Guoping, Tu, Dandan, Song, Mingcong, Si, Hongjie, Hou, Tianqi, Yan, Junchi, Deng, Zhijie

    Pubblicazione 2025
    Sommario: “... inference. However, current confidence-driven decoding strategies are constrained by limited parallelism...”
    Testo
    Preprint
  12. 1272

    Martingale Foresight Sampling: A Principled Approach to Inference-Time LLM Decoding di Li, Huayu, He, ZhengXiao, Tian, Siyuan, Wen, Jinghao, Li, Ao

    Pubblicazione 2026
    Sommario: “...Standard autoregressive decoding in large language models (LLMs) is inherently short-sighted, often...”
    Testo
    Preprint
  13. 1273
  14. 1274

    Reasoning Cache: Continual Improvement Over Long Horizons via Short-Horizon RL di Wu, Ian, Qu, Yuxiao, Setlur, Amrith, Kumar, Aviral

    Pubblicazione 2026
    Sommario: “... RC, an iterative decoding algorithm that replaces standard autoregressive decoding during both...”
    Testo
    Preprint
  15. 1275

    Listen to the Layers: Mitigating Hallucinations with Inter-Layer Disagreement di Subbalakshmi, Koduvayur, Ujjal, Sabbir Hossain, Mangichetty, Venkata Krishna Teja, Soofi, Nastaran Jamalipour

    Pubblicazione 2026
    Sommario: “... Aware) decoder, a novel, training-free decoding algorithm that mitigates hallucinations at inference...”
    Testo
    Preprint
  16. 1276

    ViterbiPlanNet: Injecting Procedural Knowledge via Differentiable Viterbi for Planning in Instructional Videos di Seminara, Luigi, Moltisanti, Davide, Furnari, Antonino

    Pubblicazione 2026
    Sommario: “... with the Viterbi decoding algorithm, replacing non-differentiable operations with smooth relaxations that enable...”
    Testo
    Preprint
  17. 1277

    $\nabla$-Reasoner: LLM Reasoning via Test-Time Gradient Descent in Latent Space di Wang, Peihao, Cai, Ruisi, Wang, Zhen, Mei, Hongyuan, Liu, Qiang, Li, Pan, Wang, Zhangyang

    Pubblicazione 2026
    Sommario: “... and suboptimal discrete search algorithms or trial-and-error prompting to improve the online policy...”
    Testo
    Preprint
  18. 1278

    The Model Knows, the Decoder Finds: Future Value Guided Particle Power Sampling di Nguyen, Tu, Zimmer, Matthieu, Tutunov, Rasul, Ji, Xiaotong, Ammar, Haitham Bou

    Pubblicazione 2026
    Sommario: “... at inference time. Power sampling provides a principled way to bias decoding toward such modes by targeting...”
    Testo
    Preprint
  19. 1279

    Mix-Quant: Quantized Prefilling, Precise Decoding for Agentic LLMs di Lu, Haiquan, Chen, Zigeng, Fang, Gongfan, Ma, Xinyin, Wang, Xinchao

    Pubblicazione 2026
    Sommario: “... preserving BF16 precision for decoding. By decoupling prefilling acceleration from decoding quality, Mix...”
    Testo
    Preprint
  20. 1280

    Best-First Ordered Statistics Decoding of Quantum LDPC Codes di Banfi, Michele, Ferrari, Marco, Favano, Antonino, Tarable, Alberto, Barletta, Luca

    Pubblicazione 2026
    Sommario: “... standard for decoding quantum low-density parity-check (QLDPC) codes. Recent advancements in this field...”
    Testo
    Preprint