Risultati della ricerca - decoding (algorithm OR (Algorithmsssic OR algorithmsssic))

  1. 1081
  2. 1082

    Leveraging Small LLMs for Argument Mining in Education: Argument Component Identification, Classification, and Assessment di Favero, Lucile, Pérez-Ortiz, Juan Antonio, Käser, Tanja, Oliver, Nuria

    Pubblicazione 2025
    Sommario: “...Argument mining algorithms analyze the argumentative structure of essays, making them a valuable...”
    Testo
    Preprint
  3. 1083

    STAR-RIS-aided NOMA for Secured xURLLC di Lulu, Song, Di, Zhang, Shaobo, Jia, Pengcheng, Zhu, Yonghui, Li

    Pubblicazione 2025
    Sommario: “...), which provides additional degree of freedom to enable a secure and fair decoding order and achieve a...”
    Testo
    Preprint
  4. 1084

    SigChord: Sniffing Wide Non-sparse Multiband Signals for Terrestrial and Non-terrestrial Wireless Networks di Peng, Jinbo, Duan, Junwen, Lin, Zheng, Yuan, Haoxuan, Gao, Yue, Chen, Zhe

    Pubblicazione 2025
    Sommario: “... on low-cost and sub-Nyquist sampling radios. We first design a deep and rule-based interleaving algorithm...”
    Testo
    Preprint
  5. 1085

    SpecEE: Accelerating Large Language Model Inference with Speculative Early Exiting di Xu, Jiaming, Pan, Jiayi, Zhou, Yongkang, Chen, Siming, Li, Jinhao, Lian, Yaoxiu, Wu, Junyi, Dai, Guohao

    Pubblicazione 2025
    Sommario: “... SpecEE, a fast LLM inference engine with speculative early exiting. (1) At the algorithm level, we...”
    Testo
    Preprint
  6. 1086

    LMDepth: Lightweight Mamba-based Monocular Depth Estimation for Real-World Deployment di Long, Jiahuan, Zhou, Xin

    Pubblicazione 2025
    Sommario: “..., existing depth estimation algorithms often struggle to effectively balance performance and computational...”
    Testo
    Preprint
  7. 1087

    Antenna Activation and Resource Allocation in Multi-Waveguide Pinching-Antenna Systems di Wang, Kaidi, Ding, Zhiguo, Karagiannidis, George K.

    Pubblicazione 2025
    Sommario: “...-theoretic algorithm is developed, in which the optimal decoding order is derived and incorporated. For power...”
    Testo
    Preprint
  8. 1088

    One Trigger Token Is Enough: A Defense Strategy for Balancing Safety and Usability in Large Language Models di Gu, Haoran, Wang, Handing, Mei, Yi, Zhang, Mengjie, Jin, Yaochu

    Pubblicazione 2025
    Sommario: “... simple yet effective defense algorithm that identifies and explicitly decodes safety trigger tokens...”
    Testo
    Preprint
  9. 1089

    Semi-Clairvoyant Scheduling of Speculative Decoding Requests to Minimize LLM Inference Latency di Li, Ruixiao, Chen, Fahao, Li, Peng

    Pubblicazione 2025
    Sommario: “...Speculative decoding accelerates Large Language Model (LLM) inference by employing a small...”
    Testo
    Preprint
  10. 1090

    An Exploratory Framework for Future SETI Applications: Detecting Generative Reactivity via Language Models di Yu, Po-Chieh

    Pubblicazione 2025
    Sommario: “... responses in language models. Instead of assuming that extraterrestrial signals must be decoded, we evaluate...”
    Testo
    Preprint
  11. 1091

    Towards Bridging the Reward-Generation Gap in Direct Alignment Algorithms di Xiao, Zeguan, Chen, Yun, Chen, Guanhua, Tang, Ke

    Pubblicazione 2025
    Sommario: “... Feedback (RLHF) algorithms for aligning large language models (LLMs) with human preferences. However, DAAs...”
    Testo
    Preprint
  12. 1092

    ABS: Enforcing Constraint Satisfaction On Generated Sequences Via Automata-Guided Beam Search di Collura, Vincenzo, Tit, Karim, Bussi, Laura, Giunchiglia, Eleonora, Cordy, Maxime

    Pubblicazione 2025
    Sommario: “... ones, and beam search is commonly used to balance exploration and fluency during decoding. While deep...”
    Testo
    Preprint
  13. 1093

    A Study of Large Language Models for Patient Information Extraction: Model Architecture, Fine-Tuning Strategy, and Multi-task Instruction Tuning di Peng, Cheng, Dong, Xinyu, Lyu, Mengxian, Paredes, Daniel, Zhang, Yaoyun, Wu, Yonghui

    Pubblicazione 2025
    Sommario: “..., including: (1) encoder-only or decoder-only LLMs, (2) prompt-based parameter-efficient fine-tuning (PEFT...”
    Testo
    Preprint
  14. 1094

    Reading Images Like Texts: Sequential Image Understanding in Vision-Language Models di Li, Yueyan, Zhao, Chenggong, Zang, Zeyuan, Yuan, Caixia, Wang, Xiaojie

    Pubblicazione 2025
    Sommario: “... an instruction-agnostic token compression algorithm based on a plug-and-play visual decoder to improve decoding...”
    Testo
    Preprint
  15. 1095

    STARS: Synchronous Token Alignment for Robust Supervision in Large Language Models di Quamar, Mohammad Atif, Areeb, Mohammad, Kuznetsov, Mikhail, Ozmen, Muslum Ozgur, Celik, Z. Berkay

    Pubblicazione 2025
    Sommario: “..., we introduce STARS, Synchronous Token Alignment for Robust Supervision, a decoding-time algorithm...”
    Testo
    Preprint
  16. 1096

    An Information-Minimal Geometry for Qubit-Efficient Optimization di Ma, Gordon, Angelakis, Dimitris G.

    Pubblicazione 2025
    Sommario: “... algorithms explore exponentially large Hilbert spaces. We recast qubit-efficient optimization as a geometric...”
    Testo
    Preprint
  17. 1097

    Evolving Diffusion and Flow Matching Policies for Online Reinforcement Learning di Zhang, Chubin, Wan, Zhenglin, Chen, Feng, Yang, Fuchao, Feng, Lang, Zhou, Yaxin, Yu, Xingrui, You, Yang, Tsang, Ivor, An, Bo

    Pubblicazione 2025
    Sommario: “... generation. Building on this, we introduce GoRL (Generative Online Reinforcement Learning), an algorithm...”
    Testo
    Preprint
  18. 1098

    Short-Context Dominance: How Much Local Context Natural Language Actually Needs? di Vakilian, Vala, Wang, Zimeng, Rawat, Ankit Singh, Thrampoulidis, Christos

    Pubblicazione 2025
    Sommario: “... with sampling strategies beyond greedy decoding. Our experiments validate that simple thresholding of the metric...”
    Testo
    Preprint
  19. 1099

    Hold Onto That Thought: Assessing KV Cache Compression On Reasoning di Liu, Minghui, Palnitkar, Aadi, Rabbani, Tahseen, Jae, Hyunwoo, Sang, Kyle Rui, Yao, Dixi, Shabihi, Shayan, Zhao, Fuheng, Li, Tian, Zhang, Ce, Huang, Furong, Zhang, Kunpeng

    Pubblicazione 2025
    Sommario: “... attention computations, grows linearly with context length. A suite of compression algorithms has been...”
    Testo
    Preprint
  20. 1100

    LoPA: Scaling dLLM Inference via Lookahead Parallel Decoding di Xu, Chenkai, Jin, Yijie, Li, Jiajun, Tu, Yi, Long, Guoping, Tu, Dandan, Song, Mingcong, Si, Hongjie, Hou, Tianqi, Yan, Junchi, Deng, Zhijie

    Pubblicazione 2025
    Sommario: “... inference. However, current confidence-driven decoding strategies are constrained by limited parallelism...”
    Testo
    Preprint