Résultats de la recherche - decoding (Algorithmssic OR algorithm)

  1. 1081
  2. 1082

    Leveraging Small LLMs for Argument Mining in Education: Argument Component Identification, Classification, and Assessment par Favero, Lucile, Pérez-Ortiz, Juan Antonio, Käser, Tanja, Oliver, Nuria

    Publié 2025
    Table des matières: “…Argument mining algorithms analyze the argumentative structure of essays, making them a valuable…”
    Accéder au texte intégral
    Preprint
  3. 1083

    STAR-RIS-aided NOMA for Secured xURLLC par Lulu, Song, Di, Zhang, Shaobo, Jia, Pengcheng, Zhu, Yonghui, Li

    Publié 2025
    Table des matières: “…), which provides additional degree of freedom to enable a secure and fair decoding order and achieve a…”
    Accéder au texte intégral
    Preprint
  4. 1084

    SigChord: Sniffing Wide Non-sparse Multiband Signals for Terrestrial and Non-terrestrial Wireless Networks par Peng, Jinbo, Duan, Junwen, Lin, Zheng, Yuan, Haoxuan, Gao, Yue, Chen, Zhe

    Publié 2025
    Table des matières: “… on low-cost and sub-Nyquist sampling radios. We first design a deep and rule-based interleaving algorithm…”
    Accéder au texte intégral
    Preprint
  5. 1085

    SpecEE: Accelerating Large Language Model Inference with Speculative Early Exiting par Xu, Jiaming, Pan, Jiayi, Zhou, Yongkang, Chen, Siming, Li, Jinhao, Lian, Yaoxiu, Wu, Junyi, Dai, Guohao

    Publié 2025
    Table des matières: “… SpecEE, a fast LLM inference engine with speculative early exiting. (1) At the algorithm level, we…”
    Accéder au texte intégral
    Preprint
  6. 1086

    LMDepth: Lightweight Mamba-based Monocular Depth Estimation for Real-World Deployment par Long, Jiahuan, Zhou, Xin

    Publié 2025
    Table des matières: “…, existing depth estimation algorithms often struggle to effectively balance performance and computational…”
    Accéder au texte intégral
    Preprint
  7. 1087

    Antenna Activation and Resource Allocation in Multi-Waveguide Pinching-Antenna Systems par Wang, Kaidi, Ding, Zhiguo, Karagiannidis, George K.

    Publié 2025
    Table des matières: “…-theoretic algorithm is developed, in which the optimal decoding order is derived and incorporated. For power…”
    Accéder au texte intégral
    Preprint
  8. 1088

    One Trigger Token Is Enough: A Defense Strategy for Balancing Safety and Usability in Large Language Models par Gu, Haoran, Wang, Handing, Mei, Yi, Zhang, Mengjie, Jin, Yaochu

    Publié 2025
    Table des matières: “… simple yet effective defense algorithm that identifies and explicitly decodes safety trigger tokens…”
    Accéder au texte intégral
    Preprint
  9. 1089

    Semi-Clairvoyant Scheduling of Speculative Decoding Requests to Minimize LLM Inference Latency par Li, Ruixiao, Chen, Fahao, Li, Peng

    Publié 2025
    Table des matières: “…Speculative decoding accelerates Large Language Model (LLM) inference by employing a small…”
    Accéder au texte intégral
    Preprint
  10. 1090

    An Exploratory Framework for Future SETI Applications: Detecting Generative Reactivity via Language Models par Yu, Po-Chieh

    Publié 2025
    Table des matières: “… responses in language models. Instead of assuming that extraterrestrial signals must be decoded, we evaluate…”
    Accéder au texte intégral
    Preprint
  11. 1091

    Towards Bridging the Reward-Generation Gap in Direct Alignment Algorithms par Xiao, Zeguan, Chen, Yun, Chen, Guanhua, Tang, Ke

    Publié 2025
    Table des matières: “… Feedback (RLHF) algorithms for aligning large language models (LLMs) with human preferences. However, DAAs…”
    Accéder au texte intégral
    Preprint
  12. 1092

    ABS: Enforcing Constraint Satisfaction On Generated Sequences Via Automata-Guided Beam Search par Collura, Vincenzo, Tit, Karim, Bussi, Laura, Giunchiglia, Eleonora, Cordy, Maxime

    Publié 2025
    Table des matières: “… ones, and beam search is commonly used to balance exploration and fluency during decoding. While deep…”
    Accéder au texte intégral
    Preprint
  13. 1093
  14. 1094

    Reading Images Like Texts: Sequential Image Understanding in Vision-Language Models par Li, Yueyan, Zhao, Chenggong, Zang, Zeyuan, Yuan, Caixia, Wang, Xiaojie

    Publié 2025
    Table des matières: “… an instruction-agnostic token compression algorithm based on a plug-and-play visual decoder to improve decoding…”
    Accéder au texte intégral
    Preprint
  15. 1095

    STARS: Synchronous Token Alignment for Robust Supervision in Large Language Models par Quamar, Mohammad Atif, Areeb, Mohammad, Kuznetsov, Mikhail, Ozmen, Muslum Ozgur, Celik, Z. Berkay

    Publié 2025
    Table des matières: “…, we introduce STARS, Synchronous Token Alignment for Robust Supervision, a decoding-time algorithm…”
    Accéder au texte intégral
    Preprint
  16. 1096

    An Information-Minimal Geometry for Qubit-Efficient Optimization par Ma, Gordon, Angelakis, Dimitris G.

    Publié 2025
    Table des matières: “… algorithms explore exponentially large Hilbert spaces. We recast qubit-efficient optimization as a geometric…”
    Accéder au texte intégral
    Preprint
  17. 1097

    Evolving Diffusion and Flow Matching Policies for Online Reinforcement Learning par Zhang, Chubin, Wan, Zhenglin, Chen, Feng, Yang, Fuchao, Feng, Lang, Zhou, Yaxin, Yu, Xingrui, You, Yang, Tsang, Ivor, An, Bo

    Publié 2025
    Table des matières: “… generation. Building on this, we introduce GoRL (Generative Online Reinforcement Learning), an algorithm…”
    Accéder au texte intégral
    Preprint
  18. 1098

    Short-Context Dominance: How Much Local Context Natural Language Actually Needs? par Vakilian, Vala, Wang, Zimeng, Rawat, Ankit Singh, Thrampoulidis, Christos

    Publié 2025
    Table des matières: “… with sampling strategies beyond greedy decoding. Our experiments validate that simple thresholding of the metric…”
    Accéder au texte intégral
    Preprint
  19. 1099

    Hold Onto That Thought: Assessing KV Cache Compression On Reasoning par Liu, Minghui, Palnitkar, Aadi, Rabbani, Tahseen, Jae, Hyunwoo, Sang, Kyle Rui, Yao, Dixi, Shabihi, Shayan, Zhao, Fuheng, Li, Tian, Zhang, Ce, Huang, Furong, Zhang, Kunpeng

    Publié 2025
    Table des matières: “… attention computations, grows linearly with context length. A suite of compression algorithms has been…”
    Accéder au texte intégral
    Preprint
  20. 1100

    LoPA: Scaling dLLM Inference via Lookahead Parallel Decoding par Xu, Chenkai, Jin, Yijie, Li, Jiajun, Tu, Yi, Long, Guoping, Tu, Dandan, Song, Mingcong, Si, Hongjie, Hou, Tianqi, Yan, Junchi, Deng, Zhijie

    Publié 2025
    Table des matières: “… inference. However, current confidence-driven decoding strategies are constrained by limited parallelism…”
    Accéder au texte intégral
    Preprint