Resultados de búsqueda - decoding (Algorithmsssic OR (algorithm OR Algorithmsssic))

  1. 1081
  2. 1082

    Leveraging Small LLMs for Argument Mining in Education: Argument Component Identification, Classification, and Assessment por Favero, Lucile, Pérez-Ortiz, Juan Antonio, Käser, Tanja, Oliver, Nuria

    Publicado 2025
    Tabla de Contenidos: “…Argument mining algorithms analyze the argumentative structure of essays, making them a valuable…”
    Enlace del recurso
    Preprint
  3. 1083

    STAR-RIS-aided NOMA for Secured xURLLC por Lulu, Song, Di, Zhang, Shaobo, Jia, Pengcheng, Zhu, Yonghui, Li

    Publicado 2025
    Tabla de Contenidos: “…), which provides additional degree of freedom to enable a secure and fair decoding order and achieve a…”
    Enlace del recurso
    Preprint
  4. 1084

    SigChord: Sniffing Wide Non-sparse Multiband Signals for Terrestrial and Non-terrestrial Wireless Networks por Peng, Jinbo, Duan, Junwen, Lin, Zheng, Yuan, Haoxuan, Gao, Yue, Chen, Zhe

    Publicado 2025
    Tabla de Contenidos: “… on low-cost and sub-Nyquist sampling radios. We first design a deep and rule-based interleaving algorithm…”
    Enlace del recurso
    Preprint
  5. 1085

    SpecEE: Accelerating Large Language Model Inference with Speculative Early Exiting por Xu, Jiaming, Pan, Jiayi, Zhou, Yongkang, Chen, Siming, Li, Jinhao, Lian, Yaoxiu, Wu, Junyi, Dai, Guohao

    Publicado 2025
    Tabla de Contenidos: “… SpecEE, a fast LLM inference engine with speculative early exiting. (1) At the algorithm level, we…”
    Enlace del recurso
    Preprint
  6. 1086

    LMDepth: Lightweight Mamba-based Monocular Depth Estimation for Real-World Deployment por Long, Jiahuan, Zhou, Xin

    Publicado 2025
    Tabla de Contenidos: “…, existing depth estimation algorithms often struggle to effectively balance performance and computational…”
    Enlace del recurso
    Preprint
  7. 1087

    Antenna Activation and Resource Allocation in Multi-Waveguide Pinching-Antenna Systems por Wang, Kaidi, Ding, Zhiguo, Karagiannidis, George K.

    Publicado 2025
    Tabla de Contenidos: “…-theoretic algorithm is developed, in which the optimal decoding order is derived and incorporated. For power…”
    Enlace del recurso
    Preprint
  8. 1088

    One Trigger Token Is Enough: A Defense Strategy for Balancing Safety and Usability in Large Language Models por Gu, Haoran, Wang, Handing, Mei, Yi, Zhang, Mengjie, Jin, Yaochu

    Publicado 2025
    Tabla de Contenidos: “… simple yet effective defense algorithm that identifies and explicitly decodes safety trigger tokens…”
    Enlace del recurso
    Preprint
  9. 1089

    Semi-Clairvoyant Scheduling of Speculative Decoding Requests to Minimize LLM Inference Latency por Li, Ruixiao, Chen, Fahao, Li, Peng

    Publicado 2025
    Tabla de Contenidos: “…Speculative decoding accelerates Large Language Model (LLM) inference by employing a small…”
    Enlace del recurso
    Preprint
  10. 1090

    An Exploratory Framework for Future SETI Applications: Detecting Generative Reactivity via Language Models por Yu, Po-Chieh

    Publicado 2025
    Tabla de Contenidos: “… responses in language models. Instead of assuming that extraterrestrial signals must be decoded, we evaluate…”
    Enlace del recurso
    Preprint
  11. 1091

    Towards Bridging the Reward-Generation Gap in Direct Alignment Algorithms por Xiao, Zeguan, Chen, Yun, Chen, Guanhua, Tang, Ke

    Publicado 2025
    Tabla de Contenidos: “… Feedback (RLHF) algorithms for aligning large language models (LLMs) with human preferences. However, DAAs…”
    Enlace del recurso
    Preprint
  12. 1092

    ABS: Enforcing Constraint Satisfaction On Generated Sequences Via Automata-Guided Beam Search por Collura, Vincenzo, Tit, Karim, Bussi, Laura, Giunchiglia, Eleonora, Cordy, Maxime

    Publicado 2025
    Tabla de Contenidos: “… ones, and beam search is commonly used to balance exploration and fluency during decoding. While deep…”
    Enlace del recurso
    Preprint
  13. 1093
  14. 1094

    Reading Images Like Texts: Sequential Image Understanding in Vision-Language Models por Li, Yueyan, Zhao, Chenggong, Zang, Zeyuan, Yuan, Caixia, Wang, Xiaojie

    Publicado 2025
    Tabla de Contenidos: “… an instruction-agnostic token compression algorithm based on a plug-and-play visual decoder to improve decoding…”
    Enlace del recurso
    Preprint
  15. 1095

    STARS: Synchronous Token Alignment for Robust Supervision in Large Language Models por Quamar, Mohammad Atif, Areeb, Mohammad, Kuznetsov, Mikhail, Ozmen, Muslum Ozgur, Celik, Z. Berkay

    Publicado 2025
    Tabla de Contenidos: “…, we introduce STARS, Synchronous Token Alignment for Robust Supervision, a decoding-time algorithm…”
    Enlace del recurso
    Preprint
  16. 1096

    An Information-Minimal Geometry for Qubit-Efficient Optimization por Ma, Gordon, Angelakis, Dimitris G.

    Publicado 2025
    Tabla de Contenidos: “… algorithms explore exponentially large Hilbert spaces. We recast qubit-efficient optimization as a geometric…”
    Enlace del recurso
    Preprint
  17. 1097

    Evolving Diffusion and Flow Matching Policies for Online Reinforcement Learning por Zhang, Chubin, Wan, Zhenglin, Chen, Feng, Yang, Fuchao, Feng, Lang, Zhou, Yaxin, Yu, Xingrui, You, Yang, Tsang, Ivor, An, Bo

    Publicado 2025
    Tabla de Contenidos: “… generation. Building on this, we introduce GoRL (Generative Online Reinforcement Learning), an algorithm…”
    Enlace del recurso
    Preprint
  18. 1098

    Short-Context Dominance: How Much Local Context Natural Language Actually Needs? por Vakilian, Vala, Wang, Zimeng, Rawat, Ankit Singh, Thrampoulidis, Christos

    Publicado 2025
    Tabla de Contenidos: “… with sampling strategies beyond greedy decoding. Our experiments validate that simple thresholding of the metric…”
    Enlace del recurso
    Preprint
  19. 1099

    Hold Onto That Thought: Assessing KV Cache Compression On Reasoning por Liu, Minghui, Palnitkar, Aadi, Rabbani, Tahseen, Jae, Hyunwoo, Sang, Kyle Rui, Yao, Dixi, Shabihi, Shayan, Zhao, Fuheng, Li, Tian, Zhang, Ce, Huang, Furong, Zhang, Kunpeng

    Publicado 2025
    Tabla de Contenidos: “… attention computations, grows linearly with context length. A suite of compression algorithms has been…”
    Enlace del recurso
    Preprint
  20. 1100

    LoPA: Scaling dLLM Inference via Lookahead Parallel Decoding por Xu, Chenkai, Jin, Yijie, Li, Jiajun, Tu, Yi, Long, Guoping, Tu, Dandan, Song, Mingcong, Si, Hongjie, Hou, Tianqi, Yan, Junchi, Deng, Zhijie

    Publicado 2025
    Tabla de Contenidos: “… inference. However, current confidence-driven decoding strategies are constrained by limited parallelism…”
    Enlace del recurso
    Preprint