Risultati della ricerca - decoding (algorithm OR Algorithms)

  1. 861

    DOPRA: Decoding Over-accumulation Penalization and Re-allocation in Specific Weighting Layer di Wei, Jinfeng, Zhang, Xiaofeng

    Pubblicazione 2024
    Sommario: “... hallucinations by decoding specific weighted layer penalties and redistribution, offering an economical...”
    Testo
    Preprint
  2. 862

    Continuous Speculative Decoding for Autoregressive Image Generation di Wang, Zili, Zhang, Robert, Ding, Kun, Yang, Qi, Li, Fei, Xiang, Shiming

    Pubblicazione 2024
    Sommario: “... Language Models (LLMs), speculative decoding has effectively accelerated discrete autoregressive inference...”
    Testo
    Preprint
  3. 863

    PLD+: Accelerating LLM inference by leveraging Language Model Artifacts di Somasundaram, Shwetha, Phukan, Anirudh, Saxena, Apoorv

    Pubblicazione 2024
    Sommario: “...To reduce the latency associated with autoretrogressive LLM inference, speculative decoding has...”
    Testo
    Preprint
  4. 864

    SPIN: Accelerating Large Language Model Inference with Heterogeneous Speculative Models di Chen, Fahao, Li, Peng, Luan, Tom H., Su, Zhou, Deng, Jing

    Pubblicazione 2025
    Sommario: “...Speculative decoding has been shown as an effective way to accelerate Large Language Model (LLM...”
    Testo
    Preprint
  5. 865

    CLaSp: In-Context Layer Skip for Self-Speculative Decoding di Chen, Longze, Shan, Renke, Wang, Huiming, Wang, Lu, Liu, Ziqiang, Luo, Run, Wang, Jiawei, Alinejad-Rokny, Hamid, Yang, Min

    Pubblicazione 2025
    Sommario: “...Speculative decoding (SD) is a promising method for accelerating the decoding process of Large...”
    Testo
    Preprint
  6. 866

    Radiation-Induced Fault Detection in Superconducting Quantum Devices di Vallero, Marzio, Casagranda, Gioele, Vella, Flavio, Rech, Paolo

    Pubblicazione 2025
    Sommario: “... propose the first algorithm to effectively exploit syndrome information for the efficient detection...”
    Testo
    Preprint
  7. 867

    Federated Learning in Offline and Online EMG Decoding: A Privacy and Performance Perspective di Malcolm, Kai, Uribe, César, Yamagami, Momona

    Pubblicazione 2025
    Sommario: “... neural decoding using high-dimensional electromyography (EMG) across both offline simulations and a real...”
    Testo
    Preprint
  8. 868

    Efficient Decoding Methods for Language Models on Encrypted Data di Avitan, Matan, Baruch, Moran, Drucker, Nir, Zimerman, Itamar, Goldberg, Yoav

    Pubblicazione 2025
    Sommario: “... for secure inference. However, neural text generation requires decoding methods like argmax and sampling...”
    Testo
    Preprint
  9. 869

    SpecAttn: Speculating Sparse Attention di Shah, Harsh

    Pubblicazione 2025
    Sommario: “... SpecAttn, a novel training-free approach that seamlessly integrates with existing speculative decoding...”
    Testo
    Preprint
  10. 870

    Conserved Kinematic Representations enable Zero-Shot Decoding in Handwriting BCIs di Ravishankar, Srinivas, de Sa, Virginia

    Pubblicazione 2026
    Sommario: “... machine learning algorithm for decoding unseen characters. Our model achieves 64% hits@3 retrieval...”
    Testo
    Preprint
  11. 871

    Joint Coordinate Regression and Association For Multi-Person Pose Estimation, A Pure Neural Network Approach di Yu, Dongyang, Xie, Yunshi, An, Wangpeng, Zhang, Li, Yao, Yufeng

    Pubblicazione 2023
    Sommario: “...We introduce a novel one-stage end-to-end multi-person 2D pose estimation algorithm, known as Joint...”
    Testo
    Preprint
  12. 872

    InversionView: A General-Purpose Method for Reading Information from Neural Activations di Huang, Xinting, Panwar, Madhur, Goyal, Navin, Hahn, Michael

    Pubblicazione 2024
    Sommario: “... to practically inspect this subset by sampling from a trained decoder model conditioned on activations...”
    Testo
    Preprint
  13. 873

    Soft Partitioning of Latent Space for Semantic Channel Equalization di Hüttebräucker, Tomás, Sana, Mohamed, Strinati, Emilio Calvanese

    Pubblicazione 2024
    Sommario: “... semantic communications. This approach aims to align the latent spaces of an encoder and a decoder which...”
    Testo
    Preprint
  14. 874

    Breaking the Attention Bottleneck di Hilsenbek, Kalle

    Pubblicazione 2024
    Sommario: “... in the transformer architecture. This algorithm is only uni-directional in the decoder and converges to a static...”
    Testo
    Preprint
  15. 875

    Tree Attention: Topology-aware Decoding for Long-Context Attention on GPU clusters di Shyam, Vasudev, Pilault, Jonathan, Shepperd, Emily, Anthony, Quentin, Millidge, Beren

    Pubblicazione 2024
    Sommario: “... in parallel through a tree reduction. Our algorithm, called Tree Attention, for parallelizing exact attention...”
    Testo
    Preprint
  16. 876

    Lightweight Transducer Based on Frame-Level Criterion di Wan, Genshun, Wang, Mengzhi, Mao, Tingzhi, Chen, Hang, Ye, Zhongfu

    Pubblicazione 2024
    Sommario: “... criterion, which uses the results of the CTC forced alignment algorithm to determine the label for each...”
    Testo
    Preprint
  17. 877

    Accelerating Codec-based Speech Synthesis with Multi-Token Prediction and Speculative Decoding di Nguyen, Tan Dat, Kim, Ji-Hoon, Choi, Jeongsoo, Choi, Shukjae, Park, Jinseok, Lee, Younglo, Chung, Joon Son

    Pubblicazione 2024
    Sommario: “... decoding technique that utilises a Viterbi-based algorithm to select the optimal sequence of generated...”
    Testo
    Preprint
  18. 878

    Token-Driven GammaTune: Adaptive Calibration for Enhanced Speculative Decoding di Gautam, Aayush, Shrestha, Susav, Reddy, Narasimha

    Pubblicazione 2025
    Sommario: “...Speculative decoding accelerates large language model (LLM) inference by using a smaller draft...”
    Testo
    Preprint
  19. 879

    Efficient and Asymptotically Unbiased Constrained Decoding for Large Language Models di Ye, Haotian, Jain, Himanshu, You, Chong, Suresh, Ananda Theertha, Lin, Haowei, Zou, James, Yu, Felix

    Pubblicazione 2025
    Sommario: “... standards, or conforming to specialized formatting styles. To control the generation, constrained decoding...”
    Testo
    Preprint
  20. 880

    Diffusion Models are Secretly Exchangeable: Parallelizing DDPMs via Autospeculation di Hu, Hengyuan, Das, Aniket, Sadigh, Dorsa, Anari, Nima

    Pubblicazione 2025
    Sommario: “...} (ASD), an extension of the widely used speculative decoding algorithm to DDPMs that does not require...”
    Testo
    Preprint