Risultati della ricerca - decoding (algorithm OR (Algorithmsssic OR algorithmsssic))
Cerca anche:
- decoding »
- algorithm »
- algorithmsssic »
-
1081
Self-Correcting Decoding with Generative Feedback for Mitigating Hallucinations in Large Vision-Language Models
Pubblicazione 2025Sommario: “... Decoding with Generative Feedback (DeGF), a novel training-free algorithm that incorporates feedback from...”
Testo
Preprint -
1082
Leveraging Small LLMs for Argument Mining in Education: Argument Component Identification, Classification, and Assessment
Pubblicazione 2025Sommario: “...Argument mining algorithms analyze the argumentative structure of essays, making them a valuable...”
Testo
Preprint -
1083
STAR-RIS-aided NOMA for Secured xURLLC
Pubblicazione 2025Sommario: “...), which provides additional degree of freedom to enable a secure and fair decoding order and achieve a...”
Testo
Preprint -
1084
SigChord: Sniffing Wide Non-sparse Multiband Signals for Terrestrial and Non-terrestrial Wireless Networks
Pubblicazione 2025Sommario: “... on low-cost and sub-Nyquist sampling radios. We first design a deep and rule-based interleaving algorithm...”
Testo
Preprint -
1085
SpecEE: Accelerating Large Language Model Inference with Speculative Early Exiting
Pubblicazione 2025Sommario: “... SpecEE, a fast LLM inference engine with speculative early exiting. (1) At the algorithm level, we...”
Testo
Preprint -
1086
LMDepth: Lightweight Mamba-based Monocular Depth Estimation for Real-World Deployment
Pubblicazione 2025Sommario: “..., existing depth estimation algorithms often struggle to effectively balance performance and computational...”
Testo
Preprint -
1087
Antenna Activation and Resource Allocation in Multi-Waveguide Pinching-Antenna Systems
Pubblicazione 2025Sommario: “...-theoretic algorithm is developed, in which the optimal decoding order is derived and incorporated. For power...”
Testo
Preprint -
1088
One Trigger Token Is Enough: A Defense Strategy for Balancing Safety and Usability in Large Language Models
Pubblicazione 2025Sommario: “... simple yet effective defense algorithm that identifies and explicitly decodes safety trigger tokens...”
Testo
Preprint -
1089
Semi-Clairvoyant Scheduling of Speculative Decoding Requests to Minimize LLM Inference Latency
Pubblicazione 2025Sommario: “...Speculative decoding accelerates Large Language Model (LLM) inference by employing a small...”
Testo
Preprint -
1090
An Exploratory Framework for Future SETI Applications: Detecting Generative Reactivity via Language Models
Pubblicazione 2025Sommario: “... responses in language models. Instead of assuming that extraterrestrial signals must be decoded, we evaluate...”
Testo
Preprint -
1091
Towards Bridging the Reward-Generation Gap in Direct Alignment Algorithms
Pubblicazione 2025Sommario: “... Feedback (RLHF) algorithms for aligning large language models (LLMs) with human preferences. However, DAAs...”
Testo
Preprint -
1092
ABS: Enforcing Constraint Satisfaction On Generated Sequences Via Automata-Guided Beam Search
Pubblicazione 2025Sommario: “... ones, and beam search is commonly used to balance exploration and fluency during decoding. While deep...”
Testo
Preprint -
1093
A Study of Large Language Models for Patient Information Extraction: Model Architecture, Fine-Tuning Strategy, and Multi-task Instruction Tuning
Pubblicazione 2025Sommario: “..., including: (1) encoder-only or decoder-only LLMs, (2) prompt-based parameter-efficient fine-tuning (PEFT...”
Testo
Preprint -
1094
Reading Images Like Texts: Sequential Image Understanding in Vision-Language Models
Pubblicazione 2025Sommario: “... an instruction-agnostic token compression algorithm based on a plug-and-play visual decoder to improve decoding...”
Testo
Preprint -
1095
STARS: Synchronous Token Alignment for Robust Supervision in Large Language Models
Pubblicazione 2025Sommario: “..., we introduce STARS, Synchronous Token Alignment for Robust Supervision, a decoding-time algorithm...”
Testo
Preprint -
1096
An Information-Minimal Geometry for Qubit-Efficient Optimization
Pubblicazione 2025Sommario: “... algorithms explore exponentially large Hilbert spaces. We recast qubit-efficient optimization as a geometric...”
Testo
Preprint -
1097
Evolving Diffusion and Flow Matching Policies for Online Reinforcement Learning
Pubblicazione 2025Sommario: “... generation. Building on this, we introduce GoRL (Generative Online Reinforcement Learning), an algorithm...”
Testo
Preprint -
1098
Short-Context Dominance: How Much Local Context Natural Language Actually Needs?
Pubblicazione 2025Sommario: “... with sampling strategies beyond greedy decoding. Our experiments validate that simple thresholding of the metric...”
Testo
Preprint -
1099
Hold Onto That Thought: Assessing KV Cache Compression On Reasoning
Pubblicazione 2025Sommario: “... attention computations, grows linearly with context length. A suite of compression algorithms has been...”
Testo
Preprint -
1100
LoPA: Scaling dLLM Inference via Lookahead Parallel Decoding
Pubblicazione 2025Sommario: “... inference. However, current confidence-driven decoding strategies are constrained by limited parallelism...”
Testo
Preprint