Learning distributed representations with efficient SoftMax normalization
Fuente:
arXiv
Salvato in:
| Autori principali: | Dall'Amico, Lorenzo, Belliardo, Enrico Maria |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
An embedding-based distance for temporal graphs
di: Dall'Amico, Lorenzo, et al.
Pubblicazione: (2024)
di: Dall'Amico, Lorenzo, et al.
Pubblicazione: (2024)
MaxCode: A Max-Reward Reinforcement Learning Framework for Automated Code Optimization
di: Ou, Jiefu, et al.
Pubblicazione: (2026)
di: Ou, Jiefu, et al.
Pubblicazione: (2026)
Learning representations of learning representations
di: González-Márquez, Rita, et al.
Pubblicazione: (2024)
di: González-Márquez, Rita, et al.
Pubblicazione: (2024)
AmpliNetECG12: A lightweight SoftMax-based relativistic amplitude amplification architecture for 12 lead ECG classification
di: Srivastava, Shreya
Pubblicazione: (2024)
di: Srivastava, Shreya
Pubblicazione: (2024)
Learning to Translate from Soft to Hard LLM Prompts
di: Kongsomjit, Pitipat, et al.
Pubblicazione: (2026)
di: Kongsomjit, Pitipat, et al.
Pubblicazione: (2026)
Perturbation: A simple and efficient adversarial tracer for representation learning in language models
di: Rozner, Joshua, et al.
Pubblicazione: (2026)
di: Rozner, Joshua, et al.
Pubblicazione: (2026)
Learning How to Ask: Querying LMs with Mixtures of Soft Prompts
di: Qin, Guanghui, et al.
Pubblicazione: (2021)
di: Qin, Guanghui, et al.
Pubblicazione: (2021)
Can we Soft Prompt LLMs for Graph Learning Tasks?
di: Liu, Zheyuan, et al.
Pubblicazione: (2024)
di: Liu, Zheyuan, et al.
Pubblicazione: (2024)
Max-pooling Network Revisited: Analyzing the Role of Semantic Probability in Multiple Instance Learning for Hallucination Detection
di: Fujikawa, Shota, et al.
Pubblicazione: (2026)
di: Fujikawa, Shota, et al.
Pubblicazione: (2026)
Entropy Ratio Clipping as a Soft Global Constraint for Stable Reinforcement Learning
di: Su, Zhenpeng, et al.
Pubblicazione: (2025)
di: Su, Zhenpeng, et al.
Pubblicazione: (2025)
Learning New Tasks from a Few Examples with Soft-Label Prototypes
di: Singh, Avyav Kumar, et al.
Pubblicazione: (2022)
di: Singh, Avyav Kumar, et al.
Pubblicazione: (2022)
SoftMatcha 2: A Fast and Soft Pattern Matcher for Trillion-Scale Corpora
di: Yoneda, Masataka, et al.
Pubblicazione: (2026)
di: Yoneda, Masataka, et al.
Pubblicazione: (2026)
PromptAL: Sample-Aware Dynamic Soft Prompts for Few-Shot Active Learning
di: Xiang, Hui, et al.
Pubblicazione: (2025)
di: Xiang, Hui, et al.
Pubblicazione: (2025)
SLIM: Let LLM Learn More and Forget Less with Soft LoRA and Identity Mixture
di: Han, Jiayi, et al.
Pubblicazione: (2024)
di: Han, Jiayi, et al.
Pubblicazione: (2024)
SoftLMs: Efficient Adaptive Low-Rank Approximation of Language Models using Soft-Thresholding Mechanism
di: Bhatnagar, Priyansh, et al.
Pubblicazione: (2024)
di: Bhatnagar, Priyansh, et al.
Pubblicazione: (2024)
Intrinsic Dimension Correlation: uncovering nonlinear connections in multimodal representations
di: Basile, Lorenzo, et al.
Pubblicazione: (2024)
di: Basile, Lorenzo, et al.
Pubblicazione: (2024)
The MiniMax-M2 Series: Mini Activations Unleashing Max Real-World Intelligence
di: MiniMax, et al.
Pubblicazione: (2026)
di: MiniMax, et al.
Pubblicazione: (2026)
MaxPoolBERT: Enhancing BERT Classification via Layer- and Token-Wise Aggregation
di: Behrendt, Maike, et al.
Pubblicazione: (2025)
di: Behrendt, Maike, et al.
Pubblicazione: (2025)
MiniMax-M1: Scaling Test-Time Compute Efficiently with Lightning Attention
di: MiniMax, et al.
Pubblicazione: (2025)
di: MiniMax, et al.
Pubblicazione: (2025)
Min-p, Max Exaggeration: A Critical Analysis of Min-p Sampling in Language Models
di: Schaeffer, Rylan, et al.
Pubblicazione: (2025)
di: Schaeffer, Rylan, et al.
Pubblicazione: (2025)
SoftQE: Learned Representations of Queries Expanded by LLMs
di: Pimpalkhute, Varad, et al.
Pubblicazione: (2024)
di: Pimpalkhute, Varad, et al.
Pubblicazione: (2024)
Subspace Representations for Soft Set Operations and Sentence Similarities
di: Ishibashi, Yoichi, et al.
Pubblicazione: (2022)
di: Ishibashi, Yoichi, et al.
Pubblicazione: (2022)
A framework for analyzing concept representations in neural models
di: Naowarat, Burin, et al.
Pubblicazione: (2026)
di: Naowarat, Burin, et al.
Pubblicazione: (2026)
Indication Finding: a novel use case for representation learning
di: Eckhoff, Maren, et al.
Pubblicazione: (2024)
di: Eckhoff, Maren, et al.
Pubblicazione: (2024)
DRES: Fake news detection by dynamic representation and ensemble selection
di: Farhangian, Faramarz, et al.
Pubblicazione: (2025)
di: Farhangian, Faramarz, et al.
Pubblicazione: (2025)
FLARE: Task-agnostic embedding model evaluation through a normalization process
di: Jiang, Jingzhou, et al.
Pubblicazione: (2026)
di: Jiang, Jingzhou, et al.
Pubblicazione: (2026)
Max It or Miss It: Benchmarking LLM On Solving Extremal Problems
di: Gao, Binxin, et al.
Pubblicazione: (2025)
di: Gao, Binxin, et al.
Pubblicazione: (2025)
MarginSel : Max-Margin Demonstration Selection for LLMs
di: Ambati, Rajeev Bhatt, et al.
Pubblicazione: (2025)
di: Ambati, Rajeev Bhatt, et al.
Pubblicazione: (2025)
Effects of Soft-Domain Transfer and Named Entity Information on Deception Detection
di: Triplett, Steven, et al.
Pubblicazione: (2024)
di: Triplett, Steven, et al.
Pubblicazione: (2024)
Kanana: Compute-efficient Bilingual Language Models
di: Kanana LLM Team, et al.
Pubblicazione: (2025)
di: Kanana LLM Team, et al.
Pubblicazione: (2025)
On multi-token prediction for efficient LLM inference
di: Mehra, Somesh, et al.
Pubblicazione: (2025)
di: Mehra, Somesh, et al.
Pubblicazione: (2025)
Sample-efficient LLM Optimization with Reset Replay
di: Liu, Zichuan, et al.
Pubblicazione: (2025)
di: Liu, Zichuan, et al.
Pubblicazione: (2025)
An Assessment of Human vs. Model Uncertainty in Soft-Label Learning and Calibration
di: Pavlovic, Maja, et al.
Pubblicazione: (2026)
di: Pavlovic, Maja, et al.
Pubblicazione: (2026)
EEG-CLIP : Learning EEG representations from natural language descriptions
di: Ndir, Tidiane Camaret, et al.
Pubblicazione: (2025)
di: Ndir, Tidiane Camaret, et al.
Pubblicazione: (2025)
NorMuon: Making Muon more efficient and scalable
di: Li, Zichong, et al.
Pubblicazione: (2025)
di: Li, Zichong, et al.
Pubblicazione: (2025)
Unlearning as multi-task optimization: A normalized gradient difference approach with an adaptive learning rate
di: Bu, Zhiqi, et al.
Pubblicazione: (2024)
di: Bu, Zhiqi, et al.
Pubblicazione: (2024)
Linear representations in language models can change dramatically over a conversation
di: Lampinen, Andrew Kyle, et al.
Pubblicazione: (2026)
di: Lampinen, Andrew Kyle, et al.
Pubblicazione: (2026)
The representation landscape of few-shot learning and fine-tuning in large language models
di: Doimo, Diego, et al.
Pubblicazione: (2024)
di: Doimo, Diego, et al.
Pubblicazione: (2024)
Towards Lightweight Reliability: Using Soft Prompts for Hallucination Mitigation in Large Language Models
di: Siddiqui, S M Tahmid, et al.
Pubblicazione: (2026)
di: Siddiqui, S M Tahmid, et al.
Pubblicazione: (2026)
Beyond Hard and Soft: Hybrid Context Compression for Balancing Local and Global Information Retention
di: Liao, Huanxuan, et al.
Pubblicazione: (2025)
di: Liao, Huanxuan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
An embedding-based distance for temporal graphs
di: Dall'Amico, Lorenzo, et al.
Pubblicazione: (2024) -
MaxCode: A Max-Reward Reinforcement Learning Framework for Automated Code Optimization
di: Ou, Jiefu, et al.
Pubblicazione: (2026) -
Learning representations of learning representations
di: González-Márquez, Rita, et al.
Pubblicazione: (2024) -
AmpliNetECG12: A lightweight SoftMax-based relativistic amplitude amplification architecture for 12 lead ECG classification
di: Srivastava, Shreya
Pubblicazione: (2024) -
Learning to Translate from Soft to Hard LLM Prompts
di: Kongsomjit, Pitipat, et al.
Pubblicazione: (2026)