Gespeichert in:
| Hauptverfasser: | Opper, Mattia, Prokhorov, Victor, Siddharth, N. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2305.05588 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Self-StrAE at SemEval-2024 Task 1: Making Self-Structuring AutoEncoders Learn More With Less
von: Opper, Mattia, et al.
Veröffentlicht: (2024)
von: Opper, Mattia, et al.
Veröffentlicht: (2024)
Banyan: Improved Representation Learning with Explicit Structure
von: Opper, Mattia, et al.
Veröffentlicht: (2024)
von: Opper, Mattia, et al.
Veröffentlicht: (2024)
TRA: Better Length Generalisation with Threshold Relative Attention
von: Opper, Mattia, et al.
Veröffentlicht: (2025)
von: Opper, Mattia, et al.
Veröffentlicht: (2025)
Autoencoding Conditional Neural Processes for Representation Learning
von: Prokhorov, Victor, et al.
Veröffentlicht: (2023)
von: Prokhorov, Victor, et al.
Veröffentlicht: (2023)
Bootstrapping Embeddings for Low Resource Languages
von: Basoz, Merve, et al.
Veröffentlicht: (2026)
von: Basoz, Merve, et al.
Veröffentlicht: (2026)
Embedding Geometries of Contrastive Language-Image Pre-Training
von: Chou, Jason Chuan-Chih, et al.
Veröffentlicht: (2024)
von: Chou, Jason Chuan-Chih, et al.
Veröffentlicht: (2024)
AE-LLM: Adaptive Efficiency Optimization for Large Language Models
von: Tanaka, Kaito, et al.
Veröffentlicht: (2026)
von: Tanaka, Kaito, et al.
Veröffentlicht: (2026)
Training Superior Sparse Autoencoders for Instruct Models
von: Li, Jiaming, et al.
Veröffentlicht: (2025)
von: Li, Jiaming, et al.
Veröffentlicht: (2025)
Memory-efficient Energy-adaptive Inference of Pre-Trained Models on Batteryless Embedded Systems
von: Farina, Pietro, et al.
Veröffentlicht: (2024)
von: Farina, Pietro, et al.
Veröffentlicht: (2024)
DEPT: Decoupled Embeddings for Pre-training Language Models
von: Iacob, Alex, et al.
Veröffentlicht: (2024)
von: Iacob, Alex, et al.
Veröffentlicht: (2024)
CURE: Controlled Unlearning for Robust Embeddings -- Mitigating Conceptual Shortcuts in Pre-Trained Language Models
von: Kocak, Aysenur, et al.
Veröffentlicht: (2025)
von: Kocak, Aysenur, et al.
Veröffentlicht: (2025)
MedFLIP: Medical Vision-and-Language Self-supervised Fast Pre-Training with Masked Autoencoder
von: Li, Lei, et al.
Veröffentlicht: (2024)
von: Li, Lei, et al.
Veröffentlicht: (2024)
Explicitly Encoding Structural Symmetry is Key to Length Generalization in Arithmetic Tasks
von: Sabbaghi, Mahdi, et al.
Veröffentlicht: (2024)
von: Sabbaghi, Mahdi, et al.
Veröffentlicht: (2024)
ZClip: Adaptive Spike Mitigation for LLM Pre-Training
von: Kumar, Abhay, et al.
Veröffentlicht: (2025)
von: Kumar, Abhay, et al.
Veröffentlicht: (2025)
Pre-Trained Policy Discriminators are General Reward Models
von: Dou, Shihan, et al.
Veröffentlicht: (2025)
von: Dou, Shihan, et al.
Veröffentlicht: (2025)
Incorporating Hierarchical Semantics in Sparse Autoencoder Architectures
von: Muchane, Mark, et al.
Veröffentlicht: (2025)
von: Muchane, Mark, et al.
Veröffentlicht: (2025)
Efficient Knowledge Probing of Large Language Models by Adapting Pre-trained Embeddings
von: Sharma, Kartik, et al.
Veröffentlicht: (2025)
von: Sharma, Kartik, et al.
Veröffentlicht: (2025)
An Aspect Extraction Framework using Different Embedding Types, Learning Models, and Dependency Structure
von: Erkan, Ali, et al.
Veröffentlicht: (2025)
von: Erkan, Ali, et al.
Veröffentlicht: (2025)
Language Model Training Paradigms for Clinical Feature Embeddings
von: Hu, Yurong, et al.
Veröffentlicht: (2023)
von: Hu, Yurong, et al.
Veröffentlicht: (2023)
Attention with Trained Embeddings Provably Selects Important Tokens
von: Wu, Diyuan, et al.
Veröffentlicht: (2025)
von: Wu, Diyuan, et al.
Veröffentlicht: (2025)
MultiGPrompt for Multi-Task Pre-Training and Prompting on Graphs
von: Yu, Xingtong, et al.
Veröffentlicht: (2023)
von: Yu, Xingtong, et al.
Veröffentlicht: (2023)
FineInstructions: Scaling Synthetic Instructions to Pre-Training Scale
von: Patel, Ajay, et al.
Veröffentlicht: (2026)
von: Patel, Ajay, et al.
Veröffentlicht: (2026)
Domain-Adaptive Continued Pre-Training of Small Language Models
von: Faroz, Salman
Veröffentlicht: (2025)
von: Faroz, Salman
Veröffentlicht: (2025)
Development of Pre-Trained Transformer-based Models for the Nepali Language
von: Thapa, Prajwal, et al.
Veröffentlicht: (2024)
von: Thapa, Prajwal, et al.
Veröffentlicht: (2024)
Heterogeneous Self-Supervised Acoustic Pre-Training with Local Constraints
von: Cui, Xiaodong, et al.
Veröffentlicht: (2025)
von: Cui, Xiaodong, et al.
Veröffentlicht: (2025)
ProteinAE: Protein Diffusion Autoencoders for Structure Encoding
von: Li, Shaoning, et al.
Veröffentlicht: (2025)
von: Li, Shaoning, et al.
Veröffentlicht: (2025)
Reinforcement Learning on Pre-Training Data
von: Li, Siheng, et al.
Veröffentlicht: (2025)
von: Li, Siheng, et al.
Veröffentlicht: (2025)
SynAE: A Framework for Measuring the Quality of Synthetic Data for Tool-Calling Agent Evaluations
von: Wang, Shuaiqi, et al.
Veröffentlicht: (2026)
von: Wang, Shuaiqi, et al.
Veröffentlicht: (2026)
Optimizing Pre-Training Data Mixtures with Mixtures of Data Expert Models
von: Belenki, Lior, et al.
Veröffentlicht: (2025)
von: Belenki, Lior, et al.
Veröffentlicht: (2025)
MapFormer: Self-Supervised Learning of Cognitive Maps with Input-Dependent Positional Embeddings
von: Rambaud, Victor, et al.
Veröffentlicht: (2025)
von: Rambaud, Victor, et al.
Veröffentlicht: (2025)
GISTEmbed: Guided In-sample Selection of Training Negatives for Text Embedding Fine-tuning
von: Solatorio, Aivin V.
Veröffentlicht: (2024)
von: Solatorio, Aivin V.
Veröffentlicht: (2024)
PolyPythias: Stability and Outliers across Fifty Language Model Pre-Training Runs
von: van der Wal, Oskar, et al.
Veröffentlicht: (2025)
von: van der Wal, Oskar, et al.
Veröffentlicht: (2025)
Empirical Analysis of Efficient Fine-Tuning Methods for Large Pre-Trained Language Models
von: Doering, Nigel, et al.
Veröffentlicht: (2024)
von: Doering, Nigel, et al.
Veröffentlicht: (2024)
$100K or 100 Days: Trade-offs when Pre-Training with Academic Resources
von: Khandelwal, Apoorv, et al.
Veröffentlicht: (2024)
von: Khandelwal, Apoorv, et al.
Veröffentlicht: (2024)
Investigating the Pre-Training Dynamics of In-Context Learning: Task Recognition vs. Task Learning
von: Wang, Xiaolei, et al.
Veröffentlicht: (2024)
von: Wang, Xiaolei, et al.
Veröffentlicht: (2024)
Time-Aware Feature Selection: Adaptive Temporal Masking for Stable Sparse Autoencoder Training
von: Li, T. Ed, et al.
Veröffentlicht: (2025)
von: Li, T. Ed, et al.
Veröffentlicht: (2025)
Centered Masking for Language-Image Pre-Training
von: Liang, Mingliang, et al.
Veröffentlicht: (2024)
von: Liang, Mingliang, et al.
Veröffentlicht: (2024)
Interpretable Company Similarity with Sparse Autoencoders
von: Molinari, Marco, et al.
Veröffentlicht: (2024)
von: Molinari, Marco, et al.
Veröffentlicht: (2024)
Evaluating Embedding Frameworks for Scientific Domain
von: Ahmed, Nouman, et al.
Veröffentlicht: (2025)
von: Ahmed, Nouman, et al.
Veröffentlicht: (2025)
Zero2Text: Zero-Training Cross-Domain Inversion Attacks on Textual Embeddings
von: Kim, Doohyun, et al.
Veröffentlicht: (2026)
von: Kim, Doohyun, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Self-StrAE at SemEval-2024 Task 1: Making Self-Structuring AutoEncoders Learn More With Less
von: Opper, Mattia, et al.
Veröffentlicht: (2024) -
Banyan: Improved Representation Learning with Explicit Structure
von: Opper, Mattia, et al.
Veröffentlicht: (2024) -
TRA: Better Length Generalisation with Threshold Relative Attention
von: Opper, Mattia, et al.
Veröffentlicht: (2025) -
Autoencoding Conditional Neural Processes for Representation Learning
von: Prokhorov, Victor, et al.
Veröffentlicht: (2023) -
Bootstrapping Embeddings for Low Resource Languages
von: Basoz, Merve, et al.
Veröffentlicht: (2026)