The more polypersonal the better -- a short look on space geometry of fine-tuned layers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kudriashov, Sergei, Zykova, Veronika, Stepanova, Angelina, Raskind, Yakov, Klyshinsky, Eduard |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Shrink the longest: improving latent space isotropy with symplicial geometry
von: Kudriashov, Sergei, et al.
Veröffentlicht: (2025)
von: Kudriashov, Sergei, et al.
Veröffentlicht: (2025)
Complexity-aware fine-tuning
von: Goncharov, Andrey, et al.
Veröffentlicht: (2025)
von: Goncharov, Andrey, et al.
Veröffentlicht: (2025)
Replaying pre-training data improves fine-tuning
von: Kotha, Suhas, et al.
Veröffentlicht: (2026)
von: Kotha, Suhas, et al.
Veröffentlicht: (2026)
Watch the Weights: Unsupervised monitoring and control of fine-tuned LLMs
von: Zhong, Ziqian, et al.
Veröffentlicht: (2025)
von: Zhong, Ziqian, et al.
Veröffentlicht: (2025)
You can remove GPT2's LayerNorm by fine-tuning
von: Heimersheim, Stefan
Veröffentlicht: (2024)
von: Heimersheim, Stefan
Veröffentlicht: (2024)
LayerNorm: A key component in parameter-efficient fine-tuning
von: ValizadehAslani, Taha, et al.
Veröffentlicht: (2024)
von: ValizadehAslani, Taha, et al.
Veröffentlicht: (2024)
The representation landscape of few-shot learning and fine-tuning in large language models
von: Doimo, Diego, et al.
Veröffentlicht: (2024)
von: Doimo, Diego, et al.
Veröffentlicht: (2024)
Deep literature reviews: an application of fine-tuned language models to migration research
von: Iacus, Stefano M., et al.
Veröffentlicht: (2025)
von: Iacus, Stefano M., et al.
Veröffentlicht: (2025)
Does fine-tuning GPT-3 with the OpenAI API leak personally-identifiable information?
von: Sun, Albert Yu, et al.
Veröffentlicht: (2023)
von: Sun, Albert Yu, et al.
Veröffentlicht: (2023)
Rethinking harmless refusals when fine-tuning foundation models
von: Pop, Florin, et al.
Veröffentlicht: (2024)
von: Pop, Florin, et al.
Veröffentlicht: (2024)
Multi-task retriever fine-tuning for domain-specific and efficient RAG
von: Béchard, Patrice, et al.
Veröffentlicht: (2025)
von: Béchard, Patrice, et al.
Veröffentlicht: (2025)
Uncertainty quantification in fine-tuned LLMs using LoRA ensembles
von: Balabanov, Oleksandr, et al.
Veröffentlicht: (2024)
von: Balabanov, Oleksandr, et al.
Veröffentlicht: (2024)
Ask more, know better: Reinforce-Learned Prompt Questions for Decision Making with Large Language Models
von: Yan, Xue, et al.
Veröffentlicht: (2023)
von: Yan, Xue, et al.
Veröffentlicht: (2023)
Improving embedding with contrastive fine-tuning on small datasets with expert-augmented scores
von: Lu, Jun, et al.
Veröffentlicht: (2024)
von: Lu, Jun, et al.
Veröffentlicht: (2024)
FiMMIA: scaling semantic perturbation-based membership inference across modalities
von: Emelyanov, Anton, et al.
Veröffentlicht: (2025)
von: Emelyanov, Anton, et al.
Veröffentlicht: (2025)
What explains the success of cross-modal fine-tuning with ORCA?
von: García-de-Herreros, Paloma, et al.
Veröffentlicht: (2024)
von: García-de-Herreros, Paloma, et al.
Veröffentlicht: (2024)
Cyclic Ablation: Testing Concept Localization against Functional Regeneration in AI
von: Kapelko, Eduard
Veröffentlicht: (2025)
von: Kapelko, Eduard
Veröffentlicht: (2025)
Packing Analysis: Packing Is More Appropriate for Large Models or Datasets in Supervised Fine-tuning
von: Wang, Shuhe, et al.
Veröffentlicht: (2024)
von: Wang, Shuhe, et al.
Veröffentlicht: (2024)
TextGram: Towards a better domain-adaptive pretraining
von: Hiwarkhedkar, Sharayu, et al.
Veröffentlicht: (2024)
von: Hiwarkhedkar, Sharayu, et al.
Veröffentlicht: (2024)
Investigating the performance of Retrieval-Augmented Generation and fine-tuning for the development of AI-driven knowledge-based systems
von: Lakatos, Robert, et al.
Veröffentlicht: (2024)
von: Lakatos, Robert, et al.
Veröffentlicht: (2024)
Enhancing Q&A Text Retrieval with Ranking Models: Benchmarking, fine-tuning and deploying Rerankers for RAG
von: Moreira, Gabriel de Souza P., et al.
Veröffentlicht: (2024)
von: Moreira, Gabriel de Souza P., et al.
Veröffentlicht: (2024)
Memento: Fine-tuning LLM Agents without Fine-tuning LLMs
von: Zhou, Huichi, et al.
Veröffentlicht: (2025)
von: Zhou, Huichi, et al.
Veröffentlicht: (2025)
SlimSpec: Low-Rank Draft LM-Head for Accelerated Speculative Decoding
von: Plaksin, Anton, et al.
Veröffentlicht: (2026)
von: Plaksin, Anton, et al.
Veröffentlicht: (2026)
From Small to Large Language Models: Revisiting the Federalist Papers
von: Jeong, So Won, et al.
Veröffentlicht: (2025)
von: Jeong, So Won, et al.
Veröffentlicht: (2025)
Semantic similarity prediction is better than other semantic similarity measures
von: Herbold, Steffen
Veröffentlicht: (2023)
von: Herbold, Steffen
Veröffentlicht: (2023)
Prompt engineering paradigms for medical applications: scoping review and recommendations for better practices
von: Zaghir, Jamil, et al.
Veröffentlicht: (2024)
von: Zaghir, Jamil, et al.
Veröffentlicht: (2024)
LK Losses: Direct Acceptance Rate Optimization for Speculative Decoding
von: Samarin, Alexander, et al.
Veröffentlicht: (2026)
von: Samarin, Alexander, et al.
Veröffentlicht: (2026)
Do LLMs Understand Romanian Driving Laws? A Study on Multimodal and Fine-Tuned Question Answering
von: Barbu, Eduard, et al.
Veröffentlicht: (2025)
von: Barbu, Eduard, et al.
Veröffentlicht: (2025)
NorMuon: Making Muon more efficient and scalable
von: Li, Zichong, et al.
Veröffentlicht: (2025)
von: Li, Zichong, et al.
Veröffentlicht: (2025)
Two is better than one: A Collapse-free Multi-Reward RLIF Training Framework
von: Joarder, Shourov, et al.
Veröffentlicht: (2026)
von: Joarder, Shourov, et al.
Veröffentlicht: (2026)
Topic Modeling with Fine-tuning LLMs and Bag of Sentences
von: Schneider, Johannes
Veröffentlicht: (2024)
von: Schneider, Johannes
Veröffentlicht: (2024)
SEE: Continual Fine-tuning with Sequential Ensemble of Experts
von: Wang, Zhilin, et al.
Veröffentlicht: (2025)
von: Wang, Zhilin, et al.
Veröffentlicht: (2025)
On the Loss of Context-awareness in General Instruction Fine-tuning
von: Wang, Yihan, et al.
Veröffentlicht: (2024)
von: Wang, Yihan, et al.
Veröffentlicht: (2024)
RAG vs Fine-tuning: Pipelines, Tradeoffs, and a Case Study on Agriculture
von: Balaguer, Angels, et al.
Veröffentlicht: (2024)
von: Balaguer, Angels, et al.
Veröffentlicht: (2024)
Whispering in Amharic: Fine-tuning Whisper for Low-resource Language
von: Gete, Dawit Ketema, et al.
Veröffentlicht: (2025)
von: Gete, Dawit Ketema, et al.
Veröffentlicht: (2025)
Efficient Ensemble for Fine-tuning Language Models on Multiple Datasets
von: Li, Dongyue, et al.
Veröffentlicht: (2025)
von: Li, Dongyue, et al.
Veröffentlicht: (2025)
Relational Knowledge Distillation Using Fine-tuned Function Vectors
von: Kang, Andrea, et al.
Veröffentlicht: (2026)
von: Kang, Andrea, et al.
Veröffentlicht: (2026)
LoRA vs Full Fine-tuning: An Illusion of Equivalence
von: Shuttleworth, Reece, et al.
Veröffentlicht: (2024)
von: Shuttleworth, Reece, et al.
Veröffentlicht: (2024)
PMSS: Pretrained Matrices Skeleton Selection for LLM Fine-tuning
von: Wang, Qibin, et al.
Veröffentlicht: (2024)
von: Wang, Qibin, et al.
Veröffentlicht: (2024)
Advancing Parameter Efficiency in Fine-tuning via Representation Editing
von: Wu, Muling, et al.
Veröffentlicht: (2024)
von: Wu, Muling, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Shrink the longest: improving latent space isotropy with symplicial geometry
von: Kudriashov, Sergei, et al.
Veröffentlicht: (2025) -
Complexity-aware fine-tuning
von: Goncharov, Andrey, et al.
Veröffentlicht: (2025) -
Replaying pre-training data improves fine-tuning
von: Kotha, Suhas, et al.
Veröffentlicht: (2026) -
Watch the Weights: Unsupervised monitoring and control of fine-tuned LLMs
von: Zhong, Ziqian, et al.
Veröffentlicht: (2025) -
You can remove GPT2's LayerNorm by fine-tuning
von: Heimersheim, Stefan
Veröffentlicht: (2024)