Investigating the Effectiveness of HyperTuning via Gisting
Fuente:
arXiv
Salvato in:
| Autore principale: | Phang, Jason |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
GistScore: Learning Better Representations for In-Context Example Selection with Gist Bottlenecks
di: Gupta, Shivanshu, et al.
Pubblicazione: (2023)
di: Gupta, Shivanshu, et al.
Pubblicazione: (2023)
Compressing Lengthy Context With UltraGist
di: Zhang, Peitian, et al.
Pubblicazione: (2024)
di: Zhang, Peitian, et al.
Pubblicazione: (2024)
Learning to Compress Prompts with Gist Tokens
di: Mu, Jesse, et al.
Pubblicazione: (2023)
di: Mu, Jesse, et al.
Pubblicazione: (2023)
Sentence-Anchored Gist Compression for Long-Context LLMs
di: Tarasov, Dmitrii, et al.
Pubblicazione: (2025)
di: Tarasov, Dmitrii, et al.
Pubblicazione: (2025)
SciGisPy: a Novel Metric for Biomedical Text Simplification via Gist Inference Score
di: Lyu, Chen, et al.
Pubblicazione: (2024)
di: Lyu, Chen, et al.
Pubblicazione: (2024)
Say More with Less: Understanding Prompt Learning Behaviors through Gist Compression
di: Li, Xinze, et al.
Pubblicazione: (2024)
di: Li, Xinze, et al.
Pubblicazione: (2024)
UniGist: Towards General and Hardware-aligned Sequence-level Long Context Compression
di: Deng, Chenlong, et al.
Pubblicazione: (2025)
di: Deng, Chenlong, et al.
Pubblicazione: (2025)
EmoGist: Efficient In-Context Learning for Visual Emotion Understanding
di: Seoh, Ronald, et al.
Pubblicazione: (2025)
di: Seoh, Ronald, et al.
Pubblicazione: (2025)
Understand Then Memory: A Cognitive Gist-Driven RAG Framework with Global Semantic Diffusion
di: Zhou, Pengcheng, et al.
Pubblicazione: (2026)
di: Zhou, Pengcheng, et al.
Pubblicazione: (2026)
A Human-Inspired Reading Agent with Gist Memory of Very Long Contexts
di: Lee, Kuang-Huei, et al.
Pubblicazione: (2024)
di: Lee, Kuang-Huei, et al.
Pubblicazione: (2024)
A Silver Bullet or a Compromise for Full Attention? A Comprehensive Study of Gist Token-based Context Compression
di: Deng, Chenlong, et al.
Pubblicazione: (2024)
di: Deng, Chenlong, et al.
Pubblicazione: (2024)
Struc-Bench: Are Large Language Models Really Good at Generating Complex Structured Data?
di: Tang, Xiangru, et al.
Pubblicazione: (2023)
di: Tang, Xiangru, et al.
Pubblicazione: (2023)
Two Failures of Self-Consistency in the Multi-Step Reasoning of LLMs
di: Chen, Angelica, et al.
Pubblicazione: (2023)
di: Chen, Angelica, et al.
Pubblicazione: (2023)
Large Language Models as Misleading Assistants in Conversation
di: Hou, Betty Li, et al.
Pubblicazione: (2024)
di: Hou, Betty Li, et al.
Pubblicazione: (2024)
From Verbatim to Gist: Distilling Pyramidal Multimodal Memory via Semantic Information Bottleneck for Long-Horizon Video Agents
di: Lian, Niu, et al.
Pubblicazione: (2026)
di: Lian, Niu, et al.
Pubblicazione: (2026)
AutoRAG-HP: Automatic Online Hyper-Parameter Tuning for Retrieval-Augmented Generation
di: Fu, Jia, et al.
Pubblicazione: (2024)
di: Fu, Jia, et al.
Pubblicazione: (2024)
ImageInWords: Unlocking Hyper-Detailed Image Descriptions
di: Garg, Roopal, et al.
Pubblicazione: (2024)
di: Garg, Roopal, et al.
Pubblicazione: (2024)
Efficient and Effective Prompt Tuning via Prompt Decomposition and Compressed Outer Product
di: Lan, Pengxiang, et al.
Pubblicazione: (2025)
di: Lan, Pengxiang, et al.
Pubblicazione: (2025)
HyperLLaVA: Dynamic Visual and Language Expert Tuning for Multimodal Large Language Models
di: Zhang, Wenqiao, et al.
Pubblicazione: (2024)
di: Zhang, Wenqiao, et al.
Pubblicazione: (2024)
HypeLoRA: Hyper-Network-Generated LoRA Adapters for Calibrated Language Model Fine-Tuning
di: Trojan, Bartosz, et al.
Pubblicazione: (2026)
di: Trojan, Bartosz, et al.
Pubblicazione: (2026)
PAFT: A Parallel Training Paradigm for Effective LLM Fine-Tuning
di: Pentyala, Shiva Kumar, et al.
Pubblicazione: (2024)
di: Pentyala, Shiva Kumar, et al.
Pubblicazione: (2024)
Evaluating The Impact of Stimulus Quality in Investigations of LLM Language Performance
di: Pistotti, Timothy, et al.
Pubblicazione: (2025)
di: Pistotti, Timothy, et al.
Pubblicazione: (2025)
Hyper-CL: Conditioning Sentence Representations with Hypernetworks
di: Yoo, Young Hyun, et al.
Pubblicazione: (2024)
di: Yoo, Young Hyun, et al.
Pubblicazione: (2024)
Investigating Multilingual Instruction-Tuning: Do Polyglot Models Demand for Multilingual Instructions?
di: Weber, Alexander Arno, et al.
Pubblicazione: (2024)
di: Weber, Alexander Arno, et al.
Pubblicazione: (2024)
Router-Tuning: A Simple and Effective Approach for Enabling Dynamic-Depth in Transformers
di: He, Shwai, et al.
Pubblicazione: (2024)
di: He, Shwai, et al.
Pubblicazione: (2024)
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation
di: Zhu, Runchuan, et al.
Pubblicazione: (2025)
di: Zhu, Runchuan, et al.
Pubblicazione: (2025)
Critique Fine-Tuning: Learning to Critique is More Effective than Learning to Imitate
di: Wang, Yubo, et al.
Pubblicazione: (2025)
di: Wang, Yubo, et al.
Pubblicazione: (2025)
Investigating Instruction Tuning Large Language Models on Graphs
di: Zhu, Kerui, et al.
Pubblicazione: (2024)
di: Zhu, Kerui, et al.
Pubblicazione: (2024)
Hyper-Connections
di: Zhu, Defa, et al.
Pubblicazione: (2024)
di: Zhu, Defa, et al.
Pubblicazione: (2024)
LoRA-Squeeze: Simple and Effective Post-Tuning and In-Tuning Compression of LoRA Modules
di: Vulić, Ivan, et al.
Pubblicazione: (2026)
di: Vulić, Ivan, et al.
Pubblicazione: (2026)
HyperEdit: Unlocking Instruction-based Text Editing in LLMs via Hypernetworks
di: Zeng, Yiming, et al.
Pubblicazione: (2025)
di: Zeng, Yiming, et al.
Pubblicazione: (2025)
Confident in a Confidence Score: Investigating the Sensitivity of Confidence Scores to Supervised Fine-Tuning
di: Flores, Lorenzo Jaime Yu, et al.
Pubblicazione: (2026)
di: Flores, Lorenzo Jaime Yu, et al.
Pubblicazione: (2026)
Agent-FLAN: Designing Data and Methods of Effective Agent Tuning for Large Language Models
di: Chen, Zehui, et al.
Pubblicazione: (2024)
di: Chen, Zehui, et al.
Pubblicazione: (2024)
Self-Tuning: Instructing LLMs to Effectively Acquire New Knowledge through Self-Teaching
di: Zhang, Xiaoying, et al.
Pubblicazione: (2024)
di: Zhang, Xiaoying, et al.
Pubblicazione: (2024)
Investigating the Multilingual Calibration Effects of Language Model Instruction-Tuning
di: Huang, Jerry, et al.
Pubblicazione: (2026)
di: Huang, Jerry, et al.
Pubblicazione: (2026)
HyperCLOVA X Technical Report
di: Yoo, Kang Min, et al.
Pubblicazione: (2024)
di: Yoo, Kang Min, et al.
Pubblicazione: (2024)
Is Fine-Tuning an Effective Solution? Reassessing Knowledge Editing for Unstructured Data
di: Xiong, Hao, et al.
Pubblicazione: (2025)
di: Xiong, Hao, et al.
Pubblicazione: (2025)
Investigating the Impact of Language-Adaptive Fine-Tuning on Sentiment Analysis in Hausa Language Using AfriBERTa
di: Sani, Sani Abdullahi, et al.
Pubblicazione: (2025)
di: Sani, Sani Abdullahi, et al.
Pubblicazione: (2025)
Instruction Matters: A Simple yet Effective Task Selection for Optimized Instruction Tuning of Specific Tasks
di: Lee, Changho, et al.
Pubblicazione: (2024)
di: Lee, Changho, et al.
Pubblicazione: (2024)
MoDE: Effective Multi-task Parameter Efficient Fine-Tuning with a Mixture of Dyadic Experts
di: Ning, Lin, et al.
Pubblicazione: (2024)
di: Ning, Lin, et al.
Pubblicazione: (2024)
Documenti analoghi
-
GistScore: Learning Better Representations for In-Context Example Selection with Gist Bottlenecks
di: Gupta, Shivanshu, et al.
Pubblicazione: (2023) -
Compressing Lengthy Context With UltraGist
di: Zhang, Peitian, et al.
Pubblicazione: (2024) -
Learning to Compress Prompts with Gist Tokens
di: Mu, Jesse, et al.
Pubblicazione: (2023) -
Sentence-Anchored Gist Compression for Long-Context LLMs
di: Tarasov, Dmitrii, et al.
Pubblicazione: (2025) -
SciGisPy: a Novel Metric for Biomedical Text Simplification via Gist Inference Score
di: Lyu, Chen, et al.
Pubblicazione: (2024)