On the Transformations across Reward Model, Parameter Update, and In-Context Prompt
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cai, Deng, Li, Huayang, Fu, Tingchen, Li, Siheng, Xu, Weiwen, Li, Shuaiyi, Cao, Bowen, Zhang, Zhisong, Huang, Xinting, Cui, Leyang, Wang, Yan, Liu, Lemao, Watanabe, Taro, Shi, Shuming |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TextBind: Multi-turn Interleaved Multimodal Instruction-following in the Wild
von: Li, Huayang, et al.
Veröffentlicht: (2023)
von: Li, Huayang, et al.
Veröffentlicht: (2023)
Inferflow: an Efficient and Highly Configurable Inference Engine for Large Language Models
von: Shi, Shuming, et al.
Veröffentlicht: (2024)
von: Shi, Shuming, et al.
Veröffentlicht: (2024)
Disperse-Then-Merge: Pushing the Limits of Instruction Tuning via Alignment Tax Reduction
von: Fu, Tingchen, et al.
Veröffentlicht: (2024)
von: Fu, Tingchen, et al.
Veröffentlicht: (2024)
Selection-p: Self-Supervised Task-Agnostic Prompt Compression for Faithfulness and Transferability
von: Chung, Tsz Ting, et al.
Veröffentlicht: (2024)
von: Chung, Tsz Ting, et al.
Veröffentlicht: (2024)
Cross-lingual Contextualized Phrase Retrieval
von: Li, Huayang, et al.
Veröffentlicht: (2024)
von: Li, Huayang, et al.
Veröffentlicht: (2024)
SeqPE: Transformer with Sequential Position Encoding
von: Li, Huayang, et al.
Veröffentlicht: (2025)
von: Li, Huayang, et al.
Veröffentlicht: (2025)
A Silver Bullet or a Compromise for Full Attention? A Comprehensive Study of Gist Token-based Context Compression
von: Deng, Chenlong, et al.
Veröffentlicht: (2024)
von: Deng, Chenlong, et al.
Veröffentlicht: (2024)
Not All Preference Pairs Are Created Equal: A Recipe for Annotation-Efficient Iterative Preference Learning
von: Yang, Sen, et al.
Veröffentlicht: (2024)
von: Yang, Sen, et al.
Veröffentlicht: (2024)
Reasons to Reject? Aligning Language Models with Judgments
von: Xu, Weiwen, et al.
Veröffentlicht: (2023)
von: Xu, Weiwen, et al.
Veröffentlicht: (2023)
Attention Entropy is a Key Factor: An Analysis of Parallel Context Encoding with Full-attention-based Pre-trained Language Models
von: Zhang, Zhisong, et al.
Veröffentlicht: (2024)
von: Zhang, Zhisong, et al.
Veröffentlicht: (2024)
Siren's Song in the AI Ocean: A Survey on Hallucination in Large Language Models
von: Zhang, Yue, et al.
Veröffentlicht: (2023)
von: Zhang, Yue, et al.
Veröffentlicht: (2023)
BBA: Bi-Modal Behavioral Alignment for Reasoning with Large Vision-Language Models
von: Zhao, Xueliang, et al.
Veröffentlicht: (2024)
von: Zhao, Xueliang, et al.
Veröffentlicht: (2024)
Knowledge Verification to Nip Hallucination in the Bud
von: Wan, Fanqi, et al.
Veröffentlicht: (2024)
von: Wan, Fanqi, et al.
Veröffentlicht: (2024)
Same Question, Different Words: A Latent Adversarial Framework for Prompt Robustness
von: Fu, Tingchen, et al.
Veröffentlicht: (2025)
von: Fu, Tingchen, et al.
Veröffentlicht: (2025)
Retrieval is Accurate Generation
von: Cao, Bowen, et al.
Veröffentlicht: (2024)
von: Cao, Bowen, et al.
Veröffentlicht: (2024)
On the Worst Prompt Performance of Large Language Models
von: Cao, Bowen, et al.
Veröffentlicht: (2024)
von: Cao, Bowen, et al.
Veröffentlicht: (2024)
RePo: Language Models with Context Re-Positioning
von: Li, Huayang, et al.
Veröffentlicht: (2025)
von: Li, Huayang, et al.
Veröffentlicht: (2025)
A Frustratingly Simple Decoding Method for Neural Text Generation
von: Yang, Haoran, et al.
Veröffentlicht: (2023)
von: Yang, Haoran, et al.
Veröffentlicht: (2023)
UniGist: Towards General and Hardware-aligned Sequence-level Long Context Compression
von: Deng, Chenlong, et al.
Veröffentlicht: (2025)
von: Deng, Chenlong, et al.
Veröffentlicht: (2025)
Alleviating Hallucinations of Large Language Models through Induced Hallucinations
von: Zhang, Yue, et al.
Veröffentlicht: (2023)
von: Zhang, Yue, et al.
Veröffentlicht: (2023)
An Energy-based Model for Word-level AutoCompletion in Computer-aided Translation
von: Yang, Cheng, et al.
Veröffentlicht: (2024)
von: Yang, Cheng, et al.
Veröffentlicht: (2024)
Towards Generalization of Block Attention via Automatic Segmentation and Block Distillation
von: Li, Shuaiyi, et al.
Veröffentlicht: (2026)
von: Li, Shuaiyi, et al.
Veröffentlicht: (2026)
DepWiGNN: A Depth-wise Graph Neural Network for Multi-hop Spatial Reasoning in Text
von: Li, Shuaiyi, et al.
Veröffentlicht: (2023)
von: Li, Shuaiyi, et al.
Veröffentlicht: (2023)
Spotting AI's Touch: Identifying LLM-Paraphrased Spans in Text
von: Li, Yafu, et al.
Veröffentlicht: (2024)
von: Li, Yafu, et al.
Veröffentlicht: (2024)
Towards Privacy-Preserving Machine Translation at the Inference Stage: A New Task and Benchmark
von: Shao, Wei, et al.
Veröffentlicht: (2026)
von: Shao, Wei, et al.
Veröffentlicht: (2026)
Knowledge Fusion of Large Language Models
von: Wan, Fanqi, et al.
Veröffentlicht: (2024)
von: Wan, Fanqi, et al.
Veröffentlicht: (2024)
Unlocking Decoding-time Controllability: Gradient-Free Multi-Objective Alignment with Contrastive Prompts
von: Fu, Tingchen, et al.
Veröffentlicht: (2024)
von: Fu, Tingchen, et al.
Veröffentlicht: (2024)
InComeS: Integrating Compression and Selection Mechanisms into LLMs for Efficient Model Editing
von: Li, Shuaiyi, et al.
Veröffentlicht: (2025)
von: Li, Shuaiyi, et al.
Veröffentlicht: (2025)
ALR$^2$: A Retrieve-then-Reason Framework for Long-context Question Answering
von: Li, Huayang, et al.
Veröffentlicht: (2024)
von: Li, Huayang, et al.
Veröffentlicht: (2024)
Equivalence of Context and Parameter Updates in Modern Transformer Blocks
von: Goldwaser, Adrian, et al.
Veröffentlicht: (2025)
von: Goldwaser, Adrian, et al.
Veröffentlicht: (2025)
From Drafts to Answers: Unlocking LLM Potential via Aggregation Fine-Tuning
von: Li, Yafu, et al.
Veröffentlicht: (2025)
von: Li, Yafu, et al.
Veröffentlicht: (2025)
Consecutive Batch Model Editing with HooK Layers
von: Li, Shuaiyi, et al.
Veröffentlicht: (2024)
von: Li, Shuaiyi, et al.
Veröffentlicht: (2024)
InfiniteICL: Breaking the Limit of Context Window Size via Long Short-term Memory Transformation
von: Cao, Bowen, et al.
Veröffentlicht: (2025)
von: Cao, Bowen, et al.
Veröffentlicht: (2025)
Large Language Models Can Self-Improve in Long-context Reasoning
von: Li, Siheng, et al.
Veröffentlicht: (2024)
von: Li, Siheng, et al.
Veröffentlicht: (2024)
Confidence as a Reward: Transforming LLMs into Reward Models
von: Du, He, et al.
Veröffentlicht: (2025)
von: Du, He, et al.
Veröffentlicht: (2025)
Block-Decomposition for 3-Parameter Persistence Modules
von: Yi, Siheng
Veröffentlicht: (2025)
von: Yi, Siheng
Veröffentlicht: (2025)
XQ-MEval: A Dataset with Cross-lingual Parallel Quality for Benchmarking Translation Metrics
von: Liu, Jingxuan, et al.
Veröffentlicht: (2026)
von: Liu, Jingxuan, et al.
Veröffentlicht: (2026)
Scaling Reasoning, Losing Control: Evaluating Instruction Following in Large Reasoning Models
von: Fu, Tingchen, et al.
Veröffentlicht: (2025)
von: Fu, Tingchen, et al.
Veröffentlicht: (2025)
MAGE: Machine-generated Text Detection in the Wild
von: Li, Yafu, et al.
Veröffentlicht: (2023)
von: Li, Yafu, et al.
Veröffentlicht: (2023)
Poly(ether‐ester)s synthesized from furandicarboxylic acid and ethylene glycol using an acidic binary catalyst
von: Zhisong Li, et al.
Veröffentlicht: (2024)
von: Zhisong Li, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
TextBind: Multi-turn Interleaved Multimodal Instruction-following in the Wild
von: Li, Huayang, et al.
Veröffentlicht: (2023) -
Inferflow: an Efficient and Highly Configurable Inference Engine for Large Language Models
von: Shi, Shuming, et al.
Veröffentlicht: (2024) -
Disperse-Then-Merge: Pushing the Limits of Instruction Tuning via Alignment Tax Reduction
von: Fu, Tingchen, et al.
Veröffentlicht: (2024) -
Selection-p: Self-Supervised Task-Agnostic Prompt Compression for Faithfulness and Transferability
von: Chung, Tsz Ting, et al.
Veröffentlicht: (2024) -
Cross-lingual Contextualized Phrase Retrieval
von: Li, Huayang, et al.
Veröffentlicht: (2024)