TRACE: TRansformer-based Attribution using Contrastive Embeddings in LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Cheng, Lu, Xinyang, Ng, See-Kiong, Low, Bryan Kian Hsiang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Global-to-Local Support Spectrums for Language Model Explainability
by: Agussurja, Lucas, et al.
Published: (2024)
by: Agussurja, Lucas, et al.
Published: (2024)
Use Your INSTINCT: INSTruction optimization for LLMs usIng Neural bandits Coupled with Transformers
by: Lin, Xiaoqiang, et al.
Published: (2023)
by: Lin, Xiaoqiang, et al.
Published: (2023)
Source Attribution for Large Language Model-Generated Data
by: Wang, Jingtan, et al.
Published: (2023)
by: Wang, Jingtan, et al.
Published: (2023)
Dipper: Diversity in Prompts for Producing Large Language Model Ensembles in Reasoning tasks
by: Lau, Gregory Kang Ruey, et al.
Published: (2024)
by: Lau, Gregory Kang Ruey, et al.
Published: (2024)
Prompt Optimization with EASE? Efficient Ordering-aware Automated Selection of Exemplars
by: Wu, Zhaoxuan, et al.
Published: (2024)
by: Wu, Zhaoxuan, et al.
Published: (2024)
On Newton's Method to Unlearn Neural Networks
by: Bui, Nhung, et al.
Published: (2024)
by: Bui, Nhung, et al.
Published: (2024)
Test-Time Scaling in Reasoning Models Is Not Effective for Knowledge-Intensive Tasks Yet
by: Zhao, James Xu, et al.
Published: (2025)
by: Zhao, James Xu, et al.
Published: (2025)
PINNACLE: PINN Adaptive ColLocation and Experimental points selection
by: Lau, Gregory Kang Ruey, et al.
Published: (2024)
by: Lau, Gregory Kang Ruey, et al.
Published: (2024)
PIED: Physics-Informed Experimental Design for Inverse Problems
by: Hemachandra, Apivich, et al.
Published: (2025)
by: Hemachandra, Apivich, et al.
Published: (2025)
DETAIL: Task DEmonsTration Attribution for Interpretable In-context Learning
by: Zhou, Zijian, et al.
Published: (2024)
by: Zhou, Zijian, et al.
Published: (2024)
REFRAG: Rethinking RAG based Decoding
by: Lin, Xiaoqiang, et al.
Published: (2025)
by: Lin, Xiaoqiang, et al.
Published: (2025)
Integrating Time Series into LLMs via Multi-layer Steerable Embedding Fusion for Enhanced Forecasting
by: Chen, Zhuomin, et al.
Published: (2025)
by: Chen, Zhuomin, et al.
Published: (2025)
Understanding the Relationship between Prompts and Response Uncertainty in Large Language Models
by: Zhang, Ze Yu, et al.
Published: (2024)
by: Zhang, Ze Yu, et al.
Published: (2024)
Ferret: Federated Full-Parameter Tuning at Scale for Large Language Models
by: Shu, Yao, et al.
Published: (2024)
by: Shu, Yao, et al.
Published: (2024)
DUPRE: Data Utility Prediction for Efficient Data Valuation
by: Pham, Kieu Thao Nguyen, et al.
Published: (2025)
by: Pham, Kieu Thao Nguyen, et al.
Published: (2025)
How Does Response Length Affect Long-Form Factuality
by: Zhao, James Xu, et al.
Published: (2025)
by: Zhao, James Xu, et al.
Published: (2025)
Confidence Elicitation: A New Attack Vector for Large Language Models
by: Formento, Brian, et al.
Published: (2025)
by: Formento, Brian, et al.
Published: (2025)
Prompt Optimization with Human Feedback
by: Lin, Xiaoqiang, et al.
Published: (2024)
by: Lin, Xiaoqiang, et al.
Published: (2024)
ActiveDPO: Active Direct Preference Optimization for Sample-Efficient Alignment
by: Lin, Xiaoqiang, et al.
Published: (2025)
by: Lin, Xiaoqiang, et al.
Published: (2025)
Rewarding the Rare: Uniqueness-Aware RL for Creative Problem Solving in LLMs
by: Hu, Zhiyuan, et al.
Published: (2026)
by: Hu, Zhiyuan, et al.
Published: (2026)
Fine-tuning Language Models with Generative Adversarial Reward Modelling
by: Yu, Zhang Ze, et al.
Published: (2023)
by: Yu, Zhang Ze, et al.
Published: (2023)
Understanding Domain Generalization: A Noise Robustness Perspective
by: Qiao, Rui, et al.
Published: (2024)
by: Qiao, Rui, et al.
Published: (2024)
Don't Just Say "I don't know"! Self-aligning Large Language Models for Responding to Unknown Questions with Explanations
by: Deng, Yang, et al.
Published: (2024)
by: Deng, Yang, et al.
Published: (2024)
SemRoDe: Macro Adversarial Training to Learn Representations That are Robust to Word-Level Attacks
by: Formento, Brian, et al.
Published: (2024)
by: Formento, Brian, et al.
Published: (2024)
Uncertainty Quantification for Multimodal Large Language Models with Incoherence-adjusted Semantic Volume
by: Lau, Gregory Kang Ruey, et al.
Published: (2026)
by: Lau, Gregory Kang Ruey, et al.
Published: (2026)
Self-Interested Agents in Collaborative Machine Learning: An Incentivized Adaptive Data-Centric Framework
by: Vijayan, Nithia, et al.
Published: (2024)
by: Vijayan, Nithia, et al.
Published: (2024)
Uncovering Scaling Laws for Large Language Models via Inverse Problems
by: Verma, Arun, et al.
Published: (2025)
by: Verma, Arun, et al.
Published: (2025)
Inference-Time Attribute Distribution Alignment for Unconditional Diffusion
by: Luan, Hao, et al.
Published: (2026)
by: Luan, Hao, et al.
Published: (2026)
PC-MoE: Memory-Efficient and Privacy-Preserving Collaborative Training for Mixture-of-Experts LLMs
by: Zhang, Ze Yu, et al.
Published: (2025)
by: Zhang, Ze Yu, et al.
Published: (2025)
MineDraft: A Framework for Batch Parallel Speculative Decoding
by: Tang, Zhenwei, et al.
Published: (2026)
by: Tang, Zhenwei, et al.
Published: (2026)
Decentralized Sum-of-Nonconvex Optimization
by: Liu, Zhuanghua, et al.
Published: (2024)
by: Liu, Zhuanghua, et al.
Published: (2024)
When In-Distribution Gains Fail: Evaluating Weak-to-Strong Reward Models under Preference Shift
by: Le, Khoi, et al.
Published: (2026)
by: Le, Khoi, et al.
Published: (2026)
Uncertainty of Thoughts: Uncertainty-Aware Planning Enhances Information Seeking in Large Language Models
by: Hu, Zhiyuan, et al.
Published: (2024)
by: Hu, Zhiyuan, et al.
Published: (2024)
Empirical Study of Named Entity Recognition Performance Using Distribution-aware Word Embedding
by: Chen, Xin, et al.
Published: (2021)
by: Chen, Xin, et al.
Published: (2021)
Continual Multimodal Contrastive Learning
by: Liu, Xiaohao, et al.
Published: (2025)
by: Liu, Xiaohao, et al.
Published: (2025)
TreeGrad-Ranker: Feature Ranking via $O(L)$-Time Gradients for Decision Trees
by: Li, Weida, et al.
Published: (2026)
by: Li, Weida, et al.
Published: (2026)
Provably Adaptive Linear Approximation for the Shapley Value and Beyond
by: Li, Weida, et al.
Published: (2026)
by: Li, Weida, et al.
Published: (2026)
TRACE Back from the Future: A Probabilistic Reasoning Approach to Controllable Language Generation
by: Weng, Gwen Yidou, et al.
Published: (2025)
by: Weng, Gwen Yidou, et al.
Published: (2025)
WaterDrum: Watermarking for Data-centric Unlearning Metric
by: Lu, Xinyang, et al.
Published: (2025)
by: Lu, Xinyang, et al.
Published: (2025)
Helpful or Harmful Data? Fine-tuning-free Shapley Attribution for Explaining Language Model Predictions
by: Wang, Jingtan, et al.
Published: (2024)
by: Wang, Jingtan, et al.
Published: (2024)
Similar Items
-
Global-to-Local Support Spectrums for Language Model Explainability
by: Agussurja, Lucas, et al.
Published: (2024) -
Use Your INSTINCT: INSTruction optimization for LLMs usIng Neural bandits Coupled with Transformers
by: Lin, Xiaoqiang, et al.
Published: (2023) -
Source Attribution for Large Language Model-Generated Data
by: Wang, Jingtan, et al.
Published: (2023) -
Dipper: Diversity in Prompts for Producing Large Language Model Ensembles in Reasoning tasks
by: Lau, Gregory Kang Ruey, et al.
Published: (2024) -
Prompt Optimization with EASE? Efficient Ordering-aware Automated Selection of Exemplars
by: Wu, Zhaoxuan, et al.
Published: (2024)