An Exploration of Self-Supervised Mutual Information Alignment for Multi-Task Settings
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Govande, Soham V. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Self-Supervised Alignment with Mutual Information: Learning to Follow Principles without Preference Labels
von: Fränken, Jan-Philipp, et al.
Veröffentlicht: (2024)
von: Fränken, Jan-Philipp, et al.
Veröffentlicht: (2024)
Chipmunk: Training-Free Acceleration of Diffusion Transformers with Dynamic Column-Sparse Deltas
von: Silveria, Austin, et al.
Veröffentlicht: (2025)
von: Silveria, Austin, et al.
Veröffentlicht: (2025)
CARMA: Enhanced Compositionality in LLMs via Advanced Regularisation and Mutual Information Alignment
von: Aljaafari, Nura, et al.
Veröffentlicht: (2025)
von: Aljaafari, Nura, et al.
Veröffentlicht: (2025)
An Exploration of Mamba for Speech Self-Supervised Models
von: Lin, Tzu-Quan, et al.
Veröffentlicht: (2025)
von: Lin, Tzu-Quan, et al.
Veröffentlicht: (2025)
MAIN: Mutual Alignment Is Necessary for instruction tuning
von: Yang, Fanyi, et al.
Veröffentlicht: (2025)
von: Yang, Fanyi, et al.
Veröffentlicht: (2025)
$\texttt{COSMIC}$: Mutual Information for Task-Agnostic Summarization Evaluation
von: Darrin, Maxime, et al.
Veröffentlicht: (2024)
von: Darrin, Maxime, et al.
Veröffentlicht: (2024)
Shared and Private Information Learning in Multimodal Sentiment Analysis with Deep Modal Alignment and Self-supervised Multi-Task Learning
von: Lai, Songning, et al.
Veröffentlicht: (2023)
von: Lai, Songning, et al.
Veröffentlicht: (2023)
Learning Concepts, Not Tokens: Self-Supervised Semantic Alignment for Language Models
von: Zhang, Christine, et al.
Veröffentlicht: (2026)
von: Zhang, Christine, et al.
Veröffentlicht: (2026)
Data Alignment for Zero-Shot Concept Generation in Dermatology AI
von: Gadgil, Soham, et al.
Veröffentlicht: (2024)
von: Gadgil, Soham, et al.
Veröffentlicht: (2024)
Self-Supervised Visual Preference Alignment
von: Zhu, Ke, et al.
Veröffentlicht: (2024)
von: Zhu, Ke, et al.
Veröffentlicht: (2024)
Poor-Supervised Evaluation for SuperLLM via Mutual Consistency
von: Yuan, Peiwen, et al.
Veröffentlicht: (2024)
von: Yuan, Peiwen, et al.
Veröffentlicht: (2024)
Language Guided Exploration for RL Agents in Text Environments
von: Golchha, Hitesh, et al.
Veröffentlicht: (2024)
von: Golchha, Hitesh, et al.
Veröffentlicht: (2024)
STENCIL: Submodular Mutual Information Based Weak Supervision for Cold-Start Active Learning
von: Beck, Nathan, et al.
Veröffentlicht: (2024)
von: Beck, Nathan, et al.
Veröffentlicht: (2024)
SelfCite: Self-Supervised Alignment for Context Attribution in Large Language Models
von: Chuang, Yung-Sung, et al.
Veröffentlicht: (2025)
von: Chuang, Yung-Sung, et al.
Veröffentlicht: (2025)
Efficient Multi-Task Inferencing: Model Merging with Gromov-Wasserstein Feature Alignment
von: Fang, Luyang, et al.
Veröffentlicht: (2025)
von: Fang, Luyang, et al.
Veröffentlicht: (2025)
Utilizing and Calibrating Hindsight Process Rewards via Reinforcement with Mutual Information Self-Evaluation
von: Yao, Jiashu, et al.
Veröffentlicht: (2026)
von: Yao, Jiashu, et al.
Veröffentlicht: (2026)
Anti-Self-Distillation for Reasoning RL via Pointwise Mutual Information
von: Shen, Guobin, et al.
Veröffentlicht: (2026)
von: Shen, Guobin, et al.
Veröffentlicht: (2026)
EMAFusion: A Self-Optimizing System for Seamless LLM Selection and Integration
von: Shah, Soham, et al.
Veröffentlicht: (2025)
von: Shah, Soham, et al.
Veröffentlicht: (2025)
Improving the Distributional Alignment of LLMs using Supervision
von: Kambhatla, Gauri, et al.
Veröffentlicht: (2025)
von: Kambhatla, Gauri, et al.
Veröffentlicht: (2025)
Demonstrating Mutual Reinforcement Effect through Information Flow
von: Gan, Chengguang, et al.
Veröffentlicht: (2024)
von: Gan, Chengguang, et al.
Veröffentlicht: (2024)
MAMI: Multi-Attentional Mutual-Information for Long Sequence Neuron Captioning
von: Fauzulhaq, Alfirsa Damasyifa, et al.
Veröffentlicht: (2024)
von: Fauzulhaq, Alfirsa Damasyifa, et al.
Veröffentlicht: (2024)
SafeInfer: Context Adaptive Decoding Time Safety Alignment for Large Language Models
von: Banerjee, Somnath, et al.
Veröffentlicht: (2024)
von: Banerjee, Somnath, et al.
Veröffentlicht: (2024)
Breaking mBad! Supervised Fine-tuning for Cross-Lingual Detoxification
von: Beniwal, Himanshu, et al.
Veröffentlicht: (2025)
von: Beniwal, Himanshu, et al.
Veröffentlicht: (2025)
RISE: Reasoning Enhancement via Iterative Self-Exploration in Multi-hop Question Answering
von: He, Bolei, et al.
Veröffentlicht: (2025)
von: He, Bolei, et al.
Veröffentlicht: (2025)
Word Sense Induction with Hierarchical Clustering and Mutual Information Maximization
von: Abdine, Hadi, et al.
Veröffentlicht: (2022)
von: Abdine, Hadi, et al.
Veröffentlicht: (2022)
Set the Clock: Temporal Alignment of Pretrained Language Models
von: Zhao, Bowen, et al.
Veröffentlicht: (2024)
von: Zhao, Bowen, et al.
Veröffentlicht: (2024)
Dissecting the Ledger: Locating and Suppressing "Liar Circuits" in Financial Large Language Models
von: Mirajkar, Soham
Veröffentlicht: (2025)
von: Mirajkar, Soham
Veröffentlicht: (2025)
Revisiting Self-supervised Learning of Speech Representation from a Mutual Information Perspective
von: Liu, Alexander H., et al.
Veröffentlicht: (2024)
von: Liu, Alexander H., et al.
Veröffentlicht: (2024)
Self-Alignment with Instruction Backtranslation
von: Li, Xian, et al.
Veröffentlicht: (2023)
von: Li, Xian, et al.
Veröffentlicht: (2023)
Empirical Study of Mutual Reinforcement Effect and Application in Few-shot Text Classification Tasks via Prompt
von: Gan, Chengguang, et al.
Veröffentlicht: (2024)
von: Gan, Chengguang, et al.
Veröffentlicht: (2024)
Interpreting and Steering LLMs with Mutual Information-based Explanations on Sparse Autoencoders
von: Wu, Xuansheng, et al.
Veröffentlicht: (2025)
von: Wu, Xuansheng, et al.
Veröffentlicht: (2025)
VOLTA: Improving Generative Diversity by Variational Mutual Information Maximizing Autoencoder
von: Ma, Yueen, et al.
Veröffentlicht: (2023)
von: Ma, Yueen, et al.
Veröffentlicht: (2023)
From Stars to Insights: Exploration and Implementation of Unified Sentiment Analysis with Distant Supervision
von: Li, Wenchang, et al.
Veröffentlicht: (2023)
von: Li, Wenchang, et al.
Veröffentlicht: (2023)
Selection-p: Self-Supervised Task-Agnostic Prompt Compression for Faithfulness and Transferability
von: Chung, Tsz Ting, et al.
Veröffentlicht: (2024)
von: Chung, Tsz Ting, et al.
Veröffentlicht: (2024)
Learning to Maximize Mutual Information for Chain-of-Thought Distillation
von: Chen, Xin, et al.
Veröffentlicht: (2024)
von: Chen, Xin, et al.
Veröffentlicht: (2024)
Intrinsic Mutual Information as a Modulator for Preference Optimization
von: Liao, Peng, et al.
Veröffentlicht: (2026)
von: Liao, Peng, et al.
Veröffentlicht: (2026)
Combining Supervised Learning and Reinforcement Learning for Multi-Label Classification Tasks with Partial Labels
von: Jia, Zixia, et al.
Veröffentlicht: (2024)
von: Jia, Zixia, et al.
Veröffentlicht: (2024)
Diver: Large Language Model Decoding with Span-Level Mutual Information Verification
von: Lu, Jinliang, et al.
Veröffentlicht: (2024)
von: Lu, Jinliang, et al.
Veröffentlicht: (2024)
A Multilingual Dataset and Empirical Validation for the Mutual Reinforcement Effect in Information Extraction
von: Gan, Chengguang, et al.
Veröffentlicht: (2024)
von: Gan, Chengguang, et al.
Veröffentlicht: (2024)
STARS: Synchronous Token Alignment for Robust Supervision in Large Language Models
von: Quamar, Mohammad Atif, et al.
Veröffentlicht: (2025)
von: Quamar, Mohammad Atif, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Self-Supervised Alignment with Mutual Information: Learning to Follow Principles without Preference Labels
von: Fränken, Jan-Philipp, et al.
Veröffentlicht: (2024) -
Chipmunk: Training-Free Acceleration of Diffusion Transformers with Dynamic Column-Sparse Deltas
von: Silveria, Austin, et al.
Veröffentlicht: (2025) -
CARMA: Enhanced Compositionality in LLMs via Advanced Regularisation and Mutual Information Alignment
von: Aljaafari, Nura, et al.
Veröffentlicht: (2025) -
An Exploration of Mamba for Speech Self-Supervised Models
von: Lin, Tzu-Quan, et al.
Veröffentlicht: (2025) -
MAIN: Mutual Alignment Is Necessary for instruction tuning
von: Yang, Fanyi, et al.
Veröffentlicht: (2025)