Demonstrating Mutual Reinforcement Effect through Information Flow
Fuente:
arXiv
Saved in:
| Main Authors: | Gan, Chengguang, He, Xuzheng, Zhang, Qinghao, Mori, Tatsunori |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Empirical Study of Mutual Reinforcement Effect and Application in Few-shot Text Classification Tasks via Prompt
by: Gan, Chengguang, et al.
Published: (2024)
by: Gan, Chengguang, et al.
Published: (2024)
Application of LLM Agents in Recruitment: A Novel Framework for Resume Screening
by: Gan, Chengguang, et al.
Published: (2024)
by: Gan, Chengguang, et al.
Published: (2024)
M-MRE: Extending the Mutual Reinforcement Effect to Multimodal Information Extraction
by: Gan, Chengguang, et al.
Published: (2025)
by: Gan, Chengguang, et al.
Published: (2025)
A Multilingual Dataset and Empirical Validation for the Mutual Reinforcement Effect in Information Extraction
by: Gan, Chengguang, et al.
Published: (2024)
by: Gan, Chengguang, et al.
Published: (2024)
GuideWeb: A Benchmark for Automatic In-App Guide Generation on Real-World Web UIs
by: Gan, Chengguang, et al.
Published: (2026)
by: Gan, Chengguang, et al.
Published: (2026)
How Do Humans Write Code? Large Models Do It the Same Way Too
by: Li, Long, et al.
Published: (2024)
by: Li, Long, et al.
Published: (2024)
MIRL: Mutual Information-Guided Reinforcement Learning for Vision-Language Models
by: Zhang, Yin, et al.
Published: (2026)
by: Zhang, Yin, et al.
Published: (2026)
MI-PRUN: Optimize Large Language Model Pruning via Mutual Information
by: Zhang, Hao, et al.
Published: (2026)
by: Zhang, Hao, et al.
Published: (2026)
Rethinking the Understanding Ability across LLMs through Mutual Information
by: Wang, Shaojie, et al.
Published: (2025)
by: Wang, Shaojie, et al.
Published: (2025)
Utilizing and Calibrating Hindsight Process Rewards via Reinforcement with Mutual Information Self-Evaluation
by: Yao, Jiashu, et al.
Published: (2026)
by: Yao, Jiashu, et al.
Published: (2026)
Mutual Information-based Representations Disentanglement for Unaligned Multimodal Language Sequences
by: Qian, Fan, et al.
Published: (2024)
by: Qian, Fan, et al.
Published: (2024)
Mutual Enhancement of Large Language and Reinforcement Learning Models through Bi-Directional Feedback Mechanisms: A Planning Case Study
by: Gu, Shangding
Published: (2024)
by: Gu, Shangding
Published: (2024)
The Extractive-Abstractive Spectrum: Uncovering Verifiability Trade-offs in LLM Generations
by: Worledge, Theodora, et al.
Published: (2024)
by: Worledge, Theodora, et al.
Published: (2024)
FineCops-Ref: A new Dataset and Task for Fine-Grained Compositional Referring Expression Comprehension
by: Liu, Junzhuo, et al.
Published: (2024)
by: Liu, Junzhuo, et al.
Published: (2024)
Diver: Large Language Model Decoding with Span-Level Mutual Information Verification
by: Lu, Jinliang, et al.
Published: (2024)
by: Lu, Jinliang, et al.
Published: (2024)
Experience Sharing in Mutual Reinforcement Learning for Heterogeneous Language Models
by: Liu, Xiaoze, et al.
Published: (2026)
by: Liu, Xiaoze, et al.
Published: (2026)
StableMask: Refining Causal Masking in Decoder-only Transformer
by: Yin, Qingyu, et al.
Published: (2024)
by: Yin, Qingyu, et al.
Published: (2024)
Language Models with Conformal Factuality Guarantees
by: Mohri, Christopher, et al.
Published: (2024)
by: Mohri, Christopher, et al.
Published: (2024)
Aligning Language Models Using Follow-up Likelihood as Reward Signal
by: Zhang, Chen, et al.
Published: (2024)
by: Zhang, Chen, et al.
Published: (2024)
TS-Align: A Teacher-Student Collaborative Framework for Scalable Iterative Finetuning of Large Language Models
by: Zhang, Chen, et al.
Published: (2024)
by: Zhang, Chen, et al.
Published: (2024)
Countering Reward Over-optimization in LLM with Demonstration-Guided Reinforcement Learning
by: Rita, Mathieu, et al.
Published: (2024)
by: Rita, Mathieu, et al.
Published: (2024)
Word Sense Induction with Hierarchical Clustering and Mutual Information Maximization
by: Abdine, Hadi, et al.
Published: (2022)
by: Abdine, Hadi, et al.
Published: (2022)
Teaching Language Models to Self-Improve through Interactive Demonstrations
by: Yu, Xiao, et al.
Published: (2023)
by: Yu, Xiao, et al.
Published: (2023)
Demonstration Selection for In-Context Learning via Reinforcement Learning
by: Wang, Xubin, et al.
Published: (2024)
by: Wang, Xubin, et al.
Published: (2024)
FlowCompile: An Optimizing Compiler for Structured LLM Workflows
by: Li, Junyan, et al.
Published: (2026)
by: Li, Junyan, et al.
Published: (2026)
Benchmarking Distributional Alignment of Large Language Models
by: Meister, Nicole, et al.
Published: (2024)
by: Meister, Nicole, et al.
Published: (2024)
Understanding Finetuning for Factual Knowledge Extraction
by: Ghosal, Gaurav, et al.
Published: (2024)
by: Ghosal, Gaurav, et al.
Published: (2024)
Improving Pretraining Data Using Perplexity Correlations
by: Thrush, Tristan, et al.
Published: (2024)
by: Thrush, Tristan, et al.
Published: (2024)
An Exploration of Self-Supervised Mutual Information Alignment for Multi-Task Settings
by: Govande, Soham V.
Published: (2024)
by: Govande, Soham V.
Published: (2024)
Interpreting and Steering LLMs with Mutual Information-based Explanations on Sparse Autoencoders
by: Wu, Xuansheng, et al.
Published: (2025)
by: Wu, Xuansheng, et al.
Published: (2025)
VOLTA: Improving Generative Diversity by Variational Mutual Information Maximizing Autoencoder
by: Ma, Yueen, et al.
Published: (2023)
by: Ma, Yueen, et al.
Published: (2023)
Demonstrations of Integrity Attacks in Multi-Agent Systems
by: Zheng, Can, et al.
Published: (2025)
by: Zheng, Can, et al.
Published: (2025)
Efficient RLVR Training via Weighted Mutual Information Data Selection
by: Zhou, Xinyu, et al.
Published: (2026)
by: Zhou, Xinyu, et al.
Published: (2026)
Pointwise Mutual Information as a Performance Gauge for Retrieval-Augmented Generation
by: Liu, Tianyu, et al.
Published: (2024)
by: Liu, Tianyu, et al.
Published: (2024)
Learning to Maximize Mutual Information for Chain-of-Thought Distillation
by: Chen, Xin, et al.
Published: (2024)
by: Chen, Xin, et al.
Published: (2024)
Intrinsic Mutual Information as a Modulator for Preference Optimization
by: Liao, Peng, et al.
Published: (2026)
by: Liao, Peng, et al.
Published: (2026)
Comparable Demonstrations are Important in In-Context Learning: A Novel Perspective on Demonstration Selection
by: Fan, Caoyun, et al.
Published: (2023)
by: Fan, Caoyun, et al.
Published: (2023)
CARMA: Enhanced Compositionality in LLMs via Advanced Regularisation and Mutual Information Alignment
by: Aljaafari, Nura, et al.
Published: (2025)
by: Aljaafari, Nura, et al.
Published: (2025)
MAIN: Mutual Alignment Is Necessary for instruction tuning
by: Yang, Fanyi, et al.
Published: (2025)
by: Yang, Fanyi, et al.
Published: (2025)
Large Language Models are Demonstration Pre-Selectors for Themselves
by: Jin, Jiarui, et al.
Published: (2025)
by: Jin, Jiarui, et al.
Published: (2025)
Similar Items
-
Empirical Study of Mutual Reinforcement Effect and Application in Few-shot Text Classification Tasks via Prompt
by: Gan, Chengguang, et al.
Published: (2024) -
Application of LLM Agents in Recruitment: A Novel Framework for Resume Screening
by: Gan, Chengguang, et al.
Published: (2024) -
M-MRE: Extending the Mutual Reinforcement Effect to Multimodal Information Extraction
by: Gan, Chengguang, et al.
Published: (2025) -
A Multilingual Dataset and Empirical Validation for the Mutual Reinforcement Effect in Information Extraction
by: Gan, Chengguang, et al.
Published: (2024) -
GuideWeb: A Benchmark for Automatic In-App Guide Generation on Real-World Web UIs
by: Gan, Chengguang, et al.
Published: (2026)