DiPT: Enhancing LLM reasoning through diversified perspective-taking
Fuente:
arXiv
Guardado en:
| Autores principales: | Just, Hoang Anh, Dabas, Mahavir, Huang, Lifu, Jin, Ming, Jia, Ruoxi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Just Enough Shifts: Mitigating Over-Refusal in Aligned Language Models with Targeted Representation Fine-Tuning
por: Dabas, Mahavir, et al.
Publicado: (2025)
por: Dabas, Mahavir, et al.
Publicado: (2025)
Characterizing Model-Native Skills
por: Kang, Feiyang, et al.
Publicado: (2026)
por: Kang, Feiyang, et al.
Publicado: (2026)
The Signal is in the Steps: Local Scoring for Reasoning Data Selection
por: Just, Hoang Anh, et al.
Publicado: (2025)
por: Just, Hoang Anh, et al.
Publicado: (2025)
Probing Knowledge Holes in Unlearned LLMs
por: Ko, Myeongseob, et al.
Publicado: (2025)
por: Ko, Myeongseob, et al.
Publicado: (2025)
Memory-Induced Tool-Drift in LLM Agents
por: Dabas, Mahavir, et al.
Publicado: (2026)
por: Dabas, Mahavir, et al.
Publicado: (2026)
Get more for less: Principled Data Selection for Warming Up Fine-Tuning in LLMs
por: Kang, Feiyang, et al.
Publicado: (2024)
por: Kang, Feiyang, et al.
Publicado: (2024)
Data-Centric Human Preference with Rationales for Direct Preference Alignment
por: Just, Hoang Anh, et al.
Publicado: (2024)
por: Just, Hoang Anh, et al.
Publicado: (2024)
Adversarial Reward Auditing for Active Detection and Mitigation of Reward Hacking
por: Beigi, Mohammad, et al.
Publicado: (2026)
por: Beigi, Mohammad, et al.
Publicado: (2026)
IR$^3$: Contrastive Inverse Reinforcement Learning for Interpretable Detection and Mitigation of Reward Hacking
por: Beigi, Mohammad, et al.
Publicado: (2026)
por: Beigi, Mohammad, et al.
Publicado: (2026)
Hierarchical clustering that takes advantage of both density-peak and density-connectivity
por: Zhu, Ye, et al.
Publicado: (2018)
por: Zhu, Ye, et al.
Publicado: (2018)
Trust Region Preference Approximation: A simple and stable reinforcement learning algorithm for LLM reasoning
por: Su, Xuerui, et al.
Publicado: (2025)
por: Su, Xuerui, et al.
Publicado: (2025)
Skin-in-the-Game: Decision Making via Multi-Stakeholder Alignment in LLMs
por: Sel, Bilgehan, et al.
Publicado: (2024)
por: Sel, Bilgehan, et al.
Publicado: (2024)
Scout Before You Attend: Sketch-and-Walk Sparse Attention for Efficient LLM Inference
por: Le, Hoang Anh Duy, et al.
Publicado: (2026)
por: Le, Hoang Anh Duy, et al.
Publicado: (2026)
PT$^2$-LLM: Post-Training Ternarization for Large Language Models
por: Yan, Xianglong, et al.
Publicado: (2025)
por: Yan, Xianglong, et al.
Publicado: (2025)
Asymmetric Proximal Policy Optimization: mini-critics boost LLM reasoning
por: Liu, Jiashun, et al.
Publicado: (2025)
por: Liu, Jiashun, et al.
Publicado: (2025)
M$^2$PT: Multimodal Prompt Tuning for Zero-shot Instruction Learning
por: Wang, Taowen, et al.
Publicado: (2024)
por: Wang, Taowen, et al.
Publicado: (2024)
Data Valuation and Selection in a Federated Model Marketplace
por: Li, Wenqian, et al.
Publicado: (2025)
por: Li, Wenqian, et al.
Publicado: (2025)
Efficient Bilevel Optimization for Meta Label Correction in Noisy Label Learning
por: Nguyen, Ba Hoang Anh, et al.
Publicado: (2026)
por: Nguyen, Ba Hoang Anh, et al.
Publicado: (2026)
Boosting Alignment for Post-Unlearning Text-to-Image Generative Models
por: Ko, Myeongseob, et al.
Publicado: (2024)
por: Ko, Myeongseob, et al.
Publicado: (2024)
Artificial Expert Intelligence through PAC-reasoning
por: Shalev-Shwartz, Shai, et al.
Publicado: (2024)
por: Shalev-Shwartz, Shai, et al.
Publicado: (2024)
RePCS: Diagnosing Data Memorization in LLM-Powered Retrieval-Augmented Generation
por: Anh, Le Vu, et al.
Publicado: (2025)
por: Anh, Le Vu, et al.
Publicado: (2025)
Enhancing LLM Planning Capabilities through Intrinsic Self-Critique
por: Bohnet, Bernd, et al.
Publicado: (2025)
por: Bohnet, Bernd, et al.
Publicado: (2025)
Flatness-aware Sequential Learning Generates Resilient Backdoors
por: Pham, Hoang, et al.
Publicado: (2024)
por: Pham, Hoang, et al.
Publicado: (2024)
DiEC: Diffusion Embedded Clustering
por: Hu, Haidong, et al.
Publicado: (2025)
por: Hu, Haidong, et al.
Publicado: (2025)
The cognitive companion: a lightweight parallel monitoring architecture for detecting and recovering from reasoning degradation in LLM agents
por: Khan, Rafflesia, et al.
Publicado: (2026)
por: Khan, Rafflesia, et al.
Publicado: (2026)
Self-rewarding correction for mathematical reasoning
por: Xiong, Wei, et al.
Publicado: (2025)
por: Xiong, Wei, et al.
Publicado: (2025)
DiCoRe: Enhancing Zero-shot Event Detection via Divergent-Convergent LLM Reasoning
por: Parekh, Tanmay, et al.
Publicado: (2025)
por: Parekh, Tanmay, et al.
Publicado: (2025)
Enhancing LLM Steering through Sparse Autoencoder-Based Vector Refinement
por: Wang, Anyi, et al.
Publicado: (2025)
por: Wang, Anyi, et al.
Publicado: (2025)
Inverse Reinforcement Learning with Dynamic Reward Scaling for LLM Alignment
por: Cheng, Ruoxi, et al.
Publicado: (2025)
por: Cheng, Ruoxi, et al.
Publicado: (2025)
CNCast: Leveraging 3D Swin Transformer and DiT for Enhanced Regional Weather Forecasting
por: Liang, Hongli, et al.
Publicado: (2025)
por: Liang, Hongli, et al.
Publicado: (2025)
Data Value in the Age of Scaling: Understanding LLM Scaling Dynamics Under Real-Synthetic Data Mixtures
por: Wang, Haohui, et al.
Publicado: (2025)
por: Wang, Haohui, et al.
Publicado: (2025)
The Graphon Limit Hypothesis: Understanding Neural Network Pruning via Infinite Width Analysis
por: Pham, Hoang, et al.
Publicado: (2025)
por: Pham, Hoang, et al.
Publicado: (2025)
Deliberative Alignment is Deep, but Uncertainty Remains: Inference time safety improvement in reasoning via attribution of unsafe behavior to base model
por: Pathmanathan, Pankayaraj, et al.
Publicado: (2026)
por: Pathmanathan, Pankayaraj, et al.
Publicado: (2026)
Winner-takes-all for Multivariate Probabilistic Time Series Forecasting
por: Cortés, Adrien, et al.
Publicado: (2025)
por: Cortés, Adrien, et al.
Publicado: (2025)
Reinforcement Learning in hyperbolic space for multi-step reasoning
por: Xu, Tao, et al.
Publicado: (2025)
por: Xu, Tao, et al.
Publicado: (2025)
ProvMind: Provenance-grounded reasoning for materials synthesis
por: Zhang, Yiming, et al.
Publicado: (2026)
por: Zhang, Yiming, et al.
Publicado: (2026)
CluCERT: Certifying LLM Robustness via Clustering-Guided Denoising Smoothing
por: Wang, Zixia, et al.
Publicado: (2025)
por: Wang, Zixia, et al.
Publicado: (2025)
FastDiSS: Few-step Match Many-step Diffusion Language Model on Sequence-to-Sequence Generation--Full Version
por: Nguyen-Cong, Dat, et al.
Publicado: (2026)
por: Nguyen-Cong, Dat, et al.
Publicado: (2026)
Adversarial Déjà Vu: Jailbreak Dictionary Learning for Stronger Generalization to Unseen Attacks
por: Dabas, Mahavir, et al.
Publicado: (2025)
por: Dabas, Mahavir, et al.
Publicado: (2025)
Exploring Criteria of Loss Reweighting to Enhance LLM Unlearning
por: Yang, Puning, et al.
Publicado: (2025)
por: Yang, Puning, et al.
Publicado: (2025)
Ejemplares similares
-
Just Enough Shifts: Mitigating Over-Refusal in Aligned Language Models with Targeted Representation Fine-Tuning
por: Dabas, Mahavir, et al.
Publicado: (2025) -
Characterizing Model-Native Skills
por: Kang, Feiyang, et al.
Publicado: (2026) -
The Signal is in the Steps: Local Scoring for Reasoning Data Selection
por: Just, Hoang Anh, et al.
Publicado: (2025) -
Probing Knowledge Holes in Unlearned LLMs
por: Ko, Myeongseob, et al.
Publicado: (2025) -
Memory-Induced Tool-Drift in LLM Agents
por: Dabas, Mahavir, et al.
Publicado: (2026)