GuidedSampling: Steering LLMs Towards Diverse Candidate Solutions at Inference-Time
Fuente:
arXiv
Salvato in:
| Autori principali: | Handa, Divij, Parmar, Mihir, RRV, Aswin, Uddin, Md Nayem, Palangi, Hamid, Baral, Chitta |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ThinkTuning: Instilling Cognitive Reflections without Distillation
di: RRV, Aswin, et al.
Pubblicazione: (2025)
di: RRV, Aswin, et al.
Pubblicazione: (2025)
When "Competency" in Reasoning Opens the Door to Vulnerability: Jailbreaking LLMs via Novel Complex Ciphers
di: Handa, Divij, et al.
Pubblicazione: (2024)
di: Handa, Divij, et al.
Pubblicazione: (2024)
Mid-Training with Self-Generated Data Improves Reinforcement Learning in Language Models
di: RRV, Aswin, et al.
Pubblicazione: (2026)
di: RRV, Aswin, et al.
Pubblicazione: (2026)
Chaos with Keywords: Exposing Large Language Models Sycophantic Hallucination to Misleading Keywords and Evaluating Defense Strategies
di: RRV, Aswin, et al.
Pubblicazione: (2024)
di: RRV, Aswin, et al.
Pubblicazione: (2024)
Step-by-Step Reasoning to Solve Grid Puzzles: Where do LLMs Falter?
di: Tyagi, Nemika, et al.
Pubblicazione: (2024)
di: Tyagi, Nemika, et al.
Pubblicazione: (2024)
PHANTOM RECALL: When Familiar Puzzles Fool Smart Models
di: Mukhopadhyay, Souradeep, et al.
Pubblicazione: (2025)
di: Mukhopadhyay, Souradeep, et al.
Pubblicazione: (2025)
UnSeenTimeQA: Time-Sensitive Question-Answering Beyond LLMs' Memorization
di: Uddin, Md Nayem, et al.
Pubblicazione: (2024)
di: Uddin, Md Nayem, et al.
Pubblicazione: (2024)
PLAN-TUNING: Post-Training Language Models to Learn Step-by-Step Planning for Complex Problem Solving
di: Parmar, Mihir, et al.
Pubblicazione: (2025)
di: Parmar, Mihir, et al.
Pubblicazione: (2025)
ActionReasoningBench: Reasoning about Actions with and without Ramification Constraints
di: Handa, Divij, et al.
Pubblicazione: (2024)
di: Handa, Divij, et al.
Pubblicazione: (2024)
Don't Blame the Annotator: Bias Already Starts in the Annotation Instructions
di: Parmar, Mihir, et al.
Pubblicazione: (2022)
di: Parmar, Mihir, et al.
Pubblicazione: (2022)
Insights into Alignment: Evaluating DPO and its Variants Across Multiple Tasks
di: Saeidi, Amir, et al.
Pubblicazione: (2024)
di: Saeidi, Amir, et al.
Pubblicazione: (2024)
Triple Preference Optimization: Achieving Better Alignment using a Single Step Optimization
di: Saeidi, Amir, et al.
Pubblicazione: (2024)
di: Saeidi, Amir, et al.
Pubblicazione: (2024)
Multi-LogiEval: Towards Evaluating Multi-Step Logical Reasoning Ability of Large Language Models
di: Patel, Nisarg, et al.
Pubblicazione: (2024)
di: Patel, Nisarg, et al.
Pubblicazione: (2024)
Hypothesis Generation for Materials Discovery and Design Using Goal-Driven and Constraint-Guided LLM Agents
di: Kumbhar, Shrinidhi, et al.
Pubblicazione: (2025)
di: Kumbhar, Shrinidhi, et al.
Pubblicazione: (2025)
Vocabulary Dropout for Curriculum Diversity in LLM Co-Evolution
di: Dineen, Jacob, et al.
Pubblicazione: (2026)
di: Dineen, Jacob, et al.
Pubblicazione: (2026)
From Recall to Forgetting: Benchmarking Long-Term Memory for Personalized Agents
di: Uddin, Md Nayem, et al.
Pubblicazione: (2026)
di: Uddin, Md Nayem, et al.
Pubblicazione: (2026)
LogicBench: Towards Systematic Evaluation of Logical Reasoning Ability of Large Language Models
di: Parmar, Mihir, et al.
Pubblicazione: (2024)
di: Parmar, Mihir, et al.
Pubblicazione: (2024)
ScholarPeer: A Context-Aware Multi-Agent Framework for Automated Peer Review
di: Goyal, Palash, et al.
Pubblicazione: (2026)
di: Goyal, Palash, et al.
Pubblicazione: (2026)
Towards LogiGLUE: A Brief Survey and A Benchmark for Analyzing Logical Reasoning Capabilities of Language Models
di: Luo, Man, et al.
Pubblicazione: (2023)
di: Luo, Man, et al.
Pubblicazione: (2023)
Polymath: A Challenging Multi-modal Mathematical Reasoning Benchmark
di: Gupta, Himanshu, et al.
Pubblicazione: (2024)
di: Gupta, Himanshu, et al.
Pubblicazione: (2024)
Diversity of Thought Improves Reasoning Abilities of LLMs
di: Naik, Ranjita, et al.
Pubblicazione: (2023)
di: Naik, Ranjita, et al.
Pubblicazione: (2023)
PlanGEN: A Multi-Agent Framework for Generating Planning and Reasoning Trajectories for Complex Problem Solving
di: Parmar, Mihir, et al.
Pubblicazione: (2025)
di: Parmar, Mihir, et al.
Pubblicazione: (2025)
VeriGuard: Enhancing LLM Agent Safety via Verified Code Generation
di: Miculicich, Lesly, et al.
Pubblicazione: (2025)
di: Miculicich, Lesly, et al.
Pubblicazione: (2025)
OptAgent: Optimizing Query Rewriting for E-commerce via Multi-Agent Simulation
di: Handa, Divij, et al.
Pubblicazione: (2025)
di: Handa, Divij, et al.
Pubblicazione: (2025)
LLM-Based Multi-Agent Blackboard System for Information Discovery in Data Science
di: Salemi, Alireza, et al.
Pubblicazione: (2025)
di: Salemi, Alireza, et al.
Pubblicazione: (2025)
TFRBench: A Reasoning Benchmark for Evaluating Forecasting Systems
di: Ahamed, Md Atik, et al.
Pubblicazione: (2026)
di: Ahamed, Md Atik, et al.
Pubblicazione: (2026)
Investigating the Shortcomings of LLMs in Step-by-Step Legal Reasoning
di: Mishra, Venkatesh, et al.
Pubblicazione: (2025)
di: Mishra, Venkatesh, et al.
Pubblicazione: (2025)
QA-LIGN: Aligning LLMs through Constitutionally Decomposed QA
di: Dineen, Jacob, et al.
Pubblicazione: (2025)
di: Dineen, Jacob, et al.
Pubblicazione: (2025)
Towards Enhancing Coherence in Extractive Summarization: Dataset and Experiments with LLMs
di: Parmar, Mihir, et al.
Pubblicazione: (2024)
di: Parmar, Mihir, et al.
Pubblicazione: (2024)
ToW: Thoughts of Words Improve Reasoning in Large Language Models
di: Xu, Zhikun, et al.
Pubblicazione: (2024)
di: Xu, Zhikun, et al.
Pubblicazione: (2024)
When Can LLMs Learn to Reason with Weak Supervision?
di: Rahman, Salman, et al.
Pubblicazione: (2026)
di: Rahman, Salman, et al.
Pubblicazione: (2026)
InfSplign: Inference-Time Spatial Alignment of Text-to-Image Diffusion Models
di: Rastegar, Sarah, et al.
Pubblicazione: (2025)
di: Rastegar, Sarah, et al.
Pubblicazione: (2025)
MMTABREAL: Real-World Benchmark for Multimodal Table Understanding
di: Titiya, Prasham, et al.
Pubblicazione: (2025)
di: Titiya, Prasham, et al.
Pubblicazione: (2025)
Fints: Efficient Inference-Time Personalization for LLMs with Fine-Grained Instance-Tailored Steering
di: Du, Kounianhua, et al.
Pubblicazione: (2025)
di: Du, Kounianhua, et al.
Pubblicazione: (2025)
Low-Rank Adaptation of Time Series Foundational Models for Out-of-Domain Modality Forecasting
di: Gupta, Divij, et al.
Pubblicazione: (2024)
di: Gupta, Divij, et al.
Pubblicazione: (2024)
Reasoning-Aware Training for Time Series Forecasting
di: Ahamed, Md Atik, et al.
Pubblicazione: (2026)
di: Ahamed, Md Atik, et al.
Pubblicazione: (2026)
Beyond LoRA: Exploring Efficient Fine-Tuning Techniques for Time Series Foundational Models
di: Gupta, Divij, et al.
Pubblicazione: (2024)
di: Gupta, Divij, et al.
Pubblicazione: (2024)
HEART: Emotionally-Driven Test-Time Scaling of Language Models
di: Pinto, Gabriela, et al.
Pubblicazione: (2025)
di: Pinto, Gabriela, et al.
Pubblicazione: (2025)
GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs
di: Nguyen, Duy, et al.
Pubblicazione: (2025)
di: Nguyen, Duy, et al.
Pubblicazione: (2025)
Latent Reward Steering: An Adaptive Inference-Time Framework that Implicitly Promotes Cognitive Behaviors in Reasoning LLMs
di: Li, Jiakang, et al.
Pubblicazione: (2026)
di: Li, Jiakang, et al.
Pubblicazione: (2026)
Documenti analoghi
-
ThinkTuning: Instilling Cognitive Reflections without Distillation
di: RRV, Aswin, et al.
Pubblicazione: (2025) -
When "Competency" in Reasoning Opens the Door to Vulnerability: Jailbreaking LLMs via Novel Complex Ciphers
di: Handa, Divij, et al.
Pubblicazione: (2024) -
Mid-Training with Self-Generated Data Improves Reinforcement Learning in Language Models
di: RRV, Aswin, et al.
Pubblicazione: (2026) -
Chaos with Keywords: Exposing Large Language Models Sycophantic Hallucination to Misleading Keywords and Evaluating Defense Strategies
di: RRV, Aswin, et al.
Pubblicazione: (2024) -
Step-by-Step Reasoning to Solve Grid Puzzles: Where do LLMs Falter?
di: Tyagi, Nemika, et al.
Pubblicazione: (2024)