ExecTune: Effective Steering of Black-Box LLMs with Guide Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Lingam, Vijay, Golatkar, Aditya, Pal, Anwesan, Vo, Ben, Sadagopan, Narayanan, Achille, Alessandro, Huan, Jun, Deoras, Anoop, Soatto, Stefano |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Training Data Protection with Compositional Diffusion Models
por: Golatkar, Aditya, et al.
Publicado: (2023)
por: Golatkar, Aditya, et al.
Publicado: (2023)
MURPHY: Feedback-Aware GRPO with Retrospective Credit Assignment for Multi-Turn Code Generation
por: Ekbote, Chanakya, et al.
Publicado: (2025)
por: Ekbote, Chanakya, et al.
Publicado: (2025)
Re-FORC: Adaptive Reward Prediction for Efficient Chain-of-Thought Reasoning
por: Zabounidis, Renos, et al.
Publicado: (2025)
por: Zabounidis, Renos, et al.
Publicado: (2025)
CPR: Retrieval Augmented Generation for Copyright Protection
por: Golatkar, Aditya, et al.
Publicado: (2024)
por: Golatkar, Aditya, et al.
Publicado: (2024)
PICASO: Permutation-Invariant Context Composition with State Space Models
por: Liu, Tian Yu, et al.
Publicado: (2025)
por: Liu, Tian Yu, et al.
Publicado: (2025)
Tangent Transformers for Composition, Privacy and Removal
por: Liu, Tian Yu, et al.
Publicado: (2023)
por: Liu, Tian Yu, et al.
Publicado: (2023)
AI Agents as Universal Task Solvers
por: Achille, Alessandro, et al.
Publicado: (2025)
por: Achille, Alessandro, et al.
Publicado: (2025)
Critical Learning Periods Emerge Even in Deep Linear Networks
por: Kleinman, Michael, et al.
Publicado: (2023)
por: Kleinman, Michael, et al.
Publicado: (2023)
Diffusion Soup: Model Merging for Text-to-Image Diffusion Models
por: Biggs, Benjamin, et al.
Publicado: (2024)
por: Biggs, Benjamin, et al.
Publicado: (2024)
B'MOJO: Hybrid State Space Realizations of Foundation Models with Eidetic and Fading Memory
por: Zancato, Luca, et al.
Publicado: (2024)
por: Zancato, Luca, et al.
Publicado: (2024)
Expansion Span: Combining Fading Memory and Retrieval in Hybrid State Space Models
por: Nunez, Elvis, et al.
Publicado: (2024)
por: Nunez, Elvis, et al.
Publicado: (2024)
Fewer Truncations Improve Language Modeling
por: Ding, Hantian, et al.
Publicado: (2024)
por: Ding, Hantian, et al.
Publicado: (2024)
NeRF-Insert: 3D Local Editing with Multimodal Control Signals
por: Sabat, Benet Oriol, et al.
Publicado: (2024)
por: Sabat, Benet Oriol, et al.
Publicado: (2024)
Compositional Structures in Neural Embedding and Interaction Decompositions
por: Trager, Matthew, et al.
Publicado: (2024)
por: Trager, Matthew, et al.
Publicado: (2024)
e1: Learning Adaptive Control of Reasoning Effort
por: Kleinman, Michael, et al.
Publicado: (2025)
por: Kleinman, Michael, et al.
Publicado: (2025)
Descriminative-Generative Custom Tokens for Vision-Language Models
por: Perera, Pramuditha, et al.
Publicado: (2025)
por: Perera, Pramuditha, et al.
Publicado: (2025)
TerraFormer: Automated Infrastructure-as-Code with LLMs Fine-Tuned via Policy-Guided Verifier Feedback
por: Jana, Prithwish, et al.
Publicado: (2026)
por: Jana, Prithwish, et al.
Publicado: (2026)
AXCEL: Automated eXplainable Consistency Evaluation using LLMs
por: Sreekar, P Aditya, et al.
Publicado: (2024)
por: Sreekar, P Aditya, et al.
Publicado: (2024)
Linear Spaces of Meanings: Compositional Structures in Vision-Language Models
por: Trager, Matthew, et al.
Publicado: (2023)
por: Trager, Matthew, et al.
Publicado: (2023)
ExecVerify: White-Box RL with Verifiable Stepwise Rewards for Code Execution Reasoning
por: Tang, Lingxiao, et al.
Publicado: (2026)
por: Tang, Lingxiao, et al.
Publicado: (2026)
Interpretable Measures of Conceptual Similarity by Complexity-Constrained Descriptive Auto-Encoding
por: Achille, Alessandro, et al.
Publicado: (2024)
por: Achille, Alessandro, et al.
Publicado: (2024)
How to Train Your Advisor: Steering Black-Box LLMs with Advisor Models
por: Asawa, Parth, et al.
Publicado: (2025)
por: Asawa, Parth, et al.
Publicado: (2025)
Multi-Modal Hallucination Control by Visual Information Grounding
por: Favero, Alessandro, et al.
Publicado: (2024)
por: Favero, Alessandro, et al.
Publicado: (2024)
Lightweight reranking for language model generations
por: Jain, Siddhartha, et al.
Publicado: (2023)
por: Jain, Siddhartha, et al.
Publicado: (2023)
Automated Evaluation of Retrieval-Augmented Language Models with Task-Specific Exam Generation
por: Guinet, Gauthier, et al.
Publicado: (2024)
por: Guinet, Gauthier, et al.
Publicado: (2024)
Semantic Volume: Quantifying and Detecting both External and Internal Uncertainty in LLMs
por: Li, Xiaomin, et al.
Publicado: (2025)
por: Li, Xiaomin, et al.
Publicado: (2025)
The Empirical Impact of Data Sanitization on Language Models
por: Pal, Anwesan, et al.
Publicado: (2024)
por: Pal, Anwesan, et al.
Publicado: (2024)
Logic-Scaffolding: Personalized Aspect-Instructed Recommendation Explanation Generation using LLMs
por: Rahdari, Behnam, et al.
Publicado: (2023)
por: Rahdari, Behnam, et al.
Publicado: (2023)
Terminal Steiner tree problem : Complexity and Algorithms
por: S, Jyothish, et al.
Publicado: (2026)
por: S, Jyothish, et al.
Publicado: (2026)
ContextWeaver: Selective and Dependency-Structured Memory Construction for LLM Agents
por: Wu, Yating, et al.
Publicado: (2026)
por: Wu, Yating, et al.
Publicado: (2026)
Lossless Token Sequence Compression via Meta-Tokens
por: Harvill, John, et al.
Publicado: (2025)
por: Harvill, John, et al.
Publicado: (2025)
Learning When to Attend: Conditional Memory Access for Long-Context LLMs
por: Choudhary, Sakshi, et al.
Publicado: (2026)
por: Choudhary, Sakshi, et al.
Publicado: (2026)
Cycles of Thought: Measuring LLM Confidence through Stable Explanations
por: Becker, Evan, et al.
Publicado: (2024)
por: Becker, Evan, et al.
Publicado: (2024)
Robust Planning for Autonomous Driving via Mixed Adversarial Diffusion Predictions
por: Zhao, Albert, et al.
Publicado: (2025)
por: Zhao, Albert, et al.
Publicado: (2025)
Effective and Efficient Jailbreaks of Black-Box LLMs with Cross-Behavior Attacks
por: Gohil, Vasudev
Publicado: (2025)
por: Gohil, Vasudev
Publicado: (2025)
CodeAssistBench (CAB): Dataset & Benchmarking for Multi-turn Chat-Based Code Assistance
por: Kim, Myeongsoo, et al.
Publicado: (2025)
por: Kim, Myeongsoo, et al.
Publicado: (2025)
Multi-IaC-Eval: Benchmarking Cloud Infrastructure as Code Across Multiple Formats
por: Davidson, Sam, et al.
Publicado: (2025)
por: Davidson, Sam, et al.
Publicado: (2025)
ELLA: Efficient Lifelong Learning for Adapters in Large Language Models
por: Biswas, Shristi Das, et al.
Publicado: (2026)
por: Biswas, Shristi Das, et al.
Publicado: (2026)
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
por: Li, Xiaomin, et al.
Publicado: (2025)
por: Li, Xiaomin, et al.
Publicado: (2025)
Realizing an Atomtronic AQUID in a Rotating-Box Potential
por: Görg, Kaspar, et al.
Publicado: (2025)
por: Görg, Kaspar, et al.
Publicado: (2025)
Ejemplares similares
-
Training Data Protection with Compositional Diffusion Models
por: Golatkar, Aditya, et al.
Publicado: (2023) -
MURPHY: Feedback-Aware GRPO with Retrospective Credit Assignment for Multi-Turn Code Generation
por: Ekbote, Chanakya, et al.
Publicado: (2025) -
Re-FORC: Adaptive Reward Prediction for Efficient Chain-of-Thought Reasoning
por: Zabounidis, Renos, et al.
Publicado: (2025) -
CPR: Retrieval Augmented Generation for Copyright Protection
por: Golatkar, Aditya, et al.
Publicado: (2024) -
PICASO: Permutation-Invariant Context Composition with State Space Models
por: Liu, Tian Yu, et al.
Publicado: (2025)