Understanding How CodeLLMs (Mis)Predict Types with Activation Steering
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lucchetti, Francesca, Guha, Arjun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
More Than a Score: Probing the Impact of Prompt Specificity on LLM Code Generation
von: Zi, Yangtian, et al.
Veröffentlicht: (2025)
von: Zi, Yangtian, et al.
Veröffentlicht: (2025)
Knowledge Transfer from High-Resource to Low-Resource Programming Languages for Code LLMs
von: Cassano, Federico, et al.
Veröffentlicht: (2023)
von: Cassano, Federico, et al.
Veröffentlicht: (2023)
Localizing Malicious Outputs from CodeLLM
von: Borana, Mayukh, et al.
Veröffentlicht: (2025)
von: Borana, Mayukh, et al.
Veröffentlicht: (2025)
Is Next Token Prediction Sufficient for GPT? Exploration on Code Logic Comprehension
von: Qi, Mengnan, et al.
Veröffentlicht: (2024)
von: Qi, Mengnan, et al.
Veröffentlicht: (2024)
Agnostics: Learning to Code in Any Programming Language via Reinforcement with a Universal Learning Environment
von: Boruch-Gruszecki, Aleksander, et al.
Veröffentlicht: (2025)
von: Boruch-Gruszecki, Aleksander, et al.
Veröffentlicht: (2025)
Guiding Giants: Lightweight Controllers for Weighted Activation Steering in LLMs
von: Hegazy, Amr, et al.
Veröffentlicht: (2025)
von: Hegazy, Amr, et al.
Veröffentlicht: (2025)
Steering MoE LLMs via Expert (De)Activation
von: Fayyaz, Mohsen, et al.
Veröffentlicht: (2025)
von: Fayyaz, Mohsen, et al.
Veröffentlicht: (2025)
Self-Infilling Code Generation
von: Zheng, Lin, et al.
Veröffentlicht: (2023)
von: Zheng, Lin, et al.
Veröffentlicht: (2023)
Evaluating and Aligning CodeLLMs on Human Preference
von: Yang, Jian, et al.
Veröffentlicht: (2024)
von: Yang, Jian, et al.
Veröffentlicht: (2024)
Spurious Rewards Paradox: Mechanistically Understanding How RLVR Activates Memorization Shortcuts in LLMs
von: Yan, Lecheng, et al.
Veröffentlicht: (2026)
von: Yan, Lecheng, et al.
Veröffentlicht: (2026)
Uncovering Gaps in How Humans and LLMs Interpret Subjective Language
von: Jones, Erik, et al.
Veröffentlicht: (2025)
von: Jones, Erik, et al.
Veröffentlicht: (2025)
SteerConf: Steering LLMs for Confidence Elicitation
von: Zhou, Ziang, et al.
Veröffentlicht: (2025)
von: Zhou, Ziang, et al.
Veröffentlicht: (2025)
Extracting Unlearned Information from LLMs with Activation Steering
von: Seyitoğlu, Atakan, et al.
Veröffentlicht: (2024)
von: Seyitoğlu, Atakan, et al.
Veröffentlicht: (2024)
Lita: Light Agent Uncovers the Agentic Coding Capabilities of LLMs
von: Dai, Hankun, et al.
Veröffentlicht: (2025)
von: Dai, Hankun, et al.
Veröffentlicht: (2025)
CodeCloak: A Method for Evaluating and Mitigating Code Leakage by LLM Code Assistants
von: Noah, Amit Finkman, et al.
Veröffentlicht: (2024)
von: Noah, Amit Finkman, et al.
Veröffentlicht: (2024)
Type-Constrained Code Generation with Language Models
von: Mündler, Niels, et al.
Veröffentlicht: (2025)
von: Mündler, Niels, et al.
Veröffentlicht: (2025)
Steering Language Models With Activation Engineering
von: Turner, Alexander Matt, et al.
Veröffentlicht: (2023)
von: Turner, Alexander Matt, et al.
Veröffentlicht: (2023)
Beyond Steering Vector: Flow-based Activation Steering for Inference-Time Intervention
von: Jin, Zehao, et al.
Veröffentlicht: (2026)
von: Jin, Zehao, et al.
Veröffentlicht: (2026)
Towards Understanding Steering Strength
von: Taimeskhanov, Magamed, et al.
Veröffentlicht: (2026)
von: Taimeskhanov, Magamed, et al.
Veröffentlicht: (2026)
Can LLMs Compress (and Decompress)? Evaluating Code Understanding and Execution via Invertibility
von: Maveli, Nickil, et al.
Veröffentlicht: (2026)
von: Maveli, Nickil, et al.
Veröffentlicht: (2026)
Few-shot Personalization of LLMs with Mis-aligned Responses
von: Kim, Jaehyung, et al.
Veröffentlicht: (2024)
von: Kim, Jaehyung, et al.
Veröffentlicht: (2024)
Top Leaderboard Ranking = Top Coding Proficiency, Always? EvoEval: Evolving Coding Benchmarks via LLM
von: Xia, Chunqiu Steven, et al.
Veröffentlicht: (2024)
von: Xia, Chunqiu Steven, et al.
Veröffentlicht: (2024)
Neural Models for Source Code Synthesis and Completion
von: Niyogi, Mitodru
Veröffentlicht: (2024)
von: Niyogi, Mitodru
Veröffentlicht: (2024)
Code Simulation Challenges for Large Language Models
von: La Malfa, Emanuele, et al.
Veröffentlicht: (2024)
von: La Malfa, Emanuele, et al.
Veröffentlicht: (2024)
HyperSteer: Activation Steering at Scale with Hypernetworks
von: Sun, Jiuding, et al.
Veröffentlicht: (2025)
von: Sun, Jiuding, et al.
Veröffentlicht: (2025)
A Multi-Perspective Architecture for Semantic Code Search
von: Haldar, Rajarshi, et al.
Veröffentlicht: (2020)
von: Haldar, Rajarshi, et al.
Veröffentlicht: (2020)
LILO: Learning Interpretable Libraries by Compressing and Documenting Code
von: Grand, Gabriel, et al.
Veröffentlicht: (2023)
von: Grand, Gabriel, et al.
Veröffentlicht: (2023)
ROAST: Rollout-based On-distribution Activation Steering Technique
von: Su, Xuanbo, et al.
Veröffentlicht: (2026)
von: Su, Xuanbo, et al.
Veröffentlicht: (2026)
Substance Beats Style: Why Beginning Students Fail to Code with LLMs
von: Lucchetti, Francesca, et al.
Veröffentlicht: (2024)
von: Lucchetti, Francesca, et al.
Veröffentlicht: (2024)
Quokka: Accelerating Program Verification with LLMs via Invariant Synthesis
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
Enhancing the Code Reasoning Capabilities of LLMs via Consistency-based Reinforcement Learning
von: Qin, Zhanyue, et al.
Veröffentlicht: (2026)
von: Qin, Zhanyue, et al.
Veröffentlicht: (2026)
NExT: Teaching Large Language Models to Reason about Code Execution
von: Ni, Ansong, et al.
Veröffentlicht: (2024)
von: Ni, Ansong, et al.
Veröffentlicht: (2024)
A Unified Understanding and Evaluation of Steering Methods
von: Im, Shawn, et al.
Veröffentlicht: (2025)
von: Im, Shawn, et al.
Veröffentlicht: (2025)
Steer Like the LLM: Activation Steering that Mimics Prompting
von: Heyman, Geert, et al.
Veröffentlicht: (2026)
von: Heyman, Geert, et al.
Veröffentlicht: (2026)
CodeARC: Benchmarking Reasoning Capabilities of LLM Agents for Inductive Program Synthesis
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
Spherical Steering: Geometry-Aware Activation Rotation for Language Models
von: You, Zejia, et al.
Veröffentlicht: (2026)
von: You, Zejia, et al.
Veröffentlicht: (2026)
\texttt{ReMind}: Understanding Deductive Code Reasoning in LLMs
von: Gao, Jun, et al.
Veröffentlicht: (2025)
von: Gao, Jun, et al.
Veröffentlicht: (2025)
SwiftEval: Developing a Language-Specific Benchmark for LLM-generated Code Evaluation
von: Petrukha, Ivan, et al.
Veröffentlicht: (2025)
von: Petrukha, Ivan, et al.
Veröffentlicht: (2025)
Bridging the Knowledge Void: Inference-time Acquisition of Unfamiliar Programming Languages for Coding Tasks
von: Shen, Chen, et al.
Veröffentlicht: (2026)
von: Shen, Chen, et al.
Veröffentlicht: (2026)
Improving LLM Code Reasoning via Semantic Equivalence Self-Play with Formal Verification
von: Barone, Antonio Valerio Miceli, et al.
Veröffentlicht: (2026)
von: Barone, Antonio Valerio Miceli, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
More Than a Score: Probing the Impact of Prompt Specificity on LLM Code Generation
von: Zi, Yangtian, et al.
Veröffentlicht: (2025) -
Knowledge Transfer from High-Resource to Low-Resource Programming Languages for Code LLMs
von: Cassano, Federico, et al.
Veröffentlicht: (2023) -
Localizing Malicious Outputs from CodeLLM
von: Borana, Mayukh, et al.
Veröffentlicht: (2025) -
Is Next Token Prediction Sufficient for GPT? Exploration on Code Logic Comprehension
von: Qi, Mengnan, et al.
Veröffentlicht: (2024) -
Agnostics: Learning to Code in Any Programming Language via Reinforcement with a Universal Learning Environment
von: Boruch-Gruszecki, Aleksander, et al.
Veröffentlicht: (2025)