Neural FOXP2 -- Language Specific Neuron Steering for Targeted Language Improvement in LLMs
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Saha, Anusa, Joshi, Tanmay, Jain, Vinija, Chadha, Aman, Das, Amitava |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
SPINAL -- Scaling-law and Preference Integration in Neural Alignment Layers
par: Das, Arion, et autres
Publié: (2026)
par: Das, Arion, et autres
Publié: (2026)
PermaFrost-Attack: Stealth Pretraining Seeding(SPS) for planting Logic Landmines During LLM Training
par: Kumar, Harsh, et autres
Publié: (2026)
par: Kumar, Harsh, et autres
Publié: (2026)
TRACEALIGN -- Tracing the Drift: Attributing Alignment Failures to Training-Time Belief Sources in LLMs
par: Das, Amitava, et autres
Publié: (2025)
par: Das, Amitava, et autres
Publié: (2025)
Are Small Language Models Ready to Compete with Large Language Models for Practical Applications?
par: Sinha, Neelabh, et autres
Publié: (2024)
par: Sinha, Neelabh, et autres
Publié: (2024)
MENTIS: What Belief Changes Under Alignment? Measuring Multi-Scale Latent Torsion in Language Models
par: Saha, Partha Pratim, et autres
Publié: (2026)
par: Saha, Partha Pratim, et autres
Publié: (2026)
On the Relationship between Sentence Analogy Identification and Sentence Structure Encoding in Large Language Models
par: Wijesiriwardene, Thilini, et autres
Publié: (2023)
par: Wijesiriwardene, Thilini, et autres
Publié: (2023)
KnowledgePrompts: Exploring the Abilities of Large Language Models to Solve Proportional Analogies via Knowledge-Enhanced Prompting
par: Wijesiriwardene, Thilini, et autres
Publié: (2024)
par: Wijesiriwardene, Thilini, et autres
Publié: (2024)
Stochastic CHAOS: Why Deterministic Inference Kills, and Distributional Variability Is the Heartbeat of Artifical Cognition
par: Joshi, Tanmay, et autres
Publié: (2026)
par: Joshi, Tanmay, et autres
Publié: (2026)
Born With a Silver Spoon? Investigating Socioeconomic Bias in Large Language Models
par: Singh, Smriti, et autres
Publié: (2024)
par: Singh, Smriti, et autres
Publié: (2024)
SAHOO: Safeguarded Alignment for High-Order Optimization Objectives in Recursive Self-Improvement
par: Sahoo, Subramanyam, et autres
Publié: (2026)
par: Sahoo, Subramanyam, et autres
Publié: (2026)
Guiding Vision-Language Model Selection for Visual Question-Answering Across Tasks, Domains, and Knowledge Types
par: Sinha, Neelabh, et autres
Publié: (2024)
par: Sinha, Neelabh, et autres
Publié: (2024)
Exploring the Frontier of Vision-Language Models: A Survey of Current Methodologies and Future Directions
par: Ghosh, Akash, et autres
Publié: (2024)
par: Ghosh, Akash, et autres
Publié: (2024)
LLMsAgainstHate @ NLU of Devanagari Script Languages 2025: Hate Speech Detection and Target Identification in Devanagari Languages via Parameter Efficient Fine-Tuning of LLMs
par: Sidibomma, Rushendra, et autres
Publié: (2024)
par: Sidibomma, Rushendra, et autres
Publié: (2024)
A Comprehensive Survey of Accelerated Generation Techniques in Large Language Models
par: Khoshnoodi, Mahsa, et autres
Publié: (2024)
par: Khoshnoodi, Mahsa, et autres
Publié: (2024)
Multilingual State Space Models for Structured Question Answering in Indic Languages
par: Vats, Arpita, et autres
Publié: (2025)
par: Vats, Arpita, et autres
Publié: (2025)
Assessing LLM Reliability on Temporally Recent Open-Domain Questions
par: Krishnappa, Pushwitha, et autres
Publié: (2026)
par: Krishnappa, Pushwitha, et autres
Publié: (2026)
A Systematic Survey of Prompt Engineering in Large Language Models: Techniques and Applications
par: Sahoo, Pranab, et autres
Publié: (2024)
par: Sahoo, Pranab, et autres
Publié: (2024)
How Culturally Aware are Vision-Language Models?
par: Burda-Lassen, Olena, et autres
Publié: (2024)
par: Burda-Lassen, Olena, et autres
Publié: (2024)
AlignMerge - Alignment-Preserving Large Language Model Merging via Fisher-Guided Geometric Constraints
par: Roy, Aniruddha, et autres
Publié: (2025)
par: Roy, Aniruddha, et autres
Publié: (2025)
Evidence-backed Fact Checking using RAG and Few-Shot In-Context Learning with LLMs
par: Singhal, Ronit, et autres
Publié: (2024)
par: Singhal, Ronit, et autres
Publié: (2024)
IndicMMLU-Pro: Benchmarking Indic Large Language Models on Multi-Task Language Understanding
par: KJ, Sankalp, et autres
Publié: (2025)
par: KJ, Sankalp, et autres
Publié: (2025)
AlignGuard-LoRA: Alignment-Preserving Fine-Tuning via Fisher-Guided Decomposition and Riemannian-Geodesic Collision Regularization
par: Das, Amitava, et autres
Publié: (2025)
par: Das, Amitava, et autres
Publié: (2025)
When Shallow Wins: Silent Failures and the Depth-Accuracy Paradox in Latent Reasoning
par: Sahoo, Subramanyam, et autres
Publié: (2026)
par: Sahoo, Subramanyam, et autres
Publié: (2026)
The Reasoning Trap -- Logical Reasoning as a Mechanistic Pathway to Situational Awareness
par: Sahoo, Subramanyam, et autres
Publié: (2026)
par: Sahoo, Subramanyam, et autres
Publié: (2026)
A Comprehensive Survey of Hallucination in Large Language, Image, Video and Audio Foundation Models
par: Sahoo, Pranab, et autres
Publié: (2024)
par: Sahoo, Pranab, et autres
Publié: (2024)
Overview of Factify5WQA: Fact Verification through 5W Question-Answering
par: Suresh, Suryavardan, et autres
Publié: (2024)
par: Suresh, Suryavardan, et autres
Publié: (2024)
Hierarchical Prompting Taxonomy: A Universal Evaluation Framework for Large Language Models Aligned with Human Cognitive Principles
par: Budagam, Devichand, et autres
Publié: (2024)
par: Budagam, Devichand, et autres
Publié: (2024)
Decoding the Diversity: A Review of the Indic AI Research Landscape
par: KJ, Sankalp, et autres
Publié: (2024)
par: KJ, Sankalp, et autres
Publié: (2024)
MAAT: Multi-phase Adapter-Aware Targeted Unlearning
par: Yagnik, Suryash, et autres
Publié: (2026)
par: Yagnik, Suryash, et autres
Publié: (2026)
CONSCIENTIA: Can LLM Agents Learn to Strategize? Emergent Deception and Trust in a Multi-Agent NYC Simulation
par: Sinha, Aarush, et autres
Publié: (2026)
par: Sinha, Aarush, et autres
Publié: (2026)
ECLIPTICA -- A Framework for Switchable LLM Alignment via CITA - Contrastive Instruction-Tuned Alignment
par: Wanaskar, Kapil, et autres
Publié: (2026)
par: Wanaskar, Kapil, et autres
Publié: (2026)
Parameter Efficient Fine Tuning: A Comprehensive Analysis Across Applications
par: Balne, Charith Chandra Sai, et autres
Publié: (2024)
par: Balne, Charith Chandra Sai, et autres
Publié: (2024)
Cause and Effect: Can Large Language Models Truly Understand Causality?
par: Ashwani, Swagata, et autres
Publié: (2024)
par: Ashwani, Swagata, et autres
Publié: (2024)
The What, Why, and How of Context Length Extension Techniques in Large Language Models -- A Detailed Survey
par: Pawar, Saurav, et autres
Publié: (2024)
par: Pawar, Saurav, et autres
Publié: (2024)
Generative Data Augmentation using LLMs improves Distributional Robustness in Question Answering
par: Chowdhury, Arijit Ghosh, et autres
Publié: (2023)
par: Chowdhury, Arijit Ghosh, et autres
Publié: (2023)
Reasoning or Rhetoric? An Empirical Analysis of Moral Reasoning Explanations in Large Language Models
par: Kasat, Aryan, et autres
Publié: (2026)
par: Kasat, Aryan, et autres
Publié: (2026)
Exploring the Impact of Large Language Models on Recommender Systems: An Extensive Review
par: Vats, Arpita, et autres
Publié: (2024)
par: Vats, Arpita, et autres
Publié: (2024)
DPO Kernels: A Semantically-Aware, Kernel-Enhanced, and Divergence-Rich Paradigm for Direct Preference Optimization
par: Das, Amitava, et autres
Publié: (2025)
par: Das, Amitava, et autres
Publié: (2025)
QuickSilver -- Speeding up LLM Inference through Dynamic Token Halting, KV Skipping, Contextual Token Fusion, and Adaptive Matryoshka Quantization
par: Khanna, Danush, et autres
Publié: (2025)
par: Khanna, Danush, et autres
Publié: (2025)
A Comprehensive Survey of Hallucination Mitigation Techniques in Large Language Models
par: Tonmoy, S. M Towhidul Islam, et autres
Publié: (2024)
par: Tonmoy, S. M Towhidul Islam, et autres
Publié: (2024)
Documents similaires
-
SPINAL -- Scaling-law and Preference Integration in Neural Alignment Layers
par: Das, Arion, et autres
Publié: (2026) -
PermaFrost-Attack: Stealth Pretraining Seeding(SPS) for planting Logic Landmines During LLM Training
par: Kumar, Harsh, et autres
Publié: (2026) -
TRACEALIGN -- Tracing the Drift: Attributing Alignment Failures to Training-Time Belief Sources in LLMs
par: Das, Amitava, et autres
Publié: (2025) -
Are Small Language Models Ready to Compete with Large Language Models for Practical Applications?
par: Sinha, Neelabh, et autres
Publié: (2024) -
MENTIS: What Belief Changes Under Alignment? Measuring Multi-Scale Latent Torsion in Language Models
par: Saha, Partha Pratim, et autres
Publié: (2026)