Towards Generating Informative Textual Description for Neurons in Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Mondal, Shrayani, Garodia, Rishabh, Qureshi, Arbaaz, Lee, Taesung, Park, Youngja |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Empowering Large Language Models for Textual Data Augmentation
por: Li, Yichuan, et al.
Publicado: (2024)
por: Li, Yichuan, et al.
Publicado: (2024)
Neuro-RIT: Neuron-Guided Instruction Tuning for Robust Retrieval-Augmented Language Model
por: Kim, Jaemin, et al.
Publicado: (2026)
por: Kim, Jaemin, et al.
Publicado: (2026)
Towards Aligning Language Models with Textual Feedback
por: Lloret, Saüc Abadal, et al.
Publicado: (2024)
por: Lloret, Saüc Abadal, et al.
Publicado: (2024)
CoSy: Evaluating Textual Explanations of Neurons
por: Kopf, Laura, et al.
Publicado: (2024)
por: Kopf, Laura, et al.
Publicado: (2024)
Cyber-Attack Technique Classification Using Two-Stage Trained Large Language Models
por: You, Weiqiu, et al.
Publicado: (2024)
por: You, Weiqiu, et al.
Publicado: (2024)
Motion Generation from Fine-grained Textual Descriptions
por: Li, Kunhang, et al.
Publicado: (2024)
por: Li, Kunhang, et al.
Publicado: (2024)
Toward a Method to Generate Capability Ontologies from Natural Language Descriptions
por: da Silva, Luis Miguel Vieira, et al.
Publicado: (2024)
por: da Silva, Luis Miguel Vieira, et al.
Publicado: (2024)
Language-specific Neurons Do Not Facilitate Cross-Lingual Transfer
por: Mondal, Soumen Kumar, et al.
Publicado: (2025)
por: Mondal, Soumen Kumar, et al.
Publicado: (2025)
FeRG-LLM : Feature Engineering by Reason Generation Large Language Models
por: Ko, Jeonghyun, et al.
Publicado: (2025)
por: Ko, Jeonghyun, et al.
Publicado: (2025)
Language Models are Homer Simpson! Safety Re-Alignment of Fine-tuned Language Models through Task Arithmetic
por: Bhardwaj, Rishabh, et al.
Publicado: (2024)
por: Bhardwaj, Rishabh, et al.
Publicado: (2024)
CritiqueLLM: Towards an Informative Critique Generation Model for Evaluation of Large Language Model Generation
por: Ke, Pei, et al.
Publicado: (2023)
por: Ke, Pei, et al.
Publicado: (2023)
Textual Aesthetics in Large Language Models
por: Jiang, Lingjie, et al.
Publicado: (2024)
por: Jiang, Lingjie, et al.
Publicado: (2024)
Towards LLM-based Autograding for Short Textual Answers
por: Schneider, Johannes, et al.
Publicado: (2023)
por: Schneider, Johannes, et al.
Publicado: (2023)
Enhancing Traffic Prediction with Textual Data Using Large Language Models
por: Huang, Xiannan
Publicado: (2024)
por: Huang, Xiannan
Publicado: (2024)
Enhancing Cache-Augmented Generation (CAG) with Adaptive Contextual Compression for Scalable Knowledge Integration
por: Agrawal, Rishabh, et al.
Publicado: (2025)
por: Agrawal, Rishabh, et al.
Publicado: (2025)
Thunder-Tok: Minimizing Tokens per Word in Tokenizing Korean Texts for Generative Language Models
por: Cho, Gyeongje, et al.
Publicado: (2025)
por: Cho, Gyeongje, et al.
Publicado: (2025)
Black-box Model Ensembling for Textual and Visual Question Answering via Information Fusion
por: Xia, Yuxi, et al.
Publicado: (2024)
por: Xia, Yuxi, et al.
Publicado: (2024)
Conceptual Contrastive Edits in Textual and Vision-Language Retrieval
por: Lymperaiou, Maria, et al.
Publicado: (2025)
por: Lymperaiou, Maria, et al.
Publicado: (2025)
Integrating Large Language Models and Knowledge Graphs for Extraction and Validation of Textual Test Data
por: De Santis, Antonio, et al.
Publicado: (2024)
por: De Santis, Antonio, et al.
Publicado: (2024)
Does "Reasoning" with Large Language Models Improve Recognizing, Generating, and Reframing Unhelpful Thoughts?
por: Qi, Yilin, et al.
Publicado: (2025)
por: Qi, Yilin, et al.
Publicado: (2025)
Potential and Perils of Large Language Models as Judges of Unstructured Textual Data
por: Bedemariam, Rewina, et al.
Publicado: (2025)
por: Bedemariam, Rewina, et al.
Publicado: (2025)
CRANE: Causal Relevance Analysis of Language-Specific Neurons in Multilingual Large Language Models
por: Le, Yifan, et al.
Publicado: (2026)
por: Le, Yifan, et al.
Publicado: (2026)
Weighted Multi-Prompt Learning with Description-free Large Language Model Distillation
por: Lee, Sua, et al.
Publicado: (2025)
por: Lee, Sua, et al.
Publicado: (2025)
Revisiting Large Language Model Pruning using Neuron Semantic Attribution
por: Ding, Yizhuo, et al.
Publicado: (2025)
por: Ding, Yizhuo, et al.
Publicado: (2025)
Confidence Regulation Neurons in Language Models
por: Stolfo, Alessandro, et al.
Publicado: (2024)
por: Stolfo, Alessandro, et al.
Publicado: (2024)
Prompt-based Depth Pruning of Large Language Models
por: Wee, Juyun, et al.
Publicado: (2025)
por: Wee, Juyun, et al.
Publicado: (2025)
Can LLM-Generated Textual Explanations Enhance Model Classification Performance? An Empirical Study
por: Dhaini, Mahdi, et al.
Publicado: (2025)
por: Dhaini, Mahdi, et al.
Publicado: (2025)
Towards Prompt Generalization: Grammar-aware Cross-Prompt Automated Essay Scoring
por: Do, Heejin, et al.
Publicado: (2025)
por: Do, Heejin, et al.
Publicado: (2025)
TAT-LLM: A Specialized Language Model for Discrete Reasoning over Tabular and Textual Data
por: Zhu, Fengbin, et al.
Publicado: (2024)
por: Zhu, Fengbin, et al.
Publicado: (2024)
Show, Don't Tell: Evaluating Large Language Models Beyond Textual Understanding with ChildPlay
por: de Carvalho, Gonçalo Hora, et al.
Publicado: (2024)
por: de Carvalho, Gonçalo Hora, et al.
Publicado: (2024)
Mitigating Biases for Instruction-following Language Models via Bias Neurons Elimination
por: Yang, Nakyeong, et al.
Publicado: (2023)
por: Yang, Nakyeong, et al.
Publicado: (2023)
Unveiling Factual Recall Behaviors of Large Language Models through Knowledge Neurons
por: Wang, Yifei, et al.
Publicado: (2024)
por: Wang, Yifei, et al.
Publicado: (2024)
Attention Speaks Volumes: Localizing and Mitigating Bias in Language Models
por: Adiga, Rishabh, et al.
Publicado: (2024)
por: Adiga, Rishabh, et al.
Publicado: (2024)
Universal Neurons in GPT2 Language Models
por: Gurnee, Wes, et al.
Publicado: (2024)
por: Gurnee, Wes, et al.
Publicado: (2024)
Towards Explainable Job Title Matching: Leveraging Semantic Textual Relatedness and Knowledge Graphs
por: Zadykian, Vadim, et al.
Publicado: (2025)
por: Zadykian, Vadim, et al.
Publicado: (2025)
REFINE-LM: Mitigating Language Model Stereotypes via Reinforcement Learning
por: Qureshi, Rameez, et al.
Publicado: (2024)
por: Qureshi, Rameez, et al.
Publicado: (2024)
SGM: Safety Glasses for Multimodal Large Language Models via Neuron-Level Detoxification
por: Wang, Hongbo, et al.
Publicado: (2025)
por: Wang, Hongbo, et al.
Publicado: (2025)
ChroKnowledge: Unveiling Chronological Knowledge of Language Models in Multiple Domains
por: Park, Yein, et al.
Publicado: (2024)
por: Park, Yein, et al.
Publicado: (2024)
Towards Fairness Assessment of Dutch Hate Speech Detection
por: Bauer, Julie, et al.
Publicado: (2025)
por: Bauer, Julie, et al.
Publicado: (2025)
SynDec: A Synthesize-then-Decode Approach for Arbitrary Textual Style Transfer via Large Language Models
por: Sun, Han, et al.
Publicado: (2025)
por: Sun, Han, et al.
Publicado: (2025)
Ejemplares similares
-
Empowering Large Language Models for Textual Data Augmentation
por: Li, Yichuan, et al.
Publicado: (2024) -
Neuro-RIT: Neuron-Guided Instruction Tuning for Robust Retrieval-Augmented Language Model
por: Kim, Jaemin, et al.
Publicado: (2026) -
Towards Aligning Language Models with Textual Feedback
por: Lloret, Saüc Abadal, et al.
Publicado: (2024) -
CoSy: Evaluating Textual Explanations of Neurons
por: Kopf, Laura, et al.
Publicado: (2024) -
Cyber-Attack Technique Classification Using Two-Stage Trained Large Language Models
por: You, Weiqiu, et al.
Publicado: (2024)