Understanding the Effects of Domain Finetuning on LLMs
Fuente:
arXiv
Guardado en:
| Autores principales: | Tanwar, Eshaan, Nathani, Deepak, Wang, William Yang, Chakraborty, Tanmoy |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Multilingual LLMs Struggle to Link Orthography and Semantics in Bilingual Word Processing
por: Tanwar, Eshaan, et al.
Publicado: (2025)
por: Tanwar, Eshaan, et al.
Publicado: (2025)
Language Models can Exploit Cross-Task In-context Learning for Data-Scarce Novel Tasks
por: Chatterjee, Anwoy, et al.
Publicado: (2024)
por: Chatterjee, Anwoy, et al.
Publicado: (2024)
Do You Know About My Nation? Investigating Multilingual Language Models' Cultural Literacy Through Factual Knowledge
por: Tanwar, Eshaan, et al.
Publicado: (2025)
por: Tanwar, Eshaan, et al.
Publicado: (2025)
Understanding the Physics of Key-Value Cache Compression for LLMs through Attention Dynamics
por: Ananthanarayanan, Samhruth, et al.
Publicado: (2026)
por: Ananthanarayanan, Samhruth, et al.
Publicado: (2026)
WildSci: Advancing Scientific Reasoning from In-the-Wild Literature
por: Liu, Tengxiao, et al.
Publicado: (2026)
por: Liu, Tengxiao, et al.
Publicado: (2026)
Position: Enough of Scaling LLMs! Lets Focus on Downscaling
por: Goel, Yash, et al.
Publicado: (2025)
por: Goel, Yash, et al.
Publicado: (2025)
Value-Guided KV Compression for LLMs via Approximated CUR Decomposition
por: Sengupta, Ayan, et al.
Publicado: (2025)
por: Sengupta, Ayan, et al.
Publicado: (2025)
Temporal Referential Consistency: Do LLMs Favor Sequences Over Absolute Time References?
por: Bajpai, Ashutosh, et al.
Publicado: (2025)
por: Bajpai, Ashutosh, et al.
Publicado: (2025)
Can LLMs replace Neil deGrasse Tyson? Evaluating the Reliability of LLMs as Science Communicators
por: Bajpai, Prasoon, et al.
Publicado: (2024)
por: Bajpai, Prasoon, et al.
Publicado: (2024)
Locking Down the Finetuned LLMs Safety
por: Zhu, Minjun, et al.
Publicado: (2024)
por: Zhu, Minjun, et al.
Publicado: (2024)
Knowledge Planning in Large Language Models for Domain-Aligned Counseling Summarization
por: Srivastava, Aseem, et al.
Publicado: (2024)
por: Srivastava, Aseem, et al.
Publicado: (2024)
Multilingual LLMs Inherently Reward In-Language Time-Sensitive Semantic Alignment for Low-Resource Languages
por: Bajpai, Ashutosh, et al.
Publicado: (2024)
por: Bajpai, Ashutosh, et al.
Publicado: (2024)
Multilingual Test-Time Scaling via Initial Thought Transfer
por: Bajpai, Prasoon, et al.
Publicado: (2025)
por: Bajpai, Prasoon, et al.
Publicado: (2025)
Harmonizing Code-mixed Conversations: Personality-assisted Code-mixed Response Generation in Dialogues
por: Kumar, Shivani, et al.
Publicado: (2024)
por: Kumar, Shivani, et al.
Publicado: (2024)
Decoding Memes: Benchmarking Narrative Role Classification across Multilingual and Multimodal Models
por: Sharma, Shivam, et al.
Publicado: (2025)
por: Sharma, Shivam, et al.
Publicado: (2025)
Can LLMs reason over extended multilingual contexts? Towards long-context evaluation beyond retrieval and haystacks
por: Hengle, Amey, et al.
Publicado: (2025)
por: Hengle, Amey, et al.
Publicado: (2025)
Understanding Factual Recall in Transformers via Associative Memories
por: Nichani, Eshaan, et al.
Publicado: (2024)
por: Nichani, Eshaan, et al.
Publicado: (2024)
Hate Personified: Investigating the role of LLMs in content moderation
por: Masud, Sarah, et al.
Publicado: (2024)
por: Masud, Sarah, et al.
Publicado: (2024)
Text2Cypher Across Languages: Evaluating and Finetuning LLMs
por: Ozsoy, Makbule Gulcin, et al.
Publicado: (2025)
por: Ozsoy, Makbule Gulcin, et al.
Publicado: (2025)
Finetuning LLMs for Comparative Assessment Tasks
por: Raina, Vatsal, et al.
Publicado: (2024)
por: Raina, Vatsal, et al.
Publicado: (2024)
Unveiling Over-Memorization in Finetuning LLMs for Reasoning Tasks
por: Ruan, Zhiwen, et al.
Publicado: (2025)
por: Ruan, Zhiwen, et al.
Publicado: (2025)
Improving the Reliability of LLMs: Combining CoT, RAG, Self-Consistency, and Self-Verification
por: Kumar, Adarsh, et al.
Publicado: (2025)
por: Kumar, Adarsh, et al.
Publicado: (2025)
Rank, Chunk and Expand: Lineage-Oriented Reasoning for Taxonomy Expansion
por: Mishra, Sahil, et al.
Publicado: (2025)
por: Mishra, Sahil, et al.
Publicado: (2025)
The Art of Scaling Test-Time Compute for Large Language Models
por: Agarwal, Aradhye, et al.
Publicado: (2025)
por: Agarwal, Aradhye, et al.
Publicado: (2025)
First Finish Search: Efficient Test-Time Scaling in Large Language Models
por: Agarwal, Aradhye, et al.
Publicado: (2025)
por: Agarwal, Aradhye, et al.
Publicado: (2025)
Innocence in the Crossfire: Roles of Skip Connections in Jailbreaking Visual Language Models
por: Nandi, Palash, et al.
Publicado: (2025)
por: Nandi, Palash, et al.
Publicado: (2025)
Compression Laws for Large Language Models
por: Sengupta, Ayan, et al.
Publicado: (2025)
por: Sengupta, Ayan, et al.
Publicado: (2025)
Counterspeech the ultimate shield! Multi-Conditioned Counterspeech Generation through Attributed Prefix Learning
por: Kumar, Aswini, et al.
Publicado: (2025)
por: Kumar, Aswini, et al.
Publicado: (2025)
You Only Prune Once: Designing Calibration-Free Model Compression With Policy Learning
por: Sengupta, Ayan, et al.
Publicado: (2025)
por: Sengupta, Ayan, et al.
Publicado: (2025)
Markovian ODE-guided scoring can assess the quality of offline reasoning traces in language models
por: Nandi, Arghodeep, et al.
Publicado: (2026)
por: Nandi, Arghodeep, et al.
Publicado: (2026)
Information Anxiety in Large Language Models
por: Bajpai, Prasoon, et al.
Publicado: (2024)
por: Bajpai, Prasoon, et al.
Publicado: (2024)
Exposing Long-Tail Safety Failures in Large Language Models through Efficient Diverse Response Sampling
por: Hajra, Suvadeep, et al.
Publicado: (2026)
por: Hajra, Suvadeep, et al.
Publicado: (2026)
Synthesizing Privacy-Preserving Text Data via Finetuning without Finetuning Billion-Scale LLMs
por: Tan, Bowen, et al.
Publicado: (2025)
por: Tan, Bowen, et al.
Publicado: (2025)
SAFE-MEME: Structured Reasoning Framework for Robust Hate Speech Detection in Memes
por: Nandi, Palash, et al.
Publicado: (2024)
por: Nandi, Palash, et al.
Publicado: (2024)
Focal Inferential Infusion Coupled with Tractable Density Discrimination for Implicit Hate Detection
por: Masud, Sarah, et al.
Publicado: (2023)
por: Masud, Sarah, et al.
Publicado: (2023)
CSEval: Towards Automated, Multi-Dimensional, and Reference-Free Counterspeech Evaluation using Auto-Calibrated LLMs
por: Hengle, Amey, et al.
Publicado: (2025)
por: Hengle, Amey, et al.
Publicado: (2025)
SASFT: Sparse Autoencoder-guided Supervised Finetuning to Mitigate Unexpected Code-Switching in LLMs
por: Deng, Boyi, et al.
Publicado: (2025)
por: Deng, Boyi, et al.
Publicado: (2025)
Understanding Finetuning for Factual Knowledge Extraction
por: Ghosal, Gaurav, et al.
Publicado: (2024)
por: Ghosal, Gaurav, et al.
Publicado: (2024)
Evaluation of Finetuned LLMs in AMR Parsing
por: Ho, Shu Han
Publicado: (2025)
por: Ho, Shu Han
Publicado: (2025)
Through the Prism of Culture: Evaluating LLMs' Understanding of Indian Subcultures and Traditions
por: Chhikara, Garima, et al.
Publicado: (2025)
por: Chhikara, Garima, et al.
Publicado: (2025)
Ejemplares similares
-
Multilingual LLMs Struggle to Link Orthography and Semantics in Bilingual Word Processing
por: Tanwar, Eshaan, et al.
Publicado: (2025) -
Language Models can Exploit Cross-Task In-context Learning for Data-Scarce Novel Tasks
por: Chatterjee, Anwoy, et al.
Publicado: (2024) -
Do You Know About My Nation? Investigating Multilingual Language Models' Cultural Literacy Through Factual Knowledge
por: Tanwar, Eshaan, et al.
Publicado: (2025) -
Understanding the Physics of Key-Value Cache Compression for LLMs through Attention Dynamics
por: Ananthanarayanan, Samhruth, et al.
Publicado: (2026) -
WildSci: Advancing Scientific Reasoning from In-the-Wild Literature
por: Liu, Tengxiao, et al.
Publicado: (2026)