Beyond Transfer Accuracy: Faithful Circuits for Controlled Low-Resource Adaptation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nur'aini, Khumaisa, Purwarianti, Ayu, Aji, Alham Fikri, Wijaya, Derry |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Does Visual Rendering Bypass Tokenization? Investigating Script-Tokenizer Misalignment in Pixel-Based Language Models
von: Susanto, Lucky, et al.
Veröffentlicht: (2026)
von: Susanto, Lucky, et al.
Veröffentlicht: (2026)
NusaAksara: A Multimodal and Multilingual Benchmark for Preserving Indonesian Indigenous Scripts
von: Adilazuarda, Muhammad Farid, et al.
Veröffentlicht: (2025)
von: Adilazuarda, Muhammad Farid, et al.
Veröffentlicht: (2025)
QLESS: A Quantized Approach for Data Valuation and Selection in Large Language Model Fine-Tuning
von: Ananta, Moses, et al.
Veröffentlicht: (2025)
von: Ananta, Moses, et al.
Veröffentlicht: (2025)
Improving Low-Resource Machine Translation via Round-Trip Reinforcement Learning
von: Attia, Ahmed, et al.
Veröffentlicht: (2026)
von: Attia, Ahmed, et al.
Veröffentlicht: (2026)
Could We Have Had Better Multilingual LLMs If English Was Not the Central Language?
von: Diandaru, Ryandito, et al.
Veröffentlicht: (2024)
von: Diandaru, Ryandito, et al.
Veröffentlicht: (2024)
IndoToxic2024: A Demographically-Enriched Dataset of Hate Speech and Toxicity Types for Indonesian Language
von: Susanto, Lucky, et al.
Veröffentlicht: (2024)
von: Susanto, Lucky, et al.
Veröffentlicht: (2024)
MLKV: Multi-Layer Key-Value Heads for Memory Efficient Transformer Decoding
von: Zuhri, Zayd Muhammad Kawakibi, et al.
Veröffentlicht: (2024)
von: Zuhri, Zayd Muhammad Kawakibi, et al.
Veröffentlicht: (2024)
Data Laundering: Artificially Boosting Benchmark Results through Knowledge Distillation
von: Mansurov, Jonibek, et al.
Veröffentlicht: (2024)
von: Mansurov, Jonibek, et al.
Veröffentlicht: (2024)
Sense Representations Are Inducible Interfaces
von: Cruz, Jan Christian Blaise, et al.
Veröffentlicht: (2026)
von: Cruz, Jan Christian Blaise, et al.
Veröffentlicht: (2026)
LLM Olympiad: Why Model Evaluation Needs a Sealed Exam
von: Cruz, Jan Christian Blaise, et al.
Veröffentlicht: (2026)
von: Cruz, Jan Christian Blaise, et al.
Veröffentlicht: (2026)
Is Active Persona Inference Necessary for Aligning Small Models to Personal Preferences?
von: Tang, Zilu, et al.
Veröffentlicht: (2025)
von: Tang, Zilu, et al.
Veröffentlicht: (2025)
MetaMetrics-MT: Tuning Meta-Metrics for Machine Translation via Human Preference Calibration
von: Anugraha, David, et al.
Veröffentlicht: (2024)
von: Anugraha, David, et al.
Veröffentlicht: (2024)
Transformer Circuit Faithfulness Metrics are not Robust
von: Miller, Joseph, et al.
Veröffentlicht: (2024)
von: Miller, Joseph, et al.
Veröffentlicht: (2024)
Crosslingual Reasoning through Test-Time Scaling
von: Yong, Zheng-Xin, et al.
Veröffentlicht: (2025)
von: Yong, Zheng-Xin, et al.
Veröffentlicht: (2025)
Enhancing Natural Language Inference Performance with Knowledge Graph for COVID-19 Automated Fact-Checking in Indonesian Language
von: Muharram, Arief Purnama, et al.
Veröffentlicht: (2024)
von: Muharram, Arief Purnama, et al.
Veröffentlicht: (2024)
TextGames: Learning to Self-Play Text-Based Puzzle Games via Language Model Reasoning
von: Hudi, Frederikus, et al.
Veröffentlicht: (2025)
von: Hudi, Frederikus, et al.
Veröffentlicht: (2025)
LinguAlchemy: Fusing Typological and Geographical Elements for Unseen Language Generalization
von: Adilazuarda, Muhammad Farid, et al.
Veröffentlicht: (2024)
von: Adilazuarda, Muhammad Farid, et al.
Veröffentlicht: (2024)
Datasheets Aren't Enough: DataRubrics for Automated Quality Metrics and Accountability
von: Winata, Genta Indra, et al.
Veröffentlicht: (2025)
von: Winata, Genta Indra, et al.
Veröffentlicht: (2025)
mR3: Multilingual Rubric-Agnostic Reward Reasoning Models
von: Anugraha, David, et al.
Veröffentlicht: (2025)
von: Anugraha, David, et al.
Veröffentlicht: (2025)
Investigating the Transferability of Code Repair for Low-Resource Programming Languages
von: Wong, Kyle, et al.
Veröffentlicht: (2024)
von: Wong, Kyle, et al.
Veröffentlicht: (2024)
MetaMetrics: Calibrating Metrics For Generation Tasks Using Human Preferences
von: Winata, Genta Indra, et al.
Veröffentlicht: (2024)
von: Winata, Genta Indra, et al.
Veröffentlicht: (2024)
R3: Robust Rubric-Agnostic Reward Models
von: Anugraha, David, et al.
Veröffentlicht: (2025)
von: Anugraha, David, et al.
Veröffentlicht: (2025)
Generating Faithful Text From a Knowledge Graph with Noisy Reference Text
von: Hashem, Tahsina, et al.
Veröffentlicht: (2023)
von: Hashem, Tahsina, et al.
Veröffentlicht: (2023)
Recover-LoRA: Data-Free Accuracy Recovery of Degraded Language Models via Low-Rank Adaptation
von: Das, Devleena, et al.
Veröffentlicht: (2025)
von: Das, Devleena, et al.
Veröffentlicht: (2025)
BhashaSetu: Cross-Lingual Knowledge Transfer from High-Resource to Extreme Low-Resource Languages
von: Maji, Subhadip, et al.
Veröffentlicht: (2026)
von: Maji, Subhadip, et al.
Veröffentlicht: (2026)
Selection-p: Self-Supervised Task-Agnostic Prompt Compression for Faithfulness and Transferability
von: Chung, Tsz Ting, et al.
Veröffentlicht: (2024)
von: Chung, Tsz Ting, et al.
Veröffentlicht: (2024)
Predicting LLM Correctness in Prosthodontics Using Metadata and Hallucination Signals
von: Susanto, Lucky, et al.
Veröffentlicht: (2025)
von: Susanto, Lucky, et al.
Veröffentlicht: (2025)
Environmental carrying capacity and environmental capacity in the issuance of recommendation and borrow-use permit of forest area
von: Wiwin, Nur'aini
Veröffentlicht: (2019)
von: Wiwin, Nur'aini
Veröffentlicht: (2019)
Circuit Insights: Towards Interpretability Beyond Activations
von: Golimblevskaia, Elena, et al.
Veröffentlicht: (2025)
von: Golimblevskaia, Elena, et al.
Veröffentlicht: (2025)
FaithLM: Towards Faithful Explanations for Large Language Models
von: Chuang, Yu-Neng, et al.
Veröffentlicht: (2024)
von: Chuang, Yu-Neng, et al.
Veröffentlicht: (2024)
Few-Shot Cross-Lingual Transfer for Prompting Large Language Models in Low-Resource Languages
von: Toukmaji, Christopher
Veröffentlicht: (2024)
von: Toukmaji, Christopher
Veröffentlicht: (2024)
Improving Faithfulness of Abstractive Summarization by Controlling Confounding Effect of Irrelevant Sentences
von: Ghoshal, Asish, et al.
Veröffentlicht: (2022)
von: Ghoshal, Asish, et al.
Veröffentlicht: (2022)
Faithfulness Evaluation for Decoder-only LLM Attributions with Controlled Retained Information
von: Huang, Xin, et al.
Veröffentlicht: (2026)
von: Huang, Xin, et al.
Veröffentlicht: (2026)
Text2Data: Low-Resource Data Generation with Textual Control
von: Wang, Shiyu, et al.
Veröffentlicht: (2024)
von: Wang, Shiyu, et al.
Veröffentlicht: (2024)
Bridging the Bosphorus: Advancing Turkish Large Language Models through Strategies for Low-Resource Language Adaptation and Benchmarking
von: Acikgoz, Emre Can, et al.
Veröffentlicht: (2024)
von: Acikgoz, Emre Can, et al.
Veröffentlicht: (2024)
Beyond Components: Singular Vector-Based Interpretability of Transformer Circuits
von: Ahmad, Areeb, et al.
Veröffentlicht: (2025)
von: Ahmad, Areeb, et al.
Veröffentlicht: (2025)
The Expressive Power of Low-Rank Adaptation
von: Zeng, Yuchen, et al.
Veröffentlicht: (2023)
von: Zeng, Yuchen, et al.
Veröffentlicht: (2023)
FaithfulSAE: Towards Capturing Faithful Features with Sparse Autoencoders without External Dataset Dependencies
von: Cho, Seonglae, et al.
Veröffentlicht: (2025)
von: Cho, Seonglae, et al.
Veröffentlicht: (2025)
FaithEval: Can Your Language Model Stay Faithful to Context, Even If "The Moon is Made of Marshmallows"
von: Ming, Yifei, et al.
Veröffentlicht: (2024)
von: Ming, Yifei, et al.
Veröffentlicht: (2024)
Batched Low-Rank Adaptation of Foundation Models
von: Wen, Yeming, et al.
Veröffentlicht: (2023)
von: Wen, Yeming, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Does Visual Rendering Bypass Tokenization? Investigating Script-Tokenizer Misalignment in Pixel-Based Language Models
von: Susanto, Lucky, et al.
Veröffentlicht: (2026) -
NusaAksara: A Multimodal and Multilingual Benchmark for Preserving Indonesian Indigenous Scripts
von: Adilazuarda, Muhammad Farid, et al.
Veröffentlicht: (2025) -
QLESS: A Quantized Approach for Data Valuation and Selection in Large Language Model Fine-Tuning
von: Ananta, Moses, et al.
Veröffentlicht: (2025) -
Improving Low-Resource Machine Translation via Round-Trip Reinforcement Learning
von: Attia, Ahmed, et al.
Veröffentlicht: (2026) -
Could We Have Had Better Multilingual LLMs If English Was Not the Central Language?
von: Diandaru, Ryandito, et al.
Veröffentlicht: (2024)