A Universal Vibe? Finding and Controlling Language-Agnostic Informal Register with SAEs
Fuente:
arXiv
Salvato in:
| Autori principali: | Kialy, Uri Z., Shtarkberg, Avi, Klein, Ayal |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CARE: Counselor-Aligned Response Engine for Online Mental-Health Support
di: Astrin, Hagai, et al.
Pubblicazione: (2026)
di: Astrin, Hagai, et al.
Pubblicazione: (2026)
Effective QA-driven Annotation of Predicate-Argument Relations Across Languages
di: Davidov, Jonathan, et al.
Pubblicazione: (2026)
di: Davidov, Jonathan, et al.
Pubblicazione: (2026)
QA-Noun: Representing Nominal Semantics via Natural Language Question-Answer Pairs
di: Tseytlin, Maria, et al.
Pubblicazione: (2025)
di: Tseytlin, Maria, et al.
Pubblicazione: (2025)
Residual Stream Analysis with Multi-Layer SAEs
di: Lawson, Tim, et al.
Pubblicazione: (2024)
di: Lawson, Tim, et al.
Pubblicazione: (2024)
Resa: Transparent Reasoning Models via SAEs
di: Wang, Shangshang, et al.
Pubblicazione: (2025)
di: Wang, Shangshang, et al.
Pubblicazione: (2025)
Can SAEs reveal and mitigate racial biases of LLMs in healthcare?
di: Ahsan, Hiba, et al.
Pubblicazione: (2025)
di: Ahsan, Hiba, et al.
Pubblicazione: (2025)
SAEs Are Good for Steering -- If You Select the Right Features
di: Arad, Dana, et al.
Pubblicazione: (2025)
di: Arad, Dana, et al.
Pubblicazione: (2025)
Teach Old SAEs New Domain Tricks with Boosting
di: Koriagin, Nikita, et al.
Pubblicazione: (2025)
di: Koriagin, Nikita, et al.
Pubblicazione: (2025)
Curated Datasets and Neural Models for Machine Translation of Informal Registers between Mayan and Spanish Vernaculars
di: Lou, Andrés, et al.
Pubblicazione: (2024)
di: Lou, Andrés, et al.
Pubblicazione: (2024)
Position: Mechanistic Interpretability Should Prioritize Feature Consistency in SAEs
di: Song, Xiangchen, et al.
Pubblicazione: (2025)
di: Song, Xiangchen, et al.
Pubblicazione: (2025)
Vibe Coding, Interface Flattening
di: Jin, Hongrui
Pubblicazione: (2025)
di: Jin, Hongrui
Pubblicazione: (2025)
What Evidence Do Language Models Find Convincing?
di: Wan, Alexander, et al.
Pubblicazione: (2024)
di: Wan, Alexander, et al.
Pubblicazione: (2024)
Explicating the Implicit: Argument Detection Beyond Sentence Boundaries
di: Roit, Paul, et al.
Pubblicazione: (2024)
di: Roit, Paul, et al.
Pubblicazione: (2024)
QAPyramid: Fine-grained Evaluation of Content Selection for Text Summarization
di: Zhang, Shiyue, et al.
Pubblicazione: (2024)
di: Zhang, Shiyue, et al.
Pubblicazione: (2024)
VibeCheck: Discover and Quantify Qualitative Differences in Large Language Models
di: Dunlap, Lisa, et al.
Pubblicazione: (2024)
di: Dunlap, Lisa, et al.
Pubblicazione: (2024)
Internalizing Tool Knowledge in Small Language Models via QLoRA Fine-Tuning
di: Shemla, Yuval, et al.
Pubblicazione: (2026)
di: Shemla, Yuval, et al.
Pubblicazione: (2026)
Improving Informally Romanized Language Identification
di: Benton, Adrian, et al.
Pubblicazione: (2025)
di: Benton, Adrian, et al.
Pubblicazione: (2025)
Ruler: A Model-Agnostic Method to Control Generated Length for Large Language Models
di: Li, Jiaming, et al.
Pubblicazione: (2024)
di: Li, Jiaming, et al.
Pubblicazione: (2024)
Script-Agnostic Language Identification
di: Agarwal, Milind, et al.
Pubblicazione: (2024)
di: Agarwal, Milind, et al.
Pubblicazione: (2024)
A Watermark for Order-Agnostic Language Models
di: Chen, Ruibo, et al.
Pubblicazione: (2024)
di: Chen, Ruibo, et al.
Pubblicazione: (2024)
MolViBench: Evaluating LLMs on Molecular Vibe Coding
di: Li, Jiatong, et al.
Pubblicazione: (2026)
di: Li, Jiatong, et al.
Pubblicazione: (2026)
Toward Informal Language Processing: Knowledge of Slang in Large Language Models
di: Sun, Zhewei, et al.
Pubblicazione: (2024)
di: Sun, Zhewei, et al.
Pubblicazione: (2024)
GLM-5: from Vibe Coding to Agentic Engineering
di: GLM-5-Team, et al.
Pubblicazione: (2026)
di: GLM-5-Team, et al.
Pubblicazione: (2026)
Steering Large Language Models with Register Analysis for Arbitrary Style Transfer
di: Yang, Xinchen, et al.
Pubblicazione: (2025)
di: Yang, Xinchen, et al.
Pubblicazione: (2025)
CoverBench: A Challenging Benchmark for Complex Claim Verification
di: Jacovi, Alon, et al.
Pubblicazione: (2024)
di: Jacovi, Alon, et al.
Pubblicazione: (2024)
VibeVoice Technical Report
di: Peng, Zhiliang, et al.
Pubblicazione: (2025)
di: Peng, Zhiliang, et al.
Pubblicazione: (2025)
Domain Fine-Tuning vs. Retrieval-Augmented Generation for Medical Multiple-Choice Question Answering: A Controlled Comparison at the 4B-Parameter Scale
di: Buskila, Avi-ad Avraam
Pubblicazione: (2026)
di: Buskila, Avi-ad Avraam
Pubblicazione: (2026)
Language-Agnostic Analysis of Speech Depression Detection
di: Binu, Sona, et al.
Pubblicazione: (2024)
di: Binu, Sona, et al.
Pubblicazione: (2024)
Registering Source Tokens to Target Language Spaces in Multilingual Neural Machine Translation
di: Qu, Zhi, et al.
Pubblicazione: (2025)
di: Qu, Zhi, et al.
Pubblicazione: (2025)
Contrastive Perplexity for Controlled Generation: An Application in Detoxifying Large Language Models
di: Klein, Tassilo, et al.
Pubblicazione: (2024)
di: Klein, Tassilo, et al.
Pubblicazione: (2024)
SAEs $\textit{Can}$ Improve Unlearning: Dynamic Sparse Autoencoder Guardrails for Precision Unlearning in LLMs
di: Muhamed, Aashiq, et al.
Pubblicazione: (2025)
di: Muhamed, Aashiq, et al.
Pubblicazione: (2025)
From Code-Centric to Concept-Centric: Teaching NLP with LLM-Assisted "Vibe Coding"
di: Al-Khalifa, Hend
Pubblicazione: (2026)
di: Al-Khalifa, Hend
Pubblicazione: (2026)
Compiling by Proving: Language-Agnostic Automatic Optimization from Formal Semantics
di: Zhao, Jianhong, et al.
Pubblicazione: (2025)
di: Zhao, Jianhong, et al.
Pubblicazione: (2025)
Are Language Models Agnostic to Linguistically Grounded Perturbations? A Case Study of Indic Languages
di: Ghosh, Poulami, et al.
Pubblicazione: (2024)
di: Ghosh, Poulami, et al.
Pubblicazione: (2024)
Multilingual Knowledge Editing with Language-Agnostic Factual Neurons
di: Zhang, Xue, et al.
Pubblicazione: (2024)
di: Zhang, Xue, et al.
Pubblicazione: (2024)
Splintering Nonconcatenative Languages for Better Tokenization
di: Gazit, Bar, et al.
Pubblicazione: (2025)
di: Gazit, Bar, et al.
Pubblicazione: (2025)
Signs of Struggle: Spotting Cognitive Distortions across Language and Register
di: Kuber, Abhishek, et al.
Pubblicazione: (2025)
di: Kuber, Abhishek, et al.
Pubblicazione: (2025)
SecureVibeBench: Benchmarking Secure Vibe Coding of AI Agents via Reconstructing Vulnerability-Introducing Scenarios
di: Chen, Junkai, et al.
Pubblicazione: (2025)
di: Chen, Junkai, et al.
Pubblicazione: (2025)
Register Always Matters: Analysis of LLM Pretraining Data Through the Lens of Language Variation
di: Myntti, Amanda, et al.
Pubblicazione: (2025)
di: Myntti, Amanda, et al.
Pubblicazione: (2025)
Chinese Language Is Not More Efficient Than English in Vibe Coding: A Preliminary Study on Token Cost and Problem-Solving Rate
di: Ren, Simiao, et al.
Pubblicazione: (2026)
di: Ren, Simiao, et al.
Pubblicazione: (2026)
Documenti analoghi
-
CARE: Counselor-Aligned Response Engine for Online Mental-Health Support
di: Astrin, Hagai, et al.
Pubblicazione: (2026) -
Effective QA-driven Annotation of Predicate-Argument Relations Across Languages
di: Davidov, Jonathan, et al.
Pubblicazione: (2026) -
QA-Noun: Representing Nominal Semantics via Natural Language Question-Answer Pairs
di: Tseytlin, Maria, et al.
Pubblicazione: (2025) -
Residual Stream Analysis with Multi-Layer SAEs
di: Lawson, Tim, et al.
Pubblicazione: (2024) -
Resa: Transparent Reasoning Models via SAEs
di: Wang, Shangshang, et al.
Pubblicazione: (2025)