Saved in:
| Main Authors: | Reynolds, Brett, Schneider, Nathan, Arora, Aryaman |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2305.17347 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CausalGym: Benchmarking causal interpretability methods on linguistic tasks
by: Arora, Aryaman, et al.
Published: (2024)
by: Arora, Aryaman, et al.
Published: (2024)
ADAG: Automatically Describing Attribution Graphs
by: Arora, Aryaman, et al.
Published: (2026)
by: Arora, Aryaman, et al.
Published: (2026)
Language Model Circuits Are Sparse in the Neuron Basis
by: Arora, Aryaman, et al.
Published: (2026)
by: Arora, Aryaman, et al.
Published: (2026)
Improved Representation Steering for Language Models
by: Wu, Zhengxuan, et al.
Published: (2025)
by: Wu, Zhengxuan, et al.
Published: (2025)
Bayesian scaling laws for in-context learning
by: Arora, Aryaman, et al.
Published: (2024)
by: Arora, Aryaman, et al.
Published: (2024)
Predicting positive transfer for improved low-resource speech recognition using acoustic pseudo-tokens
by: San, Nay, et al.
Published: (2024)
by: San, Nay, et al.
Published: (2024)
Cross-linguistically Consistent Semantic and Syntactic Annotation of Child-directed Speech
by: Szubert, Ida, et al.
Published: (2021)
by: Szubert, Ida, et al.
Published: (2021)
Mechanistic evaluation of Transformers and state space models
by: Arora, Aryaman, et al.
Published: (2025)
by: Arora, Aryaman, et al.
Published: (2025)
GLOCON Database: Design Decisions and User Manual (v1.0)
by: Hürriyetoğlu, Ali, et al.
Published: (2024)
by: Hürriyetoğlu, Ali, et al.
Published: (2024)
AutoJudge: Judge Decoding Without Manual Annotation
by: Garipov, Roman, et al.
Published: (2025)
by: Garipov, Roman, et al.
Published: (2025)
CKBP v2: Better Annotation and Reasoning for Commonsense Knowledge Base Population
by: Fang, Tianqing, et al.
Published: (2023)
by: Fang, Tianqing, et al.
Published: (2023)
A Reply to Makelov et al. (2023)'s "Interpretability Illusion" Arguments
by: Wu, Zhengxuan, et al.
Published: (2024)
by: Wu, Zhengxuan, et al.
Published: (2024)
ReFT: Representation Finetuning for Language Models
by: Wu, Zhengxuan, et al.
Published: (2024)
by: Wu, Zhengxuan, et al.
Published: (2024)
PreFT: Prefill-only finetuning for efficient inference
by: Lanpouthakoun, Andrew, et al.
Published: (2026)
by: Lanpouthakoun, Andrew, et al.
Published: (2026)
pyvene: A Library for Understanding and Improving PyTorch Models via Interventions
by: Wu, Zhengxuan, et al.
Published: (2024)
by: Wu, Zhengxuan, et al.
Published: (2024)
Lost in Translationese? Reducing Translation Effect Using Abstract Meaning Representation
by: Wein, Shira, et al.
Published: (2023)
by: Wein, Shira, et al.
Published: (2023)
ANAH-v2: Scaling Analytical Hallucination Annotation of Large Language Models
by: Gu, Yuzhe, et al.
Published: (2024)
by: Gu, Yuzhe, et al.
Published: (2024)
Verbalizing LLMs' assumptions to explain and control sycophancy
by: Cheng, Myra, et al.
Published: (2026)
by: Cheng, Myra, et al.
Published: (2026)
UCxn: Typologically Informed Annotation of Constructions Atop Universal Dependencies
by: Weissweiler, Leonie, et al.
Published: (2024)
by: Weissweiler, Leonie, et al.
Published: (2024)
MagicBrush: A Manually Annotated Dataset for Instruction-Guided Image Editing
by: Zhang, Kai, et al.
Published: (2023)
by: Zhang, Kai, et al.
Published: (2023)
Construction Identification and Disambiguation Using BERT: A Case Study of NPN
by: Scivetti, Wesley, et al.
Published: (2025)
by: Scivetti, Wesley, et al.
Published: (2025)
Speaking of Language: Reflections on Metalanguage Research in NLP
by: Schneider, Nathan, et al.
Published: (2026)
by: Schneider, Nathan, et al.
Published: (2026)
AdParaphrase v2.0: Generating Attractive Ad Texts Using a Preference-Annotated Paraphrase Dataset
by: Murakami, Soichiro, et al.
Published: (2025)
by: Murakami, Soichiro, et al.
Published: (2025)
$\textit{BenchIE}^{FL}$ : A Manually Re-Annotated Fact-Based Open Information Extraction Benchmark
by: Lamarche, Fabrice, et al.
Published: (2024)
by: Lamarche, Fabrice, et al.
Published: (2024)
AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
by: Wu, Zhengxuan, et al.
Published: (2025)
by: Wu, Zhengxuan, et al.
Published: (2025)
Improving Adversarial Data Collection by Supporting Annotators: Lessons from GAHD, a German Hate Speech Dataset
by: Goldzycher, Janis, et al.
Published: (2024)
by: Goldzycher, Janis, et al.
Published: (2024)
DevEval: A Manually-Annotated Code Generation Benchmark Aligned with Real-World Code Repositories
by: Li, Jia, et al.
Published: (2024)
by: Li, Jia, et al.
Published: (2024)
LLM Teacher-Student Framework for Text Classification With No Manually Annotated Data: A Case Study in IPTC News Topic Classification
by: Kuzman, Taja, et al.
Published: (2024)
by: Kuzman, Taja, et al.
Published: (2024)
Predicting the Emergence of Induction Heads in Language Model Pretraining
by: Aoyama, Tatsuya, et al.
Published: (2025)
by: Aoyama, Tatsuya, et al.
Published: (2025)
Manga109-v2026: Revisiting Manga109 Annotations for Modern Manga Understanding
by: Baek, Jeonghun, et al.
Published: (2026)
by: Baek, Jeonghun, et al.
Published: (2026)
Second language Korean Universal Dependency treebank v1.2: Focus on data augmentation and annotation scheme refinement
by: Sung, Hakyung, et al.
Published: (2025)
by: Sung, Hakyung, et al.
Published: (2025)
Linguistic Frameworks Go Toe-to-Toe at Neuro-Symbolic Language Modeling
by: Prange, Jakob, et al.
Published: (2021)
by: Prange, Jakob, et al.
Published: (2021)
Natural Language Processing RELIES on Linguistics
by: Opitz, Juri, et al.
Published: (2024)
by: Opitz, Juri, et al.
Published: (2024)
Analyzing COVID-19 Vaccination Sentiments in Nigerian Cyberspace: Insights from a Manually Annotated Twitter Dataset
by: Ahmad, Ibrahim Said, et al.
Published: (2024)
by: Ahmad, Ibrahim Said, et al.
Published: (2024)
From Checklists to Clusters: A Homeostatic Account of AGI Evaluation
by: Reynolds, Brett
Published: (2025)
by: Reynolds, Brett
Published: (2025)
Mitigating Catastrophic Forgetting in Mathematical Reasoning Finetuning through Mixed Training
by: Reynolds, John Graham
Published: (2025)
by: Reynolds, John Graham
Published: (2025)
LinkTransformer: A Unified Package for Record Linkage with Transformer Language Models
by: Arora, Abhishek, et al.
Published: (2023)
by: Arora, Abhishek, et al.
Published: (2023)
BoAT v2 -- A Web-Based Dependency Annotation Tool with Focus on Agglutinative Languages
by: Akkurt, Salih Furkan, et al.
Published: (2022)
by: Akkurt, Salih Furkan, et al.
Published: (2022)
Beyond the Battlefield: Framing Analysis of Media Coverage in Conflict Reporting
by: Kaur, Avneet, et al.
Published: (2025)
by: Kaur, Avneet, et al.
Published: (2025)
Geometric Properties of the Voronoi Tessellation in Latent Semantic Manifolds of Large Language Models
by: Brett, Marshall
Published: (2026)
by: Brett, Marshall
Published: (2026)
Similar Items
-
CausalGym: Benchmarking causal interpretability methods on linguistic tasks
by: Arora, Aryaman, et al.
Published: (2024) -
ADAG: Automatically Describing Attribution Graphs
by: Arora, Aryaman, et al.
Published: (2026) -
Language Model Circuits Are Sparse in the Neuron Basis
by: Arora, Aryaman, et al.
Published: (2026) -
Improved Representation Steering for Language Models
by: Wu, Zhengxuan, et al.
Published: (2025) -
Bayesian scaling laws for in-context learning
by: Arora, Aryaman, et al.
Published: (2024)