Hitting "Probe"rty with Non-Linearity, and More
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Pal, Avik, Pawar, Madhura |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Hit-RAG: Learning to Reason with Long Contexts via Preference Alignment
par: Liu, Junming, et autres
Publié: (2026)
par: Liu, Junming, et autres
Publié: (2026)
No Answer Needed: Predicting LLM Answer Accuracy from Question-Only Linear Probes
par: Cencerrado, Iván Vicente Moreno, et autres
Publié: (2025)
par: Cencerrado, Iván Vicente Moreno, et autres
Publié: (2025)
Mechanistic Interpretability of Cognitive Complexity in LLMs via Linear Probing using Bloom's Taxonomy
par: Raimondi, Bianca, et autres
Publié: (2026)
par: Raimondi, Bianca, et autres
Publié: (2026)
When Models Know More Than They Say: Probing Analogical Reasoning in LLMs
par: McGovern, Hope, et autres
Publié: (2026)
par: McGovern, Hope, et autres
Publié: (2026)
Context-Agent: Dynamic Discourse Trees for Non-Linear Dialogue
par: Hu, Junan, et autres
Publié: (2026)
par: Hu, Junan, et autres
Publié: (2026)
Enhancing In-context Learning via Linear Probe Calibration
par: Abbas, Momin, et autres
Publié: (2024)
par: Abbas, Momin, et autres
Publié: (2024)
From Exact Hits to Close Enough: Semantic Caching for LLM Embeddings
par: Biton, Dvir David, et autres
Publié: (2026)
par: Biton, Dvir David, et autres
Publié: (2026)
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench
par: Narad, Reuben, et autres
Publié: (2025)
par: Narad, Reuben, et autres
Publié: (2025)
When Linear Attention Meets Autoregressive Decoding: Towards More Effective and Efficient Linearized Large Language Models
par: You, Haoran, et autres
Publié: (2024)
par: You, Haoran, et autres
Publié: (2024)
Debating with More Persuasive LLMs Leads to More Truthful Answers
par: Khan, Akbir, et autres
Publié: (2024)
par: Khan, Akbir, et autres
Publié: (2024)
More RLHF, More Trust? On The Impact of Preference Alignment On Trustworthiness
par: Li, Aaron J., et autres
Publié: (2024)
par: Li, Aaron J., et autres
Publié: (2024)
Rhetorical Questions in LLM Representations: A Linear Probing Study
par: Yao, Louie Hong, et autres
Publié: (2026)
par: Yao, Louie Hong, et autres
Publié: (2026)
Remembering More, Risking More: Longitudinal Safety Risks in Memory-Equipped LLM Agents
par: Al-Tawaha, Ahmad, et autres
Publié: (2026)
par: Al-Tawaha, Ahmad, et autres
Publié: (2026)
Two Stones Hit One Bird: Bilevel Positional Encoding for Better Length Extrapolation
par: He, Zhenyu, et autres
Publié: (2024)
par: He, Zhenyu, et autres
Publié: (2024)
LIMO: Less is More for Reasoning
par: Ye, Yixin, et autres
Publié: (2025)
par: Ye, Yixin, et autres
Publié: (2025)
Latent Causal Probing: A Formal Perspective on Probing with Causal Models of Data
par: Jin, Charles, et autres
Publié: (2024)
par: Jin, Charles, et autres
Publié: (2024)
SMART SLM: Structured Memory and Reasoning Transformer, A Small Language Model for Accurate Document Assistance
par: Dudeja, Divij, et autres
Publié: (2025)
par: Dudeja, Divij, et autres
Publié: (2025)
Higher Layers Need More LoRA Experts
par: Gao, Chongyang, et autres
Publié: (2024)
par: Gao, Chongyang, et autres
Publié: (2024)
Less is More: Resource-Efficient Low-Rank Adaptation
par: Tian, Chunlin, et autres
Publié: (2025)
par: Tian, Chunlin, et autres
Publié: (2025)
How Do LLMs Persuade? Linear Probes Can Uncover Persuasion Dynamics in Multi-Turn Conversations
par: Jaipersaud, Brandon, et autres
Publié: (2025)
par: Jaipersaud, Brandon, et autres
Publié: (2025)
Probing for Arithmetic Errors in Language Models
par: Sun, Yucheng, et autres
Publié: (2025)
par: Sun, Yucheng, et autres
Publié: (2025)
ConDABench: Interactive Evaluation of Language Models for Data Analysis
par: Dutta, Avik, et autres
Publié: (2025)
par: Dutta, Avik, et autres
Publié: (2025)
NDP: Next Distribution Prediction as a More Broad Target
par: Ruan, Junhao, et autres
Publié: (2024)
par: Ruan, Junhao, et autres
Publié: (2024)
Towards More Standardized AI Evaluation: From Models to Agents
par: Filali, Ali El, et autres
Publié: (2026)
par: Filali, Ali El, et autres
Publié: (2026)
More of the Same: Persistent Representational Harms Under Increased Representation
par: Mickel, Jennifer, et autres
Publié: (2025)
par: Mickel, Jennifer, et autres
Publié: (2025)
Acting Less is Reasoning More! Teaching Model to Act Efficiently
par: Wang, Hongru, et autres
Publié: (2025)
par: Wang, Hongru, et autres
Publié: (2025)
Fewer is More: Boosting LLM Reasoning with Reinforced Context Pruning
par: Huang, Xijie, et autres
Publié: (2023)
par: Huang, Xijie, et autres
Publié: (2023)
Making Metadata More FAIR Using Large Language Models
par: Sundaram, Sowmya S., et autres
Publié: (2023)
par: Sundaram, Sowmya S., et autres
Publié: (2023)
LLMs Struggle with Abstract Meaning Comprehension More Than Expected
par: Alhazmi, Hamoud, et autres
Publié: (2026)
par: Alhazmi, Hamoud, et autres
Publié: (2026)
Neural Networks Remember More: The Power of Parameter Isolation and Combination
par: Zeng, Biqing, et autres
Publié: (2025)
par: Zeng, Biqing, et autres
Publié: (2025)
Spike No More: Stabilizing the Pre-training of Large Language Models
par: Takase, Sho, et autres
Publié: (2023)
par: Takase, Sho, et autres
Publié: (2023)
Drift No More? Context Equilibria in Multi-Turn LLM Interactions
par: Dongre, Vardhan, et autres
Publié: (2025)
par: Dongre, Vardhan, et autres
Publié: (2025)
Entangled in Representations: Mechanistic Investigation of Cultural Biases in Large Language Models
par: Yu, Haeun, et autres
Publié: (2025)
par: Yu, Haeun, et autres
Publié: (2025)
LinearARD: Linear-Memory Attention Distillation for RoPE Restoration
par: Yang, Ning, et autres
Publié: (2026)
par: Yang, Ning, et autres
Publié: (2026)
ReProbe: Efficient Test-Time Scaling of Multi-Step Reasoning by Probing Internal States of Large Language Models
par: Ni, Jingwei, et autres
Publié: (2025)
par: Ni, Jingwei, et autres
Publié: (2025)
Do explanations generalize across large reasoning models?
par: Pal, Koyena, et autres
Publié: (2026)
par: Pal, Koyena, et autres
Publié: (2026)
Probing Causality Manipulation of Large Language Models
par: Zhang, Chenyang, et autres
Publié: (2024)
par: Zhang, Chenyang, et autres
Publié: (2024)
Probing and Steering Evaluation Awareness of Language Models
par: Nguyen, Jord, et autres
Publié: (2025)
par: Nguyen, Jord, et autres
Publié: (2025)
Probing Neural Topology of Large Language Models
par: Zheng, Yu, et autres
Publié: (2025)
par: Zheng, Yu, et autres
Publié: (2025)
Probing the Lack of Stable Internal Beliefs in LLMs
par: Luo, Yifan, et autres
Publié: (2026)
par: Luo, Yifan, et autres
Publié: (2026)
Documents similaires
-
Hit-RAG: Learning to Reason with Long Contexts via Preference Alignment
par: Liu, Junming, et autres
Publié: (2026) -
No Answer Needed: Predicting LLM Answer Accuracy from Question-Only Linear Probes
par: Cencerrado, Iván Vicente Moreno, et autres
Publié: (2025) -
Mechanistic Interpretability of Cognitive Complexity in LLMs via Linear Probing using Bloom's Taxonomy
par: Raimondi, Bianca, et autres
Publié: (2026) -
When Models Know More Than They Say: Probing Analogical Reasoning in LLMs
par: McGovern, Hope, et autres
Publié: (2026) -
Context-Agent: Dynamic Discourse Trees for Non-Linear Dialogue
par: Hu, Junan, et autres
Publié: (2026)