Saved in:
| Main Authors: | Qraitem, Maan, Saenko, Kate, Plummer, Bryan A. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2601.03396 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
From Fake to Real: Pretraining on Balanced Synthetic Images to Prevent Spurious Correlations in Image Recognition
by: Qraitem, Maan, et al.
Published: (2023)
by: Qraitem, Maan, et al.
Published: (2023)
Web Artifact Attacks Disrupt Vision Language Models
by: Qraitem, Maan, et al.
Published: (2025)
by: Qraitem, Maan, et al.
Published: (2025)
SLANT: Spurious Logo ANalysis Toolkit
by: Qraitem, Maan, et al.
Published: (2024)
by: Qraitem, Maan, et al.
Published: (2024)
Vision-LLMs Can Fool Themselves with Self-Generated Typographic Attacks
by: Qraitem, Maan, et al.
Published: (2024)
by: Qraitem, Maan, et al.
Published: (2024)
Tell Me What's Next: Textual Foresight for Generic UI Representations
by: Burns, Andrea, et al.
Published: (2024)
by: Burns, Andrea, et al.
Published: (2024)
KiVA: Kid-inspired Visual Analogies for Testing Large Multimodal Models
by: Yiu, Eunice, et al.
Published: (2024)
by: Yiu, Eunice, et al.
Published: (2024)
Scaling Up Temporal Domain Generalization via Temporal Experts Averaging
by: Liu, Aoming, et al.
Published: (2025)
by: Liu, Aoming, et al.
Published: (2025)
ERM++: An Improved Baseline for Domain Generalization
by: Teterwak, Piotr, et al.
Published: (2023)
by: Teterwak, Piotr, et al.
Published: (2023)
Is Large-Scale Pretraining the Secret to Good Domain Generalization?
by: Teterwak, Piotr, et al.
Published: (2024)
by: Teterwak, Piotr, et al.
Published: (2024)
OP-LoRA: The Blessing of Dimensionality
by: Teterwak, Piotr, et al.
Published: (2024)
by: Teterwak, Piotr, et al.
Published: (2024)
CLAMP: Contrastive LAnguage Model Prompt-tuning
by: Teterwak, Piotr, et al.
Published: (2023)
by: Teterwak, Piotr, et al.
Published: (2023)
Machine-Generated Text Localization
by: Zhang, Zhongping, et al.
Published: (2024)
by: Zhang, Zhongping, et al.
Published: (2024)
Single Character Perturbations Break LLM Alignment
by: Lin, Leon, et al.
Published: (2024)
by: Lin, Leon, et al.
Published: (2024)
Koala: Key frame-conditioned long video-LLM
by: Tan, Reuben, et al.
Published: (2024)
by: Tan, Reuben, et al.
Published: (2024)
A Single Character can Make or Break Your LLM Evals
by: Su, Jingtong, et al.
Published: (2025)
by: Su, Jingtong, et al.
Published: (2025)
Breaking the Mold: Nonlinear Ranking Function Synthesis Without Templates
by: Zhu, Shaowei, et al.
Published: (2024)
by: Zhu, Shaowei, et al.
Published: (2024)
Enhancing LLM Character-Level Manipulation via Divide and Conquer
by: Xiong, Zhen, et al.
Published: (2025)
by: Xiong, Zhen, et al.
Published: (2025)
Concept Arithmetics for Circumventing Concept Inhibition in Diffusion Models
by: Petsiuk, Vitali, et al.
Published: (2024)
by: Petsiuk, Vitali, et al.
Published: (2024)
C-LLM: Learn to Check Chinese Spelling Errors Character by Character
by: Li, Kunting, et al.
Published: (2024)
by: Li, Kunting, et al.
Published: (2024)
Real, Fake, or Manipulated? Detecting Machine-Influenced Text
by: Wang, Yitong, et al.
Published: (2025)
by: Wang, Yitong, et al.
Published: (2025)
SCRAMBLe : Enhancing Multimodal LLM Compositionality with Synthetic Preference Data
by: Mishra, Samarth, et al.
Published: (2025)
by: Mishra, Samarth, et al.
Published: (2025)
RoleBreak: Character Hallucination as a Jailbreak Attack in Role-Playing Systems
by: Tang, Yihong, et al.
Published: (2024)
by: Tang, Yihong, et al.
Published: (2024)
How Does Personalized Memory Shape LLM Behavior? Benchmarking Rational Preference Utilization in Personalized Assistants
by: Feng, Xueyang, et al.
Published: (2026)
by: Feng, Xueyang, et al.
Published: (2026)
PaperHelper: Knowledge-Based LLM QA Paper Reading Assistant
by: Yin, Congrui, et al.
Published: (2025)
by: Yin, Congrui, et al.
Published: (2025)
Variation is the Key: A Variation-Based Framework for LLM-Generated Text Detection
by: Li, Xuecong, et al.
Published: (2026)
by: Li, Xuecong, et al.
Published: (2026)
Open Character Training: Shaping the Persona of AI Assistants through Constitutional AI
by: Maiya, Sharan, et al.
Published: (2025)
by: Maiya, Sharan, et al.
Published: (2025)
Mull-Tokens: Modality-Agnostic Latent Thinking
by: Ray, Arijit, et al.
Published: (2025)
by: Ray, Arijit, et al.
Published: (2025)
Improving LLM-Powered EDA Assistants with RAFT
by: Shi, Luyao, et al.
Published: (2025)
by: Shi, Luyao, et al.
Published: (2025)
Proactive Conversational Assistant for a Procedural Manual Task based on Audio and IMU
by: Mahfuz, Rehana, et al.
Published: (2026)
by: Mahfuz, Rehana, et al.
Published: (2026)
A Training-free LLM-based Approach to General Chinese Character Error Correction
by: Zhou, Houquan, et al.
Published: (2025)
by: Zhou, Houquan, et al.
Published: (2025)
CharacterBench: Benchmarking Character Customization of Large Language Models
by: Zhou, Jinfeng, et al.
Published: (2024)
by: Zhou, Jinfeng, et al.
Published: (2024)
Circuit Breaking: Removing Model Behaviors with Targeted Ablation
by: Li, Maximilian, et al.
Published: (2023)
by: Li, Maximilian, et al.
Published: (2023)
SpeLLM: Character-Level Multi-Head Decoding
by: Ben-Artzy, Amit, et al.
Published: (2025)
by: Ben-Artzy, Amit, et al.
Published: (2025)
Paraphrase-Induced Output-Mode Collapse: When LLMs Break Character Under Semantically Equivalent Inputs
by: Liu, Aofan, et al.
Published: (2026)
by: Liu, Aofan, et al.
Published: (2026)
VoiceBench: Benchmarking LLM-Based Voice Assistants
by: Chen, Yiming, et al.
Published: (2024)
by: Chen, Yiming, et al.
Published: (2024)
Personality Expression Across Contexts: Linguistic and Behavioral Variation in LLM Agents
by: Han, Bin, et al.
Published: (2026)
by: Han, Bin, et al.
Published: (2026)
PeerCoPilot: A Language Model-Powered Assistant for Behavioral Health Organizations
by: Mo, Gao, et al.
Published: (2025)
by: Mo, Gao, et al.
Published: (2025)
Spotting Out-of-Character Behavior: Atomic-Level Evaluation of Persona Fidelity in Open-Ended Generation
by: Shin, Jisu, et al.
Published: (2025)
by: Shin, Jisu, et al.
Published: (2025)
A RAG-Based Institutional Assistant
by: Kuratomi, Gustavo, et al.
Published: (2025)
by: Kuratomi, Gustavo, et al.
Published: (2025)
Quantifying the Utility of User Simulators for Building Collaborative LLM Assistants
by: Suh, Joseph, et al.
Published: (2026)
by: Suh, Joseph, et al.
Published: (2026)
Similar Items
-
From Fake to Real: Pretraining on Balanced Synthetic Images to Prevent Spurious Correlations in Image Recognition
by: Qraitem, Maan, et al.
Published: (2023) -
Web Artifact Attacks Disrupt Vision Language Models
by: Qraitem, Maan, et al.
Published: (2025) -
SLANT: Spurious Logo ANalysis Toolkit
by: Qraitem, Maan, et al.
Published: (2024) -
Vision-LLMs Can Fool Themselves with Self-Generated Typographic Attacks
by: Qraitem, Maan, et al.
Published: (2024) -
Tell Me What's Next: Textual Foresight for Generic UI Representations
by: Burns, Andrea, et al.
Published: (2024)