Unnatural Languages Are Not Bugs but Features for LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Duan, Keyu, Zhao, Yiran, Feng, Zhili, Ni, Jinjie, Pang, Tianyu, Liu, Qian, Cai, Tianle, Dou, Longxu, Kawaguchi, Kenji, Goyal, Anirudh, Kolter, J. Zico, Shieh, Michael Qizhe |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Diffusion Language Models are Super Data Learners
von: Ni, Jinjie, et al.
Veröffentlicht: (2025)
von: Ni, Jinjie, et al.
Veröffentlicht: (2025)
Training Optimal Large Diffusion Language Models
von: Ni, Jinjie, et al.
Veröffentlicht: (2025)
von: Ni, Jinjie, et al.
Veröffentlicht: (2025)
NoisyRollout: Reinforcing Visual Reasoning with Data Augmentation
von: Liu, Xiangyan, et al.
Veröffentlicht: (2025)
von: Liu, Xiangyan, et al.
Veröffentlicht: (2025)
Efficient Process Reward Model Training via Active Learning
von: Duan, Keyu, et al.
Veröffentlicht: (2025)
von: Duan, Keyu, et al.
Veröffentlicht: (2025)
Accelerating Greedy Coordinate Gradient and General Prompt Optimization via Probe Sampling
von: Zhao, Yiran, et al.
Veröffentlicht: (2024)
von: Zhao, Yiran, et al.
Veröffentlicht: (2024)
In-Context Reinforcement Learning for Tool Use in Large Language Models
von: Ye, Yaoqi, et al.
Veröffentlicht: (2026)
von: Ye, Yaoqi, et al.
Veröffentlicht: (2026)
Reasoning Robustness of LLMs to Adversarial Typographical Errors
von: Gan, Esther, et al.
Veröffentlicht: (2024)
von: Gan, Esther, et al.
Veröffentlicht: (2024)
RAPID: Long-Context Inference with Retrieval-Augmented Speculative Decoding
von: Chen, Guanzheng, et al.
Veröffentlicht: (2025)
von: Chen, Guanzheng, et al.
Veröffentlicht: (2025)
TOFU: A Task of Fictitious Unlearning for LLMs
von: Maini, Pratyush, et al.
Veröffentlicht: (2024)
von: Maini, Pratyush, et al.
Veröffentlicht: (2024)
SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis
von: Wu, Zijian, et al.
Veröffentlicht: (2025)
von: Wu, Zijian, et al.
Veröffentlicht: (2025)
Self-Evaluation as a Defense Against Adversarial Attacks on LLMs
von: Brown, Hannah, et al.
Veröffentlicht: (2024)
von: Brown, Hannah, et al.
Veröffentlicht: (2024)
Gradually Compacting Large Language Models for Reasoning Like a Boiling Frog
von: Zhao, Yiran, et al.
Veröffentlicht: (2026)
von: Zhao, Yiran, et al.
Veröffentlicht: (2026)
T-MARS: Improving Visual Representations by Circumventing Text Feature Learning
von: Maini, Pratyush, et al.
Veröffentlicht: (2023)
von: Maini, Pratyush, et al.
Veröffentlicht: (2023)
An Axiomatic Approach to Model-Agnostic Concept Explanations
von: Feng, Zhili, et al.
Veröffentlicht: (2024)
von: Feng, Zhili, et al.
Veröffentlicht: (2024)
Predicting the Performance of Black-box LLMs through Follow-up Queries
von: Sam, Dylan, et al.
Veröffentlicht: (2025)
von: Sam, Dylan, et al.
Veröffentlicht: (2025)
Context-Parametric Inversion: Why Instruction Finetuning Can Worsen Context Reliance
von: Goyal, Sachin, et al.
Veröffentlicht: (2024)
von: Goyal, Sachin, et al.
Veröffentlicht: (2024)
Monte Carlo Tree Search Boosts Reasoning via Iterative Preference Learning
von: Xie, Yuxi, et al.
Veröffentlicht: (2024)
von: Xie, Yuxi, et al.
Veröffentlicht: (2024)
Adaptive Data Optimization: Dynamic Sample Selection with Scaling Laws
von: Jiang, Yiding, et al.
Veröffentlicht: (2024)
von: Jiang, Yiding, et al.
Veröffentlicht: (2024)
Mimetic Initialization of MLPs
von: Trockman, Asher, et al.
Veröffentlicht: (2026)
von: Trockman, Asher, et al.
Veröffentlicht: (2026)
AcceleratedLiNGAM: Learning Causal DAGs at the speed of GPUs
von: Akinwande, Victor, et al.
Veröffentlicht: (2024)
von: Akinwande, Victor, et al.
Veröffentlicht: (2024)
Rethinking LLM Memorization through the Lens of Adversarial Compression
von: Schwarzschild, Avi, et al.
Veröffentlicht: (2024)
von: Schwarzschild, Avi, et al.
Veröffentlicht: (2024)
ImageEdit-R1: Boosting Multi-Agent Image Editing via Reinforcement Learning
von: Zhao, Yiran, et al.
Veröffentlicht: (2026)
von: Zhao, Yiran, et al.
Veröffentlicht: (2026)
The Emergence of Abstract Thought in Large Language Models Beyond Any Language
von: Chen, Yuxin, et al.
Veröffentlicht: (2025)
von: Chen, Yuxin, et al.
Veröffentlicht: (2025)
Inference Optimal VLMs Need Fewer Visual Tokens and More Parameters
von: Li, Kevin Y., et al.
Veröffentlicht: (2024)
von: Li, Kevin Y., et al.
Veröffentlicht: (2024)
When Should We Introduce Safety Interventions During Pretraining?
von: Sam, Dylan, et al.
Veröffentlicht: (2026)
von: Sam, Dylan, et al.
Veröffentlicht: (2026)
FUSE-ing Language Models: Zero-Shot Adapter Discovery for Prompt Optimization Across Tokenizers
von: Williams, Joshua Nathaniel, et al.
Veröffentlicht: (2024)
von: Williams, Joshua Nathaniel, et al.
Veröffentlicht: (2024)
Equilibrium Reasoners: Learning Attractors Enables Scalable Reasoning
von: Huang, Benhao, et al.
Veröffentlicht: (2026)
von: Huang, Benhao, et al.
Veröffentlicht: (2026)
Why is SAM Robust to Label Noise?
von: Baek, Christina, et al.
Veröffentlicht: (2024)
von: Baek, Christina, et al.
Veröffentlicht: (2024)
Single Character Perturbations Break LLM Alignment
von: Lin, Leon, et al.
Veröffentlicht: (2024)
von: Lin, Leon, et al.
Veröffentlicht: (2024)
Scaling Laws for Data Filtering -- Data Curation cannot be Compute Agnostic
von: Goyal, Sachin, et al.
Veröffentlicht: (2024)
von: Goyal, Sachin, et al.
Veröffentlicht: (2024)
LongRLVR: Long-Context Reinforcement Learning Requires Verifiable Context Rewards
von: Chen, Guanzheng, et al.
Veröffentlicht: (2026)
von: Chen, Guanzheng, et al.
Veröffentlicht: (2026)
Improving Autoregressive Image Generation through Coarse-to-Fine Token Prediction
von: Guo, Ziyao, et al.
Veröffentlicht: (2025)
von: Guo, Ziyao, et al.
Veröffentlicht: (2025)
Think in Parallel, Answer as One: Logit Averaging for Open-Ended Reasoning
von: Wang, Haonan, et al.
Veröffentlicht: (2025)
von: Wang, Haonan, et al.
Veröffentlicht: (2025)
One-Step Diffusion Distillation via Deep Equilibrium Models
von: Geng, Zhengyang, et al.
Veröffentlicht: (2023)
von: Geng, Zhengyang, et al.
Veröffentlicht: (2023)
Diffusing Differentiable Representations
von: Savani, Yash, et al.
Veröffentlicht: (2024)
von: Savani, Yash, et al.
Veröffentlicht: (2024)
RegMix: Data Mixture as Regression for Language Model Pre-training
von: Liu, Qian, et al.
Veröffentlicht: (2024)
von: Liu, Qian, et al.
Veröffentlicht: (2024)
Generative Posterior Networks for Approximately Bayesian Epistemic Uncertainty Estimation
von: Roderick, Melrose, et al.
Veröffentlicht: (2023)
von: Roderick, Melrose, et al.
Veröffentlicht: (2023)
LongPO: Long Context Self-Evolution of Large Language Models through Short-to-Long Preference Optimization
von: Chen, Guanzheng, et al.
Veröffentlicht: (2025)
von: Chen, Guanzheng, et al.
Veröffentlicht: (2025)
Can AI Be as Creative as Humans?
von: Wang, Haonan, et al.
Veröffentlicht: (2024)
von: Wang, Haonan, et al.
Veröffentlicht: (2024)
Reasoning Does Not Necessarily Improve Role-Playing Ability
von: Feng, Xiachong, et al.
Veröffentlicht: (2025)
von: Feng, Xiachong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Diffusion Language Models are Super Data Learners
von: Ni, Jinjie, et al.
Veröffentlicht: (2025) -
Training Optimal Large Diffusion Language Models
von: Ni, Jinjie, et al.
Veröffentlicht: (2025) -
NoisyRollout: Reinforcing Visual Reasoning with Data Augmentation
von: Liu, Xiangyan, et al.
Veröffentlicht: (2025) -
Efficient Process Reward Model Training via Active Learning
von: Duan, Keyu, et al.
Veröffentlicht: (2025) -
Accelerating Greedy Coordinate Gradient and General Prompt Optimization via Probe Sampling
von: Zhao, Yiran, et al.
Veröffentlicht: (2024)