Enregistré dans:
| Auteur principal: | Rana, Shailesh |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | https://arxiv.org/abs/2601.08070 |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
The Shape of Wisdom: Decision Trajectories in Language Models
par: Rana, Shailesh
Publié: (2026)
par: Rana, Shailesh
Publié: (2026)
When Chain-of-Thought Backfires: Evaluating Prompt Sensitivity in Medical Language Models
par: Sadanandan, Binesh, et autres
Publié: (2026)
par: Sadanandan, Binesh, et autres
Publié: (2026)
Emulated Disalignment: Safety Alignment for Large Language Models May Backfire!
par: Zhou, Zhanhui, et autres
Publié: (2024)
par: Zhou, Zhanhui, et autres
Publié: (2024)
NCO: A Versatile Plug-in for Handling Negative Constraints in Decoding
par: Jin, Hyundong, et autres
Publié: (2026)
par: Jin, Hyundong, et autres
Publié: (2026)
Negation Triplet Extraction with Syntactic Dependency and Semantic Consistency
par: Shi, Yuchen, et autres
Publié: (2024)
par: Shi, Yuchen, et autres
Publié: (2024)
Enhancing Semantics in Multimodal Chain of Thought via Soft Negative Sampling
par: Zheng, Guangmin, et autres
Publié: (2024)
par: Zheng, Guangmin, et autres
Publié: (2024)
When Incentives Backfire, Data Stops Being Human
par: Santy, Sebastin, et autres
Publié: (2025)
par: Santy, Sebastin, et autres
Publié: (2025)
Semantic Adapter for Universal Text Embeddings: Diagnosing and Mitigating Negation Blindness to Enhance Universality
par: Cao, Hongliu
Publié: (2025)
par: Cao, Hongliu
Publié: (2025)
Alignment Backfire: Language-Dependent Reversal of Safety Interventions Across 16 Languages in LLM Multi-Agent Systems
par: Fukui, Hiroki
Publié: (2026)
par: Fukui, Hiroki
Publié: (2026)
Mitigating Semantic Leakage in Cross-lingual Embeddings via Orthogonality Constraint
par: Ki, Dayeon, et autres
Publié: (2024)
par: Ki, Dayeon, et autres
Publié: (2024)
Negating Negatives: Alignment with Human Negative Samples via Distributional Dispreference Optimization
par: Duan, Shitong, et autres
Publié: (2024)
par: Duan, Shitong, et autres
Publié: (2024)
Under Pressure: Emotional Framing Induces Measurable Behavioral Shifts and Structured Internal Geometry in Small Language Models
par: Usman, Rana Muhammad
Publié: (2026)
par: Usman, Rana Muhammad
Publié: (2026)
Which bird does not have wings: Negative-constrained KGQA with Schema-guided Semantic Matching and Self-directed Refinement
par: Shim, Midan, et autres
Publié: (2026)
par: Shim, Midan, et autres
Publié: (2026)
Semantic Anchors in In-Context Learning: Why Small LLMs Cannot Flip Their Labels
par: Kumar, Anantha Padmanaban Krishna
Publié: (2025)
par: Kumar, Anantha Padmanaban Krishna
Publié: (2025)
Semantic Integrity Constraints: Declarative Guardrails for AI-Augmented Data Processing Systems
par: Lee, Alexander W., et autres
Publié: (2025)
par: Lee, Alexander W., et autres
Publié: (2025)
A Pseudo-Semantic Loss for Autoregressive Models with Logical Constraints
par: Ahmed, Kareem, et autres
Publié: (2023)
par: Ahmed, Kareem, et autres
Publié: (2023)
Correcting Negative Bias in Large Language Models through Negative Attention Score Alignment
par: Yu, Sangwon, et autres
Publié: (2024)
par: Yu, Sangwon, et autres
Publié: (2024)
WellDunn: On the Robustness and Explainability of Language Models and Large Language Models in Identifying Wellness Dimensions
par: Mohammadi, Seyedali, et autres
Publié: (2024)
par: Mohammadi, Seyedali, et autres
Publié: (2024)
Why are LLMs' abilities emergent?
par: Havlík, Vladimír
Publié: (2025)
par: Havlík, Vladimír
Publié: (2025)
AraTable: Benchmarking LLMs' Reasoning and Understanding of Arabic Tabular Data
par: Alshaikh, Rana, et autres
Publié: (2025)
par: Alshaikh, Rana, et autres
Publié: (2025)
Why Slop Matters
par: Kommers, Cody, et autres
Publié: (2025)
par: Kommers, Cody, et autres
Publié: (2025)
Why Attend to Everything? Focus is the Key
par: Yao, Hengshuai, et autres
Publié: (2026)
par: Yao, Hengshuai, et autres
Publié: (2026)
Generating Diverse Negations from Affirmative Sentences
par: Vasquez, Darian Rodriguez, et autres
Publié: (2024)
par: Vasquez, Darian Rodriguez, et autres
Publié: (2024)
Fake Alignment: Are LLMs Really Aligned Well?
par: Wang, Yixu, et autres
Publié: (2023)
par: Wang, Yixu, et autres
Publié: (2023)
Reasoning Models Reason Well, Until They Don't
par: Rameshkumar, Revanth, et autres
Publié: (2025)
par: Rameshkumar, Revanth, et autres
Publié: (2025)
Large-Scale Constraint Generation -- Can LLMs Parse Hundreds of Constraints?
par: Boffa, Matteo, et autres
Publié: (2025)
par: Boffa, Matteo, et autres
Publié: (2025)
The Impact of Negated Text on Hallucination with Large Language Models
par: Seo, Jaehyung, et autres
Publié: (2025)
par: Seo, Jaehyung, et autres
Publié: (2025)
How Well Do LLMs Understand Tunisian Arabic?
par: Mahdi, Mohamed
Publié: (2025)
par: Mahdi, Mohamed
Publié: (2025)
Why Chain of Thought Fails in Clinical Text Understanding
par: Wu, Jiageng, et autres
Publié: (2025)
par: Wu, Jiageng, et autres
Publié: (2025)
Why is constrained neural language generation particularly challenging?
par: Garbacea, Cristina, et autres
Publié: (2022)
par: Garbacea, Cristina, et autres
Publié: (2022)
Why Braking? Scenario Extraction and Reasoning Utilizing LLM
par: Wu, Yin, et autres
Publié: (2025)
par: Wu, Yin, et autres
Publié: (2025)
Carrot and Stick: Inducing Self-Motivation with Positive & Negative Feedback
par: Sohn, Jimin, et autres
Publié: (2024)
par: Sohn, Jimin, et autres
Publié: (2024)
Mitigating the Negative Impact of Over-association for Conversational Query Production
par: Wang, Ante, et autres
Publié: (2024)
par: Wang, Ante, et autres
Publié: (2024)
Teaching with Lies: Curriculum DPO on Synthetic Negatives for Hallucination Detection
par: Pandit, Shrey, et autres
Publié: (2025)
par: Pandit, Shrey, et autres
Publié: (2025)
How Well Do Large Language Models Truly Ground?
par: Lee, Hyunji, et autres
Publié: (2023)
par: Lee, Hyunji, et autres
Publié: (2023)
Why Retrieval-Augmented Generation Fails: A Graph Perspective
par: Guo, Kai, et autres
Publié: (2026)
par: Guo, Kai, et autres
Publié: (2026)
Surgical Feature-Space Decomposition of LLMs: Why, When and How?
par: Chavan, Arnav, et autres
Publié: (2024)
par: Chavan, Arnav, et autres
Publié: (2024)
Look Within, Why LLMs Hallucinate: A Causal Perspective
par: Li, He, et autres
Publié: (2024)
par: Li, He, et autres
Publié: (2024)
MCJudgeBench: A Benchmark for Constraint-Level Judge Evaluation in Multi-Constraint Instruction Following
par: Lee, Jaeyun, et autres
Publié: (2026)
par: Lee, Jaeyun, et autres
Publié: (2026)
Towards Minimal Targeted Updates of Language Models with Targeted Negative Training
par: Zhang, Lily H., et autres
Publié: (2024)
par: Zhang, Lily H., et autres
Publié: (2024)
Documents similaires
-
The Shape of Wisdom: Decision Trajectories in Language Models
par: Rana, Shailesh
Publié: (2026) -
When Chain-of-Thought Backfires: Evaluating Prompt Sensitivity in Medical Language Models
par: Sadanandan, Binesh, et autres
Publié: (2026) -
Emulated Disalignment: Safety Alignment for Large Language Models May Backfire!
par: Zhou, Zhanhui, et autres
Publié: (2024) -
NCO: A Versatile Plug-in for Handling Negative Constraints in Decoding
par: Jin, Hyundong, et autres
Publié: (2026) -
Negation Triplet Extraction with Syntactic Dependency and Semantic Consistency
par: Shi, Yuchen, et autres
Publié: (2024)