Probing the Limits of Stylistic Alignment in Vision-Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Farajidizaji, Asma, Gupta, Akash, Raina, Vatsal |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Is it Possible to Modify Text to a Target Readability Level? An Initial Investigation Using Zero-Shot Large Language Models
von: Farajidizaji, Asma, et al.
Veröffentlicht: (2023)
von: Farajidizaji, Asma, et al.
Veröffentlicht: (2023)
Question Difficulty Ranking for Multiple-Choice Reading Comprehension
von: Raina, Vatsal, et al.
Veröffentlicht: (2024)
von: Raina, Vatsal, et al.
Veröffentlicht: (2024)
An Information-Theoretic Approach to Analyze NLP Classification Tasks
von: Wang, Luran, et al.
Veröffentlicht: (2024)
von: Wang, Luran, et al.
Veröffentlicht: (2024)
NTSEBENCH: Cognitive Reasoning Benchmark for Vision Language Models
von: Pandya, Pranshu, et al.
Veröffentlicht: (2024)
von: Pandya, Pranshu, et al.
Veröffentlicht: (2024)
A Survey of Prompt Engineering Methods in Large Language Models for Different NLP Tasks
von: Vatsal, Shubham, et al.
Veröffentlicht: (2024)
von: Vatsal, Shubham, et al.
Veröffentlicht: (2024)
On the Limitations of Steering in Language Model Alignment
von: Niranjan, Chebrolu, et al.
Veröffentlicht: (2025)
von: Niranjan, Chebrolu, et al.
Veröffentlicht: (2025)
Evaluating Concurrent Robustness of Language Models Across Diverse Challenge Sets
von: Gupta, Vatsal, et al.
Veröffentlicht: (2023)
von: Gupta, Vatsal, et al.
Veröffentlicht: (2023)
Style over Substance: Distilled Language Models Reason Via Stylistic Replication
von: Lippmann, Philip, et al.
Veröffentlicht: (2025)
von: Lippmann, Philip, et al.
Veröffentlicht: (2025)
Fundamental Limitations of Alignment in Large Language Models
von: Wolf, Yotam, et al.
Veröffentlicht: (2023)
von: Wolf, Yotam, et al.
Veröffentlicht: (2023)
Polarity-Aware Probing for Quantifying Latent Alignment in Language Models
von: Sadiekh, Sabrina, et al.
Veröffentlicht: (2025)
von: Sadiekh, Sabrina, et al.
Veröffentlicht: (2025)
A Role-specific Guided Large Language Model for Ophthalmic Consultation Based on Stylistic Differentiation
von: Fu, Laiyi, et al.
Veröffentlicht: (2024)
von: Fu, Laiyi, et al.
Veröffentlicht: (2024)
Adversarial Humanities Benchmark: Results on Stylistic Robustness in Frontier Model Safety
von: Galisai, Marcello, et al.
Veröffentlicht: (2026)
von: Galisai, Marcello, et al.
Veröffentlicht: (2026)
Value Augmented Sampling for Language Model Alignment and Personalization
von: Han, Seungwook, et al.
Veröffentlicht: (2024)
von: Han, Seungwook, et al.
Veröffentlicht: (2024)
VAL-Bench: Belief Consistency as a measure for Value Alignment in Language Models
von: Gupta, Aman, et al.
Veröffentlicht: (2025)
von: Gupta, Aman, et al.
Veröffentlicht: (2025)
Multilingual Prompt Engineering in Large Language Models: A Survey Across NLP Tasks
von: Vatsal, Shubham, et al.
Veröffentlicht: (2025)
von: Vatsal, Shubham, et al.
Veröffentlicht: (2025)
Panoramic Interests: Stylistic-Content Aware Personalized Headline Generation
von: Lian, Junhong, et al.
Veröffentlicht: (2025)
von: Lian, Junhong, et al.
Veröffentlicht: (2025)
MENTIS: What Belief Changes Under Alignment? Measuring Multi-Scale Latent Torsion in Language Models
von: Saha, Partha Pratim, et al.
Veröffentlicht: (2026)
von: Saha, Partha Pratim, et al.
Veröffentlicht: (2026)
RLHF: A comprehensive Survey for Cultural, Multimodal and Low Latency Alignment Methods
von: Sharma, Raghav, et al.
Veröffentlicht: (2025)
von: Sharma, Raghav, et al.
Veröffentlicht: (2025)
DeAL: Decoding-time Alignment for Large Language Models
von: Huang, James Y., et al.
Veröffentlicht: (2024)
von: Huang, James Y., et al.
Veröffentlicht: (2024)
Emotion Classification in Low and Moderate Resource Languages
von: Tafreshi, Shabnam, et al.
Veröffentlicht: (2024)
von: Tafreshi, Shabnam, et al.
Veröffentlicht: (2024)
StyleDecipher: Robust and Explainable Detection of LLM-Generated Texts with Stylistic Analysis
von: Li, Siyuan, et al.
Veröffentlicht: (2025)
von: Li, Siyuan, et al.
Veröffentlicht: (2025)
Mitigating Stylistic Biases of Machine Translation Systems via Monolingual Corpora Only
von: Gao, Xuanqi, et al.
Veröffentlicht: (2025)
von: Gao, Xuanqi, et al.
Veröffentlicht: (2025)
Probing and Inducing Combinational Creativity in Vision-Language Models
von: Peng, Yongqian, et al.
Veröffentlicht: (2025)
von: Peng, Yongqian, et al.
Veröffentlicht: (2025)
On the Robustness of Reward Models for Language Model Alignment
von: Hong, Jiwoo, et al.
Veröffentlicht: (2025)
von: Hong, Jiwoo, et al.
Veröffentlicht: (2025)
The Limited Impact of Medical Adaptation of Large Language and Vision-Language Models
von: Jeong, Daniel P., et al.
Veröffentlicht: (2024)
von: Jeong, Daniel P., et al.
Veröffentlicht: (2024)
Unraveling and Mitigating Safety Alignment Degradation of Vision-Language Models
von: Liu, Qin, et al.
Veröffentlicht: (2024)
von: Liu, Qin, et al.
Veröffentlicht: (2024)
Safe Inputs but Unsafe Output: Benchmarking Cross-modality Safety Alignment of Large Vision-Language Model
von: Wang, Siyin, et al.
Veröffentlicht: (2024)
von: Wang, Siyin, et al.
Veröffentlicht: (2024)
Probing for Arithmetic Errors in Language Models
von: Sun, Yucheng, et al.
Veröffentlicht: (2025)
von: Sun, Yucheng, et al.
Veröffentlicht: (2025)
Convergent Evolution: How Different Language Models Learn Similar Number Representations
von: Fu, Deqing, et al.
Veröffentlicht: (2026)
von: Fu, Deqing, et al.
Veröffentlicht: (2026)
Languages are Modalities: Cross-Lingual Alignment via Encoder Injection
von: Agarwal, Rajan, et al.
Veröffentlicht: (2025)
von: Agarwal, Rajan, et al.
Veröffentlicht: (2025)
Exploring the Frontier of Vision-Language Models: A Survey of Current Methodologies and Future Directions
von: Ghosh, Akash, et al.
Veröffentlicht: (2024)
von: Ghosh, Akash, et al.
Veröffentlicht: (2024)
Probing and Steering Evaluation Awareness of Language Models
von: Nguyen, Jord, et al.
Veröffentlicht: (2025)
von: Nguyen, Jord, et al.
Veröffentlicht: (2025)
Probing Neural Topology of Large Language Models
von: Zheng, Yu, et al.
Veröffentlicht: (2025)
von: Zheng, Yu, et al.
Veröffentlicht: (2025)
Probing Causality Manipulation of Large Language Models
von: Zhang, Chenyang, et al.
Veröffentlicht: (2024)
von: Zhang, Chenyang, et al.
Veröffentlicht: (2024)
Probing Persona-Dependent Preferences in Language Models
von: Gilg, Oscar, et al.
Veröffentlicht: (2026)
von: Gilg, Oscar, et al.
Veröffentlicht: (2026)
Can GPT Redefine Medical Understanding? Evaluating GPT on Biomedical Machine Reading Comprehension
von: Vatsal, Shubham, et al.
Veröffentlicht: (2024)
von: Vatsal, Shubham, et al.
Veröffentlicht: (2024)
Benchmarking Distributional Alignment of Large Language Models
von: Meister, Nicole, et al.
Veröffentlicht: (2024)
von: Meister, Nicole, et al.
Veröffentlicht: (2024)
Latent Concept Disentanglement in Transformer-based Language Models
von: Hong, Guan Zhe, et al.
Veröffentlicht: (2025)
von: Hong, Guan Zhe, et al.
Veröffentlicht: (2025)
MoSEs: Uncertainty-Aware AI-Generated Text Detection via Mixture of Stylistics Experts with Conditional Thresholds
von: Wu, Junxi, et al.
Veröffentlicht: (2025)
von: Wu, Junxi, et al.
Veröffentlicht: (2025)
Probing the Difficulty Perception Mechanism of Large Language Models
von: Lee, Sunbowen, et al.
Veröffentlicht: (2025)
von: Lee, Sunbowen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Is it Possible to Modify Text to a Target Readability Level? An Initial Investigation Using Zero-Shot Large Language Models
von: Farajidizaji, Asma, et al.
Veröffentlicht: (2023) -
Question Difficulty Ranking for Multiple-Choice Reading Comprehension
von: Raina, Vatsal, et al.
Veröffentlicht: (2024) -
An Information-Theoretic Approach to Analyze NLP Classification Tasks
von: Wang, Luran, et al.
Veröffentlicht: (2024) -
NTSEBENCH: Cognitive Reasoning Benchmark for Vision Language Models
von: Pandya, Pranshu, et al.
Veröffentlicht: (2024) -
A Survey of Prompt Engineering Methods in Large Language Models for Different NLP Tasks
von: Vatsal, Shubham, et al.
Veröffentlicht: (2024)