Steering Large Language Models to Evaluate and Amplify Creativity
Fuente:
arXiv
Salvato in:
| Autori principali: | Olson, Matthew Lyle, Ratzlaff, Neale, Hinck, Musashi, Tseng, Shao-yen, Lal, Vasudev |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
LieCraft: A Multi-Agent Framework for Evaluating Deceptive Capabilities in Language Models
di: Olson, Matthew Lyle, et al.
Pubblicazione: (2026)
di: Olson, Matthew Lyle, et al.
Pubblicazione: (2026)
Probing Semantic Routing in Large Mixture-of-Expert Models
di: Olson, Matthew Lyle, et al.
Pubblicazione: (2025)
di: Olson, Matthew Lyle, et al.
Pubblicazione: (2025)
Debiasing Large Vision-Language Models by Ablating Protected Attribute Representations
di: Ratzlaff, Neale, et al.
Pubblicazione: (2024)
di: Ratzlaff, Neale, et al.
Pubblicazione: (2024)
LLaVA-Gemma: Accelerating Multimodal Foundation Models with a Compact Language Model
di: Hinck, Musashi, et al.
Pubblicazione: (2024)
di: Hinck, Musashi, et al.
Pubblicazione: (2024)
Debias your Large Multi-Modal Model at Test-Time via Non-Contrastive Visual Attribute Steering
di: Ratzlaff, Neale, et al.
Pubblicazione: (2024)
di: Ratzlaff, Neale, et al.
Pubblicazione: (2024)
Probing the Representational Power of Sparse Autoencoders in Vision Models
di: Olson, Matthew Lyle, et al.
Pubblicazione: (2025)
di: Olson, Matthew Lyle, et al.
Pubblicazione: (2025)
Analyzing Hierarchical Structure in Vision Models with Sparse Autoencoders
di: Olson, Matthew Lyle, et al.
Pubblicazione: (2025)
di: Olson, Matthew Lyle, et al.
Pubblicazione: (2025)
Why do LLaVA Vision-Language Models Reply to Images in English?
di: Hinck, Musashi, et al.
Pubblicazione: (2024)
di: Hinck, Musashi, et al.
Pubblicazione: (2024)
Training-Free Mitigation of Language Reasoning Degradation After Multimodal Instruction Tuning
di: Ratzlaff, Neale, et al.
Pubblicazione: (2024)
di: Ratzlaff, Neale, et al.
Pubblicazione: (2024)
Political Compass or Spinning Arrow? Towards More Meaningful Evaluations for Values and Opinions in Large Language Models
di: Röttger, Paul, et al.
Pubblicazione: (2024)
di: Röttger, Paul, et al.
Pubblicazione: (2024)
BILLY: Steering Large Language Models via Merging Persona Vectors for Creative Generation
di: Pai, Tsung-Min, et al.
Pubblicazione: (2025)
di: Pai, Tsung-Min, et al.
Pubblicazione: (2025)
ICSVR: Investigating Compositional and Syntactic Understanding in Video Retrieval Models
di: Madasu, Avinash, et al.
Pubblicazione: (2023)
di: Madasu, Avinash, et al.
Pubblicazione: (2023)
Cultural Awareness in Vision-Language Models: A Cross-Country Exploration
di: Madasu, Avinash, et al.
Pubblicazione: (2025)
di: Madasu, Avinash, et al.
Pubblicazione: (2025)
Divergent Creativity in Humans and Large Language Models
di: Bellemare-Pepin, Antoine, et al.
Pubblicazione: (2024)
di: Bellemare-Pepin, Antoine, et al.
Pubblicazione: (2024)
CreativityPrism: A Holistic Evaluation Framework for Large Language Model Creativity
di: Hou, Zhaoyi Joey, et al.
Pubblicazione: (2025)
di: Hou, Zhaoyi Joey, et al.
Pubblicazione: (2025)
Evaluating Creative Short Story Generation in Humans and Large Language Models
di: Ismayilzada, Mete, et al.
Pubblicazione: (2024)
di: Ismayilzada, Mete, et al.
Pubblicazione: (2024)
Steering When Necessary: Flexible Steering Large Language Models with Backtracking
di: Cheng, Zifeng, et al.
Pubblicazione: (2025)
di: Cheng, Zifeng, et al.
Pubblicazione: (2025)
ASRU: Activation Steering Meets Reinforcement Unlearning for Multimodal Large Language Models
di: Guang, Jiahui, et al.
Pubblicazione: (2026)
di: Guang, Jiahui, et al.
Pubblicazione: (2026)
On the Creativity of Large Language Models
di: Franceschelli, Giorgio, et al.
Pubblicazione: (2023)
di: Franceschelli, Giorgio, et al.
Pubblicazione: (2023)
Democratizing Diplomacy: A Harness for Evaluating Any Large Language Model on Full-Press Diplomacy
di: Duffy, Alexander, et al.
Pubblicazione: (2025)
di: Duffy, Alexander, et al.
Pubblicazione: (2025)
Automated Creativity Evaluation for Large Language Models: A Reference-Based Approach
di: Li, Ruizhe, et al.
Pubblicazione: (2025)
di: Li, Ruizhe, et al.
Pubblicazione: (2025)
Probing and Steering Evaluation Awareness of Language Models
di: Nguyen, Jord, et al.
Pubblicazione: (2025)
di: Nguyen, Jord, et al.
Pubblicazione: (2025)
Pruning the Paradox: How CLIP's Most Informative Heads Enhance Performance While Amplifying Bias
di: Madasu, Avinash, et al.
Pubblicazione: (2025)
di: Madasu, Avinash, et al.
Pubblicazione: (2025)
xGen-MM (BLIP-3): A Family of Open Large Multimodal Models
di: Xue, Le, et al.
Pubblicazione: (2024)
di: Xue, Le, et al.
Pubblicazione: (2024)
Is Your Paper Being Reviewed by an LLM? Investigating AI Text Detectability in Peer Review
di: Yu, Sungduk, et al.
Pubblicazione: (2024)
di: Yu, Sungduk, et al.
Pubblicazione: (2024)
Is Your Paper Being Reviewed by an LLM? Benchmarking AI Text Detection in Peer Review
di: Yu, Sungduk, et al.
Pubblicazione: (2025)
di: Yu, Sungduk, et al.
Pubblicazione: (2025)
Assessing and Understanding Creativity in Large Language Models
di: Zhao, Yunpu, et al.
Pubblicazione: (2024)
di: Zhao, Yunpu, et al.
Pubblicazione: (2024)
Is Temperature the Creativity Parameter of Large Language Models?
di: Peeperkorn, Max, et al.
Pubblicazione: (2024)
di: Peeperkorn, Max, et al.
Pubblicazione: (2024)
AI Steerability 360: A Toolkit for Steering Large Language Models
di: Miehling, Erik, et al.
Pubblicazione: (2026)
di: Miehling, Erik, et al.
Pubblicazione: (2026)
Compositional Steering of Large Language Models with Steering Tokens
di: Radevski, Gorjan, et al.
Pubblicazione: (2026)
di: Radevski, Gorjan, et al.
Pubblicazione: (2026)
Side-by-side Comparison Amplifies Dialect Bias in Language Models
di: Kondapally, Kritee, et al.
Pubblicazione: (2026)
di: Kondapally, Kritee, et al.
Pubblicazione: (2026)
Prompt-Based Value Steering of Large Language Models
di: Abbo, Giulio Antonio, et al.
Pubblicazione: (2025)
di: Abbo, Giulio Antonio, et al.
Pubblicazione: (2025)
CogSteer: Cognition-Inspired Selective Layer Intervention for Efficiently Steering Large Language Models
di: Wang, Xintong, et al.
Pubblicazione: (2024)
di: Wang, Xintong, et al.
Pubblicazione: (2024)
LLM Discussion: Enhancing the Creativity of Large Language Models via Discussion Framework and Role-Play
di: Lu, Li-Chun, et al.
Pubblicazione: (2024)
di: Lu, Li-Chun, et al.
Pubblicazione: (2024)
LVLM-Compress-Bench: Benchmarking the Broader Impact of Large Vision-Language Model Compression
di: Kundu, Souvik, et al.
Pubblicazione: (2025)
di: Kundu, Souvik, et al.
Pubblicazione: (2025)
Using Imperfect Surrogates for Downstream Inference: Design-based Supervised Learning for Social Science Applications of Large Language Models
di: Egami, Naoki, et al.
Pubblicazione: (2023)
di: Egami, Naoki, et al.
Pubblicazione: (2023)
Steering Evaluation-Aware Language Models to Act Like They Are Deployed
di: Hua, Tim Tian, et al.
Pubblicazione: (2025)
di: Hua, Tim Tian, et al.
Pubblicazione: (2025)
On Effects of Steering Latent Representation for Large Language Model Unlearning
di: Huu-Tien, Dang, et al.
Pubblicazione: (2024)
di: Huu-Tien, Dang, et al.
Pubblicazione: (2024)
The Effectiveness of Style Vectors for Steering Large Language Models: A Human Evaluation
di: Diallo, Diaoulé, et al.
Pubblicazione: (2026)
di: Diallo, Diaoulé, et al.
Pubblicazione: (2026)
Deep Associations, High Creativity: A Simple yet Effective Metric for Evaluating Large Language Models
di: Qiu, Ziliang, et al.
Pubblicazione: (2025)
di: Qiu, Ziliang, et al.
Pubblicazione: (2025)
Documenti analoghi
-
LieCraft: A Multi-Agent Framework for Evaluating Deceptive Capabilities in Language Models
di: Olson, Matthew Lyle, et al.
Pubblicazione: (2026) -
Probing Semantic Routing in Large Mixture-of-Expert Models
di: Olson, Matthew Lyle, et al.
Pubblicazione: (2025) -
Debiasing Large Vision-Language Models by Ablating Protected Attribute Representations
di: Ratzlaff, Neale, et al.
Pubblicazione: (2024) -
LLaVA-Gemma: Accelerating Multimodal Foundation Models with a Compact Language Model
di: Hinck, Musashi, et al.
Pubblicazione: (2024) -
Debias your Large Multi-Modal Model at Test-Time via Non-Contrastive Visual Attribute Steering
di: Ratzlaff, Neale, et al.
Pubblicazione: (2024)