Enregistré dans:
| Auteurs principaux: | Feiglin, Erin, Hutnik, Nir, Lapid, Raz |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | https://arxiv.org/abs/2601.08490 |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Activation Steering for Masked Diffusion Language Models
par: Shnaidman, Adi, et autres
Publié: (2025)
par: Shnaidman, Adi, et autres
Publié: (2025)
SastBench: A Benchmark for Testing Agentic SAST Triage
par: Feiglin, Jake, et autres
Publié: (2026)
par: Feiglin, Jake, et autres
Publié: (2026)
Few-shot Name Entity Recognition on StackOverflow
par: Chen, Xinwei, et autres
Publié: (2024)
par: Chen, Xinwei, et autres
Publié: (2024)
Can ChatGPT replace StackOverflow? A Study on Robustness and Reliability of Large Language Model Code Generation
par: Zhong, Li, et autres
Publié: (2023)
par: Zhong, Li, et autres
Publié: (2023)
Evaluating Privacy Questions From Stack Overflow: Can ChatGPT Compete?
par: Delile, Zack, et autres
Publié: (2023)
par: Delile, Zack, et autres
Publié: (2023)
CardiffNLP at CLEARS-2025: Prompting Large Language Models for Plain Language and Easy-to-Read Text Rewriting
par: Ayesh, Mutaz, et autres
Publié: (2025)
par: Ayesh, Mutaz, et autres
Publié: (2025)
Backdoors in Conditional Diffusion: Threats to Responsible Synthetic Data Pipelines
par: Lapid, Raz, et autres
Publié: (2025)
par: Lapid, Raz, et autres
Publié: (2025)
Open Sesame! Universal Black Box Jailbreaking of Large Language Models
par: Lapid, Raz, et autres
Publié: (2023)
par: Lapid, Raz, et autres
Publié: (2023)
Breaking Audio Large Language Models by Attacking Only the Encoder: A Universal Targeted Latent-Space Audio Attack
par: Ziv, Roee, et autres
Publié: (2025)
par: Ziv, Roee, et autres
Publié: (2025)
PromptBench: A Unified Library for Evaluation of Large Language Models
par: Zhu, Kaijie, et autres
Publié: (2023)
par: Zhu, Kaijie, et autres
Publié: (2023)
VLind-Bench: Measuring Language Priors in Large Vision-Language Models
par: Lee, Kang-il, et autres
Publié: (2024)
par: Lee, Kang-il, et autres
Publié: (2024)
On the Robustness of Diffusion-Based Image Compression to Bit-Flip Errors
par: Vaisman, Amit, et autres
Publié: (2026)
par: Vaisman, Amit, et autres
Publié: (2026)
Fortify the Guardian, Not the Treasure: Resilient Adversarial Detectors
par: Lapid, Raz, et autres
Publié: (2024)
par: Lapid, Raz, et autres
Publié: (2024)
Patch of Invisibility: Naturalistic Physical Black-Box Adversarial Attacks on Object Detectors
par: Lapid, Raz, et autres
Publié: (2023)
par: Lapid, Raz, et autres
Publié: (2023)
DETAIL Matters: Measuring the Impact of Prompt Specificity on Reasoning in Large Language Models
par: Kim, Olivia
Publié: (2025)
par: Kim, Olivia
Publié: (2025)
Revisiting Prompt Sensitivity in Large Language Models for Text Classification: The Role of Prompt Underspecification
par: Pecher, Branislav, et autres
Publié: (2026)
par: Pecher, Branislav, et autres
Publié: (2026)
An Evaluation of Large Language Models on Text Summarization Tasks Using Prompt Engineering Techniques
par: Aly, Walid Mohamed, et autres
Publié: (2025)
par: Aly, Walid Mohamed, et autres
Publié: (2025)
T2S-Bench & Structure-of-Thought: Benchmarking and Prompting Comprehensive Text-to-Structure Reasoning
par: Wang, Qinsi, et autres
Publié: (2026)
par: Wang, Qinsi, et autres
Publié: (2026)
Are Large Language Models a Threat to Digital Public Goods? Evidence from Activity on Stack Overflow
par: del Rio-Chanona, Maria, et autres
Publié: (2023)
par: del Rio-Chanona, Maria, et autres
Publié: (2023)
MEMO-Bench: A Multiple Benchmark for Text-to-Image and Multimodal Large Language Models on Human Emotion Analysis
par: Zhou, Yingjie, et autres
Publié: (2024)
par: Zhou, Yingjie, et autres
Publié: (2024)
Text Adaptation to Plain Language and Easy Read via Automatic Post-Editing Cycles
par: Calleja, Jesús, et autres
Publié: (2025)
par: Calleja, Jesús, et autres
Publié: (2025)
Prompt Selection Matters: Enhancing Text Annotations for Social Sciences with Large Language Models
par: Abraham, Louis, et autres
Publié: (2024)
par: Abraham, Louis, et autres
Publié: (2024)
OR-Bench: An Over-Refusal Benchmark for Large Language Models
par: Cui, Justin, et autres
Publié: (2024)
par: Cui, Justin, et autres
Publié: (2024)
PromptBridge: Cross-Model Prompt Transfer for Large Language Models
par: Wang, Yaxuan, et autres
Publié: (2025)
par: Wang, Yaxuan, et autres
Publié: (2025)
Integrating Chemistry Knowledge in Large Language Models via Prompt Engineering
par: Liu, Hongxuan, et autres
Publié: (2024)
par: Liu, Hongxuan, et autres
Publié: (2024)
A Universal Prompting Strategy for Extracting Process Model Information from Natural Language Text using Large Language Models
par: Neuberger, Julian, et autres
Publié: (2024)
par: Neuberger, Julian, et autres
Publié: (2024)
Is Stack Overflow Obsolete? An Empirical Study of the Characteristics of ChatGPT Answers to Stack Overflow Questions
par: Kabir, Samia, et autres
Publié: (2023)
par: Kabir, Samia, et autres
Publié: (2023)
TaskBench: Benchmarking Large Language Models for Task Automation
par: Shen, Yongliang, et autres
Publié: (2023)
par: Shen, Yongliang, et autres
Publié: (2023)
Protecting Privacy in Multimodal Large Language Models with MLLMU-Bench
par: Liu, Zheyuan, et autres
Publié: (2024)
par: Liu, Zheyuan, et autres
Publié: (2024)
EmoBench: Evaluating the Emotional Intelligence of Large Language Models
par: Sabour, Sahand, et autres
Publié: (2024)
par: Sabour, Sahand, et autres
Publié: (2024)
Debiasing Large Language Models via Adaptive Causal Prompting with Sketch-of-Thought
par: Li, Bowen, et autres
Publié: (2026)
par: Li, Bowen, et autres
Publié: (2026)
Are Large Language Models Good Prompt Optimizers?
par: Ma, Ruotian, et autres
Publié: (2024)
par: Ma, Ruotian, et autres
Publié: (2024)
Large Language Models Prompting With Episodic Memory
par: Do, Dai, et autres
Publié: (2024)
par: Do, Dai, et autres
Publié: (2024)
On the Worst Prompt Performance of Large Language Models
par: Cao, Bowen, et autres
Publié: (2024)
par: Cao, Bowen, et autres
Publié: (2024)
ReadBench: Measuring the Dense Text Visual Reading Ability of Vision-Language Models
par: Clavié, Benjamin, et autres
Publié: (2025)
par: Clavié, Benjamin, et autres
Publié: (2025)
KnowledgePrompts: Exploring the Abilities of Large Language Models to Solve Proportional Analogies via Knowledge-Enhanced Prompting
par: Wijesiriwardene, Thilini, et autres
Publié: (2024)
par: Wijesiriwardene, Thilini, et autres
Publié: (2024)
Robustness of Prompting: Enhancing Robustness of Large Language Models Against Prompting Attacks
par: Mu, Lin, et autres
Publié: (2025)
par: Mu, Lin, et autres
Publié: (2025)
Diverse Prompts: Illuminating the Prompt Space of Large Language Models with MAP-Elites
par: Santos, Gabriel Machado, et autres
Publié: (2025)
par: Santos, Gabriel Machado, et autres
Publié: (2025)
TurkBench: A Benchmark for Evaluating Turkish Large Language Models
par: Toraman, Çağrı, et autres
Publié: (2026)
par: Toraman, Çağrı, et autres
Publié: (2026)
BrainBench: Exposing the Commonsense Reasoning Gap in Large Language Models
par: Tang, Yuzhe
Publié: (2026)
par: Tang, Yuzhe
Publié: (2026)
Documents similaires
-
Activation Steering for Masked Diffusion Language Models
par: Shnaidman, Adi, et autres
Publié: (2025) -
SastBench: A Benchmark for Testing Agentic SAST Triage
par: Feiglin, Jake, et autres
Publié: (2026) -
Few-shot Name Entity Recognition on StackOverflow
par: Chen, Xinwei, et autres
Publié: (2024) -
Can ChatGPT replace StackOverflow? A Study on Robustness and Reliability of Large Language Model Code Generation
par: Zhong, Li, et autres
Publié: (2023) -
Evaluating Privacy Questions From Stack Overflow: Can ChatGPT Compete?
par: Delile, Zack, et autres
Publié: (2023)