Salvato in:
| Autori principali: | Choi, Junhyuk, Park, Sohhyung, Cho, Chanhee, Park, Hyeonchu, Kim, Bugeun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2602.00521 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Pay What LLM Wants: Can LLM Simulate Economics Experiment with 522 Real-human Persona?
di: Choi, Junhyuk, et al.
Pubblicazione: (2025)
di: Choi, Junhyuk, et al.
Pubblicazione: (2025)
DART: An AIGT Detector using AMR of Rephrased Text
di: Park, Hyeonchu, et al.
Pubblicazione: (2024)
di: Park, Hyeonchu, et al.
Pubblicazione: (2024)
People will agree what I think: Investigating LLM's False Consensus Effect
di: Choi, Junhyuk, et al.
Pubblicazione: (2024)
di: Choi, Junhyuk, et al.
Pubblicazione: (2024)
A Stereotype Content Analysis on Color-related Social Bias in Large Vision Language Models
di: Choi, Junhyuk, et al.
Pubblicazione: (2025)
di: Choi, Junhyuk, et al.
Pubblicazione: (2025)
Belief in Authority: Impact of Authority in Multi-Agent Evaluation Framework
di: Choi, Junhyuk, et al.
Pubblicazione: (2026)
di: Choi, Junhyuk, et al.
Pubblicazione: (2026)
PHISH in MESH: Korean Adversarial Phonetic Substitution and Phonetic-Semantic Feature Integration Defense
di: Kim, Byungjun, et al.
Pubblicazione: (2025)
di: Kim, Byungjun, et al.
Pubblicazione: (2025)
Acoustic-based Gender Differentiation in Speech-aware Language Models
di: Choi, Junhyuk, et al.
Pubblicazione: (2025)
di: Choi, Junhyuk, et al.
Pubblicazione: (2025)
KITE: A Benchmark for Evaluating Korean Instruction-Following Abilities in Large Language Models
di: Kim, Dongjun, et al.
Pubblicazione: (2025)
di: Kim, Dongjun, et al.
Pubblicazione: (2025)
SPAM: Style Prompt Adherence Metric for Prompt-based TTS
di: Cho, Chanhee, et al.
Pubblicazione: (2026)
di: Cho, Chanhee, et al.
Pubblicazione: (2026)
MIRAGE: A Metric-Intensive Benchmark for Retrieval-Augmented Generation Evaluation
di: Park, Chanhee, et al.
Pubblicazione: (2025)
di: Park, Chanhee, et al.
Pubblicazione: (2025)
Examining Identity Drift in Conversations of LLM Agents
di: Choi, Junhyuk, et al.
Pubblicazione: (2024)
di: Choi, Junhyuk, et al.
Pubblicazione: (2024)
Diagnosing LLM Judge Reliability: Conformal Prediction Sets and Transitivity Violations
di: Gupta, Manan, et al.
Pubblicazione: (2026)
di: Gupta, Manan, et al.
Pubblicazione: (2026)
Learning Compact Representations of LLM Abilities via Item Response Theory
di: Chen, Jianhao, et al.
Pubblicazione: (2025)
di: Chen, Jianhao, et al.
Pubblicazione: (2025)
Judge Reliability Harness: Stress Testing the Reliability of LLM Judges
di: Dev, Sunishchal, et al.
Pubblicazione: (2026)
di: Dev, Sunishchal, et al.
Pubblicazione: (2026)
Token-Efficient Item Representation via Images for LLM Recommender Systems
di: Kim, Kibum, et al.
Pubblicazione: (2025)
di: Kim, Kibum, et al.
Pubblicazione: (2025)
Beyond Task-Oriented and Chitchat Dialogues: Proactive and Transition-Aware Conversational Agents
di: Yoon, Yejin, et al.
Pubblicazione: (2025)
di: Yoon, Yejin, et al.
Pubblicazione: (2025)
IRT-Router: Effective and Interpretable Multi-LLM Routing via Item Response Theory
di: Song, Wei, et al.
Pubblicazione: (2025)
di: Song, Wei, et al.
Pubblicazione: (2025)
Analysis of Utterance Embeddings and Clustering Methods Related to Intent Induction for Task-Oriented Dialogue
di: Park, Jeiyoon, et al.
Pubblicazione: (2022)
di: Park, Jeiyoon, et al.
Pubblicazione: (2022)
ARGORA: Orchestrated Argumentation for Causally Grounded LLM Reasoning and Decision Making
di: Jin, Youngjin, et al.
Pubblicazione: (2026)
di: Jin, Youngjin, et al.
Pubblicazione: (2026)
FRDiff : Feature Reuse for Universal Training-free Acceleration of Diffusion Models
di: So, Junhyuk, et al.
Pubblicazione: (2023)
di: So, Junhyuk, et al.
Pubblicazione: (2023)
LLMServingSim2.0: A Unified Simulator for Heterogeneous Hardware and Serving Techniques in LLM Infrastructure
di: Cho, Jaehong, et al.
Pubblicazione: (2025)
di: Cho, Jaehong, et al.
Pubblicazione: (2025)
Fine-Grained and Thematic Evaluation of LLMs in Social Deduction Game
di: Kim, Byungjun, et al.
Pubblicazione: (2024)
di: Kim, Byungjun, et al.
Pubblicazione: (2024)
Leveraging Large Language Models for Active Merchant Non-player Characters
di: Kim, Byungjun, et al.
Pubblicazione: (2024)
di: Kim, Byungjun, et al.
Pubblicazione: (2024)
Judgment-of-Thought Prompting: A Courtroom-Inspired Framework for Binary Logical Reasoning with Large Language Models
di: Park, Sungjune, et al.
Pubblicazione: (2024)
di: Park, Sungjune, et al.
Pubblicazione: (2024)
JudgeFlow: Agentic Workflow Optimization via Block Judge
di: Ma, Zihan, et al.
Pubblicazione: (2026)
di: Ma, Zihan, et al.
Pubblicazione: (2026)
Can LLMs and humans be friends? Uncovering factors affecting human-AI intimacy formation
di: Hong, Yeseon, et al.
Pubblicazione: (2025)
di: Hong, Yeseon, et al.
Pubblicazione: (2025)
Towards Trustworthy LLM-Based Recommendation via Rationale Integration
di: Park, Chung, et al.
Pubblicazione: (2025)
di: Park, Chung, et al.
Pubblicazione: (2025)
Towards Privacy-Preserving Large Language Model: Text-free Inference Through Alignment and Adaptation
di: Yoon, Jeongho, et al.
Pubblicazione: (2026)
di: Yoon, Jeongho, et al.
Pubblicazione: (2026)
Evidential Transformation Network: Turning Pretrained Models into Evidential Models for Post-hoc Uncertainty Estimation
di: Chun, Yongchan, et al.
Pubblicazione: (2026)
di: Chun, Yongchan, et al.
Pubblicazione: (2026)
LLMServingSim: A HW/SW Co-Simulation Infrastructure for LLM Inference Serving at Scale
di: Cho, Jaehong, et al.
Pubblicazione: (2024)
di: Cho, Jaehong, et al.
Pubblicazione: (2024)
LLM-guided Plan and Retrieval: A Strategic Alignment for Interpretable User Satisfaction Estimation in Dialogue
di: Kim, Sangyeop, et al.
Pubblicazione: (2025)
di: Kim, Sangyeop, et al.
Pubblicazione: (2025)
VorTEX: Various overlap ratio for Target speech EXtraction
di: Oh, Ro-hoon, et al.
Pubblicazione: (2026)
di: Oh, Ro-hoon, et al.
Pubblicazione: (2026)
Self-HarmLLM: Can Large Language Model Harm Itself?
di: Kim, Heehwan, et al.
Pubblicazione: (2025)
di: Kim, Heehwan, et al.
Pubblicazione: (2025)
AIS-LLM: A Unified Framework for Maritime Trajectory Prediction, Anomaly Detection, and Collision Risk Assessment with Explainable Forecasting
di: Park, Hyobin, et al.
Pubblicazione: (2025)
di: Park, Hyobin, et al.
Pubblicazione: (2025)
VIRO: Robust and Efficient Neuro-Symbolic Reasoning with Verification for Referring Expression Comprehension
di: Park, Hyejin, et al.
Pubblicazione: (2026)
di: Park, Hyejin, et al.
Pubblicazione: (2026)
Expanding Search Space with Diverse Prompting Agents: An Efficient Sampling Approach for LLM Mathematical Reasoning
di: Lee, Gisang, et al.
Pubblicazione: (2024)
di: Lee, Gisang, et al.
Pubblicazione: (2024)
LLMServingSim 2.0: A Unified Simulator for Heterogeneous and Disaggregated LLM Serving Infrastructure
di: Cho, Jaehong, et al.
Pubblicazione: (2026)
di: Cho, Jaehong, et al.
Pubblicazione: (2026)
MMTB: Evaluating Terminal Agents on Multimedia-File Tasks
di: Heo, Chiyeong, et al.
Pubblicazione: (2026)
di: Heo, Chiyeong, et al.
Pubblicazione: (2026)
TRUEBench: Can LLM Response Meet Real-world Constraints as Productivity Assistant?
di: Park, Jiho, et al.
Pubblicazione: (2025)
di: Park, Jiho, et al.
Pubblicazione: (2025)
PsyProbe: Proactive and Interpretable Dialogue through User State Modeling for Exploratory Counseling
di: Park, Sohhyung, et al.
Pubblicazione: (2026)
di: Park, Sohhyung, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Pay What LLM Wants: Can LLM Simulate Economics Experiment with 522 Real-human Persona?
di: Choi, Junhyuk, et al.
Pubblicazione: (2025) -
DART: An AIGT Detector using AMR of Rephrased Text
di: Park, Hyeonchu, et al.
Pubblicazione: (2024) -
People will agree what I think: Investigating LLM's False Consensus Effect
di: Choi, Junhyuk, et al.
Pubblicazione: (2024) -
A Stereotype Content Analysis on Color-related Social Bias in Large Vision Language Models
di: Choi, Junhyuk, et al.
Pubblicazione: (2025) -
Belief in Authority: Impact of Authority in Multi-Agent Evaluation Framework
di: Choi, Junhyuk, et al.
Pubblicazione: (2026)