IDGen: Item Discrimination Induced Prompt Generation for LLM Evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Lin, Fan, Xie, Shuyi, Dai, Yong, Yao, Wenlin, Lang, Tianjiao, Xu, Zishan, Hu, Zhichao, Xiao, Xiao, Liu, Yuhong, Zhang, Yu |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Diagnosing Failures in Large Language Models' Answers: Integrating Error Attribution into Evaluation Framework
by: Xu, Zishan, et al.
Published: (2025)
by: Xu, Zishan, et al.
Published: (2025)
IDGenRec: LLM-RecSys Alignment with Textual ID Learning
by: Tan, Juntao, et al.
Published: (2024)
by: Tan, Juntao, et al.
Published: (2024)
HDFlow: Enhancing LLM Complex Problem-Solving with Hybrid Thinking and Dynamic Workflows
by: Yao, Wenlin, et al.
Published: (2024)
by: Yao, Wenlin, et al.
Published: (2024)
Insights Into Solvent‐Polarity‐Related Photo‐Induced Excited State Behaviors for H2BP‐(OH)2DC Compound: A Theoretical Study
by: Junping Xiao, et al.
Published: (2025)
by: Junping Xiao, et al.
Published: (2025)
Do Physicians Know How to Prompt? The Need for Automatic Prompt Optimization Help in Clinical Note Generation
by: Yao, Zonghai, et al.
Published: (2023)
by: Yao, Zonghai, et al.
Published: (2023)
Numerical approximation based on deep convolutional neural network for high‐dimensional fully nonlinear merged PDEs and 2BSDEs
by: Xu Xiao, et al.
Published: (2024)
by: Xu Xiao, et al.
Published: (2024)
LongT2IBench: A Benchmark for Evaluating Long Text-to-Image Generation with Graph-structured Annotations
by: Yang, Zhichao, et al.
Published: (2025)
by: Yang, Zhichao, et al.
Published: (2025)
Generative Students: Using LLM-Simulated Student Profiles to Support Question Item Evaluation
by: Lu, Xinyi, et al.
Published: (2024)
by: Lu, Xinyi, et al.
Published: (2024)
Skill-Conditioned Gated Self-Distillation for LLM Reasoning
by: Huang, Jiazhen, et al.
Published: (2026)
by: Huang, Jiazhen, et al.
Published: (2026)
The Impact of Item-Writing Flaws on Difficulty and Discrimination in Item Response Theory
by: Schmucker, Robin, et al.
Published: (2025)
by: Schmucker, Robin, et al.
Published: (2025)
GenRecEdit: Adapting Model Editing for Generative Recommendation with Cold-Start Items
by: Shen, Chenglei, et al.
Published: (2026)
by: Shen, Chenglei, et al.
Published: (2026)
LLM-I2I: Boost Your Small Item2Item Recommendation Model with Large Language Model
by: Feng, Yinfu, et al.
Published: (2025)
by: Feng, Yinfu, et al.
Published: (2025)
Neural Retrievers are Biased Towards LLM-Generated Content
by: Dai, Sunhao, et al.
Published: (2023)
by: Dai, Sunhao, et al.
Published: (2023)
Prompt-Induced Linguistic Fingerprints for LLM-Generated Fake News Detection
by: Wang, Chi, et al.
Published: (2025)
by: Wang, Chi, et al.
Published: (2025)
An Item is Worth a Prompt: Versatile Image Editing with Disentangled Control
by: Feng, Aosong, et al.
Published: (2024)
by: Feng, Aosong, et al.
Published: (2024)
Efficient Prompting for LLM-based Generative Internet of Things
by: Xiao, Bin, et al.
Published: (2024)
by: Xiao, Bin, et al.
Published: (2024)
The Fabrication of Reality and Fantasy: Scene Generation with LLM-Assisted Prompt Interpretation
by: Yao, Yi, et al.
Published: (2024)
by: Yao, Yi, et al.
Published: (2024)
An Automatic Pathway Searching Strategy in Enzyme Catalysis: A Case Study of Lm CpfC
by: Yuhong Lin, et al.
Published: (2025)
by: Yuhong Lin, et al.
Published: (2025)
Tokenize Once, Recommend Anywhere: Unified Item Tokenization for Multi-domain LLM-based Recommendation
by: Hou, Yu, et al.
Published: (2025)
by: Hou, Yu, et al.
Published: (2025)
Finite Difference Method for Nonlinear Damped Viscoelastic Euler‐Bernoulli Beam Model
by: Wenlin Qiu, et al.
Published: (2025)
by: Wenlin Qiu, et al.
Published: (2025)
Finite difference method for nonlinear damped viscoelastic Euler-Bernoulli beam model
by: Qiu, Wenlin, et al.
Published: (2025)
by: Qiu, Wenlin, et al.
Published: (2025)
AI Evaluation Should Require Standardized Item-Level Data Releases
by: Jiang, Han, et al.
Published: (2026)
by: Jiang, Han, et al.
Published: (2026)
Unveiling LLM Evaluation Focused on Metrics: Challenges and Solutions
by: Hu, Taojun, et al.
Published: (2024)
by: Hu, Taojun, et al.
Published: (2024)
MVR-cache: Optimizing Semantic Caching via Multi-Vector Retrieval and Learned Prompt Segmentation
by: Noshad, Ali, et al.
Published: (2026)
by: Noshad, Ali, et al.
Published: (2026)
Prompt Tuning for Item Cold-start Recommendation
by: Jiang, Yuezihan, et al.
Published: (2024)
by: Jiang, Yuezihan, et al.
Published: (2024)
ItemRAG: Item-Based Retrieval-Augmented Generation for LLM-Based Recommendation
by: Kim, Sunwoo, et al.
Published: (2025)
by: Kim, Sunwoo, et al.
Published: (2025)
DPIC: Decoupling Prompt and Intrinsic Characteristics for LLM Generated Text Detection
by: Yu, Xiao, et al.
Published: (2023)
by: Yu, Xiao, et al.
Published: (2023)
A Versatile Multimodal Agent for Multimedia Content Generation
by: Zhang, Daoan, et al.
Published: (2026)
by: Zhang, Daoan, et al.
Published: (2026)
CAPO: Towards Enhancing LLM Reasoning through Generative Credit Assignment
by: Xie, Guofu, et al.
Published: (2025)
by: Xie, Guofu, et al.
Published: (2025)
Control at Stake: Evaluating the Security Landscape of LLM-Driven Email Agents
by: Wu, Jiangrong, et al.
Published: (2025)
by: Wu, Jiangrong, et al.
Published: (2025)
Learning Multi-Aspect Item Palette: A Semantic Tokenization Framework for Generative Recommendation
by: Liu, Qijiong, et al.
Published: (2024)
by: Liu, Qijiong, et al.
Published: (2024)
LLM-Empowered Representation Learning for Emerging Item Recommendation
by: Zhang, Ziying, et al.
Published: (2025)
by: Zhang, Ziying, et al.
Published: (2025)
CLAVE: An Adaptive Framework for Evaluating Values of LLM Generated Responses
by: Yao, Jing, et al.
Published: (2024)
by: Yao, Jing, et al.
Published: (2024)
IR-Flow: Bridging Discriminative and Generative Image Restoration via Rectified Flow
by: Fan, Zihao, et al.
Published: (2026)
by: Fan, Zihao, et al.
Published: (2026)
Geometric Understanding of Discriminability and Transferability for Visual Domain Adaptation
by: Luo, You-Wei, et al.
Published: (2024)
by: Luo, You-Wei, et al.
Published: (2024)
Computational Explorations About Photoinduced Behaviors for 2‐(2‐hydroxyphenyl)benzoxazole Derivatives by Substituted Alkyl Groups: A TDDFT Study
by: Junping Xiao, et al.
Published: (2025)
by: Junping Xiao, et al.
Published: (2025)
Insights into solvent‐polarity‐dependent excited state behaviors for EDBT fluorophore: A computational study
by: Junping Xiao, et al.
Published: (2025)
by: Junping Xiao, et al.
Published: (2025)
Correlative and Discriminative Label Grouping for Multi-Label Visual Prompt Tuning
by: Ma, LeiLei, et al.
Published: (2025)
by: Ma, LeiLei, et al.
Published: (2025)
Soundness-Aware Level: A Microscopic Signature that Predicts LLM Reasoning Potential
by: Wu, Xuansheng, et al.
Published: (2025)
by: Wu, Xuansheng, et al.
Published: (2025)
Q-Save: Towards Scoring and Attribution for Generated Video Evaluation
by: Wu, Xiele, et al.
Published: (2025)
by: Wu, Xiele, et al.
Published: (2025)
Similar Items
-
Diagnosing Failures in Large Language Models' Answers: Integrating Error Attribution into Evaluation Framework
by: Xu, Zishan, et al.
Published: (2025) -
IDGenRec: LLM-RecSys Alignment with Textual ID Learning
by: Tan, Juntao, et al.
Published: (2024) -
HDFlow: Enhancing LLM Complex Problem-Solving with Hybrid Thinking and Dynamic Workflows
by: Yao, Wenlin, et al.
Published: (2024) -
Insights Into Solvent‐Polarity‐Related Photo‐Induced Excited State Behaviors for H2BP‐(OH)2DC Compound: A Theoretical Study
by: Junping Xiao, et al.
Published: (2025) -
Do Physicians Know How to Prompt? The Need for Automatic Prompt Optimization Help in Clinical Note Generation
by: Yao, Zonghai, et al.
Published: (2023)