Gespeichert in:
| 1. Verfasser: | Iourovitski, Dmitri |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2406.12043 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Hide and Seek: Fingerprinting Large Language Models with Evolutionary Learning
von: Iourovitski, Dmitri, et al.
Veröffentlicht: (2024)
von: Iourovitski, Dmitri, et al.
Veröffentlicht: (2024)
LLMs May Perform MCQA by Selecting the Least Incorrect Option
von: Wang, Haochun, et al.
Veröffentlicht: (2024)
von: Wang, Haochun, et al.
Veröffentlicht: (2024)
Mitigating Selection Bias with Node Pruning and Auxiliary Options
von: Choi, Hyeong Kyu, et al.
Veröffentlicht: (2024)
von: Choi, Hyeong Kyu, et al.
Veröffentlicht: (2024)
Flow-of-Options: Diversified and Improved LLM Reasoning by Thinking Through Options
von: Nair, Lakshmi, et al.
Veröffentlicht: (2025)
von: Nair, Lakshmi, et al.
Veröffentlicht: (2025)
Retention Score: Quantifying Jailbreak Risks for Vision Language Models
von: Li, Zaitang, et al.
Veröffentlicht: (2024)
von: Li, Zaitang, et al.
Veröffentlicht: (2024)
From Parameter Dynamics to Risk Scoring : Quantifying Sample-Level Safety Degradation in LLM Fine-tuning
von: Wang, Xiao, et al.
Veröffentlicht: (2026)
von: Wang, Xiao, et al.
Veröffentlicht: (2026)
Silicon Showdown: Performance, Efficiency, and Ecosystem Barriers in Consumer-Grade LLM Inference
von: Javat, Abdurrahman, et al.
Veröffentlicht: (2026)
von: Javat, Abdurrahman, et al.
Veröffentlicht: (2026)
OptionZero: Planning with Learned Options
von: Huang, Po-Wei, et al.
Veröffentlicht: (2025)
von: Huang, Po-Wei, et al.
Veröffentlicht: (2025)
LLM-assisted Semantic Option Discovery for Facilitating Adaptive Deep Reinforcement Learning
von: Yao, Chang, et al.
Veröffentlicht: (2026)
von: Yao, Chang, et al.
Veröffentlicht: (2026)
Option Discovery Using LLM-guided Semantic Hierarchical Reinforcement Learning
von: Shek, Chak Lam, et al.
Veröffentlicht: (2025)
von: Shek, Chak Lam, et al.
Veröffentlicht: (2025)
Data Compressibility Quantifies LLM Memorization
von: Huang, Yizhan, et al.
Veröffentlicht: (2025)
von: Huang, Yizhan, et al.
Veröffentlicht: (2025)
Pipeline for Verifying LLM-Generated Mathematical Solutions
von: Sazonova, Varvara, et al.
Veröffentlicht: (2026)
von: Sazonova, Varvara, et al.
Veröffentlicht: (2026)
From Scores to Steps: Diagnosing and Improving LLM Performance in Evidence-Based Medical Calculations
von: Wang, Benlu, et al.
Veröffentlicht: (2025)
von: Wang, Benlu, et al.
Veröffentlicht: (2025)
Not All Options Are Created Equal: Textual Option Weighting for Token-Efficient LLM-Based Knowledge Tracing
von: Kim, JongWoo, et al.
Veröffentlicht: (2024)
von: Kim, JongWoo, et al.
Veröffentlicht: (2024)
Epistemic Reject Option Prediction
von: Franc, Vojtech, et al.
Veröffentlicht: (2025)
von: Franc, Vojtech, et al.
Veröffentlicht: (2025)
Crucible: Quantifying the Potential of Control Algorithms through LLM Agents
von: Jia, Lianchen, et al.
Veröffentlicht: (2025)
von: Jia, Lianchen, et al.
Veröffentlicht: (2025)
Quantifying Cross-Query Contradictions in Multi-Query LLM Reasoning
von: Salla, Rohit Kumar, et al.
Veröffentlicht: (2026)
von: Salla, Rohit Kumar, et al.
Veröffentlicht: (2026)
LLM-Guided Quantified SMT Solving over Uninterpreted Functions
von: Lv, Kunhang, et al.
Veröffentlicht: (2026)
von: Lv, Kunhang, et al.
Veröffentlicht: (2026)
GradingAttack: Exposing Security Vulnerabilities in LLM Based Educational Grading Agents
von: Li, Xueyi, et al.
Veröffentlicht: (2026)
von: Li, Xueyi, et al.
Veröffentlicht: (2026)
BotzoneBench: Scalable LLM Evaluation via Graded AI Anchors
von: Li, Lingfeng, et al.
Veröffentlicht: (2026)
von: Li, Lingfeng, et al.
Veröffentlicht: (2026)
From Description to Score: Can LLMs Quantify Vulnerabilities?
von: Jafarikhah, Sima, et al.
Veröffentlicht: (2025)
von: Jafarikhah, Sima, et al.
Veröffentlicht: (2025)
Optimizing In-Context Demonstrations for LLM-based Automated Grading
von: Chu, Yucheng, et al.
Veröffentlicht: (2026)
von: Chu, Yucheng, et al.
Veröffentlicht: (2026)
Human-in-the-Loop LLM Grading for Handwritten Mathematics Assessments
von: Vanhoyweghen, Arne, et al.
Veröffentlicht: (2026)
von: Vanhoyweghen, Arne, et al.
Veröffentlicht: (2026)
Language Ranker: A Metric for Quantifying LLM Performance Across High and Low-Resource Languages
von: Li, Zihao, et al.
Veröffentlicht: (2024)
von: Li, Zihao, et al.
Veröffentlicht: (2024)
Grading Scale Impact on LLM-as-a-Judge: Human-LLM Alignment Is Highest on 0-5 Grading Scale
von: Li, Weiyue, et al.
Veröffentlicht: (2026)
von: Li, Weiyue, et al.
Veröffentlicht: (2026)
Justice or Prejudice? Quantifying Biases in LLM-as-a-Judge
von: Ye, Jiayi, et al.
Veröffentlicht: (2024)
von: Ye, Jiayi, et al.
Veröffentlicht: (2024)
Quantifying Frontier LLM Capabilities for Container Sandbox Escape
von: Marchand, Rahul, et al.
Veröffentlicht: (2026)
von: Marchand, Rahul, et al.
Veröffentlicht: (2026)
GLIDER: Grading LLM Interactions and Decisions using Explainable Ranking
von: Deshpande, Darshan, et al.
Veröffentlicht: (2024)
von: Deshpande, Darshan, et al.
Veröffentlicht: (2024)
Confusion-Aware Rubric Optimization for LLM-based Automated Grading
von: Chu, Yucheng, et al.
Veröffentlicht: (2026)
von: Chu, Yucheng, et al.
Veröffentlicht: (2026)
Accelerate Scaling of LLM Finetuning via Quantifying the Coverage and Depth of Instruction Set
von: Wu, Chengwei, et al.
Veröffentlicht: (2025)
von: Wu, Chengwei, et al.
Veröffentlicht: (2025)
Ran Score: a LLM-based Evaluation Score for Radiology Report Generation
von: Zhang, Ran, et al.
Veröffentlicht: (2026)
von: Zhang, Ran, et al.
Veröffentlicht: (2026)
How Uncertain Is the Grade? A Benchmark of Uncertainty Metrics for LLM-Based Automatic Assessment
von: Li, Hang, et al.
Veröffentlicht: (2026)
von: Li, Hang, et al.
Veröffentlicht: (2026)
Speech Emotion Recognition via Entropy-Aware Score Selection
von: Chua, ChenYi, et al.
Veröffentlicht: (2025)
von: Chua, ChenYi, et al.
Veröffentlicht: (2025)
The Signal is in the Steps: Local Scoring for Reasoning Data Selection
von: Just, Hoang Anh, et al.
Veröffentlicht: (2025)
von: Just, Hoang Anh, et al.
Veröffentlicht: (2025)
OLLM: Options-based Large Language Models
von: Sharma, Shashank, et al.
Veröffentlicht: (2026)
von: Sharma, Shashank, et al.
Veröffentlicht: (2026)
Unveiling Options with Neural Decomposition
von: Alikhasi, Mahdi, et al.
Veröffentlicht: (2024)
von: Alikhasi, Mahdi, et al.
Veröffentlicht: (2024)
Diversity-Enriched Option-Critic
von: Kamat, Anand, et al.
Veröffentlicht: (2020)
von: Kamat, Anand, et al.
Veröffentlicht: (2020)
RankLLM: Weighted Ranking of LLMs by Quantifying Question Difficulty
von: Zhang, Ziqian, et al.
Veröffentlicht: (2026)
von: Zhang, Ziqian, et al.
Veröffentlicht: (2026)
Quantifying Loss Aversion in Cyber Adversaries via LLM Analysis
von: Hans, Soham, et al.
Veröffentlicht: (2025)
von: Hans, Soham, et al.
Veröffentlicht: (2025)
Quantifying LLM Attention-Head Stability: Implications for Circuit Universality
von: Bali, Karan, et al.
Veröffentlicht: (2026)
von: Bali, Karan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Hide and Seek: Fingerprinting Large Language Models with Evolutionary Learning
von: Iourovitski, Dmitri, et al.
Veröffentlicht: (2024) -
LLMs May Perform MCQA by Selecting the Least Incorrect Option
von: Wang, Haochun, et al.
Veröffentlicht: (2024) -
Mitigating Selection Bias with Node Pruning and Auxiliary Options
von: Choi, Hyeong Kyu, et al.
Veröffentlicht: (2024) -
Flow-of-Options: Diversified and Improved LLM Reasoning by Thinking Through Options
von: Nair, Lakshmi, et al.
Veröffentlicht: (2025) -
Retention Score: Quantifying Jailbreak Risks for Vision Language Models
von: Li, Zaitang, et al.
Veröffentlicht: (2024)