Salvato in:
| Autore principale: | Iourovitski, Dmitri |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2406.12043 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Hide and Seek: Fingerprinting Large Language Models with Evolutionary Learning
di: Iourovitski, Dmitri, et al.
Pubblicazione: (2024)
di: Iourovitski, Dmitri, et al.
Pubblicazione: (2024)
LLMs May Perform MCQA by Selecting the Least Incorrect Option
di: Wang, Haochun, et al.
Pubblicazione: (2024)
di: Wang, Haochun, et al.
Pubblicazione: (2024)
Mitigating Selection Bias with Node Pruning and Auxiliary Options
di: Choi, Hyeong Kyu, et al.
Pubblicazione: (2024)
di: Choi, Hyeong Kyu, et al.
Pubblicazione: (2024)
Flow-of-Options: Diversified and Improved LLM Reasoning by Thinking Through Options
di: Nair, Lakshmi, et al.
Pubblicazione: (2025)
di: Nair, Lakshmi, et al.
Pubblicazione: (2025)
Retention Score: Quantifying Jailbreak Risks for Vision Language Models
di: Li, Zaitang, et al.
Pubblicazione: (2024)
di: Li, Zaitang, et al.
Pubblicazione: (2024)
From Parameter Dynamics to Risk Scoring : Quantifying Sample-Level Safety Degradation in LLM Fine-tuning
di: Wang, Xiao, et al.
Pubblicazione: (2026)
di: Wang, Xiao, et al.
Pubblicazione: (2026)
Silicon Showdown: Performance, Efficiency, and Ecosystem Barriers in Consumer-Grade LLM Inference
di: Javat, Abdurrahman, et al.
Pubblicazione: (2026)
di: Javat, Abdurrahman, et al.
Pubblicazione: (2026)
OptionZero: Planning with Learned Options
di: Huang, Po-Wei, et al.
Pubblicazione: (2025)
di: Huang, Po-Wei, et al.
Pubblicazione: (2025)
LLM-assisted Semantic Option Discovery for Facilitating Adaptive Deep Reinforcement Learning
di: Yao, Chang, et al.
Pubblicazione: (2026)
di: Yao, Chang, et al.
Pubblicazione: (2026)
Option Discovery Using LLM-guided Semantic Hierarchical Reinforcement Learning
di: Shek, Chak Lam, et al.
Pubblicazione: (2025)
di: Shek, Chak Lam, et al.
Pubblicazione: (2025)
Data Compressibility Quantifies LLM Memorization
di: Huang, Yizhan, et al.
Pubblicazione: (2025)
di: Huang, Yizhan, et al.
Pubblicazione: (2025)
Pipeline for Verifying LLM-Generated Mathematical Solutions
di: Sazonova, Varvara, et al.
Pubblicazione: (2026)
di: Sazonova, Varvara, et al.
Pubblicazione: (2026)
From Scores to Steps: Diagnosing and Improving LLM Performance in Evidence-Based Medical Calculations
di: Wang, Benlu, et al.
Pubblicazione: (2025)
di: Wang, Benlu, et al.
Pubblicazione: (2025)
Not All Options Are Created Equal: Textual Option Weighting for Token-Efficient LLM-Based Knowledge Tracing
di: Kim, JongWoo, et al.
Pubblicazione: (2024)
di: Kim, JongWoo, et al.
Pubblicazione: (2024)
Epistemic Reject Option Prediction
di: Franc, Vojtech, et al.
Pubblicazione: (2025)
di: Franc, Vojtech, et al.
Pubblicazione: (2025)
Crucible: Quantifying the Potential of Control Algorithms through LLM Agents
di: Jia, Lianchen, et al.
Pubblicazione: (2025)
di: Jia, Lianchen, et al.
Pubblicazione: (2025)
Quantifying Cross-Query Contradictions in Multi-Query LLM Reasoning
di: Salla, Rohit Kumar, et al.
Pubblicazione: (2026)
di: Salla, Rohit Kumar, et al.
Pubblicazione: (2026)
LLM-Guided Quantified SMT Solving over Uninterpreted Functions
di: Lv, Kunhang, et al.
Pubblicazione: (2026)
di: Lv, Kunhang, et al.
Pubblicazione: (2026)
GradingAttack: Exposing Security Vulnerabilities in LLM Based Educational Grading Agents
di: Li, Xueyi, et al.
Pubblicazione: (2026)
di: Li, Xueyi, et al.
Pubblicazione: (2026)
BotzoneBench: Scalable LLM Evaluation via Graded AI Anchors
di: Li, Lingfeng, et al.
Pubblicazione: (2026)
di: Li, Lingfeng, et al.
Pubblicazione: (2026)
From Description to Score: Can LLMs Quantify Vulnerabilities?
di: Jafarikhah, Sima, et al.
Pubblicazione: (2025)
di: Jafarikhah, Sima, et al.
Pubblicazione: (2025)
Optimizing In-Context Demonstrations for LLM-based Automated Grading
di: Chu, Yucheng, et al.
Pubblicazione: (2026)
di: Chu, Yucheng, et al.
Pubblicazione: (2026)
Human-in-the-Loop LLM Grading for Handwritten Mathematics Assessments
di: Vanhoyweghen, Arne, et al.
Pubblicazione: (2026)
di: Vanhoyweghen, Arne, et al.
Pubblicazione: (2026)
Language Ranker: A Metric for Quantifying LLM Performance Across High and Low-Resource Languages
di: Li, Zihao, et al.
Pubblicazione: (2024)
di: Li, Zihao, et al.
Pubblicazione: (2024)
Grading Scale Impact on LLM-as-a-Judge: Human-LLM Alignment Is Highest on 0-5 Grading Scale
di: Li, Weiyue, et al.
Pubblicazione: (2026)
di: Li, Weiyue, et al.
Pubblicazione: (2026)
Justice or Prejudice? Quantifying Biases in LLM-as-a-Judge
di: Ye, Jiayi, et al.
Pubblicazione: (2024)
di: Ye, Jiayi, et al.
Pubblicazione: (2024)
Quantifying Frontier LLM Capabilities for Container Sandbox Escape
di: Marchand, Rahul, et al.
Pubblicazione: (2026)
di: Marchand, Rahul, et al.
Pubblicazione: (2026)
GLIDER: Grading LLM Interactions and Decisions using Explainable Ranking
di: Deshpande, Darshan, et al.
Pubblicazione: (2024)
di: Deshpande, Darshan, et al.
Pubblicazione: (2024)
Confusion-Aware Rubric Optimization for LLM-based Automated Grading
di: Chu, Yucheng, et al.
Pubblicazione: (2026)
di: Chu, Yucheng, et al.
Pubblicazione: (2026)
Accelerate Scaling of LLM Finetuning via Quantifying the Coverage and Depth of Instruction Set
di: Wu, Chengwei, et al.
Pubblicazione: (2025)
di: Wu, Chengwei, et al.
Pubblicazione: (2025)
Ran Score: a LLM-based Evaluation Score for Radiology Report Generation
di: Zhang, Ran, et al.
Pubblicazione: (2026)
di: Zhang, Ran, et al.
Pubblicazione: (2026)
How Uncertain Is the Grade? A Benchmark of Uncertainty Metrics for LLM-Based Automatic Assessment
di: Li, Hang, et al.
Pubblicazione: (2026)
di: Li, Hang, et al.
Pubblicazione: (2026)
Speech Emotion Recognition via Entropy-Aware Score Selection
di: Chua, ChenYi, et al.
Pubblicazione: (2025)
di: Chua, ChenYi, et al.
Pubblicazione: (2025)
The Signal is in the Steps: Local Scoring for Reasoning Data Selection
di: Just, Hoang Anh, et al.
Pubblicazione: (2025)
di: Just, Hoang Anh, et al.
Pubblicazione: (2025)
OLLM: Options-based Large Language Models
di: Sharma, Shashank, et al.
Pubblicazione: (2026)
di: Sharma, Shashank, et al.
Pubblicazione: (2026)
Unveiling Options with Neural Decomposition
di: Alikhasi, Mahdi, et al.
Pubblicazione: (2024)
di: Alikhasi, Mahdi, et al.
Pubblicazione: (2024)
Diversity-Enriched Option-Critic
di: Kamat, Anand, et al.
Pubblicazione: (2020)
di: Kamat, Anand, et al.
Pubblicazione: (2020)
RankLLM: Weighted Ranking of LLMs by Quantifying Question Difficulty
di: Zhang, Ziqian, et al.
Pubblicazione: (2026)
di: Zhang, Ziqian, et al.
Pubblicazione: (2026)
Quantifying Loss Aversion in Cyber Adversaries via LLM Analysis
di: Hans, Soham, et al.
Pubblicazione: (2025)
di: Hans, Soham, et al.
Pubblicazione: (2025)
Quantifying LLM Attention-Head Stability: Implications for Circuit Universality
di: Bali, Karan, et al.
Pubblicazione: (2026)
di: Bali, Karan, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Hide and Seek: Fingerprinting Large Language Models with Evolutionary Learning
di: Iourovitski, Dmitri, et al.
Pubblicazione: (2024) -
LLMs May Perform MCQA by Selecting the Least Incorrect Option
di: Wang, Haochun, et al.
Pubblicazione: (2024) -
Mitigating Selection Bias with Node Pruning and Auxiliary Options
di: Choi, Hyeong Kyu, et al.
Pubblicazione: (2024) -
Flow-of-Options: Diversified and Improved LLM Reasoning by Thinking Through Options
di: Nair, Lakshmi, et al.
Pubblicazione: (2025) -
Retention Score: Quantifying Jailbreak Risks for Vision Language Models
di: Li, Zaitang, et al.
Pubblicazione: (2024)