Guardado en:
| Autor principal: | Iourovitski, Dmitri |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2406.12043 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Hide and Seek: Fingerprinting Large Language Models with Evolutionary Learning
por: Iourovitski, Dmitri, et al.
Publicado: (2024)
por: Iourovitski, Dmitri, et al.
Publicado: (2024)
LLMs May Perform MCQA by Selecting the Least Incorrect Option
por: Wang, Haochun, et al.
Publicado: (2024)
por: Wang, Haochun, et al.
Publicado: (2024)
Mitigating Selection Bias with Node Pruning and Auxiliary Options
por: Choi, Hyeong Kyu, et al.
Publicado: (2024)
por: Choi, Hyeong Kyu, et al.
Publicado: (2024)
Flow-of-Options: Diversified and Improved LLM Reasoning by Thinking Through Options
por: Nair, Lakshmi, et al.
Publicado: (2025)
por: Nair, Lakshmi, et al.
Publicado: (2025)
Retention Score: Quantifying Jailbreak Risks for Vision Language Models
por: Li, Zaitang, et al.
Publicado: (2024)
por: Li, Zaitang, et al.
Publicado: (2024)
From Parameter Dynamics to Risk Scoring : Quantifying Sample-Level Safety Degradation in LLM Fine-tuning
por: Wang, Xiao, et al.
Publicado: (2026)
por: Wang, Xiao, et al.
Publicado: (2026)
Silicon Showdown: Performance, Efficiency, and Ecosystem Barriers in Consumer-Grade LLM Inference
por: Javat, Abdurrahman, et al.
Publicado: (2026)
por: Javat, Abdurrahman, et al.
Publicado: (2026)
OptionZero: Planning with Learned Options
por: Huang, Po-Wei, et al.
Publicado: (2025)
por: Huang, Po-Wei, et al.
Publicado: (2025)
LLM-assisted Semantic Option Discovery for Facilitating Adaptive Deep Reinforcement Learning
por: Yao, Chang, et al.
Publicado: (2026)
por: Yao, Chang, et al.
Publicado: (2026)
Option Discovery Using LLM-guided Semantic Hierarchical Reinforcement Learning
por: Shek, Chak Lam, et al.
Publicado: (2025)
por: Shek, Chak Lam, et al.
Publicado: (2025)
Data Compressibility Quantifies LLM Memorization
por: Huang, Yizhan, et al.
Publicado: (2025)
por: Huang, Yizhan, et al.
Publicado: (2025)
Pipeline for Verifying LLM-Generated Mathematical Solutions
por: Sazonova, Varvara, et al.
Publicado: (2026)
por: Sazonova, Varvara, et al.
Publicado: (2026)
From Scores to Steps: Diagnosing and Improving LLM Performance in Evidence-Based Medical Calculations
por: Wang, Benlu, et al.
Publicado: (2025)
por: Wang, Benlu, et al.
Publicado: (2025)
Not All Options Are Created Equal: Textual Option Weighting for Token-Efficient LLM-Based Knowledge Tracing
por: Kim, JongWoo, et al.
Publicado: (2024)
por: Kim, JongWoo, et al.
Publicado: (2024)
Epistemic Reject Option Prediction
por: Franc, Vojtech, et al.
Publicado: (2025)
por: Franc, Vojtech, et al.
Publicado: (2025)
Crucible: Quantifying the Potential of Control Algorithms through LLM Agents
por: Jia, Lianchen, et al.
Publicado: (2025)
por: Jia, Lianchen, et al.
Publicado: (2025)
Quantifying Cross-Query Contradictions in Multi-Query LLM Reasoning
por: Salla, Rohit Kumar, et al.
Publicado: (2026)
por: Salla, Rohit Kumar, et al.
Publicado: (2026)
LLM-Guided Quantified SMT Solving over Uninterpreted Functions
por: Lv, Kunhang, et al.
Publicado: (2026)
por: Lv, Kunhang, et al.
Publicado: (2026)
GradingAttack: Exposing Security Vulnerabilities in LLM Based Educational Grading Agents
por: Li, Xueyi, et al.
Publicado: (2026)
por: Li, Xueyi, et al.
Publicado: (2026)
BotzoneBench: Scalable LLM Evaluation via Graded AI Anchors
por: Li, Lingfeng, et al.
Publicado: (2026)
por: Li, Lingfeng, et al.
Publicado: (2026)
From Description to Score: Can LLMs Quantify Vulnerabilities?
por: Jafarikhah, Sima, et al.
Publicado: (2025)
por: Jafarikhah, Sima, et al.
Publicado: (2025)
Optimizing In-Context Demonstrations for LLM-based Automated Grading
por: Chu, Yucheng, et al.
Publicado: (2026)
por: Chu, Yucheng, et al.
Publicado: (2026)
Human-in-the-Loop LLM Grading for Handwritten Mathematics Assessments
por: Vanhoyweghen, Arne, et al.
Publicado: (2026)
por: Vanhoyweghen, Arne, et al.
Publicado: (2026)
Language Ranker: A Metric for Quantifying LLM Performance Across High and Low-Resource Languages
por: Li, Zihao, et al.
Publicado: (2024)
por: Li, Zihao, et al.
Publicado: (2024)
Grading Scale Impact on LLM-as-a-Judge: Human-LLM Alignment Is Highest on 0-5 Grading Scale
por: Li, Weiyue, et al.
Publicado: (2026)
por: Li, Weiyue, et al.
Publicado: (2026)
Justice or Prejudice? Quantifying Biases in LLM-as-a-Judge
por: Ye, Jiayi, et al.
Publicado: (2024)
por: Ye, Jiayi, et al.
Publicado: (2024)
Quantifying Frontier LLM Capabilities for Container Sandbox Escape
por: Marchand, Rahul, et al.
Publicado: (2026)
por: Marchand, Rahul, et al.
Publicado: (2026)
GLIDER: Grading LLM Interactions and Decisions using Explainable Ranking
por: Deshpande, Darshan, et al.
Publicado: (2024)
por: Deshpande, Darshan, et al.
Publicado: (2024)
Confusion-Aware Rubric Optimization for LLM-based Automated Grading
por: Chu, Yucheng, et al.
Publicado: (2026)
por: Chu, Yucheng, et al.
Publicado: (2026)
Accelerate Scaling of LLM Finetuning via Quantifying the Coverage and Depth of Instruction Set
por: Wu, Chengwei, et al.
Publicado: (2025)
por: Wu, Chengwei, et al.
Publicado: (2025)
Ran Score: a LLM-based Evaluation Score for Radiology Report Generation
por: Zhang, Ran, et al.
Publicado: (2026)
por: Zhang, Ran, et al.
Publicado: (2026)
How Uncertain Is the Grade? A Benchmark of Uncertainty Metrics for LLM-Based Automatic Assessment
por: Li, Hang, et al.
Publicado: (2026)
por: Li, Hang, et al.
Publicado: (2026)
Speech Emotion Recognition via Entropy-Aware Score Selection
por: Chua, ChenYi, et al.
Publicado: (2025)
por: Chua, ChenYi, et al.
Publicado: (2025)
The Signal is in the Steps: Local Scoring for Reasoning Data Selection
por: Just, Hoang Anh, et al.
Publicado: (2025)
por: Just, Hoang Anh, et al.
Publicado: (2025)
OLLM: Options-based Large Language Models
por: Sharma, Shashank, et al.
Publicado: (2026)
por: Sharma, Shashank, et al.
Publicado: (2026)
Unveiling Options with Neural Decomposition
por: Alikhasi, Mahdi, et al.
Publicado: (2024)
por: Alikhasi, Mahdi, et al.
Publicado: (2024)
Diversity-Enriched Option-Critic
por: Kamat, Anand, et al.
Publicado: (2020)
por: Kamat, Anand, et al.
Publicado: (2020)
RankLLM: Weighted Ranking of LLMs by Quantifying Question Difficulty
por: Zhang, Ziqian, et al.
Publicado: (2026)
por: Zhang, Ziqian, et al.
Publicado: (2026)
Quantifying Loss Aversion in Cyber Adversaries via LLM Analysis
por: Hans, Soham, et al.
Publicado: (2025)
por: Hans, Soham, et al.
Publicado: (2025)
Quantifying LLM Attention-Head Stability: Implications for Circuit Universality
por: Bali, Karan, et al.
Publicado: (2026)
por: Bali, Karan, et al.
Publicado: (2026)
Ejemplares similares
-
Hide and Seek: Fingerprinting Large Language Models with Evolutionary Learning
por: Iourovitski, Dmitri, et al.
Publicado: (2024) -
LLMs May Perform MCQA by Selecting the Least Incorrect Option
por: Wang, Haochun, et al.
Publicado: (2024) -
Mitigating Selection Bias with Node Pruning and Auxiliary Options
por: Choi, Hyeong Kyu, et al.
Publicado: (2024) -
Flow-of-Options: Diversified and Improved LLM Reasoning by Thinking Through Options
por: Nair, Lakshmi, et al.
Publicado: (2025) -
Retention Score: Quantifying Jailbreak Risks for Vision Language Models
por: Li, Zaitang, et al.
Publicado: (2024)