Saved in:
| Main Authors: | Uto, Masaki, Ito, Yuma |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2506.20119 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MARS: Meaning-Aware Response Scoring for Uncertainty Estimation in Generative LLMs
by: Bakman, Yavuz Faruk, et al.
Published: (2024)
by: Bakman, Yavuz Faruk, et al.
Published: (2024)
Difficulty-Controllable Multiple-Choice Question Generation Using Large Language Models and Direct Preference Optimization
by: Tomikawa, Yuto, et al.
Published: (2025)
by: Tomikawa, Yuto, et al.
Published: (2025)
Pensieve Grader: An AI-Powered, Ready-to-Use Platform for Effortless Handwritten STEM Grading
by: Yang, Yoonseok, et al.
Published: (2025)
by: Yang, Yoonseok, et al.
Published: (2025)
A Baseline Analysis of Reward Models' Ability To Accurately Analyze Foundation Models Under Distribution Shift
by: LeVine, Will, et al.
Published: (2023)
by: LeVine, Will, et al.
Published: (2023)
Constructing Synthetic Instruction Datasets for Improving Reasoning in Domain-Specific LLMs: A Case Study in the Japanese Financial Domain
by: Okochi, Yuma, et al.
Published: (2026)
by: Okochi, Yuma, et al.
Published: (2026)
CEQuest: Benchmarking Large Language Models for Construction Estimation
by: Wu, Yanzhao, et al.
Published: (2025)
by: Wu, Yanzhao, et al.
Published: (2025)
LLM-Forest: Ensemble Learning of LLMs with Graph-Augmented Prompts for Data Imputation
by: He, Xinrui, et al.
Published: (2024)
by: He, Xinrui, et al.
Published: (2024)
A Context-Aware Approach for Enhancing Data Imputation with Pre-trained Language Models
by: Hayat, Ahatsham, et al.
Published: (2024)
by: Hayat, Ahatsham, et al.
Published: (2024)
LPCD: Unified Framework from Layer-Wise to Submodule Quantization
by: Ichikawa, Yuma, et al.
Published: (2025)
by: Ichikawa, Yuma, et al.
Published: (2025)
Still More Shades of Null: An Evaluation Suite for Responsible Missing Value Imputation
by: Khan, Falaah Arif, et al.
Published: (2024)
by: Khan, Falaah Arif, et al.
Published: (2024)
Emergent Abilities in Reduced-Scale Generative Language Models
by: Muckatira, Sherin, et al.
Published: (2024)
by: Muckatira, Sherin, et al.
Published: (2024)
Scaling Test-Time Compute to Achieve IOI Gold Medal with Open-Weight Models
by: Samadi, Mehrzad, et al.
Published: (2025)
by: Samadi, Mehrzad, et al.
Published: (2025)
Sign Lock-In: Randomly Initialized Weight Signs Persist and Bottleneck Sub-Bit Model Compression
by: Sakai, Akira, et al.
Published: (2026)
by: Sakai, Akira, et al.
Published: (2026)
Has Automated Essay Scoring Reached Sufficient Accuracy? Deriving Achievable QWK Ceilings from Classical Test Theory
by: Uto, Masaki
Published: (2026)
by: Uto, Masaki
Published: (2026)
Social Genome: Grounded Social Reasoning Abilities of Multimodal Models
by: Mathur, Leena, et al.
Published: (2025)
by: Mathur, Leena, et al.
Published: (2025)
An ASR-Based Tutor for Learning to Read: How to Optimize Feedback to First Graders
by: Bai, Yu, et al.
Published: (2023)
by: Bai, Yu, et al.
Published: (2023)
On the Ability of Transformers to Verify Plans
by: Sarrof, Yash, et al.
Published: (2026)
by: Sarrof, Yash, et al.
Published: (2026)
Automated Text Scoring in the Age of Generative AI for the GPU-poor
by: Ormerod, Christopher Michael, et al.
Published: (2024)
by: Ormerod, Christopher Michael, et al.
Published: (2024)
Does Continued Pretraining on a Learner Corpus Improve Automated Essay Scoring on English Proficiency Tests? Evidence from EFCAMDAT
by: Nguyen, Duy Anh
Published: (2026)
by: Nguyen, Duy Anh
Published: (2026)
Berezinskii--Kosterlitz--Thouless transition in a context-sensitive random language model
by: Toji, Yuma, et al.
Published: (2024)
by: Toji, Yuma, et al.
Published: (2024)
Phase transition on a context-sensitive random language model with short range interactions
by: Toji, Yuma, et al.
Published: (2026)
by: Toji, Yuma, et al.
Published: (2026)
Cyclic Ablation: Testing Concept Localization against Functional Regeneration in AI
by: Kapelko, Eduard
Published: (2025)
by: Kapelko, Eduard
Published: (2025)
Missing-by-Design: Certifiable Modality Deletion for Revocable Multimodal Sentiment Analysis
by: Fu, Rong, et al.
Published: (2026)
by: Fu, Rong, et al.
Published: (2026)
Multicalibration for Confidence Scoring in LLMs
by: Detommaso, Gianluca, et al.
Published: (2024)
by: Detommaso, Gianluca, et al.
Published: (2024)
Born a Transformer -- Always a Transformer? On the Effect of Pretraining on Architectural Abilities
by: Jobanputra, Mayank, et al.
Published: (2025)
by: Jobanputra, Mayank, et al.
Published: (2025)
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models
by: Zhao, Zekai, et al.
Published: (2025)
by: Zhao, Zekai, et al.
Published: (2025)
Towards Better Understanding of In-Context Learning Ability from In-Context Uncertainty Quantification
by: Liu, Shang, et al.
Published: (2024)
by: Liu, Shang, et al.
Published: (2024)
The Missing Half: Unveiling Training-time Implicit Safety Risks Beyond Deployment
by: Zhang, Zhexin, et al.
Published: (2026)
by: Zhang, Zhexin, et al.
Published: (2026)
Missed Causes and Ambiguous Effects: Counterfactuals Pose Challenges for Interpreting Neural Networks
by: Mueller, Aaron
Published: (2024)
by: Mueller, Aaron
Published: (2024)
Geometric Stability: The Missing Axis of Representations
by: Raju, Prashant C.
Published: (2026)
by: Raju, Prashant C.
Published: (2026)
Speaking the Same Language: Leveraging LLMs in Standardizing Clinical Data for AI
by: Sett, Arindam, et al.
Published: (2024)
by: Sett, Arindam, et al.
Published: (2024)
SumRec: A Framework for Recommendation using Open-Domain Dialogue
by: Asahara, Ryutaro, et al.
Published: (2024)
by: Asahara, Ryutaro, et al.
Published: (2024)
Maximum Score Routing For Mixture-of-Experts
by: Dong, Bowen, et al.
Published: (2025)
by: Dong, Bowen, et al.
Published: (2025)
On the Reasoning Abilities of Masked Diffusion Language Models
by: Svete, Anej, et al.
Published: (2025)
by: Svete, Anej, et al.
Published: (2025)
Towards Reasoning Ability of Small Language Models
by: Srivastava, Gaurav, et al.
Published: (2025)
by: Srivastava, Gaurav, et al.
Published: (2025)
Unlocking Continual Learning Abilities in Language Models
by: Du, Wenyu, et al.
Published: (2024)
by: Du, Wenyu, et al.
Published: (2024)
Deception Abilities Emerged in Large Language Models
by: Hagendorff, Thilo
Published: (2023)
by: Hagendorff, Thilo
Published: (2023)
CLaSP: Learning Concepts for Time-Series Signals from Natural Language Supervision
by: Ito, Aoi, et al.
Published: (2024)
by: Ito, Aoi, et al.
Published: (2024)
Leverage Unlearning to Sanitize LLMs
by: Boutet, Antoine, et al.
Published: (2025)
by: Boutet, Antoine, et al.
Published: (2025)
Leveraging the true depth of LLMs
by: González, Ramón Calvo, et al.
Published: (2025)
by: González, Ramón Calvo, et al.
Published: (2025)
Similar Items
-
MARS: Meaning-Aware Response Scoring for Uncertainty Estimation in Generative LLMs
by: Bakman, Yavuz Faruk, et al.
Published: (2024) -
Difficulty-Controllable Multiple-Choice Question Generation Using Large Language Models and Direct Preference Optimization
by: Tomikawa, Yuto, et al.
Published: (2025) -
Pensieve Grader: An AI-Powered, Ready-to-Use Platform for Effortless Handwritten STEM Grading
by: Yang, Yoonseok, et al.
Published: (2025) -
A Baseline Analysis of Reward Models' Ability To Accurately Analyze Foundation Models Under Distribution Shift
by: LeVine, Will, et al.
Published: (2023) -
Constructing Synthetic Instruction Datasets for Improving Reasoning in Domain-Specific LLMs: A Case Study in the Japanese Financial Domain
by: Okochi, Yuma, et al.
Published: (2026)