Semiparametric Token-Sequence Co-Supervision
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lee, Hyunji, Kim, Doyoung, Jun, Jihoon, Joo, Sejune, Jang, Joel, On, Kyoung-Woon, Seo, Minjoon |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
How Well Do Large Language Models Truly Ground?
von: Lee, Hyunji, et al.
Veröffentlicht: (2023)
von: Lee, Hyunji, et al.
Veröffentlicht: (2023)
Knowledge Entropy Decay during Language Model Pretraining Hinders New Knowledge Acquisition
von: Kim, Jiyeon, et al.
Veröffentlicht: (2024)
von: Kim, Jiyeon, et al.
Veröffentlicht: (2024)
Differential Information Distribution: A Bayesian Perspective on Direct Preference Optimization
von: Won, Yunjae, et al.
Veröffentlicht: (2025)
von: Won, Yunjae, et al.
Veröffentlicht: (2025)
FLASK: Fine-grained Language Model Evaluation based on Alignment Skill Sets
von: Ye, Seonghyeon, et al.
Veröffentlicht: (2023)
von: Ye, Seonghyeon, et al.
Veröffentlicht: (2023)
Rethinking the Role of Proxy Rewards in Language Model Alignment
von: Kim, Sungdong, et al.
Veröffentlicht: (2024)
von: Kim, Sungdong, et al.
Veröffentlicht: (2024)
TSLM: Tree-Structured Language Modeling for Divergent Thinking
von: Kim, Doyoung, et al.
Veröffentlicht: (2026)
von: Kim, Doyoung, et al.
Veröffentlicht: (2026)
Exploiting the Potential of Seq2Seq Models as Robust Few-Shot Learners
von: Lee, Jihyeon, et al.
Veröffentlicht: (2023)
von: Lee, Jihyeon, et al.
Veröffentlicht: (2023)
Lost in the Noise: How Reasoning Models Fail with Contextual Distractors
von: Lee, Seongyun, et al.
Veröffentlicht: (2026)
von: Lee, Seongyun, et al.
Veröffentlicht: (2026)
Early Decisions Matter: Proximity Bias and Initial Trajectory Shaping in Non-Autoregressive Diffusion Language Models
von: Kim, Jiyeon, et al.
Veröffentlicht: (2026)
von: Kim, Jiyeon, et al.
Veröffentlicht: (2026)
How language models extrapolate outside the training data: A case study in Textualized Gridworld
von: Kim, Doyoung, et al.
Veröffentlicht: (2024)
von: Kim, Doyoung, et al.
Veröffentlicht: (2024)
Improving Probability-based Prompt Selection Through Unified Evaluation and Analysis
von: Yang, Sohee, et al.
Veröffentlicht: (2023)
von: Yang, Sohee, et al.
Veröffentlicht: (2023)
Instruction Tuning with and without Context: Behavioral Shifts and Downstream Impact
von: Lee, Hyunji, et al.
Veröffentlicht: (2025)
von: Lee, Hyunji, et al.
Veröffentlicht: (2025)
Exploring the Practicality of Generative Retrieval on Dynamic Corpora
von: Kim, Chaeeun, et al.
Veröffentlicht: (2023)
von: Kim, Chaeeun, et al.
Veröffentlicht: (2023)
The CoT Encyclopedia: Analyzing, Predicting, and Controlling how a Reasoning Model will Think
von: Lee, Seongyun, et al.
Veröffentlicht: (2025)
von: Lee, Seongyun, et al.
Veröffentlicht: (2025)
Evaluating Legal Reasoning Traces with Legal Issue Tree Rubrics
von: Lee, Jinu, et al.
Veröffentlicht: (2025)
von: Lee, Jinu, et al.
Veröffentlicht: (2025)
Hierarchical Deconstruction of LLM Reasoning: A Graph-Based Framework for Analyzing Knowledge Utilization
von: Ko, Miyoung, et al.
Veröffentlicht: (2024)
von: Ko, Miyoung, et al.
Veröffentlicht: (2024)
Generative Prompt Internalization
von: Shin, Haebin, et al.
Veröffentlicht: (2024)
von: Shin, Haebin, et al.
Veröffentlicht: (2024)
Confidence-guided Refinement Reasoning for Zero-shot Question Answering
von: Jang, Youwon, et al.
Veröffentlicht: (2025)
von: Jang, Youwon, et al.
Veröffentlicht: (2025)
Multi-View Attention Multiple-Instance Learning Enhanced by LLM Reasoning for Cognitive Distortion Detection
von: Kim, Jun Seo, et al.
Veröffentlicht: (2025)
von: Kim, Jun Seo, et al.
Veröffentlicht: (2025)
Aligning Large Language Models by On-Policy Self-Judgment
von: Lee, Sangkyu, et al.
Veröffentlicht: (2024)
von: Lee, Sangkyu, et al.
Veröffentlicht: (2024)
LangBridge: Multilingual Reasoning Without Multilingual Supervision
von: Yoon, Dongkeun, et al.
Veröffentlicht: (2024)
von: Yoon, Dongkeun, et al.
Veröffentlicht: (2024)
Self-Explore: Enhancing Mathematical Reasoning in Language Models with Fine-grained Rewards
von: Hwang, Hyeonbin, et al.
Veröffentlicht: (2024)
von: Hwang, Hyeonbin, et al.
Veröffentlicht: (2024)
Verbalized Confidence Triggers Self-Verification: Emergent Behavior Without Explicit Reasoning Supervision
von: Jang, Chaeyun, et al.
Veröffentlicht: (2025)
von: Jang, Chaeyun, et al.
Veröffentlicht: (2025)
Reasoning Models Better Express Their Confidence
von: Yoon, Dongkeun, et al.
Veröffentlicht: (2025)
von: Yoon, Dongkeun, et al.
Veröffentlicht: (2025)
Latent Reasoning via Sentence Embedding Prediction
von: Hwang, Hyeonbin, et al.
Veröffentlicht: (2025)
von: Hwang, Hyeonbin, et al.
Veröffentlicht: (2025)
Prime the search: Using large language models for guiding geometric task and motion planning by warm-starting tree search
von: Lee, Dongryung, et al.
Veröffentlicht: (2025)
von: Lee, Dongryung, et al.
Veröffentlicht: (2025)
Incorporating Domain Knowledge into Materials Tokenization
von: Oh, Yerim, et al.
Veröffentlicht: (2025)
von: Oh, Yerim, et al.
Veröffentlicht: (2025)
Mitigating LLM biases toward spurious social contexts using direct preference optimization
von: Nam, Hyunji, et al.
Veröffentlicht: (2026)
von: Nam, Hyunji, et al.
Veröffentlicht: (2026)
EHRSQL: A Practical Text-to-SQL Benchmark for Electronic Health Records
von: Lee, Gyubok, et al.
Veröffentlicht: (2023)
von: Lee, Gyubok, et al.
Veröffentlicht: (2023)
Extracting and Steering Emotion Representations in Small Language Models: A Methodological Comparison
von: Jeong, Jihoon
Veröffentlicht: (2026)
von: Jeong, Jihoon
Veröffentlicht: (2026)
Shared Emotion Geometry Across Small Language Models: A Cross-Architecture Study of Representation, Behavior, and Methodological Confounds
von: Jeong, Jihoon
Veröffentlicht: (2026)
von: Jeong, Jihoon
Veröffentlicht: (2026)
MTI: A Behavior-Based Temperament Profiling System for AI Agents
von: Jeong, Jihoon
Veröffentlicht: (2026)
von: Jeong, Jihoon
Veröffentlicht: (2026)
INSTRUCTIR: A Benchmark for Instruction Following of Information Retrieval Models
von: Oh, Hanseok, et al.
Veröffentlicht: (2024)
von: Oh, Hanseok, et al.
Veröffentlicht: (2024)
Binary Classifier Optimization for Large Language Model Alignment
von: Jung, Seungjae, et al.
Veröffentlicht: (2024)
von: Jung, Seungjae, et al.
Veröffentlicht: (2024)
CORG: Generating Answers from Complex, Interrelated Contexts
von: Lee, Hyunji, et al.
Veröffentlicht: (2025)
von: Lee, Hyunji, et al.
Veröffentlicht: (2025)
Control Token with Dense Passage Retrieval
von: Lee, Juhwan, et al.
Veröffentlicht: (2024)
von: Lee, Juhwan, et al.
Veröffentlicht: (2024)
Semantic Tokens in Retrieval Augmented Generation
von: Suro, Joel
Veröffentlicht: (2024)
von: Suro, Joel
Veröffentlicht: (2024)
Token-Supervised Value Models for Enhancing Mathematical Problem-Solving Capabilities of Large Language Models
von: Lee, Jung Hyun, et al.
Veröffentlicht: (2024)
von: Lee, Jung Hyun, et al.
Veröffentlicht: (2024)
KoACD: The First Korean Adolescent Dataset for Cognitive Distortion Analysis via Role-Switching Multi-LLM Negotiation
von: Kim, JunSeo, et al.
Veröffentlicht: (2025)
von: Kim, JunSeo, et al.
Veröffentlicht: (2025)
HPU: High-Bandwidth Processing Unit for Scalable, Cost-effective LLM Inference via GPU Co-processing
von: Rhee, Myunghyun, et al.
Veröffentlicht: (2025)
von: Rhee, Myunghyun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
How Well Do Large Language Models Truly Ground?
von: Lee, Hyunji, et al.
Veröffentlicht: (2023) -
Knowledge Entropy Decay during Language Model Pretraining Hinders New Knowledge Acquisition
von: Kim, Jiyeon, et al.
Veröffentlicht: (2024) -
Differential Information Distribution: A Bayesian Perspective on Direct Preference Optimization
von: Won, Yunjae, et al.
Veröffentlicht: (2025) -
FLASK: Fine-grained Language Model Evaluation based on Alignment Skill Sets
von: Ye, Seonghyeon, et al.
Veröffentlicht: (2023) -
Rethinking the Role of Proxy Rewards in Language Model Alignment
von: Kim, Sungdong, et al.
Veröffentlicht: (2024)