Saved in:
| Main Authors: | Zhang, Tianhao, Sheng, Zhecheng, Lin, Zhexiao, Jiang, Chen, Kang, Dongyeop |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2405.17764 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BBScore: A Brownian Bridge Based Metric for Assessing Text Coherence
by: Sheng, Zhecheng, et al.
Published: (2023)
by: Sheng, Zhecheng, et al.
Published: (2023)
Transformers as Decision Makers: Provable In-Context Reinforcement Learning via Supervised Pretraining
by: Lin, Licong, et al.
Published: (2023)
by: Lin, Licong, et al.
Published: (2023)
Balancing Complexity and Informativeness in LLM-Based Clustering: Finding the Goldilocks Zone
by: Miller, Justin, et al.
Published: (2025)
by: Miller, Justin, et al.
Published: (2025)
Towards Efficient Online Exploration for Reinforcement Learning with Human Feedback
by: Li, Gen, et al.
Published: (2025)
by: Li, Gen, et al.
Published: (2025)
Training Dynamics of Multi-Head Softmax Attention for In-Context Learning: Emergence, Convergence, and Optimality
by: Chen, Siyu, et al.
Published: (2024)
by: Chen, Siyu, et al.
Published: (2024)
Is a Good Foundation Necessary for Efficient Reinforcement Learning? The Computational Role of the Base Model in Exploration
by: Foster, Dylan J., et al.
Published: (2025)
by: Foster, Dylan J., et al.
Published: (2025)
Unveiling the Statistical Foundations of Chain-of-Thought Prompting Methods
by: Hu, Xinyang, et al.
Published: (2024)
by: Hu, Xinyang, et al.
Published: (2024)
Causal Sufficiency and Necessity Improves Chain-of-Thought Reasoning
by: Yu, Xiangning, et al.
Published: (2025)
by: Yu, Xiangning, et al.
Published: (2025)
Reject, Resample, Repeat: Understanding Parallel Reasoning in Language Model Inference
by: Golowich, Noah, et al.
Published: (2026)
by: Golowich, Noah, et al.
Published: (2026)
The Coverage Principle: How Pre-Training Enables Post-Training
by: Chen, Fan, et al.
Published: (2025)
by: Chen, Fan, et al.
Published: (2025)
Stochastic Direct Search Method for Blind Resource Allocation
by: Achddou, Juliette, et al.
Published: (2022)
by: Achddou, Juliette, et al.
Published: (2022)
Near-Optimal Learning and Planning in Separated Latent MDPs
by: Chen, Fan, et al.
Published: (2024)
by: Chen, Fan, et al.
Published: (2024)
Counterfactual reasoning: an analysis of in-context emergence
by: Miller, Moritz, et al.
Published: (2025)
by: Miller, Moritz, et al.
Published: (2025)
Don't Pass@k: A Bayesian Framework for Large Language Model Evaluation
by: Hariri, Mohsen, et al.
Published: (2025)
by: Hariri, Mohsen, et al.
Published: (2025)
Reasoning with Sampling: Cutting at Decision Points
by: Zhou, Felix, et al.
Published: (2026)
by: Zhou, Felix, et al.
Published: (2026)
Unifying regression-based and design-based causal inference in time-series experiments
by: Lin, Zhexiao, et al.
Published: (2025)
by: Lin, Zhexiao, et al.
Published: (2025)
Retrieval-Augmented Generation as Noisy In-Context Learning: A Unified Theory and Risk Bounds
by: Guo, Yang, et al.
Published: (2025)
by: Guo, Yang, et al.
Published: (2025)
Learning Interpretable Concepts: Unifying Causal Representation Learning and Foundation Models
by: Rajendran, Goutham, et al.
Published: (2024)
by: Rajendran, Goutham, et al.
Published: (2024)
Foundations of Structural Causal Models with Latent Selection
by: Chen, Leihao, et al.
Published: (2024)
by: Chen, Leihao, et al.
Published: (2024)
Limit theorems of Chatterjee's rank correlation
by: Lin, Zhexiao, et al.
Published: (2022)
by: Lin, Zhexiao, et al.
Published: (2022)
Becoming Experienced Judges: Selective Test-Time Learning for Evaluators
by: Jwa, Seungyeon, et al.
Published: (2025)
by: Jwa, Seungyeon, et al.
Published: (2025)
Canonical Representations of Markovian Structural Causal Models: A Framework for Counterfactual Reasoning
by: de Lara, Lucas
Published: (2025)
by: de Lara, Lucas
Published: (2025)
Abstain-R1: Calibrated Abstention and Post-Refusal Clarification via Verifiable RL
by: Zhai, Skylar, et al.
Published: (2026)
by: Zhai, Skylar, et al.
Published: (2026)
L$^2$M: Mutual Information Scaling Law for Long-Context Language Modeling
by: Chen, Zhuo, et al.
Published: (2025)
by: Chen, Zhuo, et al.
Published: (2025)
A Diffusion Analysis of Policy Gradient for Stochastic Bandits
by: Lattimore, Tor
Published: (2026)
by: Lattimore, Tor
Published: (2026)
Complete Characterization for Adjustment in Summary Causal Graphs of Time Series
by: Yvernes, Clément, et al.
Published: (2025)
by: Yvernes, Clément, et al.
Published: (2025)
Learning from Aggregate responses: Instance Level versus Bag Level Loss Functions
by: Javanmard, Adel, et al.
Published: (2024)
by: Javanmard, Adel, et al.
Published: (2024)
Stein-Rule Shrinkage for Stochastic Gradient Estimation in High Dimensions
by: Arashi, M., et al.
Published: (2026)
by: Arashi, M., et al.
Published: (2026)
A Statistical Hypothesis Testing Framework for Data Misappropriation Detection in Large Language Models
by: Cai, Yinpeng, et al.
Published: (2025)
by: Cai, Yinpeng, et al.
Published: (2025)
Nearest-Neighbor Radii under Dependent Sampling
by: Gao, Yuanyuan, et al.
Published: (2026)
by: Gao, Yuanyuan, et al.
Published: (2026)
Low-Dimensional Adaptation of Rectified Flow: A Diffusion and Stochastic Localization Perspective
by: Roy, Saptarshi, et al.
Published: (2026)
by: Roy, Saptarshi, et al.
Published: (2026)
MESSY Estimation: Maximum-Entropy based Stochastic and Symbolic densitY Estimation
by: Tohme, Tony, et al.
Published: (2023)
by: Tohme, Tony, et al.
Published: (2023)
On Rosenbaum's Rank-based Matching Estimator
by: Cattaneo, Matias D., et al.
Published: (2023)
by: Cattaneo, Matias D., et al.
Published: (2023)
Do LLMs Recognize Your Latent Preferences? A Benchmark for Latent Information Discovery in Personalized Interaction
by: Tsaknakis, Ioannis, et al.
Published: (2025)
by: Tsaknakis, Ioannis, et al.
Published: (2025)
From Spikes to Heavy Tails: Unveiling the Spectral Evolution of Neural Networks
by: Kothapalli, Vignesh, et al.
Published: (2024)
by: Kothapalli, Vignesh, et al.
Published: (2024)
A Statistical Case Against Empirical Human-AI Alignment
by: Rodemann, Julian, et al.
Published: (2025)
by: Rodemann, Julian, et al.
Published: (2025)
A Unified Pair-GRPO Family: From Implicit to Explicit Preference Constraints for Stable and General RL Alignment
by: Yu, Hao
Published: (2026)
by: Yu, Hao
Published: (2026)
Revisiting Incremental Stochastic Majorization-Minimization Algorithms with Applications to Mixture of Experts
by: Tran, TrungKhang, et al.
Published: (2026)
by: Tran, TrungKhang, et al.
Published: (2026)
Representation learning with a transformer by contrastive learning for money laundering detection
by: Guéneau, Harold, et al.
Published: (2025)
by: Guéneau, Harold, et al.
Published: (2025)
A2P-Vis: an Analyzer-to-Presenter Agentic Pipeline for Visual Insights Generation and Reporting
by: Gan, Shuyu, et al.
Published: (2025)
by: Gan, Shuyu, et al.
Published: (2025)
Similar Items
-
BBScore: A Brownian Bridge Based Metric for Assessing Text Coherence
by: Sheng, Zhecheng, et al.
Published: (2023) -
Transformers as Decision Makers: Provable In-Context Reinforcement Learning via Supervised Pretraining
by: Lin, Licong, et al.
Published: (2023) -
Balancing Complexity and Informativeness in LLM-Based Clustering: Finding the Goldilocks Zone
by: Miller, Justin, et al.
Published: (2025) -
Towards Efficient Online Exploration for Reinforcement Learning with Human Feedback
by: Li, Gen, et al.
Published: (2025) -
Training Dynamics of Multi-Head Softmax Attention for In-Context Learning: Emergence, Convergence, and Optimality
by: Chen, Siyu, et al.
Published: (2024)