Risultati della ricerca - Lin, F.
- Mostra 1 - 20 risultati su 73
- Vai alla pagina seguente
-
1
Winning Gold at IMO 2025 with a Model-Agnostic Verification-and-Refinement Pipeline di Huang, Yichen, Yang, Lin F.
Pubblicazione 2025Testo
Preprint -
2
Confident Natural Policy Gradient for Local Planning in $q_π$-realizable Constrained MDPs di Tian, Tian, Yang, Lin F., Szepesvári, Csaba
Pubblicazione 2024Testo
Preprint -
3
Sample Complexity Bounds for Linear Constrained MDPs with a Generative Model di Liu, Xingtu, Yang, Lin F., Vaswani, Sharan
Pubblicazione 2025Testo
Preprint -
4
Near-Optimal Sample Complexity Bounds for Constrained Average-Reward MDPs di Wei, Yukuan, Li, Xudong, Yang, Lin F.
Pubblicazione 2025Testo
Preprint -
5
Near-Optimal Sample Complexity for Online Constrained MDPs di Liu, Chang, Li, Yunfan, Yang, Lin F.
Pubblicazione 2026Testo
Preprint -
6
Misspecified $Q$-Learning with Sparse Linear Function Approximation: Tight Bounds on Approximation Error di Du, Ally Yalei, Yang, Lin F., Wang, Ruosong
Pubblicazione 2024Testo
Preprint -
7
-
8
Uniform Last-Iterate Guarantee for Bandits and Reinforcement Learning di Liu, Junyan, Li, Yunfan, Wang, Ruosong, Yang, Lin F.
Pubblicazione 2024Testo
Preprint -
9
Learning for Bandits under Action Erasures di Hanna, Osama, Karakas, Merve, Yang, Lin F., Fragouli, Christina
Pubblicazione 2024Testo
Preprint -
10
Does Feedback Help in Bandits with Arm Erasures? di Karakas, Merve, Hanna, Osama, Yang, Lin F., Fragouli, Christina
Pubblicazione 2025Testo
Preprint -
11
On the optimal regret of collaborative personalized linear bandits di Huang, Bruce, Zhou, Ruida, Yang, Lin F., Diggavi, Suhas
Pubblicazione 2025Testo
Preprint -
12
Best-Arm Identification with Noisy Actuation di Karakas, Merve, Hanna, Osama, Yang, Lin F., Fragouli, Christina
Pubblicazione 2026Testo
Preprint -
13
Multi-Agent Bandit Learning through Heterogeneous Action Erasure Channels di Hanna, Osama A., Karakas, Merve, Yang, Lin F., Fragouli, Christina
Pubblicazione 2023Testo
Preprint -
14
Don't Forget to Connect! Improving RAG with Graph-based Reranking di Dong, Jialin, Fatemi, Bahare, Perozzi, Bryan, Yang, Lin F., Tsitsulin, Anton
Pubblicazione 2024Testo
Preprint -
15
ARMOR: High-Performance Semi-Structured Pruning via Adaptive Matrix Factorization di Liu, Lawrence, Liu, Alexander, Wang, Mengdi, Zhao, Tuo, Yang, Lin F.
Pubblicazione 2025Testo
Preprint -
16
A geometric distortion solution specifically for historical observations and its implementation di Lin, F. R., Peng, Q. Y., Zheng, Z. J., Guo, B. F.
Pubblicazione 2024Testo
Preprint -
17
Precision premium transformation -- a high-precision astrometric solution based on the precision premium curve di Zheng, Z. J., Peng, Q. Y., Lin, F. R., Li, D., Zheng, Y.
Pubblicazione 2024Testo
Preprint -
18
Hyper: Hyperparameter Robust Efficient Exploration in Reinforcement Learning di Wang, Yiran, Liu, Chenshu, Li, Yunfan, Amani, Sanae, Zhou, Bolei, Yang, Lin F.
Pubblicazione 2024Testo
Preprint -
19
NoWag: A Unified Framework for Shape Preserving Compression of Large Language Models di Liu, Lawrence, Chakrabarti, Inesh, Li, Yixiao, Wang, Mengdi, Zhao, Tuo, Yang, Lin F.
Pubblicazione 2025Testo
Preprint -
20
LACONIC: Length-Aware Constrained Reinforcement Learning for LLM di Liu, Chang, Zhao, Yiran, Liu, Lawrence, Ye, Yaoqi, Szepesvári, Csaba, Yang, Lin F.
Pubblicazione 2026Testo
Preprint