Cross-Entropy Games and Frost Training
Fuente:
arXiv
Guardado en:
| Autores principales: | Renard, Arthur, Gabriel, Franck, Hartmann, Valentin, Hongler, Clément |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Cognitive Training for Language Models: Towards General Capabilities via Cross-Entropy Games
por: Hongler, Clément, et al.
Publicado: (2026)
por: Hongler, Clément, et al.
Publicado: (2026)
Cross-Entropy Games for Language Models: From Implicit Knowledge to General Capability Measures
por: Hongler, Clément, et al.
Publicado: (2025)
por: Hongler, Clément, et al.
Publicado: (2025)
Arrows of Time for Large Language Models
por: Papadopoulos, Vassilis, et al.
Publicado: (2024)
por: Papadopoulos, Vassilis, et al.
Publicado: (2024)
Visualizing the Structure of Lenia Parameter Space
por: Hudcová, Barbora, et al.
Publicado: (2026)
por: Hudcová, Barbora, et al.
Publicado: (2026)
Looking for Complexity at Phase Boundaries in Continuous Cellular Automata
por: Papadopoulos, Vassilis, et al.
Publicado: (2024)
por: Papadopoulos, Vassilis, et al.
Publicado: (2024)
PermaFrost-Attack: Stealth Pretraining Seeding(SPS) for planting Logic Landmines During LLM Training
por: Kumar, Harsh, et al.
Publicado: (2026)
por: Kumar, Harsh, et al.
Publicado: (2026)
Deduction Game Framework and Information Set Entropy Search
por: Meng, Fandi, et al.
Publicado: (2024)
por: Meng, Fandi, et al.
Publicado: (2024)
Federated Learning with Workload Reduction through Partial Training of Client Models and Entropy-Based Data Selection
por: Shi, Hongrui, et al.
Publicado: (2024)
por: Shi, Hongrui, et al.
Publicado: (2024)
Cross-Entropy Optimization for Hyperparameter Optimization in Stochastic Gradient-based Approaches to Train Deep Neural Networks
por: Li, Kevin, et al.
Publicado: (2024)
por: Li, Kevin, et al.
Publicado: (2024)
MixCE: Training Autoregressive Language Models by Mixing Forward and Reverse Cross-Entropies
por: Zhang, Shiyue, et al.
Publicado: (2023)
por: Zhang, Shiyue, et al.
Publicado: (2023)
$α$-GAN by Rényi Cross Entropy
por: Ding, Ni, et al.
Publicado: (2025)
por: Ding, Ni, et al.
Publicado: (2025)
Trajectory Entropy: Modeling Game State Stability from Multimodality Trajectory Prediction
por: Zhang, Yesheng, et al.
Publicado: (2025)
por: Zhang, Yesheng, et al.
Publicado: (2025)
The Token Games: Evaluating Language Model Reasoning with Puzzle Duels
por: Henniger, Simon, et al.
Publicado: (2026)
por: Henniger, Simon, et al.
Publicado: (2026)
Deep MMD Gradient Flow without adversarial training
por: Galashov, Alexandre, et al.
Publicado: (2024)
por: Galashov, Alexandre, et al.
Publicado: (2024)
Understanding Prediction Discrepancies in Machine Learning Classifiers
por: Renard, Xavier, et al.
Publicado: (2021)
por: Renard, Xavier, et al.
Publicado: (2021)
Sparse Training of Discrete Diffusion Models for Graph Generation
por: Qin, Yiming, et al.
Publicado: (2023)
por: Qin, Yiming, et al.
Publicado: (2023)
Cross-Dataset Generalization For Retinal Lesions Segmentation
por: Playout, Clément, et al.
Publicado: (2024)
por: Playout, Clément, et al.
Publicado: (2024)
Efficient Benchmarking of AI Agents
por: Ndzomga, Franck
Publicado: (2026)
por: Ndzomga, Franck
Publicado: (2026)
Model-based reinforcement learning for protein backbone design
por: Renard, Frederic, et al.
Publicado: (2024)
por: Renard, Frederic, et al.
Publicado: (2024)
No-Regret Strategy Solving in Imperfect-Information Games via Pre-Trained Embedding
por: Fu, Yanchang, et al.
Publicado: (2025)
por: Fu, Yanchang, et al.
Publicado: (2025)
Label-Free Reinforcement Learning via Cross-Model Entropy
por: Gorbett, Matt, et al.
Publicado: (2026)
por: Gorbett, Matt, et al.
Publicado: (2026)
Online Training and Pruning of Deep Reinforcement Learning Networks
por: Guenter, Valentin Frank Ingmar, et al.
Publicado: (2025)
por: Guenter, Valentin Frank Ingmar, et al.
Publicado: (2025)
SimCE: Simplifying Cross-Entropy Loss for Collaborative Filtering
por: Yang, Xiaodong, et al.
Publicado: (2024)
por: Yang, Xiaodong, et al.
Publicado: (2024)
Rethinking Independent Cross-Entropy Loss For Graph-Structured Data
por: Miao, Rui, et al.
Publicado: (2024)
por: Miao, Rui, et al.
Publicado: (2024)
Rejuvenating Cross-Entropy Loss in Knowledge Distillation for Recommender Systems
por: Zhu, Zhangchi, et al.
Publicado: (2025)
por: Zhu, Zhangchi, et al.
Publicado: (2025)
Beyond Shapley Values: Cooperative Games for the Interpretation of Machine Learning Models
por: Idrissi, Marouane Il, et al.
Publicado: (2025)
por: Idrissi, Marouane Il, et al.
Publicado: (2025)
Orak: A Foundational Benchmark for Training and Evaluating LLM Agents on Diverse Video Games
por: Park, Dongmin, et al.
Publicado: (2025)
por: Park, Dongmin, et al.
Publicado: (2025)
Training Interactive Agent in Large FPS Game Map with Rule-enhanced Reinforcement Learning
por: Zhang, Chen, et al.
Publicado: (2024)
por: Zhang, Chen, et al.
Publicado: (2024)
EntroCoT: Enhancing Chain-of-Thought via Adaptive Entropy-Guided Segmentation
por: Li, Zihang, et al.
Publicado: (2026)
por: Li, Zihang, et al.
Publicado: (2026)
Learned Hallucination Detection in Black-Box LLMs using Token-level Entropy Production Rate
por: Moslonka, Charles, et al.
Publicado: (2025)
por: Moslonka, Charles, et al.
Publicado: (2025)
What Scales in Cross-Entropy Scaling Law?
por: Yan, Junxi, et al.
Publicado: (2025)
por: Yan, Junxi, et al.
Publicado: (2025)
Bregman Centroid Guided Cross-Entropy Method
por: Gu, Yuliang, et al.
Publicado: (2025)
por: Gu, Yuliang, et al.
Publicado: (2025)
Convergence of a model-free entropy-regularized inverse reinforcement learning algorithm
por: Renard, Titouan, et al.
Publicado: (2024)
por: Renard, Titouan, et al.
Publicado: (2024)
Market-Driven Subset Selection for Budgeted Training
por: Jha, Ashish, et al.
Publicado: (2025)
por: Jha, Ashish, et al.
Publicado: (2025)
Entropy-Guided Data-Efficient Training for Multimodal Reasoning Reward Models
por: Yang, Shidong, et al.
Publicado: (2026)
por: Yang, Shidong, et al.
Publicado: (2026)
EntropyLong: Effective Long-Context Training via Predictive Uncertainty
por: Jia, Junlong, et al.
Publicado: (2025)
por: Jia, Junlong, et al.
Publicado: (2025)
Entropy-Driven Curriculum for Multi-Task Training in Human Mobility Prediction
por: Fang, Tianye, et al.
Publicado: (2025)
por: Fang, Tianye, et al.
Publicado: (2025)
Gradient Dynamics of Attention: How Cross-Entropy Sculpts Bayesian Manifolds
por: Agarwal, Naman, et al.
Publicado: (2025)
por: Agarwal, Naman, et al.
Publicado: (2025)
Cross-Entropy Is All You Need To Invert the Data Generating Process
por: Reizinger, Patrik, et al.
Publicado: (2024)
por: Reizinger, Patrik, et al.
Publicado: (2024)
How Useful Is Cross-Domain Generalization for Training LLM Monitors?
por: Martin, Sam, et al.
Publicado: (2026)
por: Martin, Sam, et al.
Publicado: (2026)
Ejemplares similares
-
Cognitive Training for Language Models: Towards General Capabilities via Cross-Entropy Games
por: Hongler, Clément, et al.
Publicado: (2026) -
Cross-Entropy Games for Language Models: From Implicit Knowledge to General Capability Measures
por: Hongler, Clément, et al.
Publicado: (2025) -
Arrows of Time for Large Language Models
por: Papadopoulos, Vassilis, et al.
Publicado: (2024) -
Visualizing the Structure of Lenia Parameter Space
por: Hudcová, Barbora, et al.
Publicado: (2026) -
Looking for Complexity at Phase Boundaries in Continuous Cellular Automata
por: Papadopoulos, Vassilis, et al.
Publicado: (2024)