Transformer Based Planning in the Observation Space with Applications to Trick Taking Card Games
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Rebstock, Douglas, Solinas, Christopher, Sturtevant, Nathan R., Buro, Michael |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Neural Bayesian Filtering
von: Solinas, Christopher, et al.
Veröffentlicht: (2025)
von: Solinas, Christopher, et al.
Veröffentlicht: (2025)
Learning Admissible Heuristics for A*: Theory and Practice
von: Futuhi, Ehsan, et al.
Veröffentlicht: (2025)
von: Futuhi, Ehsan, et al.
Veröffentlicht: (2025)
Approximating Nash Equilibria in General-Sum Games via Meta-Learning
von: Sychrovský, David, et al.
Veröffentlicht: (2025)
von: Sychrovský, David, et al.
Veröffentlicht: (2025)
Outer-Learning Framework for Playing Multi-Player Trick-Taking Card Games: A Case Study in Skat
von: Edelkamp, Stefan
Veröffentlicht: (2025)
von: Edelkamp, Stefan
Veröffentlicht: (2025)
Transformer-Based Approach to Optimal Sensor Placement for Structural Health Monitoring of Probe Cards
von: Bejani, Mehdi, et al.
Veröffentlicht: (2025)
von: Bejani, Mehdi, et al.
Veröffentlicht: (2025)
Credit Card Fraud Detection Using Advanced Transformer Model
von: Yu, Chang, et al.
Veröffentlicht: (2024)
von: Yu, Chang, et al.
Veröffentlicht: (2024)
Enhancing LLM Evaluations: The Garbling Trick
von: Bradley, William F.
Veröffentlicht: (2024)
von: Bradley, William F.
Veröffentlicht: (2024)
Speeding Up MACE: Low-Precision Tricks for Equivarient Force Fields
von: Benoit, Alexandre
Veröffentlicht: (2025)
von: Benoit, Alexandre
Veröffentlicht: (2025)
Teach Me to Trick: Exploring Adversarial Transferability via Knowledge Distillation
von: Pradhan, Siddhartha, et al.
Veröffentlicht: (2025)
von: Pradhan, Siddhartha, et al.
Veröffentlicht: (2025)
Optimization of Latent-Space Compression using Game-Theoretic Techniques for Transformer-Based Vector Search
von: Agrawal, Kushagra, et al.
Veröffentlicht: (2025)
von: Agrawal, Kushagra, et al.
Veröffentlicht: (2025)
Credit Card Fraud Detection
von: Popova, Iva, et al.
Veröffentlicht: (2025)
von: Popova, Iva, et al.
Veröffentlicht: (2025)
Heterogeneous Graph Auto-Encoder for CreditCard Fraud Detection
von: Singh, Moirangthem Tiken, et al.
Veröffentlicht: (2024)
von: Singh, Moirangthem Tiken, et al.
Veröffentlicht: (2024)
Teach Old SAEs New Domain Tricks with Boosting
von: Koriagin, Nikita, et al.
Veröffentlicht: (2025)
von: Koriagin, Nikita, et al.
Veröffentlicht: (2025)
Causal Reinforcement Learning for Complex Card Games: A Magic The Gathering Benchmark
von: Cunha, Cristiano da Costa, et al.
Veröffentlicht: (2026)
von: Cunha, Cristiano da Costa, et al.
Veröffentlicht: (2026)
Set-Based Retrograde Analysis: Precomputing the Solution to 24-card Bridge Double Dummy Deals
von: Stone, Isaac, et al.
Veröffentlicht: (2024)
von: Stone, Isaac, et al.
Veröffentlicht: (2024)
Math Takes Two: A test for emergent mathematical reasoning in communication
von: Cooper, Michael, et al.
Veröffentlicht: (2026)
von: Cooper, Michael, et al.
Veröffentlicht: (2026)
Can We Rely on LLM Agents to Draft Long-Horizon Plans? Let's Take TravelPlanner as an Example
von: Chen, Yanan, et al.
Veröffentlicht: (2024)
von: Chen, Yanan, et al.
Veröffentlicht: (2024)
Goal-Space Planning with Subgoal Models
von: Lo, Chunlok, et al.
Veröffentlicht: (2022)
von: Lo, Chunlok, et al.
Veröffentlicht: (2022)
It Just Takes Two: Scaling Amortized Inference to Large Sets
von: Wehenkel, Antoine, et al.
Veröffentlicht: (2026)
von: Wehenkel, Antoine, et al.
Veröffentlicht: (2026)
On the Completeness of Conflict-Based Search: Temporally-Relative Duplicate Pruning
von: Walker, Thayne T, et al.
Veröffentlicht: (2024)
von: Walker, Thayne T, et al.
Veröffentlicht: (2024)
Low-Rank MDPs with Continuous Action Spaces
von: Bennett, Andrew, et al.
Veröffentlicht: (2023)
von: Bennett, Andrew, et al.
Veröffentlicht: (2023)
Taking the GP Out of the Loop
von: Bafna, Mehul, et al.
Veröffentlicht: (2025)
von: Bafna, Mehul, et al.
Veröffentlicht: (2025)
On the Ability of Transformers to Verify Plans
von: Sarrof, Yash, et al.
Veröffentlicht: (2026)
von: Sarrof, Yash, et al.
Veröffentlicht: (2026)
Bridging Local and Global Knowledge via Transformer in Board Games
von: Ju, Yan-Ru, et al.
Veröffentlicht: (2024)
von: Ju, Yan-Ru, et al.
Veröffentlicht: (2024)
Towards Sustainability Model Cards
von: Jouneaux, Gwendal, et al.
Veröffentlicht: (2025)
von: Jouneaux, Gwendal, et al.
Veröffentlicht: (2025)
Symmetry-Aware Transformer Training for Automated Planning
von: Fritzsche, Markus, et al.
Veröffentlicht: (2025)
von: Fritzsche, Markus, et al.
Veröffentlicht: (2025)
Report Cards: Qualitative Evaluation of Language Models Using Natural Language Summaries
von: Yang, Blair, et al.
Veröffentlicht: (2024)
von: Yang, Blair, et al.
Veröffentlicht: (2024)
BERT4Traj: Transformer Based Trajectory Reconstruction for Sparse Mobility Data
von: Yang, Hao, et al.
Veröffentlicht: (2025)
von: Yang, Hao, et al.
Veröffentlicht: (2025)
A Parallel CPU-GPU Framework for Batching Heuristic Operations in Depth-First Heuristic Search
von: Futuhi, Ehsan, et al.
Veröffentlicht: (2025)
von: Futuhi, Ehsan, et al.
Veröffentlicht: (2025)
SR-Reward: Taking The Path More Traveled
von: Azad, Seyed Mahdi B., et al.
Veröffentlicht: (2025)
von: Azad, Seyed Mahdi B., et al.
Veröffentlicht: (2025)
House of Cards: Massive Weights in LLMs
von: Oh, Jaehoon, et al.
Veröffentlicht: (2024)
von: Oh, Jaehoon, et al.
Veröffentlicht: (2024)
A Latent Space Metric for Enhancing Prediction Confidence in Earth Observation Data
von: Pitsiorlas, Ioannis, et al.
Veröffentlicht: (2024)
von: Pitsiorlas, Ioannis, et al.
Veröffentlicht: (2024)
Mastering Board Games by External and Internal Planning with Language Models
von: Schultz, John, et al.
Veröffentlicht: (2024)
von: Schultz, John, et al.
Veröffentlicht: (2024)
Read to Play (R2-Play): Decision Transformer with Multimodal Game Instruction
von: Jin, Yonggang, et al.
Veröffentlicht: (2024)
von: Jin, Yonggang, et al.
Veröffentlicht: (2024)
Target Return Optimizer for Multi-Game Decision Transformer
von: Tatematsu, Kensuke, et al.
Veröffentlicht: (2025)
von: Tatematsu, Kensuke, et al.
Veröffentlicht: (2025)
On the Design Space Between Transformers and Recursive Neural Nets
von: Chowdhury, Jishnu Ray, et al.
Veröffentlicht: (2024)
von: Chowdhury, Jishnu Ray, et al.
Veröffentlicht: (2024)
Hypformer: Exploring Efficient Transformer Fully in Hyperbolic Space
von: Yang, Menglin, et al.
Veröffentlicht: (2024)
von: Yang, Menglin, et al.
Veröffentlicht: (2024)
How Transformers Learn to Plan via Multi-Token Prediction
von: Huang, Jianhao, et al.
Veröffentlicht: (2026)
von: Huang, Jianhao, et al.
Veröffentlicht: (2026)
On Measuring Unnoticeability of Graph Adversarial Attacks: Observations, New Measure, and Applications
von: Jo, Hyeonsoo, et al.
Veröffentlicht: (2025)
von: Jo, Hyeonsoo, et al.
Veröffentlicht: (2025)
Planning Under Observation Mismatch for Traffic Signal Control via Adaptive Modular World Models
von: Huang, Zherui, et al.
Veröffentlicht: (2025)
von: Huang, Zherui, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Neural Bayesian Filtering
von: Solinas, Christopher, et al.
Veröffentlicht: (2025) -
Learning Admissible Heuristics for A*: Theory and Practice
von: Futuhi, Ehsan, et al.
Veröffentlicht: (2025) -
Approximating Nash Equilibria in General-Sum Games via Meta-Learning
von: Sychrovský, David, et al.
Veröffentlicht: (2025) -
Outer-Learning Framework for Playing Multi-Player Trick-Taking Card Games: A Case Study in Skat
von: Edelkamp, Stefan
Veröffentlicht: (2025) -
Transformer-Based Approach to Optimal Sensor Placement for Structural Health Monitoring of Probe Cards
von: Bejani, Mehdi, et al.
Veröffentlicht: (2025)