Improving Generalization on the ProcGen Benchmark with Simple Architectural Changes and Scale
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jesson, Andrew, Jiang, Yiding |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ProcGen3D: Learning Neural Procedural Graph Representations for Image-to-3D Reconstruction
von: Zhang, Xinyi, et al.
Veröffentlicht: (2025)
von: Zhang, Xinyi, et al.
Veröffentlicht: (2025)
Can Generative AI Solve Your In-Context Learning Problem? A Martingale Perspective
von: Jesson, Andrew, et al.
Veröffentlicht: (2024)
von: Jesson, Andrew, et al.
Veröffentlicht: (2024)
ProcBench: Benchmark for Multi-Step Reasoning and Following Procedure
von: Fujisawa, Ippei, et al.
Veröffentlicht: (2024)
von: Fujisawa, Ippei, et al.
Veröffentlicht: (2024)
ReLU to the Rescue: Improve Your On-Policy Actor-Critic with Positive Advantages
von: Jesson, Andrew, et al.
Veröffentlicht: (2023)
von: Jesson, Andrew, et al.
Veröffentlicht: (2023)
JULI: Jailbreak Large Language Models by Self-Introspection
von: Wang, Jesson, et al.
Veröffentlicht: (2025)
von: Wang, Jesson, et al.
Veröffentlicht: (2025)
ProcVLM: Learning Procedure-Grounded Progress Rewards for Robotic Manipulation
von: Feng, Youhe, et al.
Veröffentlicht: (2026)
von: Feng, Youhe, et al.
Veröffentlicht: (2026)
DialectGen: Benchmarking and Improving Dialect Robustness in Multimodal Generation
von: Zhou, Yu, et al.
Veröffentlicht: (2025)
von: Zhou, Yu, et al.
Veröffentlicht: (2025)
Estimating the Hallucination Rate of Generative AI
von: Jesson, Andrew, et al.
Veröffentlicht: (2024)
von: Jesson, Andrew, et al.
Veröffentlicht: (2024)
Adaptive Data Optimization: Dynamic Sample Selection with Scaling Laws
von: Jiang, Yiding, et al.
Veröffentlicht: (2024)
von: Jiang, Yiding, et al.
Veröffentlicht: (2024)
Scaling Law for Language Models Training Considering Batch Size
von: Shuai, Xian, et al.
Veröffentlicht: (2024)
von: Shuai, Xian, et al.
Veröffentlicht: (2024)
FairLangProc: A Python package for fairness in NLP
von: Pérez-Peralta, Arturo, et al.
Veröffentlicht: (2025)
von: Pérez-Peralta, Arturo, et al.
Veröffentlicht: (2025)
ArcGen: Generalizing Neural Backdoor Detection Across Diverse Architectures
von: Yang, Zhonghao, et al.
Veröffentlicht: (2025)
von: Yang, Zhonghao, et al.
Veröffentlicht: (2025)
GLAD: Improving Latent Graph Generative Modeling with Simple Quantization
von: Nguyen, Van Khoa, et al.
Veröffentlicht: (2024)
von: Nguyen, Van Khoa, et al.
Veröffentlicht: (2024)
Branch Scaling Manifests as Implicit Architectural Regularization for Improving Generalization in Overparameterized ResNets
von: Yu, Zixiong, et al.
Veröffentlicht: (2024)
von: Yu, Zixiong, et al.
Veröffentlicht: (2024)
Bridging Quantum and Classical Computing in Drug Design: Architecture Principles for Improved Molecule Generation
von: Smith, Andrew, et al.
Veröffentlicht: (2025)
von: Smith, Andrew, et al.
Veröffentlicht: (2025)
Pruning Increases Orderedness in Recurrent Computation
von: Song, Yiding
Veröffentlicht: (2025)
von: Song, Yiding
Veröffentlicht: (2025)
Are Your Generated Instances Truly Useful? GenBench-MILP: A Benchmark Suite for MILP Instance Generation
von: Luo, Yidong, et al.
Veröffentlicht: (2025)
von: Luo, Yidong, et al.
Veröffentlicht: (2025)
Detecting Blinks in Healthy and Parkinson's EEG: A Deep Learning Perspective
von: Lensky, Artem, et al.
Veröffentlicht: (2025)
von: Lensky, Artem, et al.
Veröffentlicht: (2025)
Model Capacity Determines Grokking through Competing Memorisation and Generalisation Speeds
von: Song, Yiding, et al.
Veröffentlicht: (2026)
von: Song, Yiding, et al.
Veröffentlicht: (2026)
Convergence Of Consistency Model With Multistep Sampling Under General Data Assumptions
von: Chen, Yiding, et al.
Veröffentlicht: (2025)
von: Chen, Yiding, et al.
Veröffentlicht: (2025)
AppellateGen: A Benchmark for Appellate Legal Judgment Generation
von: Yang, Hongkun, et al.
Veröffentlicht: (2026)
von: Yang, Hongkun, et al.
Veröffentlicht: (2026)
GenTS: A Comprehensive Benchmark Library for Generative Time Series Models
von: Wang, Chenxi, et al.
Veröffentlicht: (2026)
von: Wang, Chenxi, et al.
Veröffentlicht: (2026)
Scaling Laws and Representation Learning in Simple Hierarchical Languages: Transformers vs. Convolutional Architectures
von: Cagnetta, Francesco, et al.
Veröffentlicht: (2025)
von: Cagnetta, Francesco, et al.
Veröffentlicht: (2025)
From Entropy to Epiplexity: Rethinking Information for Computationally Bounded Intelligence
von: Finzi, Marc, et al.
Veröffentlicht: (2026)
von: Finzi, Marc, et al.
Veröffentlicht: (2026)
Bencher: Simple and Reproducible Benchmarking for Black-Box Optimization
von: Papenmeier, Leonard, et al.
Veröffentlicht: (2025)
von: Papenmeier, Leonard, et al.
Veröffentlicht: (2025)
Improved Robust Estimation for Erdős-Rényi Graphs: The Sparse Regime and Optimal Breakdown Point
von: Chen, Hongjie, et al.
Veröffentlicht: (2025)
von: Chen, Hongjie, et al.
Veröffentlicht: (2025)
Vision Transformer Neural Architecture Search for Out-of-Distribution Generalization: Benchmark and Insights
von: Ho, Sy-Tuyen, et al.
Veröffentlicht: (2025)
von: Ho, Sy-Tuyen, et al.
Veröffentlicht: (2025)
StiefelGen: A Simple, Model Agnostic Approach for Time Series Data Augmentation over Riemannian Manifolds
von: Cheema, Prasad, et al.
Veröffentlicht: (2024)
von: Cheema, Prasad, et al.
Veröffentlicht: (2024)
It's Not That Simple. An Analysis of Simple Test-Time Scaling
von: Wu, Guojun
Veröffentlicht: (2025)
von: Wu, Guojun
Veröffentlicht: (2025)
EvolveGen: Algorithmic Level Hardware Model Checking Benchmark Generation through Reinforcement Learning
von: Hu, Guangyu, et al.
Veröffentlicht: (2026)
von: Hu, Guangyu, et al.
Veröffentlicht: (2026)
Non-Stationary Online Resource Allocation: Learning from a Single Sample
von: Feng, Yiding, et al.
Veröffentlicht: (2026)
von: Feng, Yiding, et al.
Veröffentlicht: (2026)
Active Context Selection Improves Simple Regret in Contextual Bandits
von: Shahverdikondori, Mohammad, et al.
Veröffentlicht: (2026)
von: Shahverdikondori, Mohammad, et al.
Veröffentlicht: (2026)
ScribbleGen: Generative Data Augmentation Improves Scribble-supervised Semantic Segmentation
von: Schnell, Jacob, et al.
Veröffentlicht: (2023)
von: Schnell, Jacob, et al.
Veröffentlicht: (2023)
RadioGen3D: 3D Radio Map Generation via Adversarial Learning on Large-Scale Synthetic Data
von: Chen, Junshen, et al.
Veröffentlicht: (2026)
von: Chen, Junshen, et al.
Veröffentlicht: (2026)
Hypothesis Testing the Circuit Hypothesis in LLMs
von: Shi, Claudia, et al.
Veröffentlicht: (2024)
von: Shi, Claudia, et al.
Veröffentlicht: (2024)
Improve Cross-Architecture Generalization on Dataset Distillation
von: Zhou, Binglin, et al.
Veröffentlicht: (2024)
von: Zhou, Binglin, et al.
Veröffentlicht: (2024)
Generative Medical Event Models Improve with Scale
von: Waxler, Shane, et al.
Veröffentlicht: (2025)
von: Waxler, Shane, et al.
Veröffentlicht: (2025)
Principled Architecture-aware Scaling of Hyperparameters
von: Chen, Wuyang, et al.
Veröffentlicht: (2024)
von: Chen, Wuyang, et al.
Veröffentlicht: (2024)
Mechanistic Design and Scaling of Hybrid Architectures
von: Poli, Michael, et al.
Veröffentlicht: (2024)
von: Poli, Michael, et al.
Veröffentlicht: (2024)
SteinGen: Generating Fidelitous and Diverse Graph Samples
von: Reinert, Gesine, et al.
Veröffentlicht: (2024)
von: Reinert, Gesine, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
ProcGen3D: Learning Neural Procedural Graph Representations for Image-to-3D Reconstruction
von: Zhang, Xinyi, et al.
Veröffentlicht: (2025) -
Can Generative AI Solve Your In-Context Learning Problem? A Martingale Perspective
von: Jesson, Andrew, et al.
Veröffentlicht: (2024) -
ProcBench: Benchmark for Multi-Step Reasoning and Following Procedure
von: Fujisawa, Ippei, et al.
Veröffentlicht: (2024) -
ReLU to the Rescue: Improve Your On-Policy Actor-Critic with Positive Advantages
von: Jesson, Andrew, et al.
Veröffentlicht: (2023) -
JULI: Jailbreak Large Language Models by Self-Introspection
von: Wang, Jesson, et al.
Veröffentlicht: (2025)