LCDB 1.1: A Database Illustrating Learning Curves Are More Ill-Behaved Than Previously Thought
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Yan, Cheng, Mohr, Felix, Viering, Tom |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Transformers can do Bayesian Clustering
par: Bhaskaran, Prajit, et autres
Publié: (2025)
par: Bhaskaran, Prajit, et autres
Publié: (2025)
The Unreasonable Effectiveness Of Early Discarding After One Epoch In Neural Network Hyperparameter Optimization
par: Egele, Romain, et autres
Publié: (2024)
par: Egele, Romain, et autres
Publié: (2024)
More Than Routing: Joint GPS and Route Modeling for Refine Trajectory Representation Learning
par: Ma, Zhipeng, et autres
Publié: (2024)
par: Ma, Zhipeng, et autres
Publié: (2024)
A Framework for Mining Collectively-Behaving Bots in MMORPGs
par: Kim, Hyunsoo, et autres
Publié: (2025)
par: Kim, Hyunsoo, et autres
Publié: (2025)
Return-to-Go Is More Than a Number: Q-Guided Alignment for Return-Conditioned Supervised Learning
par: Yang, Yuxiao, et autres
Publié: (2026)
par: Yang, Yuxiao, et autres
Publié: (2026)
Why DDIM Hallucinates More Than DDPM: A Theoretical Analysis of Reverse Dynamics
par: Ashiq, Muhammad H., et autres
Publié: (2026)
par: Ashiq, Muhammad H., et autres
Publié: (2026)
Graph Neural Networks Are More Than Filters: Revisiting and Benchmarking from A Spectral Perspective
par: Dong, Yushun, et autres
Publié: (2024)
par: Dong, Yushun, et autres
Publié: (2024)
Curriculum Is More Influential Than Haptic Information During Reinforcement Learning of Object Manipulation Against Gravity
par: Ojaghi, Pegah, et autres
Publié: (2024)
par: Ojaghi, Pegah, et autres
Publié: (2024)
MiCA Learns More Knowledge Than LoRA and Full Fine-Tuning
par: Rüdiger, Sten, et autres
Publié: (2026)
par: Rüdiger, Sten, et autres
Publié: (2026)
Labels Matter More Than Models: Rethinking the Unsupervised Paradigm in Time Series Anomaly Detection
par: Zhong, Zhijie, et autres
Publié: (2025)
par: Zhong, Zhijie, et autres
Publié: (2025)
Shape of Thought: When Distribution Matters More than Correctness in Reasoning Tasks
par: Chandra, Abhranil, et autres
Publié: (2025)
par: Chandra, Abhranil, et autres
Publié: (2025)
More Than Irrational: Modeling Belief-Biased Agents
par: Zhu, Yifan, et autres
Publié: (2025)
par: Zhu, Yifan, et autres
Publié: (2025)
Is Q-learning an Ill-posed Problem?
par: Wissmann, Philipp, et autres
Publié: (2025)
par: Wissmann, Philipp, et autres
Publié: (2025)
VLA Models Are More Generalizable Than You Think: Revisiting Physical and Spatial Modeling
par: Li, Weiqi, et autres
Publié: (2025)
par: Li, Weiqi, et autres
Publié: (2025)
Feedback Over Form: Why Execution Feedback Matters More Than Pipeline Topology in 1-3B Code Generation
par: McAndrews, Charles Junichi
Publié: (2026)
par: McAndrews, Charles Junichi
Publié: (2026)
More Than Bits: Multi-Envelope Double Binary Factorization for Extreme Quantization
par: Ichikawa, Yuma, et autres
Publié: (2025)
par: Ichikawa, Yuma, et autres
Publié: (2025)
When Models Know More Than They Say: Probing Analogical Reasoning in LLMs
par: McGovern, Hope, et autres
Publié: (2026)
par: McGovern, Hope, et autres
Publié: (2026)
When More is Less: Understanding Chain-of-Thought Length in LLMs
par: Wu, Yuyang, et autres
Publié: (2025)
par: Wu, Yuyang, et autres
Publié: (2025)
Group and Exclusive Sparse Regularization-based Continual Learning of CNNs
par: Tousside, Basile, et autres
Publié: (2026)
par: Tousside, Basile, et autres
Publié: (2026)
More Than One Teacher: Adaptive Multi-Guidance Policy Optimization for Diverse Exploration
par: Yuan, Xiaoyang, et autres
Publié: (2025)
par: Yuan, Xiaoyang, et autres
Publié: (2025)
Rethinking Layer Redundancy: Calibration Matters More Than Search in LLM Depth Pruning
par: Kim, Minkyu, et autres
Publié: (2026)
par: Kim, Minkyu, et autres
Publié: (2026)
CoDAR: Continuous Diffusion Language Models are More Powerful Than You Think
par: Shen, Junzhe, et autres
Publié: (2026)
par: Shen, Junzhe, et autres
Publié: (2026)
GRILL: Restoring Gradient Signal in Ill-Conditioned Layers for More Effective Adversarial Attacks on Autoencoders
par: Ramanaik, Chethan Krishnamurthy, et autres
Publié: (2025)
par: Ramanaik, Chethan Krishnamurthy, et autres
Publié: (2025)
Angle Domain Guidance: Latent Diffusion Requires Rotation Rather Than Extrapolation
par: Jin, Cheng, et autres
Publié: (2025)
par: Jin, Cheng, et autres
Publié: (2025)
ChemFixer: Correcting Invalid Molecules to Unlock Previously Unseen Chemical Space
par: Park, Jun-Hyoung, et autres
Publié: (2025)
par: Park, Jun-Hyoung, et autres
Publié: (2025)
Assessing Social Alignment: Do Personality-Prompted Large Language Models Behave Like Humans?
par: Zakazov, Ivan, et autres
Publié: (2024)
par: Zakazov, Ivan, et autres
Publié: (2024)
Only the Curve Shape Matters: Training Foundation Models for Zero-Shot Multivariate Time Series Forecasting through Next Curve Shape Prediction
par: Feng, Cheng, et autres
Publié: (2024)
par: Feng, Cheng, et autres
Publié: (2024)
The Recipe Matters More Than the Kitchen:Mathematical Foundations of the AI Weather Prediction Pipeline
par: Garg, Piyush, et autres
Publié: (2026)
par: Garg, Piyush, et autres
Publié: (2026)
Fine, I'll Merge It Myself: A Multi-Fidelity Framework for Automated Model Merging
par: Su, Guinan, et autres
Publié: (2025)
par: Su, Guinan, et autres
Publié: (2025)
Atoms of Thought: Universal EEG Representation Learning with Microstates
par: Tian, Xinyang, et autres
Publié: (2026)
par: Tian, Xinyang, et autres
Publié: (2026)
Are Transformers More Robust? Towards Exact Robustness Verification for Transformers
par: Liao, Brian Hsuan-Cheng, et autres
Publié: (2022)
par: Liao, Brian Hsuan-Cheng, et autres
Publié: (2022)
Learning to Correct: Calibrated Reinforcement Learning for Multi-Attempt Chain-of-Thought
par: Ildiz, Muhammed Emrullah, et autres
Publié: (2026)
par: Ildiz, Muhammed Emrullah, et autres
Publié: (2026)
Retrieval-of-Thought: Efficient Reasoning via Reusing Thoughts
par: Ahmed, Ammar, et autres
Publié: (2025)
par: Ahmed, Ammar, et autres
Publié: (2025)
Rel-MOSS: Towards Imbalanced Relational Deep Learning on Relational Databases
par: Yin, Jun, et autres
Publié: (2026)
par: Yin, Jun, et autres
Publié: (2026)
Everything of Thoughts: Defying the Law of Penrose Triangle for Thought Generation
par: Ding, Ruomeng, et autres
Publié: (2023)
par: Ding, Ruomeng, et autres
Publié: (2023)
The Kinetics of Reasoning: How Chain-of-Thought Shapes Learning in Transformers?
par: Pengmei, Zihan, et autres
Publié: (2025)
par: Pengmei, Zihan, et autres
Publié: (2025)
Solving for X and Beyond: Can Large Language Models Solve Complex Math Problems with More-Than-Two Unknowns?
par: Kao, Kuei-Chun, et autres
Publié: (2024)
par: Kao, Kuei-Chun, et autres
Publié: (2024)
ResNets Are Deeper Than You Think
par: Mehmeti-Göpel, Christian H. X. Ali, et autres
Publié: (2025)
par: Mehmeti-Göpel, Christian H. X. Ali, et autres
Publié: (2025)
Transformers Learn to Implement Multi-step Gradient Descent with Chain of Thought
par: Huang, Jianhao, et autres
Publié: (2025)
par: Huang, Jianhao, et autres
Publié: (2025)
Thought Cloning: Learning to Think while Acting by Imitating Human Thinking
par: Hu, Shengran, et autres
Publié: (2023)
par: Hu, Shengran, et autres
Publié: (2023)
Documents similaires
-
Transformers can do Bayesian Clustering
par: Bhaskaran, Prajit, et autres
Publié: (2025) -
The Unreasonable Effectiveness Of Early Discarding After One Epoch In Neural Network Hyperparameter Optimization
par: Egele, Romain, et autres
Publié: (2024) -
More Than Routing: Joint GPS and Route Modeling for Refine Trajectory Representation Learning
par: Ma, Zhipeng, et autres
Publié: (2024) -
A Framework for Mining Collectively-Behaving Bots in MMORPGs
par: Kim, Hyunsoo, et autres
Publié: (2025) -
Return-to-Go Is More Than a Number: Q-Guided Alignment for Return-Conditioned Supervised Learning
par: Yang, Yuxiao, et autres
Publié: (2026)