Pretrained Hybrids with MAD Skills
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Roberts, Nicholas, Guo, Samuel, Gao, Zhiqi, GNVV, Satya Sai Srinath Namburi, Cromp, Sonia, Wu, Chengjun, Duan, Chengyu, Sala, Frederic |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Tabby: A Language Model Architecture for Tabular and Structured Data Synthesis
von: Cromp, Sonia, et al.
Veröffentlicht: (2025)
von: Cromp, Sonia, et al.
Veröffentlicht: (2025)
CARE: Confounder-Aware Aggregation for Reliable LLM Evaluation
von: Zhao, Jitian, et al.
Veröffentlicht: (2026)
von: Zhao, Jitian, et al.
Veröffentlicht: (2026)
R&B: Domain Regrouping and Data Mixture Balancing for Efficient Foundation Model Training
von: Ge, Albert, et al.
Veröffentlicht: (2025)
von: Ge, Albert, et al.
Veröffentlicht: (2025)
LETS Forecast: Learning Embedology for Time Series Forecasting
von: Majeedi, Abrar, et al.
Veröffentlicht: (2025)
von: Majeedi, Abrar, et al.
Veröffentlicht: (2025)
RICA2: Rubric-Informed, Calibrated Assessment of Actions
von: Majeedi, Abrar, et al.
Veröffentlicht: (2024)
von: Majeedi, Abrar, et al.
Veröffentlicht: (2024)
Pearls from Pebbles: Improved Confidence Functions for Auto-labeling
von: Vishwakarma, Harit, et al.
Veröffentlicht: (2024)
von: Vishwakarma, Harit, et al.
Veröffentlicht: (2024)
Linguistic properties and model scale in brain encoding: from small to compressed language models
von: Oota, Subba Reddy, et al.
Veröffentlicht: (2026)
von: Oota, Subba Reddy, et al.
Veröffentlicht: (2026)
OTTER: Effortless Label Distribution Adaptation of Zero-shot Models
von: Shin, Changho, et al.
Veröffentlicht: (2024)
von: Shin, Changho, et al.
Veröffentlicht: (2024)
Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli
von: Oota, Subba Reddy, et al.
Veröffentlicht: (2025)
von: Oota, Subba Reddy, et al.
Veröffentlicht: (2025)
MoRe Fine-Tuning with 10x Fewer Parameters
von: Tan, Wenxuan, et al.
Veröffentlicht: (2024)
von: Tan, Wenxuan, et al.
Veröffentlicht: (2024)
Brain alignment of reasoning and action representations from vision-language and action models during naturalistic gameplay
von: Oota, Subba Reddy, et al.
Veröffentlicht: (2026)
von: Oota, Subba Reddy, et al.
Veröffentlicht: (2026)
Compute Optimal Scaling of Skills: Knowledge vs Reasoning
von: Roberts, Nicholas, et al.
Veröffentlicht: (2025)
von: Roberts, Nicholas, et al.
Veröffentlicht: (2025)
Pareto Optimal Code Generation
von: Orlanski, Gabriel, et al.
Veröffentlicht: (2025)
von: Orlanski, Gabriel, et al.
Veröffentlicht: (2025)
Evaluating Language Model Context Windows: A "Working Memory" Test and Inference-time Correction
von: Dsouza, Amanda, et al.
Veröffentlicht: (2024)
von: Dsouza, Amanda, et al.
Veröffentlicht: (2024)
MASS: Mathematical Data Selection via Skill Graphs for Pretraining Large Language Models
von: Li, Jiazheng, et al.
Veröffentlicht: (2025)
von: Li, Jiazheng, et al.
Veröffentlicht: (2025)
Test-Time Scaling Makes Overtraining Compute-Optimal
von: Roberts, Nicholas, et al.
Veröffentlicht: (2026)
von: Roberts, Nicholas, et al.
Veröffentlicht: (2026)
SlopCodeBench: Benchmarking How Coding Agents Degrade Over Long-Horizon Iterative Tasks
von: Orlanski, Gabriel, et al.
Veröffentlicht: (2026)
von: Orlanski, Gabriel, et al.
Veröffentlicht: (2026)
Correlating instruction-tuning (in multimodal models) with vision-language processing (in the brain)
von: Oota, Subba Reddy, et al.
Veröffentlicht: (2025)
von: Oota, Subba Reddy, et al.
Veröffentlicht: (2025)
Beyond URLs: Metadata Diversity and Position for Efficient LLM Pretraining
von: Fan, Dongyang, et al.
Veröffentlicht: (2025)
von: Fan, Dongyang, et al.
Veröffentlicht: (2025)
Harnessing the Intrinsic Knowledge of Pretrained Language Models for Challenging Text Classification Settings
von: Gao, Lingyu
Veröffentlicht: (2024)
von: Gao, Lingyu
Veröffentlicht: (2024)
Generalizable and Stable Finetuning of Pretrained Language Models on Low-Resource Texts
von: Somayajula, Sai Ashish, et al.
Veröffentlicht: (2024)
von: Somayajula, Sai Ashish, et al.
Veröffentlicht: (2024)
The ALCHEmist: Automated Labeling 500x CHEaper Than LLM Data Annotators
von: Huang, Tzu-Heng, et al.
Veröffentlicht: (2024)
von: Huang, Tzu-Heng, et al.
Veröffentlicht: (2024)
From Pragmas to Partners: A Symbiotic Evolution of Agentic High-Level Synthesis
von: Zhang, Niansong, et al.
Veröffentlicht: (2026)
von: Zhang, Niansong, et al.
Veröffentlicht: (2026)
Preference Curriculum: LLMs Should Always Be Pretrained on Their Preferred Data
von: Zhang, Xuemiao, et al.
Veröffentlicht: (2025)
von: Zhang, Xuemiao, et al.
Veröffentlicht: (2025)
MMG2Skill: Can Agents Distill In-the-Wild Guides into Self-Evolving Skills?
von: Che, Xinyu, et al.
Veröffentlicht: (2026)
von: Che, Xinyu, et al.
Veröffentlicht: (2026)
daVinci-LLM:Towards the Science of Pretraining
von: Qin, Yiwei, et al.
Veröffentlicht: (2026)
von: Qin, Yiwei, et al.
Veröffentlicht: (2026)
M-MAD: Multidimensional Multi-Agent Debate for Advanced Machine Translation Evaluation
von: Feng, Zhaopeng, et al.
Veröffentlicht: (2024)
von: Feng, Zhaopeng, et al.
Veröffentlicht: (2024)
WRAP++: Web discoveRy Amplified Pretraining
von: Zhou, Jiang, et al.
Veröffentlicht: (2026)
von: Zhou, Jiang, et al.
Veröffentlicht: (2026)
Riddle Quest : The Enigma of Words
von: Parasa, Niharika Sri, et al.
Veröffentlicht: (2026)
von: Parasa, Niharika Sri, et al.
Veröffentlicht: (2026)
ESC-Skills: Discovering and Self-Evolving Skills for Emotional Support Conversations
von: Zhu, Jie, et al.
Veröffentlicht: (2026)
von: Zhu, Jie, et al.
Veröffentlicht: (2026)
Knowledgeable In-Context Tuning: Exploring and Exploiting Factual Knowledge for In-Context Learning
von: Wang, Jianing, et al.
Veröffentlicht: (2023)
von: Wang, Jianing, et al.
Veröffentlicht: (2023)
SkillOpt: Executive Strategy for Self-Evolving Agent Skills
von: Yang, Yifan, et al.
Veröffentlicht: (2026)
von: Yang, Yifan, et al.
Veröffentlicht: (2026)
Back to the Barn with LLAMAs: Evolving Pretrained LLM Backbones in Finetuning Vision Language Models
von: Horawalavithana, Sameera, et al.
Veröffentlicht: (2026)
von: Horawalavithana, Sameera, et al.
Veröffentlicht: (2026)
Towards Goal-oriented Prompt Engineering for Large Language Models: A Survey
von: Li, Haochen, et al.
Veröffentlicht: (2024)
von: Li, Haochen, et al.
Veröffentlicht: (2024)
Generative Pretrained Structured Transformers: Unsupervised Syntactic Language Models at Scale
von: Hu, Xiang, et al.
Veröffentlicht: (2024)
von: Hu, Xiang, et al.
Veröffentlicht: (2024)
Group of Skills: Group-Structured Skill Retrieval for Agent Skill Libraries
von: Zeng, Kun, et al.
Veröffentlicht: (2026)
von: Zeng, Kun, et al.
Veröffentlicht: (2026)
Alignment Pretraining: AI Discourse Causes Self-Fulfilling (Mis)alignment
von: Tice, Cameron, et al.
Veröffentlicht: (2026)
von: Tice, Cameron, et al.
Veröffentlicht: (2026)
Memory Decoder: A Pretrained, Plug-and-Play Memory for Large Language Models
von: Cao, Jiaqi, et al.
Veröffentlicht: (2025)
von: Cao, Jiaqi, et al.
Veröffentlicht: (2025)
S$^2$-MAD: Breaking the Token Barrier to Enhance Multi-Agent Debate Efficiency
von: Zeng, Yuting, et al.
Veröffentlicht: (2025)
von: Zeng, Yuting, et al.
Veröffentlicht: (2025)
DCRM: A Heuristic to Measure Response Pair Quality in Preference Optimization
von: Huang, Chengyu, et al.
Veröffentlicht: (2025)
von: Huang, Chengyu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Tabby: A Language Model Architecture for Tabular and Structured Data Synthesis
von: Cromp, Sonia, et al.
Veröffentlicht: (2025) -
CARE: Confounder-Aware Aggregation for Reliable LLM Evaluation
von: Zhao, Jitian, et al.
Veröffentlicht: (2026) -
R&B: Domain Regrouping and Data Mixture Balancing for Efficient Foundation Model Training
von: Ge, Albert, et al.
Veröffentlicht: (2025) -
LETS Forecast: Learning Embedology for Time Series Forecasting
von: Majeedi, Abrar, et al.
Veröffentlicht: (2025) -
RICA2: Rubric-Informed, Calibrated Assessment of Actions
von: Majeedi, Abrar, et al.
Veröffentlicht: (2024)