Skill Set Optimization: Reinforcing Language Model Behavior via Transferable Skills
Fuente:
arXiv
Guardado en:
| Autores principales: | Nottingham, Kolby, Majumder, Bodhisattwa Prasad, Mishra, Bhavana Dalvi, Singh, Sameer, Clark, Peter, Fox, Roy |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
DiscoveryBench: Towards Data-Driven Discovery with Large Language Models
por: Majumder, Bodhisattwa Prasad, et al.
Publicado: (2024)
por: Majumder, Bodhisattwa Prasad, et al.
Publicado: (2024)
Latent Factor Models Meets Instructions: Goal-conditioned Latent Factor Discovery without Task Supervision
por: Xie, Zhouhang, et al.
Publicado: (2025)
por: Xie, Zhouhang, et al.
Publicado: (2025)
AutoDiscovery: Open-ended Scientific Discovery via Bayesian Surprise
por: Agarwal, Dhruv, et al.
Publicado: (2025)
por: Agarwal, Dhruv, et al.
Publicado: (2025)
DISCOVERYWORLD: A Virtual Environment for Developing and Evaluating Automated Scientific Discovery Agents
por: Jansen, Peter, et al.
Publicado: (2024)
por: Jansen, Peter, et al.
Publicado: (2024)
CodeScientist: End-to-End Semi-Automated Scientific Discovery with Code-based Experimentation
por: Jansen, Peter, et al.
Publicado: (2025)
por: Jansen, Peter, et al.
Publicado: (2025)
To Tell The Truth: Language of Deception and Language Models
por: Hazra, Sanchaita, et al.
Publicado: (2023)
por: Hazra, Sanchaita, et al.
Publicado: (2023)
BaRDa: A Belief and Reasoning Dataset that Separates Factual Accuracy and Reasoning Ability
por: Clark, Peter, et al.
Publicado: (2023)
por: Clark, Peter, et al.
Publicado: (2023)
Data-driven Discovery with Large Generative Models
por: Majumder, Bodhisattwa Prasad, et al.
Publicado: (2024)
por: Majumder, Bodhisattwa Prasad, et al.
Publicado: (2024)
Sotopia-RL: Reward Design for Social Intelligence
por: Yu, Haofei, et al.
Publicado: (2025)
por: Yu, Haofei, et al.
Publicado: (2025)
Tell, Don't Show!: Language Guidance Eases Transfer Across Domains in Images and Videos
por: Kalluri, Tarun, et al.
Publicado: (2024)
por: Kalluri, Tarun, et al.
Publicado: (2024)
The Good, the Bad, and the Ugly: The Role of AI Quality Disclosure in Lie Detection
por: Bhattacharya, Haimanti, et al.
Publicado: (2024)
por: Bhattacharya, Haimanti, et al.
Publicado: (2024)
ArtifactLinker: Linking Scientific Artifacts for Automatic State-of-the-Art Discovery
por: Yu, Haofei, et al.
Publicado: (2026)
por: Yu, Haofei, et al.
Publicado: (2026)
HypER: Literature-grounded Hypothesis Generation and Distillation with Provenance
por: Vasu, Rosni, et al.
Publicado: (2025)
por: Vasu, Rosni, et al.
Publicado: (2025)
AI Safety Should Prioritize the Future of Work
por: Hazra, Sanchaita, et al.
Publicado: (2025)
por: Hazra, Sanchaita, et al.
Publicado: (2025)
SkillOS: Learning Skill Curation for Self-Evolving Agents
por: Ouyang, Siru, et al.
Publicado: (2026)
por: Ouyang, Siru, et al.
Publicado: (2026)
Tailoring with Targeted Precision: Edit-Based Agents for Open-Domain Procedure Customization
por: Lal, Yash Kumar, et al.
Publicado: (2023)
por: Lal, Yash Kumar, et al.
Publicado: (2023)
HARPA: A Testability-Driven, Literature-Grounded Framework for Research Ideation
por: Vasu, Rosni, et al.
Publicado: (2025)
por: Vasu, Rosni, et al.
Publicado: (2025)
Can Language Models Compose Skills In-Context?
por: Liu, Zidong, et al.
Publicado: (2025)
por: Liu, Zidong, et al.
Publicado: (2025)
Maestro: Reinforcement Learning to Orchestrate Hierarchical Model-Skill Ensembles
por: Wu, Jinyang, et al.
Publicado: (2026)
por: Wu, Jinyang, et al.
Publicado: (2026)
LiSA: Lifelong Safety Adaptation via Conservative Policy Induction
por: Kim, Minbeom, et al.
Publicado: (2026)
por: Kim, Minbeom, et al.
Publicado: (2026)
From Models to Microtheories: Distilling a Model's Topical Knowledge for Grounded Question Answering
por: Weir, Nathaniel, et al.
Publicado: (2024)
por: Weir, Nathaniel, et al.
Publicado: (2024)
Dynamic Skill Lifecycle Management for Agentic Reinforcement Learning
por: Shen, Junhao, et al.
Publicado: (2026)
por: Shen, Junhao, et al.
Publicado: (2026)
Accepted with Minor Revisions: Value of AI-Assisted Scientific Writing
por: Hazra, Sanchaita, et al.
Publicado: (2025)
por: Hazra, Sanchaita, et al.
Publicado: (2025)
Merge to Learn: Efficiently Adding Skills to Language Models with Model Merging
por: Morrison, Jacob, et al.
Publicado: (2024)
por: Morrison, Jacob, et al.
Publicado: (2024)
Measuring Vision-Language STEM Skills of Neural Models
por: Shen, Jianhao, et al.
Publicado: (2024)
por: Shen, Jianhao, et al.
Publicado: (2024)
Language Guided Skill Discovery
por: Rho, Seungeun, et al.
Publicado: (2024)
por: Rho, Seungeun, et al.
Publicado: (2024)
SkillGraph: Skill-Augmented Reinforcement Learning for Agents via Evolving Skill Graphs
por: Li, Xiaoyuan, et al.
Publicado: (2026)
por: Li, Xiaoyuan, et al.
Publicado: (2026)
Skill-Based Mixture-of-Experts: Adaptive Routing for Heterogeneous Reasoning via Inferred Skills
por: Chen, Justin Chih-Yao, et al.
Publicado: (2025)
por: Chen, Justin Chih-Yao, et al.
Publicado: (2025)
Leveraging Parameter Space Symmetries for Reasoning Skill Transfer in LLMs
por: Horoi, Stefan, et al.
Publicado: (2025)
por: Horoi, Stefan, et al.
Publicado: (2025)
Put Your Money Where Your Mouth Is: Evaluating Strategic Planning and Execution of LLM Agents in an Auction Arena
por: Chen, Jiangjie, et al.
Publicado: (2023)
por: Chen, Jiangjie, et al.
Publicado: (2023)
Characterizing Model-Native Skills
por: Kang, Feiyang, et al.
Publicado: (2026)
por: Kang, Feiyang, et al.
Publicado: (2026)
Skill Availability and Presentation Granularity in Large-Language-Model Agents: A Controlled SkillsBench Study
por: Xu, Xiaonan, et al.
Publicado: (2026)
por: Xu, Xiaonan, et al.
Publicado: (2026)
SoMi-ToM: Evaluating Multi-Perspective Theory of Mind in Embodied Social Interactions
por: Fan, Xianzhe, et al.
Publicado: (2025)
por: Fan, Xianzhe, et al.
Publicado: (2025)
Knowledge Fusion of Large Language Models Via Modular SkillPacks
por: Du, Guodong, et al.
Publicado: (2025)
por: Du, Guodong, et al.
Publicado: (2025)
MemSkill: Learning and Evolving Memory Skills for Self-Evolving Agents
por: Zhang, Haozhen, et al.
Publicado: (2026)
por: Zhang, Haozhen, et al.
Publicado: (2026)
Parallel Token Prediction for Language Models
por: Draxler, Felix, et al.
Publicado: (2025)
por: Draxler, Felix, et al.
Publicado: (2025)
Language-guided Skill Learning with Temporal Variational Inference
por: Fu, Haotian, et al.
Publicado: (2024)
por: Fu, Haotian, et al.
Publicado: (2024)
TIDE: Textual Identity Detection for Evaluating and Augmenting Classification and Language Models
por: Klu, Emmanuel, et al.
Publicado: (2023)
por: Klu, Emmanuel, et al.
Publicado: (2023)
Enhancing Quantitative Reasoning Skills of Large Language Models through Dimension Perception
por: Huang, Yuncheng, et al.
Publicado: (2023)
por: Huang, Yuncheng, et al.
Publicado: (2023)
Enhancing Systematic Decompositional Natural Language Inference Using Informal Logic
por: Weir, Nathaniel, et al.
Publicado: (2024)
por: Weir, Nathaniel, et al.
Publicado: (2024)
Ejemplares similares
-
DiscoveryBench: Towards Data-Driven Discovery with Large Language Models
por: Majumder, Bodhisattwa Prasad, et al.
Publicado: (2024) -
Latent Factor Models Meets Instructions: Goal-conditioned Latent Factor Discovery without Task Supervision
por: Xie, Zhouhang, et al.
Publicado: (2025) -
AutoDiscovery: Open-ended Scientific Discovery via Bayesian Surprise
por: Agarwal, Dhruv, et al.
Publicado: (2025) -
DISCOVERYWORLD: A Virtual Environment for Developing and Evaluating Automated Scientific Discovery Agents
por: Jansen, Peter, et al.
Publicado: (2024) -
CodeScientist: End-to-End Semi-Automated Scientific Discovery with Code-based Experimentation
por: Jansen, Peter, et al.
Publicado: (2025)