ScaLearn: Simple and Highly Parameter-Efficient Task Transfer by Learning to Scale
Fuente:
arXiv
Guardado en:
| Autores principales: | Frohmann, Markus, Holtermann, Carolin, Masoudian, Shahed, Lauscher, Anne, Rekabsaz, Navid |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
What the Weight?! A Unified Framework for Zero-Shot Knowledge Composition
por: Holtermann, Carolin, et al.
Publicado: (2024)
por: Holtermann, Carolin, et al.
Publicado: (2024)
Unlabeled Debiasing in Downstream Tasks via Class-wise Low Variance Regularization
por: Masoudian, Shahed, et al.
Publicado: (2024)
por: Masoudian, Shahed, et al.
Publicado: (2024)
Effective Controllable Bias Mitigation for Classification and Retrieval using Gate Adapters
por: Masoudian, Shahed, et al.
Publicado: (2024)
por: Masoudian, Shahed, et al.
Publicado: (2024)
SoS: Analysis of Surface over Semantics in Multilingual Text-To-Image Generation
por: Holtermann, Carolin, et al.
Publicado: (2026)
por: Holtermann, Carolin, et al.
Publicado: (2026)
TempViz: On the Evaluation of Temporal Knowledge in Text-to-Image Models
por: Holtermann, Carolin, et al.
Publicado: (2026)
por: Holtermann, Carolin, et al.
Publicado: (2026)
Evaluating the Elementary Multilingual Capabilities of Large Language Models with MultiQ
por: Holtermann, Carolin, et al.
Publicado: (2024)
por: Holtermann, Carolin, et al.
Publicado: (2024)
Investigating Gender Bias in LLM-Generated Stories via Psychological Stereotypes
por: Masoudian, Shahed, et al.
Publicado: (2025)
por: Masoudian, Shahed, et al.
Publicado: (2025)
Segment Any Text: A Universal Approach for Robust, Efficient and Adaptable Sentence Segmentation
por: Frohmann, Markus, et al.
Publicado: (2024)
por: Frohmann, Markus, et al.
Publicado: (2024)
Around the World in 24 Hours: Probing LLM Knowledge of Time and Place
por: Holtermann, Carolin, et al.
Publicado: (2025)
por: Holtermann, Carolin, et al.
Publicado: (2025)
Synthetic Lyrics Detection Across Languages and Genres
por: Labrak, Yanis, et al.
Publicado: (2024)
por: Labrak, Yanis, et al.
Publicado: (2024)
Large Language Models for Human-Machine Collaborative Particle Accelerator Tuning through Natural Language
por: Kaiser, Jan, et al.
Publicado: (2024)
por: Kaiser, Jan, et al.
Publicado: (2024)
GIMMICK -- Globally Inclusive Multimodal Multitask Cultural Knowledge Benchmarking
por: Schneider, Florian, et al.
Publicado: (2025)
por: Schneider, Florian, et al.
Publicado: (2025)
MoSLD: An Extremely Parameter-Efficient Mixture-of-Shared LoRAs for Multi-Task Learning
por: Zhao, Lulu, et al.
Publicado: (2024)
por: Zhao, Lulu, et al.
Publicado: (2024)
Parameter Efficient Reinforcement Learning from Human Feedback
por: Sidahmed, Hakim, et al.
Publicado: (2024)
por: Sidahmed, Hakim, et al.
Publicado: (2024)
Increasing Model Capacity for Free: A Simple Strategy for Parameter Efficient Fine-tuning
por: Song, Haobo, et al.
Publicado: (2024)
por: Song, Haobo, et al.
Publicado: (2024)
It's Not That Simple. An Analysis of Simple Test-Time Scaling
por: Wu, Guojun
Publicado: (2025)
por: Wu, Guojun
Publicado: (2025)
SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning
por: Li, Haozhan, et al.
Publicado: (2025)
por: Li, Haozhan, et al.
Publicado: (2025)
Always Learning, Always Mixing: Efficient and Simple Data Mixing All The Time
por: Hu, Michael Y., et al.
Publicado: (2026)
por: Hu, Michael Y., et al.
Publicado: (2026)
Large Scale Transfer Learning for Tabular Data via Language Modeling
por: Gardner, Josh, et al.
Publicado: (2024)
por: Gardner, Josh, et al.
Publicado: (2024)
Learning Rate Scaling across LoRA Ranks and Transfer to Full Finetuning
por: Chen, Nan, et al.
Publicado: (2026)
por: Chen, Nan, et al.
Publicado: (2026)
Frontier LLMs Still Struggle with Simple Reasoning Tasks
por: Malek, Alan, et al.
Publicado: (2025)
por: Malek, Alan, et al.
Publicado: (2025)
Language Control Diffusion: Efficiently Scaling through Space, Time, and Tasks
por: Zhang, Edwin, et al.
Publicado: (2022)
por: Zhang, Edwin, et al.
Publicado: (2022)
Batched Contextual Reinforcement: A Task-Scaling Law for Efficient Reasoning
por: Yang, Bangji, et al.
Publicado: (2026)
por: Yang, Bangji, et al.
Publicado: (2026)
Unified Multi-Task Learning & Model Fusion for Efficient Language Model Guardrailing
por: Neill, James O', et al.
Publicado: (2025)
por: Neill, James O', et al.
Publicado: (2025)
SEMFED: Semantic-Aware Resource-Efficient Federated Learning for Heterogeneous NLP Tasks
por: Hussain, Sajid, et al.
Publicado: (2025)
por: Hussain, Sajid, et al.
Publicado: (2025)
OrchMoE: Efficient Multi-Adapter Learning with Task-Skill Synergy
por: Wang, Haowen, et al.
Publicado: (2024)
por: Wang, Haowen, et al.
Publicado: (2024)
AI-Generated Song Detection via Lyrics Transcripts
por: Frohmann, Markus, et al.
Publicado: (2025)
por: Frohmann, Markus, et al.
Publicado: (2025)
Decoding Multilingual Moral Preferences: Unveiling LLM's Biases Through the Moral Machine Experiment
por: Vida, Karina, et al.
Publicado: (2024)
por: Vida, Karina, et al.
Publicado: (2024)
Sensitivity, Performance, Robustness: Deconstructing the Effect of Sociodemographic Prompting
por: Beck, Tilman, et al.
Publicado: (2023)
por: Beck, Tilman, et al.
Publicado: (2023)
Multi-Task Reinforcement Learning Enables Parameter Scaling
por: McLean, Reginald, et al.
Publicado: (2025)
por: McLean, Reginald, et al.
Publicado: (2025)
The Curious Case of Factual (Mis)Alignment between LLMs' Short- and Long-Form Answers
por: Islam, Saad Obaid ul, et al.
Publicado: (2025)
por: Islam, Saad Obaid ul, et al.
Publicado: (2025)
How Much Do LLMs Hallucinate across Languages? On Realistic Multilingual Estimation of LLM Hallucination
por: Islam, Saad Obaid ul, et al.
Publicado: (2025)
por: Islam, Saad Obaid ul, et al.
Publicado: (2025)
Donors and Recipients: On Asymmetric Transfer Across Tasks and Languages with Parameter-Efficient Fine-Tuning
por: Dymkiewicz, Kajetan, et al.
Publicado: (2025)
por: Dymkiewicz, Kajetan, et al.
Publicado: (2025)
Ability Transfer and Recovery via Modularized Parameters Localization
por: Jin, Songyao, et al.
Publicado: (2026)
por: Jin, Songyao, et al.
Publicado: (2026)
MT$^{3}$: Scaling MLLM-based Text Image Machine Translation via Multi-Task Reinforcement Learning
por: Feng, Zhaopeng, et al.
Publicado: (2025)
por: Feng, Zhaopeng, et al.
Publicado: (2025)
Efficient Systematic Reviews: Literature Filtering with Transformers & Transfer Learning
por: Hawkins, John, et al.
Publicado: (2024)
por: Hawkins, John, et al.
Publicado: (2024)
SSR-Zero: Simple Self-Rewarding Reinforcement Learning for Machine Translation
por: Yang, Wenjie, et al.
Publicado: (2025)
por: Yang, Wenjie, et al.
Publicado: (2025)
Double Entendre: Robust Audio-Based AI-Generated Lyrics Detection via Multi-View Fusion
por: Frohmann, Markus, et al.
Publicado: (2025)
por: Frohmann, Markus, et al.
Publicado: (2025)
Facet-Level Tracing of Evidence Uncertainty and Hallucination in RAG
por: Elchafei, Passant, et al.
Publicado: (2026)
por: Elchafei, Passant, et al.
Publicado: (2026)
Leveraging Parameter Space Symmetries for Reasoning Skill Transfer in LLMs
por: Horoi, Stefan, et al.
Publicado: (2025)
por: Horoi, Stefan, et al.
Publicado: (2025)
Ejemplares similares
-
What the Weight?! A Unified Framework for Zero-Shot Knowledge Composition
por: Holtermann, Carolin, et al.
Publicado: (2024) -
Unlabeled Debiasing in Downstream Tasks via Class-wise Low Variance Regularization
por: Masoudian, Shahed, et al.
Publicado: (2024) -
Effective Controllable Bias Mitigation for Classification and Retrieval using Gate Adapters
por: Masoudian, Shahed, et al.
Publicado: (2024) -
SoS: Analysis of Surface over Semantics in Multilingual Text-To-Image Generation
por: Holtermann, Carolin, et al.
Publicado: (2026) -
TempViz: On the Evaluation of Temporal Knowledge in Text-to-Image Models
por: Holtermann, Carolin, et al.
Publicado: (2026)