Salvato in:
| Autori principali: | Vysogorets, Artem, Gopal, Achintya |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2402.15613 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
DRoP: Distributionally Robust Data Pruning
di: Vysogorets, Artem, et al.
Pubblicazione: (2024)
di: Vysogorets, Artem, et al.
Pubblicazione: (2024)
Reassessing Active Learning Adoption in Contemporary NLP: A Community Survey
di: Romberg, Julia, et al.
Pubblicazione: (2025)
di: Romberg, Julia, et al.
Pubblicazione: (2025)
Does Differential Privacy Impact Bias in Pretrained NLP Models?
di: Islam, Md. Khairul, et al.
Pubblicazione: (2024)
di: Islam, Md. Khairul, et al.
Pubblicazione: (2024)
NeuralFactors: A Novel Factor Learning Approach to Generative Modeling of Equities
di: Gopal, Achintya
Pubblicazione: (2024)
di: Gopal, Achintya
Pubblicazione: (2024)
Deconstructing the Goldilocks Zone of Neural Network Initialization
di: Vysogorets, Artem, et al.
Pubblicazione: (2024)
di: Vysogorets, Artem, et al.
Pubblicazione: (2024)
Improving Language Plasticity via Pretraining with Active Forgetting
di: Chen, Yihong, et al.
Pubblicazione: (2023)
di: Chen, Yihong, et al.
Pubblicazione: (2023)
Efficient Stagewise Pretraining via Progressive Subnetworks
di: Panigrahi, Abhishek, et al.
Pubblicazione: (2024)
di: Panigrahi, Abhishek, et al.
Pubblicazione: (2024)
Pretraining Language Models with Subword Regularization: An Empirical Study of BPE Dropout in Low-Resource NLP
di: Visser, Ruan, et al.
Pubblicazione: (2026)
di: Visser, Ruan, et al.
Pubblicazione: (2026)
A Novel Recurrent Neural Network Framework for Prediction and Treatment of Oncogenic Mutation Progression
di: Parthasarathy, Rishab, et al.
Pubblicazione: (2025)
di: Parthasarathy, Rishab, et al.
Pubblicazione: (2025)
Efficient Causal Discovery for Autoregressive Time Series
di: Fesanghary, Mohammad, et al.
Pubblicazione: (2025)
di: Fesanghary, Mohammad, et al.
Pubblicazione: (2025)
On Importance of Pruning and Distillation for Efficient Low Resource NLP
di: Mirashi, Aishwarya, et al.
Pubblicazione: (2024)
di: Mirashi, Aishwarya, et al.
Pubblicazione: (2024)
Filling in Missing FX Implied Volatilities with Uncertainties: Improving VAE-Based Volatility Imputation
di: Gopal, Achintya
Pubblicazione: (2024)
di: Gopal, Achintya
Pubblicazione: (2024)
NLP-ADBench: NLP Anomaly Detection Benchmark
di: Li, Yuangang, et al.
Pubblicazione: (2024)
di: Li, Yuangang, et al.
Pubblicazione: (2024)
TiME: Tiny Monolingual Encoders for Efficient NLP Pipelines
di: Schulmeister, David, et al.
Pubblicazione: (2025)
di: Schulmeister, David, et al.
Pubblicazione: (2025)
Efficiently Distilling LLMs for Edge Applications
di: Kundu, Achintya, et al.
Pubblicazione: (2024)
di: Kundu, Achintya, et al.
Pubblicazione: (2024)
Group-Level Data Selection for Efficient Pretraining
di: Yu, Zichun, et al.
Pubblicazione: (2025)
di: Yu, Zichun, et al.
Pubblicazione: (2025)
SEMFED: Semantic-Aware Resource-Efficient Federated Learning for Heterogeneous NLP Tasks
di: Hussain, Sajid, et al.
Pubblicazione: (2025)
di: Hussain, Sajid, et al.
Pubblicazione: (2025)
Don't Pay Attention, PLANT It: Pretraining Attention via Learning-to-Rank
di: Roy, Debjyoti Saha, et al.
Pubblicazione: (2024)
di: Roy, Debjyoti Saha, et al.
Pubblicazione: (2024)
Aviation Safety Enhancement via NLP & Deep Learning: Classifying Flight Phases in ATSB Safety Reports
di: Nanyonga, Aziida, et al.
Pubblicazione: (2025)
di: Nanyonga, Aziida, et al.
Pubblicazione: (2025)
NLP Verification: Towards a General Methodology for Certifying Robustness
di: Casadio, Marco, et al.
Pubblicazione: (2024)
di: Casadio, Marco, et al.
Pubblicazione: (2024)
Federated Learning with Layer Skipping: Efficient Training of Large Language Models for Healthcare NLP
di: Zhang, Lihong, et al.
Pubblicazione: (2025)
di: Zhang, Lihong, et al.
Pubblicazione: (2025)
AdaLRS: Loss-Guided Adaptive Learning Rate Search for Efficient Foundation Model Pretraining
di: Dong, Hongyuan, et al.
Pubblicazione: (2025)
di: Dong, Hongyuan, et al.
Pubblicazione: (2025)
Dense vs Sparse Pretraining at Tiny Scale: Active-Parameter vs Total-Parameter Matching
di: Wael, Abdalrahman
Pubblicazione: (2026)
di: Wael, Abdalrahman
Pubblicazione: (2026)
Detecting AI Generated Text Based on NLP and Machine Learning Approaches
di: Prova, Nuzhat
Pubblicazione: (2024)
di: Prova, Nuzhat
Pubblicazione: (2024)
Megalodon: Efficient LLM Pretraining and Inference with Unlimited Context Length
di: Ma, Xuezhe, et al.
Pubblicazione: (2024)
di: Ma, Xuezhe, et al.
Pubblicazione: (2024)
ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining
di: Bal, Melis Ilayda, et al.
Pubblicazione: (2025)
di: Bal, Melis Ilayda, et al.
Pubblicazione: (2025)
Privacy Evaluation Benchmarks for NLP Models
di: Huang, Wei, et al.
Pubblicazione: (2024)
di: Huang, Wei, et al.
Pubblicazione: (2024)
Efficient and Flexible Topic Modeling using Pretrained Embeddings and Bag of Sentences
di: Schneider, Johannes
Pubblicazione: (2023)
di: Schneider, Johannes
Pubblicazione: (2023)
Language Model-Driven Data Pruning Enables Efficient Active Learning
di: Azeemi, Abdul Hameed, et al.
Pubblicazione: (2024)
di: Azeemi, Abdul Hameed, et al.
Pubblicazione: (2024)
AnchorAL: Computationally Efficient Active Learning for Large and Imbalanced Datasets
di: Lesci, Pietro, et al.
Pubblicazione: (2024)
di: Lesci, Pietro, et al.
Pubblicazione: (2024)
HOP to the Next Tasks and Domains for Continual Learning in NLP
di: Michieli, Umberto, et al.
Pubblicazione: (2024)
di: Michieli, Umberto, et al.
Pubblicazione: (2024)
The NLP Task Effectiveness of Long-Range Transformers
di: Qin, Guanghui, et al.
Pubblicazione: (2022)
di: Qin, Guanghui, et al.
Pubblicazione: (2022)
APT: Adaptive Pruning and Tuning Pretrained Language Models for Efficient Training and Inference
di: Zhao, Bowen, et al.
Pubblicazione: (2024)
di: Zhao, Bowen, et al.
Pubblicazione: (2024)
MATES: Model-Aware Data Selection for Efficient Pretraining with Data Influence Models
di: Yu, Zichun, et al.
Pubblicazione: (2024)
di: Yu, Zichun, et al.
Pubblicazione: (2024)
Towards Safer Pretraining: Analyzing and Filtering Harmful Content in Webscale datasets for Responsible LLMs
di: Mendu, Sai Krishna, et al.
Pubblicazione: (2025)
di: Mendu, Sai Krishna, et al.
Pubblicazione: (2025)
Soup-of-Experts: Pretraining Specialist Models via Parameters Averaging
di: Ablin, Pierre, et al.
Pubblicazione: (2025)
di: Ablin, Pierre, et al.
Pubblicazione: (2025)
Not All Tokens Matter: Towards Efficient LLM Reasoning via Token Significance in Reinforcement Learning
di: Liu, Hanbing, et al.
Pubblicazione: (2025)
di: Liu, Hanbing, et al.
Pubblicazione: (2025)
Tracing the Representation Geometry of Language Models from Pretraining to Post-training
di: Li, Melody Zixuan, et al.
Pubblicazione: (2025)
di: Li, Melody Zixuan, et al.
Pubblicazione: (2025)
ActiveUltraFeedback: Efficient Preference Data Generation using Active Learning
di: Melikidze, Davit, et al.
Pubblicazione: (2026)
di: Melikidze, Davit, et al.
Pubblicazione: (2026)
Patent Representation Learning via Self-supervision
di: Zuo, You, et al.
Pubblicazione: (2025)
di: Zuo, You, et al.
Pubblicazione: (2025)
Documenti analoghi
-
DRoP: Distributionally Robust Data Pruning
di: Vysogorets, Artem, et al.
Pubblicazione: (2024) -
Reassessing Active Learning Adoption in Contemporary NLP: A Community Survey
di: Romberg, Julia, et al.
Pubblicazione: (2025) -
Does Differential Privacy Impact Bias in Pretrained NLP Models?
di: Islam, Md. Khairul, et al.
Pubblicazione: (2024) -
NeuralFactors: A Novel Factor Learning Approach to Generative Modeling of Equities
di: Gopal, Achintya
Pubblicazione: (2024) -
Deconstructing the Goldilocks Zone of Neural Network Initialization
di: Vysogorets, Artem, et al.
Pubblicazione: (2024)