Salvato in:
| Autori principali: | Nguyen, Trung, Leng, Yan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2502.16385 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Offline Preference Optimization via Maximum Marginal Likelihood Estimation
di: Najafi, Saeed, et al.
Pubblicazione: (2025)
di: Najafi, Saeed, et al.
Pubblicazione: (2025)
The Linear Representation Hypothesis and the Geometry of Large Language Models
di: Park, Kiho, et al.
Pubblicazione: (2023)
di: Park, Kiho, et al.
Pubblicazione: (2023)
Lizard: An Efficient Linearization Framework for Large Language Models
di: Van Nguyen, Chien, et al.
Pubblicazione: (2025)
di: Van Nguyen, Chien, et al.
Pubblicazione: (2025)
Linear Representation Transferability Hypothesis: Leveraging Small Models to Steer Large Models
di: Bello, Femi, et al.
Pubblicazione: (2025)
di: Bello, Femi, et al.
Pubblicazione: (2025)
Structured Pruning for Diverse Best-of-N Reasoning Optimization
di: Nguyen, Hieu Trung, et al.
Pubblicazione: (2025)
di: Nguyen, Hieu Trung, et al.
Pubblicazione: (2025)
Task-driven Layerwise Additive Activation Intervention
di: Nguyen, Hieu Trung, et al.
Pubblicazione: (2025)
di: Nguyen, Hieu Trung, et al.
Pubblicazione: (2025)
Exploring the Linear Subspace Hypothesis in Gender Bias Mitigation
di: Vargas, Francisco, et al.
Pubblicazione: (2020)
di: Vargas, Francisco, et al.
Pubblicazione: (2020)
How Many Features Can a Language Model Store Under the Linear Representation Hypothesis?
di: Garg, Nikhil, et al.
Pubblicazione: (2026)
di: Garg, Nikhil, et al.
Pubblicazione: (2026)
Towards Understanding Multi-Task Learning (Generalization) of LLMs via Detecting and Exploring Task-Specific Neurons
di: Leng, Yongqi, et al.
Pubblicazione: (2024)
di: Leng, Yongqi, et al.
Pubblicazione: (2024)
Explainable Disentangled Representation Learning for Generalizable Authorship Attribution in the Era of Generative AI
di: Man, Hieu, et al.
Pubblicazione: (2026)
di: Man, Hieu, et al.
Pubblicazione: (2026)
Optimizing Multi-Stage Language Models for Effective Text Retrieval
di: Trung, Quang Hoang, et al.
Pubblicazione: (2024)
di: Trung, Quang Hoang, et al.
Pubblicazione: (2024)
Adaptive Two-Phase Finetuning LLMs for Japanese Legal Text Retrieval
di: Trung, Quang Hoang, et al.
Pubblicazione: (2024)
di: Trung, Quang Hoang, et al.
Pubblicazione: (2024)
On the Origins of Linear Representations in Large Language Models
di: Jiang, Yibo, et al.
Pubblicazione: (2024)
di: Jiang, Yibo, et al.
Pubblicazione: (2024)
Reasoning Planning for Language Models
di: Nguyen, Bao, et al.
Pubblicazione: (2025)
di: Nguyen, Bao, et al.
Pubblicazione: (2025)
HypoBench: Towards Systematic and Principled Benchmarking for Hypothesis Generation
di: Liu, Haokun, et al.
Pubblicazione: (2025)
di: Liu, Haokun, et al.
Pubblicazione: (2025)
Momentum SVGD-EM for Accelerated Maximum Marginal Likelihood Estimation
di: Rozzio, Adam, et al.
Pubblicazione: (2026)
di: Rozzio, Adam, et al.
Pubblicazione: (2026)
Large Language Model-Enhanced Algorithm Selection: Towards Comprehensive Algorithm Representation
di: Wu, Xingyu, et al.
Pubblicazione: (2023)
di: Wu, Xingyu, et al.
Pubblicazione: (2023)
Don't Read Everything: A Curvature-Conditioned Query for Linear Attention
di: Le, Dong, et al.
Pubblicazione: (2026)
di: Le, Dong, et al.
Pubblicazione: (2026)
Mixture-of-Personas Language Models for Population Simulation
di: Bui, Ngoc, et al.
Pubblicazione: (2025)
di: Bui, Ngoc, et al.
Pubblicazione: (2025)
Adaptive Rollout Allocation for Online Reinforcement Learning with Verifiable Rewards
di: Nguyen, Hieu Trung, et al.
Pubblicazione: (2026)
di: Nguyen, Hieu Trung, et al.
Pubblicazione: (2026)
Towards Generalising Neural Topical Representations
di: Yang, Xiaohao, et al.
Pubblicazione: (2023)
di: Yang, Xiaohao, et al.
Pubblicazione: (2023)
NoveltyRank: A Retrieval-Augmented Framework for Conceptual Novelty Estimation in AI Research
di: Yan, Zhengxu, et al.
Pubblicazione: (2025)
di: Yan, Zhengxu, et al.
Pubblicazione: (2025)
Interacting Particle Langevin Algorithm for Maximum Marginal Likelihood Estimation
di: Akyildiz, Ö. Deniz, et al.
Pubblicazione: (2023)
di: Akyildiz, Ö. Deniz, et al.
Pubblicazione: (2023)
SEA: Sparse Linear Attention with Estimated Attention Mask
di: Lee, Heejun, et al.
Pubblicazione: (2023)
di: Lee, Heejun, et al.
Pubblicazione: (2023)
MIC: Maximizing Informational Capacity in Adaptive Representations via Isotropic Subspace Alignment
di: Hong, Dang Nguyen, et al.
Pubblicazione: (2026)
di: Hong, Dang Nguyen, et al.
Pubblicazione: (2026)
Large Language Models Encode Semantics and Alignment in Linearly Separable Representations
di: Saglam, Baturay, et al.
Pubblicazione: (2025)
di: Saglam, Baturay, et al.
Pubblicazione: (2025)
Cross-Tokenizer Likelihood Scoring Algorithms for Language Model Distillation
di: Phan, Buu, et al.
Pubblicazione: (2025)
di: Phan, Buu, et al.
Pubblicazione: (2025)
VITRO: Vocabulary Inversion for Time-series Representation Optimization
di: Bellos, Filippos, et al.
Pubblicazione: (2024)
di: Bellos, Filippos, et al.
Pubblicazione: (2024)
Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning
di: Xie, Wanyun, et al.
Pubblicazione: (2025)
di: Xie, Wanyun, et al.
Pubblicazione: (2025)
Scientific Hypothesis Generation and Validation: Methods, Datasets, and Future Directions
di: Kulkarni, Adithya, et al.
Pubblicazione: (2025)
di: Kulkarni, Adithya, et al.
Pubblicazione: (2025)
Exploiting LLMs for Automatic Hypothesis Assessment via a Logit-Based Calibrated Prior
di: Gong, Yue, et al.
Pubblicazione: (2025)
di: Gong, Yue, et al.
Pubblicazione: (2025)
Towards Efficient Active Learning in NLP via Pretrained Representations
di: Vysogorets, Artem, et al.
Pubblicazione: (2024)
di: Vysogorets, Artem, et al.
Pubblicazione: (2024)
ReCaLL: Membership Inference via Relative Conditional Log-Likelihoods
di: Xie, Roy, et al.
Pubblicazione: (2024)
di: Xie, Roy, et al.
Pubblicazione: (2024)
Revisiting the Superficial Alignment Hypothesis
di: Raghavendra, Mohit, et al.
Pubblicazione: (2024)
di: Raghavendra, Mohit, et al.
Pubblicazione: (2024)
Maximum Score Routing For Mixture-of-Experts
di: Dong, Bowen, et al.
Pubblicazione: (2025)
di: Dong, Bowen, et al.
Pubblicazione: (2025)
Failure Modes of Maximum Entropy RLHF
di: Çağatan, Ömer Veysel, et al.
Pubblicazione: (2025)
di: Çağatan, Ömer Veysel, et al.
Pubblicazione: (2025)
Disentangling the Roles of Representation and Selection in Data Pruning
di: Du, Yupei, et al.
Pubblicazione: (2025)
di: Du, Yupei, et al.
Pubblicazione: (2025)
Memories Retrieved from Many Paths: A Multi-Prefix Framework for Robust Detection of Training Data Leakage in Large Language Models
di: Dang, Trung Cuong, et al.
Pubblicazione: (2025)
di: Dang, Trung Cuong, et al.
Pubblicazione: (2025)
What Do Language Models Learn in Context? The Structured Task Hypothesis
di: Li, Jiaoda, et al.
Pubblicazione: (2024)
di: Li, Jiaoda, et al.
Pubblicazione: (2024)
Learning State-Tracking from Code Using Linear RNNs
di: Siems, Julien, et al.
Pubblicazione: (2026)
di: Siems, Julien, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Offline Preference Optimization via Maximum Marginal Likelihood Estimation
di: Najafi, Saeed, et al.
Pubblicazione: (2025) -
The Linear Representation Hypothesis and the Geometry of Large Language Models
di: Park, Kiho, et al.
Pubblicazione: (2023) -
Lizard: An Efficient Linearization Framework for Large Language Models
di: Van Nguyen, Chien, et al.
Pubblicazione: (2025) -
Linear Representation Transferability Hypothesis: Leveraging Small Models to Steer Large Models
di: Bello, Femi, et al.
Pubblicazione: (2025) -
Structured Pruning for Diverse Best-of-N Reasoning Optimization
di: Nguyen, Hieu Trung, et al.
Pubblicazione: (2025)