Gespeichert in:
| Hauptverfasser: | Qin, Yulu, Wang, Wentao, Lake, Brenden M. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2402.07899 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Rapid Word Learning Through Meta In-Context Learning
von: Wang, Wentao, et al.
Veröffentlicht: (2025)
von: Wang, Wentao, et al.
Veröffentlicht: (2025)
On the robustness of modeling grounded word learning through a child's egocentric input
von: Vong, Wai Keen, et al.
Veröffentlicht: (2025)
von: Vong, Wai Keen, et al.
Veröffentlicht: (2025)
An explainable transformer circuit for compositional generalization
von: Tang, Cheng, et al.
Veröffentlicht: (2025)
von: Tang, Cheng, et al.
Veröffentlicht: (2025)
Do different prompting methods yield a common task representation in language models?
von: Davidson, Guy, et al.
Veröffentlicht: (2025)
von: Davidson, Guy, et al.
Veröffentlicht: (2025)
Self-supervised learning of video representations from a child's perspective
von: Orhan, A. Emin, et al.
Veröffentlicht: (2024)
von: Orhan, A. Emin, et al.
Veröffentlicht: (2024)
Vocabulary shapes cross-lingual variation of word-order learnability in language models
von: Martins, Jonas Mayer, et al.
Veröffentlicht: (2026)
von: Martins, Jonas Mayer, et al.
Veröffentlicht: (2026)
Compositional learning of functions in humans and machines
von: Zhou, Yanli, et al.
Veröffentlicht: (2024)
von: Zhou, Yanli, et al.
Veröffentlicht: (2024)
From learnable objects to learnable random objects
von: Anderson, Aaron, et al.
Veröffentlicht: (2025)
von: Anderson, Aaron, et al.
Veröffentlicht: (2025)
Are they human? Detecting large language models by probing human memory constraints
von: Schug, Simon, et al.
Veröffentlicht: (2026)
von: Schug, Simon, et al.
Veröffentlicht: (2026)
From Distributional to Overton Pluralism: Investigating Large Language Model Alignment
von: Lake, Thom, et al.
Veröffentlicht: (2024)
von: Lake, Thom, et al.
Veröffentlicht: (2024)
Multiple output samples per input in a single-output Gaussian process
von: Wong, Jeremy H. M., et al.
Veröffentlicht: (2023)
von: Wong, Jeremy H. M., et al.
Veröffentlicht: (2023)
Injecting linguistic knowledge into BERT for Dialogue State Tracking
von: Feng, Xiaohan, et al.
Veröffentlicht: (2023)
von: Feng, Xiaohan, et al.
Veröffentlicht: (2023)
Overcoming classic challenges for artificial neural networks by providing incentives and practice
von: Irie, Kazuki, et al.
Veröffentlicht: (2024)
von: Irie, Kazuki, et al.
Veröffentlicht: (2024)
LoRA-SP: Streamlined Partial Parameter Adaptation for Resource-Efficient Fine-Tuning of Large Language Models
von: Wu, Yichao, et al.
Veröffentlicht: (2024)
von: Wu, Yichao, et al.
Veröffentlicht: (2024)
Probing self-attention in self-supervised speech models for cross-linguistic differences
von: Gopinath, Sai, et al.
Veröffentlicht: (2024)
von: Gopinath, Sai, et al.
Veröffentlicht: (2024)
CoLLEGe: Concept Embedding Generation for Large Language Models
von: Teehan, Ryan, et al.
Veröffentlicht: (2024)
von: Teehan, Ryan, et al.
Veröffentlicht: (2024)
Dimensional Collapse in Transformer Attention Outputs: A Challenge for Sparse Dictionary Learning
von: Wang, Junxuan, et al.
Veröffentlicht: (2025)
von: Wang, Junxuan, et al.
Veröffentlicht: (2025)
Talking with Oompa Loompas: A novel framework for evaluating linguistic acquisition of LLM agents
von: Swain, Sankalp Tattwadarshi, et al.
Veröffentlicht: (2025)
von: Swain, Sankalp Tattwadarshi, et al.
Veröffentlicht: (2025)
Deep Learning Approaches for Improving Question Answering Systems in Hepatocellular Carcinoma Research
von: Huo, Shuning, et al.
Veröffentlicht: (2024)
von: Huo, Shuning, et al.
Veröffentlicht: (2024)
Direct Multi-Turn Preference Optimization for Language Agents
von: Shi, Wentao, et al.
Veröffentlicht: (2024)
von: Shi, Wentao, et al.
Veröffentlicht: (2024)
Classification errors distort findings in automated speech processing: examples and solutions from child-development research
von: Gautheron, Lucas, et al.
Veröffentlicht: (2025)
von: Gautheron, Lucas, et al.
Veröffentlicht: (2025)
SUS backprop: linear backpropagation algorithm for long inputs in transformers
von: Pankov, Sergey, et al.
Veröffentlicht: (2025)
von: Pankov, Sergey, et al.
Veröffentlicht: (2025)
Building Korean linguistic resource for NLU data generation of banking app CS dialog system
von: Yoon, Jeongwoo, et al.
Veröffentlicht: (2026)
von: Yoon, Jeongwoo, et al.
Veröffentlicht: (2026)
Automatically Identifying Local and Global Circuits with Linear Computation Graphs
von: Ge, Xuyang, et al.
Veröffentlicht: (2024)
von: Ge, Xuyang, et al.
Veröffentlicht: (2024)
CAMPHOR: Collaborative Agents for Multi-input Planning and High-Order Reasoning On Device
von: Fu, Yicheng, et al.
Veröffentlicht: (2024)
von: Fu, Yicheng, et al.
Veröffentlicht: (2024)
Overcoming linguistic barriers in code assistants: creating a QLoRA adapter to improve support for Russian-language code writing instructions
von: Pronin, C. B., et al.
Veröffentlicht: (2024)
von: Pronin, C. B., et al.
Veröffentlicht: (2024)
True Knowledge Comes from Practice: Aligning LLMs with Embodied Environments via Reinforcement Learning
von: Tan, Weihao, et al.
Veröffentlicht: (2024)
von: Tan, Weihao, et al.
Veröffentlicht: (2024)
Self-Improvement Towards Pareto Optimality: Mitigating Preference Conflicts in Multi-Objective Alignment
von: Li, Moxin, et al.
Veröffentlicht: (2025)
von: Li, Moxin, et al.
Veröffentlicht: (2025)
From Parameters to Data: A Task-Parameter-Guided Fine-Tuning Pipeline for Efficient LLM Alignment
von: Chen, Hao, et al.
Veröffentlicht: (2026)
von: Chen, Hao, et al.
Veröffentlicht: (2026)
Inference time LLM alignment in single and multidomain preference spectrum
von: Shahriar, Sadat, et al.
Veröffentlicht: (2024)
von: Shahriar, Sadat, et al.
Veröffentlicht: (2024)
Towards Understanding the Nature of Attention with Low-Rank Sparse Decomposition
von: He, Zhengfu, et al.
Veröffentlicht: (2025)
von: He, Zhengfu, et al.
Veröffentlicht: (2025)
From Chat Logs to Collective Insights: Aggregative Question Answering
von: Zhang, Wentao, et al.
Veröffentlicht: (2025)
von: Zhang, Wentao, et al.
Veröffentlicht: (2025)
Learning Multiplex Representations on Text-Attributed Graphs with One Language Model Encoder
von: Jin, Bowen, et al.
Veröffentlicht: (2023)
von: Jin, Bowen, et al.
Veröffentlicht: (2023)
pQuant: Towards Effective Low-Bit Language Models via Decoupled Linear Quantization-Aware Training
von: Zhang, Wenzheng, et al.
Veröffentlicht: (2026)
von: Zhang, Wenzheng, et al.
Veröffentlicht: (2026)
Bridging the Gap between 2D and 3D Visual Question Answering: A Fusion Approach for 3D VQA
von: Mo, Wentao, et al.
Veröffentlicht: (2024)
von: Mo, Wentao, et al.
Veröffentlicht: (2024)
Density estimation with LLMs: a geometric investigation of in-context learning trajectories
von: Liu, Toni J. B., et al.
Veröffentlicht: (2024)
von: Liu, Toni J. B., et al.
Veröffentlicht: (2024)
Supervised Fine-Tuning Needs to Unlock the Potential of Token Priority
von: Shen, Zhanming, et al.
Veröffentlicht: (2026)
von: Shen, Zhanming, et al.
Veröffentlicht: (2026)
Interactive Training: Feedback-Driven Neural Network Optimization
von: Zhang, Wentao, et al.
Veröffentlicht: (2025)
von: Zhang, Wentao, et al.
Veröffentlicht: (2025)
BEATS: Optimizing LLM Mathematical Capabilities with BackVerify and Adaptive Disambiguate based Efficient Tree Search
von: Sun, Linzhuang, et al.
Veröffentlicht: (2024)
von: Sun, Linzhuang, et al.
Veröffentlicht: (2024)
Machine-assisted writing evaluation: Exploring pre-trained language models in analyzing argumentative moves
von: Qin, Wenjuan, et al.
Veröffentlicht: (2025)
von: Qin, Wenjuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Rapid Word Learning Through Meta In-Context Learning
von: Wang, Wentao, et al.
Veröffentlicht: (2025) -
On the robustness of modeling grounded word learning through a child's egocentric input
von: Vong, Wai Keen, et al.
Veröffentlicht: (2025) -
An explainable transformer circuit for compositional generalization
von: Tang, Cheng, et al.
Veröffentlicht: (2025) -
Do different prompting methods yield a common task representation in language models?
von: Davidson, Guy, et al.
Veröffentlicht: (2025) -
Self-supervised learning of video representations from a child's perspective
von: Orhan, A. Emin, et al.
Veröffentlicht: (2024)