Gespeichert in:
| Hauptverfasser: | Yang, Ivy Yuqian, Zhang, David Yu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2511.14630 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multi-step retrieval and reasoning improves radiology question answering with large language models
von: Wind, Sebastian, et al.
Veröffentlicht: (2025)
von: Wind, Sebastian, et al.
Veröffentlicht: (2025)
CaLMQA: Exploring culturally specific long-form question answering across 23 languages
von: Arora, Shane, et al.
Veröffentlicht: (2024)
von: Arora, Shane, et al.
Veröffentlicht: (2024)
Neural operators struggle to learn complex PDEs in pedestrian mobility: Hughes model case study
von: Chauhan, Prajwal, et al.
Veröffentlicht: (2025)
von: Chauhan, Prajwal, et al.
Veröffentlicht: (2025)
Policy alone is probably not the solution: A large-scale experiment on how developers struggle to design meaningful end-user explanations
von: Nahar, Nadia, et al.
Veröffentlicht: (2025)
von: Nahar, Nadia, et al.
Veröffentlicht: (2025)
How predictable is language model benchmark performance?
von: Owen, David
Veröffentlicht: (2024)
von: Owen, David
Veröffentlicht: (2024)
Large language models as uncertainty-calibrated optimizers for experimental discovery
von: Ranković, Bojana, et al.
Veröffentlicht: (2025)
von: Ranković, Bojana, et al.
Veröffentlicht: (2025)
Agentic retrieval-augmented reasoning reshapes collective reliability under model variability in radiology question answering
von: Farajiamiri, Mina, et al.
Veröffentlicht: (2026)
von: Farajiamiri, Mina, et al.
Veröffentlicht: (2026)
Revisiting On-Policy Distillation: Empirical Failure Modes and Simple Fixes
von: Fu, Yuqian, et al.
Veröffentlicht: (2026)
von: Fu, Yuqian, et al.
Veröffentlicht: (2026)
Efficiency optimization of large-scale language models based on deep learning in natural language processing tasks
von: Mei, Taiyuan, et al.
Veröffentlicht: (2024)
von: Mei, Taiyuan, et al.
Veröffentlicht: (2024)
Agribot: agriculture-specific question answer system
von: Jain, Naman, et al.
Veröffentlicht: (2025)
von: Jain, Naman, et al.
Veröffentlicht: (2025)
MicroMix: Efficient Mixed-Precision Quantization with Microscaling Formats for Large Language Models
von: Liu, Wenyuan, et al.
Veröffentlicht: (2025)
von: Liu, Wenyuan, et al.
Veröffentlicht: (2025)
Select to Perfect: Imitating desired behavior from large multi-agent data
von: Franzmeyer, Tim, et al.
Veröffentlicht: (2024)
von: Franzmeyer, Tim, et al.
Veröffentlicht: (2024)
Latent label distribution grid representation for modeling uncertainty
von: Sun, ShuNing, et al.
Veröffentlicht: (2025)
von: Sun, ShuNing, et al.
Veröffentlicht: (2025)
AnyAttack: Towards Large-scale Self-supervised Adversarial Attacks on Vision-language Models
von: Zhang, Jiaming, et al.
Veröffentlicht: (2024)
von: Zhang, Jiaming, et al.
Veröffentlicht: (2024)
Large language models struggle with ethnographic text annotation
von: Goodall, Leonardo S., et al.
Veröffentlicht: (2026)
von: Goodall, Leonardo S., et al.
Veröffentlicht: (2026)
Deep Feature Embedding for Tabular Data
von: Wu, Yuqian, et al.
Veröffentlicht: (2024)
von: Wu, Yuqian, et al.
Veröffentlicht: (2024)
But what is your honest answer? Aiding LLM-judges with honest alternatives using steering vectors
von: Eshuijs, Leon, et al.
Veröffentlicht: (2025)
von: Eshuijs, Leon, et al.
Veröffentlicht: (2025)
Implicit meta-learning may lead language models to trust more reliable sources
von: Krasheninnikov, Dmitrii, et al.
Veröffentlicht: (2023)
von: Krasheninnikov, Dmitrii, et al.
Veröffentlicht: (2023)
Latent Space Data Fusion Outperforms Early Fusion in Multimodal Mental Health Digital Phenotyping Data
von: Barkat, Youcef, et al.
Veröffentlicht: (2025)
von: Barkat, Youcef, et al.
Veröffentlicht: (2025)
Context information can be more important than reasoning for time series forecasting with a large language model
von: Yang, Janghoon
Veröffentlicht: (2025)
von: Yang, Janghoon
Veröffentlicht: (2025)
Large language models can learn and generalize steganographic chain-of-thought under process supervision
von: Skaf, Joey, et al.
Veröffentlicht: (2025)
von: Skaf, Joey, et al.
Veröffentlicht: (2025)
Agentic reinforcement learning empowers next-generation chemical language models for molecular design and synthesis
von: Li, Hao, et al.
Veröffentlicht: (2026)
von: Li, Hao, et al.
Veröffentlicht: (2026)
The language of time: a language model perspective on time-series foundation models
von: Xie, Yi, et al.
Veröffentlicht: (2025)
von: Xie, Yi, et al.
Veröffentlicht: (2025)
Safety challenges of AI in medicine in the era of large language models
von: Wang, Xiaoye, et al.
Veröffentlicht: (2024)
von: Wang, Xiaoye, et al.
Veröffentlicht: (2024)
Quantifying construct validity in large language model evaluations
von: Kearns, Ryan Othniel
Veröffentlicht: (2026)
von: Kearns, Ryan Othniel
Veröffentlicht: (2026)
Applying sparse autoencoders to unlearn knowledge in language models
von: Farrell, Eoin, et al.
Veröffentlicht: (2024)
von: Farrell, Eoin, et al.
Veröffentlicht: (2024)
Text-guided multi-property molecular optimization with a diffusion language model
von: Xiong, Yida, et al.
Veröffentlicht: (2024)
von: Xiong, Yida, et al.
Veröffentlicht: (2024)
Alignment faking in large language models
von: Greenblatt, Ryan, et al.
Veröffentlicht: (2024)
von: Greenblatt, Ryan, et al.
Veröffentlicht: (2024)
Large Language Model Reasoning Failures
von: Song, Peiyang, et al.
Veröffentlicht: (2026)
von: Song, Peiyang, et al.
Veröffentlicht: (2026)
Can large language models explore in-context?
von: Krishnamurthy, Akshay, et al.
Veröffentlicht: (2024)
von: Krishnamurthy, Akshay, et al.
Veröffentlicht: (2024)
Representation in large language models
von: Yetman, Cameron
Veröffentlicht: (2025)
von: Yetman, Cameron
Veröffentlicht: (2025)
Long-form factuality in large language models
von: Wei, Jerry, et al.
Veröffentlicht: (2024)
von: Wei, Jerry, et al.
Veröffentlicht: (2024)
MergeIT: From Selection to Merging for Efficient Instruction Tuning
von: Cai, Hongyi, et al.
Veröffentlicht: (2025)
von: Cai, Hongyi, et al.
Veröffentlicht: (2025)
Improving training time and GPU utilization in geo-distributed language model training
von: Palak, et al.
Veröffentlicht: (2024)
von: Palak, et al.
Veröffentlicht: (2024)
Multimodal large language model for wheat breeding: a new exploration of smart breeding
von: Yang, Guofeng, et al.
Veröffentlicht: (2024)
von: Yang, Guofeng, et al.
Veröffentlicht: (2024)
Large language models can accurately predict searcher preferences
von: Thomas, Paul, et al.
Veröffentlicht: (2023)
von: Thomas, Paul, et al.
Veröffentlicht: (2023)
Replacing thinking with tool usage enables reasoning in small language models
von: Rainone, Corrado, et al.
Veröffentlicht: (2025)
von: Rainone, Corrado, et al.
Veröffentlicht: (2025)
Fresh in memory: Training-order recency is linearly encoded in language model activations
von: Krasheninnikov, Dmitrii, et al.
Veröffentlicht: (2025)
von: Krasheninnikov, Dmitrii, et al.
Veröffentlicht: (2025)
Conditional Equivalence of DPO and RLHF: Implicit Assumption, Failure Modes, and Provable Alignment
von: Yang, Zhiqin, et al.
Veröffentlicht: (2026)
von: Yang, Zhiqin, et al.
Veröffentlicht: (2026)
Auditing language models for hidden objectives
von: Marks, Samuel, et al.
Veröffentlicht: (2025)
von: Marks, Samuel, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Multi-step retrieval and reasoning improves radiology question answering with large language models
von: Wind, Sebastian, et al.
Veröffentlicht: (2025) -
CaLMQA: Exploring culturally specific long-form question answering across 23 languages
von: Arora, Shane, et al.
Veröffentlicht: (2024) -
Neural operators struggle to learn complex PDEs in pedestrian mobility: Hughes model case study
von: Chauhan, Prajwal, et al.
Veröffentlicht: (2025) -
Policy alone is probably not the solution: A large-scale experiment on how developers struggle to design meaningful end-user explanations
von: Nahar, Nadia, et al.
Veröffentlicht: (2025) -
How predictable is language model benchmark performance?
von: Owen, David
Veröffentlicht: (2024)