Position: Avoid Overstretching LLMs for every Enterprise Task
Fuente:
arXiv
Salvato in:
| Autori principali: | Singh, Kuldeep, Bastos, Anson, Mulang', Isaiah Onando |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SALT-KG: A Benchmark for Semantics-Aware Learning on Enterprise Tables
di: Mulang, Isaiah Onando, et al.
Pubblicazione: (2026)
di: Mulang, Isaiah Onando, et al.
Pubblicazione: (2026)
Position is Power: System Prompts as a Mechanism of Bias in Large Language Models (LLMs)
di: Neumann, Anna, et al.
Pubblicazione: (2025)
di: Neumann, Anna, et al.
Pubblicazione: (2025)
Hybrid OCR-LLM Framework for Enterprise-Scale Document Information Extraction Under Copy-heavy Task
di: Wang, Zilong, et al.
Pubblicazione: (2025)
di: Wang, Zilong, et al.
Pubblicazione: (2025)
Difficult Task Yes but Simple Task No: Unveiling the Laziness in Multimodal LLMs
di: Zhao, Sihang, et al.
Pubblicazione: (2024)
di: Zhao, Sihang, et al.
Pubblicazione: (2024)
PositionID: LLMs can Control Lengths, Copy and Paste with Explicit Positional Awareness
di: Wang, Zekun, et al.
Pubblicazione: (2024)
di: Wang, Zekun, et al.
Pubblicazione: (2024)
Beyond Fine-Tuning: Effective Strategies for Mitigating Hallucinations in Large Language Models for Data Analytics
di: Rumiantsau, Mikhail, et al.
Pubblicazione: (2024)
di: Rumiantsau, Mikhail, et al.
Pubblicazione: (2024)
Enterprise Deep Research: Steerable Multi-Agent Deep Research for Enterprise Analytics
di: Prabhakar, Akshara, et al.
Pubblicazione: (2025)
di: Prabhakar, Akshara, et al.
Pubblicazione: (2025)
Position: LLMs Can be Good Tutors in English Education
di: Ye, Jingheng, et al.
Pubblicazione: (2025)
di: Ye, Jingheng, et al.
Pubblicazione: (2025)
Extending the Context of Pretrained LLMs by Dropping Their Positional Embeddings
di: Gelberg, Yoav, et al.
Pubblicazione: (2025)
di: Gelberg, Yoav, et al.
Pubblicazione: (2025)
ArgBench: Benchmarking LLMs on Computational Argumentation Tasks
di: Ajjour, Yamen, et al.
Pubblicazione: (2026)
di: Ajjour, Yamen, et al.
Pubblicazione: (2026)
Autonomous Evaluation of LLMs for Truth Maintenance and Reasoning Tasks
di: Karia, Rushang, et al.
Pubblicazione: (2024)
di: Karia, Rushang, et al.
Pubblicazione: (2024)
Are Long-LLMs A Necessity For Long-Context Tasks?
di: Qian, Hongjin, et al.
Pubblicazione: (2024)
di: Qian, Hongjin, et al.
Pubblicazione: (2024)
Beyond Text-to-SQL: An Agentic LLM System for Governed Enterprise Analytics APIs
di: Singh, Gundeep, et al.
Pubblicazione: (2026)
di: Singh, Gundeep, et al.
Pubblicazione: (2026)
Can LLMs Capture Human Preferences?
di: Goli, Ali, et al.
Pubblicazione: (2023)
di: Goli, Ali, et al.
Pubblicazione: (2023)
Concept Attractors in LLMs and their Applications
di: Chytas, Sotirios Panagiotis, et al.
Pubblicazione: (2025)
di: Chytas, Sotirios Panagiotis, et al.
Pubblicazione: (2025)
MAFA: A Multi-Agent Framework for Enterprise-Scale Annotation with Configurable Task Adaptation
di: Hegazy, Mahmood, et al.
Pubblicazione: (2025)
di: Hegazy, Mahmood, et al.
Pubblicazione: (2025)
Identifying Good and Bad Neurons for Task-Level Controllable LLMs
di: Li, Wenjie, et al.
Pubblicazione: (2026)
di: Li, Wenjie, et al.
Pubblicazione: (2026)
Magis-Bench: Evaluating LLMs on Magistrate-Level Legal Tasks
di: Pires, Ramon, et al.
Pubblicazione: (2026)
di: Pires, Ramon, et al.
Pubblicazione: (2026)
Evaluating the Evaluator: Measuring LLMs' Adherence to Task Evaluation Instructions
di: Murugadoss, Bhuvanashree, et al.
Pubblicazione: (2024)
di: Murugadoss, Bhuvanashree, et al.
Pubblicazione: (2024)
Improving Task Diversity in Label Efficient Supervised Finetuning of LLMs
di: Arabelly, Abhinav, et al.
Pubblicazione: (2025)
di: Arabelly, Abhinav, et al.
Pubblicazione: (2025)
Can LLMs Generate High-Quality Task-Specific Conversations?
di: Li, Shengqi, et al.
Pubblicazione: (2025)
di: Li, Shengqi, et al.
Pubblicazione: (2025)
Incomplete Tasks Induce Shutdown Resistance in Some Frontier LLMs
di: Schlatter, Jeremy, et al.
Pubblicazione: (2025)
di: Schlatter, Jeremy, et al.
Pubblicazione: (2025)
PsychiatryBench: A Multi-Task Benchmark for LLMs in Psychiatry
di: Fouda, Aya E., et al.
Pubblicazione: (2025)
di: Fouda, Aya E., et al.
Pubblicazione: (2025)
SELT: Self-Evaluation Tree Search for LLMs with Task Decomposition
di: Wu, Mengsong, et al.
Pubblicazione: (2025)
di: Wu, Mengsong, et al.
Pubblicazione: (2025)
Moral Lenses, Political Coordinates: Towards Ideological Positioning of Morally Conditioned LLMs
di: Yuan, Chenchen, et al.
Pubblicazione: (2026)
di: Yuan, Chenchen, et al.
Pubblicazione: (2026)
Multi-Task Learning with LLMs for Implicit Sentiment Analysis: Data-level and Task-level Automatic Weight Learning
di: Lai, Wenna, et al.
Pubblicazione: (2024)
di: Lai, Wenna, et al.
Pubblicazione: (2024)
Benchmarking Complex Multimodal Document Processing Pipelines: A Unified Evaluation Framework for Enterprise AI
di: Singh, Saurabh K., et al.
Pubblicazione: (2026)
di: Singh, Saurabh K., et al.
Pubblicazione: (2026)
Finding the Translation Switch: Discovering and Exploiting the Task-Initiation Features in LLMs
di: Wu, Xinwei, et al.
Pubblicazione: (2026)
di: Wu, Xinwei, et al.
Pubblicazione: (2026)
Teaching and Evaluating LLMs to Reason About Polymer Design Related Tasks
di: Mohanty, Dikshya, et al.
Pubblicazione: (2026)
di: Mohanty, Dikshya, et al.
Pubblicazione: (2026)
CulturALL: Benchmarking Multilingual and Multicultural Competence of LLMs on Grounded Tasks
di: Lin, Peiqin, et al.
Pubblicazione: (2026)
di: Lin, Peiqin, et al.
Pubblicazione: (2026)
Structured Thinking Matters: Improving LLMs Generalization in Causal Inference Tasks
di: Sun, Wentao, et al.
Pubblicazione: (2025)
di: Sun, Wentao, et al.
Pubblicazione: (2025)
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning
di: Fei, Zhaoye, et al.
Pubblicazione: (2025)
di: Fei, Zhaoye, et al.
Pubblicazione: (2025)
Pseudocode-Injection Magic: Enabling LLMs to Tackle Graph Computational Tasks
di: Gong, Chang, et al.
Pubblicazione: (2025)
di: Gong, Chang, et al.
Pubblicazione: (2025)
Active Task Disambiguation with LLMs
di: Kobalczyk, Katarzyna, et al.
Pubblicazione: (2025)
di: Kobalczyk, Katarzyna, et al.
Pubblicazione: (2025)
Avoiding Knowledge Edit Skipping in Multi-hop Question Answering with Guided Decomposition
di: Liu, Yi, et al.
Pubblicazione: (2025)
di: Liu, Yi, et al.
Pubblicazione: (2025)
Position: LLMs Must Use Functor-Based and RAG-Driven Bias Mitigation for Fairness
di: Ranjan, Ravi, et al.
Pubblicazione: (2026)
di: Ranjan, Ravi, et al.
Pubblicazione: (2026)
Disambiguation-Centric Finetuning Makes Enterprise Tool-Calling LLMs More Realistic and Less Risky
di: Hathidara, Ashutosh, et al.
Pubblicazione: (2025)
di: Hathidara, Ashutosh, et al.
Pubblicazione: (2025)
Benchmarking Deep Search over Heterogeneous Enterprise Data
di: Choubey, Prafulla Kumar, et al.
Pubblicazione: (2025)
di: Choubey, Prafulla Kumar, et al.
Pubblicazione: (2025)
STaD: Scaffolded Task Design for Identifying Compositional Skill Gaps in LLMs
di: An, Sungeun, et al.
Pubblicazione: (2026)
di: An, Sungeun, et al.
Pubblicazione: (2026)
ATACompressor: Adaptive Task-Aware Compression for Efficient Long-Context Processing in LLMs
di: Li, Xuancheng, et al.
Pubblicazione: (2026)
di: Li, Xuancheng, et al.
Pubblicazione: (2026)
Documenti analoghi
-
SALT-KG: A Benchmark for Semantics-Aware Learning on Enterprise Tables
di: Mulang, Isaiah Onando, et al.
Pubblicazione: (2026) -
Position is Power: System Prompts as a Mechanism of Bias in Large Language Models (LLMs)
di: Neumann, Anna, et al.
Pubblicazione: (2025) -
Hybrid OCR-LLM Framework for Enterprise-Scale Document Information Extraction Under Copy-heavy Task
di: Wang, Zilong, et al.
Pubblicazione: (2025) -
Difficult Task Yes but Simple Task No: Unveiling the Laziness in Multimodal LLMs
di: Zhao, Sihang, et al.
Pubblicazione: (2024) -
PositionID: LLMs can Control Lengths, Copy and Paste with Explicit Positional Awareness
di: Wang, Zekun, et al.
Pubblicazione: (2024)