Gespeichert in:
| Hauptverfasser: | Tohar, Vered, Hayat, Tsahi, Leshem, Amir |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2602.16578 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Stance Reasoner: Zero-Shot Stance Detection on Social Media with Explicit Reasoning
von: Taranukhin, Maksym, et al.
Veröffentlicht: (2024)
von: Taranukhin, Maksym, et al.
Veröffentlicht: (2024)
A Hitchhiker's Guide to Scaling Law Estimation
von: Choshen, Leshem, et al.
Veröffentlicht: (2024)
von: Choshen, Leshem, et al.
Veröffentlicht: (2024)
Fuse to Forget: Bias Reduction and Selective Memorization through Model Fusion
von: Zaman, Kerem, et al.
Veröffentlicht: (2023)
von: Zaman, Kerem, et al.
Veröffentlicht: (2023)
Beneath the Surface of Consistency: Exploring Cross-lingual Knowledge Representation Sharing in LLMs
von: Ifergan, Maxim, et al.
Veröffentlicht: (2024)
von: Ifergan, Maxim, et al.
Veröffentlicht: (2024)
Empowering Air Travelers: A Chatbot for Canadian Air Passenger Rights
von: Taranukhin, Maksym, et al.
Veröffentlicht: (2024)
von: Taranukhin, Maksym, et al.
Veröffentlicht: (2024)
Do LLMs Benefit From Their Own Words?
von: Huang, Jenny Y., et al.
Veröffentlicht: (2026)
von: Huang, Jenny Y., et al.
Veröffentlicht: (2026)
ComPEFT: Compression for Communicating Parameter Efficient Updates via Sparsification and Quantization
von: Yadav, Prateek, et al.
Veröffentlicht: (2023)
von: Yadav, Prateek, et al.
Veröffentlicht: (2023)
ErrorMap and ErrorAtlas: Charting the Failure Landscape of Large Language Models
von: Ashury-Tahan, Shir, et al.
Veröffentlicht: (2026)
von: Ashury-Tahan, Shir, et al.
Veröffentlicht: (2026)
Is It Bad to Work All the Time? Cross-Cultural Evaluation of Social Norm Biases in GPT-4
von: Liu, Zhuozhuo Joy, et al.
Veröffentlicht: (2025)
von: Liu, Zhuozhuo Joy, et al.
Veröffentlicht: (2025)
CRISP: Complex Reasoning with Interpretable Step-based Plans
von: Vetzler, Matan, et al.
Veröffentlicht: (2025)
von: Vetzler, Matan, et al.
Veröffentlicht: (2025)
Small But Funny: A Feedback-Driven Approach to Humor Distillation
von: Ravi, Sahithya, et al.
Veröffentlicht: (2024)
von: Ravi, Sahithya, et al.
Veröffentlicht: (2024)
Bridging Information Gaps with Comprehensive Answers: Improving the Diversity and Informativeness of Follow-Up Questions
von: Liu, Zhe, et al.
Veröffentlicht: (2025)
von: Liu, Zhe, et al.
Veröffentlicht: (2025)
From Local Concepts to Universals: Evaluating the Multicultural Understanding of Vision-Language Models
von: Bhatia, Mehar, et al.
Veröffentlicht: (2024)
von: Bhatia, Mehar, et al.
Veröffentlicht: (2024)
Robustness as an Emergent Property of Task Performance
von: Ashury-Tahan, Shir, et al.
Veröffentlicht: (2026)
von: Ashury-Tahan, Shir, et al.
Veröffentlicht: (2026)
Creating an AI Observer: Generative Semantic Workspaces
von: Holur, Pavan, et al.
Veröffentlicht: (2024)
von: Holur, Pavan, et al.
Veröffentlicht: (2024)
tinyBenchmarks: evaluating LLMs with fewer examples
von: Polo, Felipe Maia, et al.
Veröffentlicht: (2024)
von: Polo, Felipe Maia, et al.
Veröffentlicht: (2024)
Generative Ontology: When Structured Knowledge Learns to Create
von: Cheung, Benny
Veröffentlicht: (2026)
von: Cheung, Benny
Veröffentlicht: (2026)
Beyond Binary Rewards: Training LMs to Reason About Their Uncertainty
von: Damani, Mehul, et al.
Veröffentlicht: (2025)
von: Damani, Mehul, et al.
Veröffentlicht: (2025)
Resolving Interference (RI): Disentangling Models for Improved Model Merging
von: Ramesh, Pratik, et al.
Veröffentlicht: (2026)
von: Ramesh, Pratik, et al.
Veröffentlicht: (2026)
Games That Teach, Chats That Convince: Comparing Interactive and Static Formats for Persuasive Learning
von: Alavi, Seyed Hossein, et al.
Veröffentlicht: (2026)
von: Alavi, Seyed Hossein, et al.
Veröffentlicht: (2026)
Bridging the Data Gap: Creating a Hindi Text Summarization Dataset from the English XSUM
von: Katwe, Praveenkumar, et al.
Veröffentlicht: (2026)
von: Katwe, Praveenkumar, et al.
Veröffentlicht: (2026)
SiNFluD: Creating and Evaluating Figurative Language Dataset for Sindhi
von: Ali, Wazir, et al.
Veröffentlicht: (2026)
von: Ali, Wazir, et al.
Veröffentlicht: (2026)
Past Meets Present: Creating Historical Analogy with Large Language Models
von: Li, Nianqi, et al.
Veröffentlicht: (2024)
von: Li, Nianqi, et al.
Veröffentlicht: (2024)
Why Does New Knowledge Create Messy Ripple Effects in LLMs?
von: Qin, Jiaxin, et al.
Veröffentlicht: (2024)
von: Qin, Jiaxin, et al.
Veröffentlicht: (2024)
Genie: Achieving Human Parity in Content-Grounded Datasets Generation
von: Yehudai, Asaf, et al.
Veröffentlicht: (2024)
von: Yehudai, Asaf, et al.
Veröffentlicht: (2024)
Development and Evaluation of a Retrieval-Augmented Generation Tool for Creating SAPPhIRE Models of Artificial Systems
von: Majumder, Anubhab, et al.
Veröffentlicht: (2024)
von: Majumder, Anubhab, et al.
Veröffentlicht: (2024)
Can Large Language Models Create New Knowledge for Spatial Reasoning Tasks?
von: Greatrix, Thomas, et al.
Veröffentlicht: (2024)
von: Greatrix, Thomas, et al.
Veröffentlicht: (2024)
MCPDial: A Minecraft Persona-driven Dialogue Dataset
von: Alavi, Seyed Hossein, et al.
Veröffentlicht: (2024)
von: Alavi, Seyed Hossein, et al.
Veröffentlicht: (2024)
TextArena
von: Guertler, Leon, et al.
Veröffentlicht: (2025)
von: Guertler, Leon, et al.
Veröffentlicht: (2025)
A Critical Review of the Need for Knowledge-Centric Evaluation of Quranic Recitation
von: Al-Kharusi, Mohammed Hilal, et al.
Veröffentlicht: (2025)
von: Al-Kharusi, Mohammed Hilal, et al.
Veröffentlicht: (2025)
Bridging Writing Manner Gap in Visual Instruction Tuning by Creating LLM-aligned Instructions
von: Jing, Dong, et al.
Veröffentlicht: (2025)
von: Jing, Dong, et al.
Veröffentlicht: (2025)
From Polyester Girlfriends to Blind Mice: Creating the First Pragmatics Understanding Benchmarks for Slovene
von: Brglez, Mojca, et al.
Veröffentlicht: (2025)
von: Brglez, Mojca, et al.
Veröffentlicht: (2025)
Redefining "Hallucination" in LLMs: Towards a psychology-informed framework for mitigating misinformation
von: Berberette, Elijah, et al.
Veröffentlicht: (2024)
von: Berberette, Elijah, et al.
Veröffentlicht: (2024)
A Survey on Model MoErging: Recycling and Routing Among Specialized Experts for Collaborative Learning
von: Yadav, Prateek, et al.
Veröffentlicht: (2024)
von: Yadav, Prateek, et al.
Veröffentlicht: (2024)
ParaNames 1.0: Creating an Entity Name Corpus for 400+ Languages using Wikidata
von: Sälevä, Jonne, et al.
Veröffentlicht: (2024)
von: Sälevä, Jonne, et al.
Veröffentlicht: (2024)
Pastiche Novel Generation Creating: Fan Fiction You Love in Your Favorite Author's Style
von: Han, Xueran, et al.
Veröffentlicht: (2025)
von: Han, Xueran, et al.
Veröffentlicht: (2025)
How Teachers Can Use Large Language Models and Bloom's Taxonomy to Create Educational Quizzes
von: Elkins, Sabina, et al.
Veröffentlicht: (2024)
von: Elkins, Sabina, et al.
Veröffentlicht: (2024)
Token Statistics Reveal Conversational Drift in Multi-turn LLM Interaction
von: Hafez, Wael, et al.
Veröffentlicht: (2026)
von: Hafez, Wael, et al.
Veröffentlicht: (2026)
Exploring the Potential of Machine Translation for Generating Named Entity Datasets: A Case Study between Persian and English
von: Sartipi, Amir, et al.
Veröffentlicht: (2023)
von: Sartipi, Amir, et al.
Veröffentlicht: (2023)
Creating General User Models from Computer Use
von: Shaikh, Omar, et al.
Veröffentlicht: (2025)
von: Shaikh, Omar, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Stance Reasoner: Zero-Shot Stance Detection on Social Media with Explicit Reasoning
von: Taranukhin, Maksym, et al.
Veröffentlicht: (2024) -
A Hitchhiker's Guide to Scaling Law Estimation
von: Choshen, Leshem, et al.
Veröffentlicht: (2024) -
Fuse to Forget: Bias Reduction and Selective Memorization through Model Fusion
von: Zaman, Kerem, et al.
Veröffentlicht: (2023) -
Beneath the Surface of Consistency: Exploring Cross-lingual Knowledge Representation Sharing in LLMs
von: Ifergan, Maxim, et al.
Veröffentlicht: (2024) -
Empowering Air Travelers: A Chatbot for Canadian Air Passenger Rights
von: Taranukhin, Maksym, et al.
Veröffentlicht: (2024)