Gespeichert in:
| Hauptverfasser: | Shayegh, Behzad, Ahmed, Mohamed Osama, Tung, Fred, Feng, Leo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2605.06919 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Tree-Averaging Algorithms for Ensemble-Based Unsupervised Discontinuous Constituency Parsing
von: Shayegh, Behzad, et al.
Veröffentlicht: (2024)
von: Shayegh, Behzad, et al.
Veröffentlicht: (2024)
Ensemble Distillation for Unsupervised Constituency Parsing
von: Shayegh, Behzad, et al.
Veröffentlicht: (2023)
von: Shayegh, Behzad, et al.
Veröffentlicht: (2023)
Error Diversity Matters: An Error-Resistant Ensemble Method for Unsupervised Dependency Parsing
von: Shayegh, Behzad, et al.
Veröffentlicht: (2024)
von: Shayegh, Behzad, et al.
Veröffentlicht: (2024)
EBBS: An Ensemble with Bi-Level Beam Search for Zero-Shot Machine Translation
von: Wen, Yuqiao, et al.
Veröffentlicht: (2024)
von: Wen, Yuqiao, et al.
Veröffentlicht: (2024)
Feeding Two Birds or Favoring One? Adequacy-Fluency Tradeoffs in Evaluation and Meta-Evaluation of Machine Translation
von: Shayegh, Behzad, et al.
Veröffentlicht: (2025)
von: Shayegh, Behzad, et al.
Veröffentlicht: (2025)
Memory Efficient Neural Processes via Constant Memory Attention Block
von: Feng, Leo, et al.
Veröffentlicht: (2023)
von: Feng, Leo, et al.
Veröffentlicht: (2023)
GPT-HateCheck: Can LLMs Write Better Functional Tests for Hate Speech Detection?
von: Jin, Yiping, et al.
Veröffentlicht: (2024)
von: Jin, Yiping, et al.
Veröffentlicht: (2024)
A Dual-Path Architecture for Scaling Compute and Capacity in LLMs
von: Frey, Markus, et al.
Veröffentlicht: (2026)
von: Frey, Markus, et al.
Veröffentlicht: (2026)
Signal in the Noise: Decoding the Reality of Airline Service Quality with Large Language Models
von: Dawoud, Ahmed, et al.
Veröffentlicht: (2026)
von: Dawoud, Ahmed, et al.
Veröffentlicht: (2026)
Fingerprinting LLMs via Prompt Injection
von: Hu, Yuepeng, et al.
Veröffentlicht: (2025)
von: Hu, Yuepeng, et al.
Veröffentlicht: (2025)
SLIM-LLMs: Modeling of Style-Sensory Language RelationshipsThrough Low-Dimensional Representations
von: Khalid, Osama, et al.
Veröffentlicht: (2025)
von: Khalid, Osama, et al.
Veröffentlicht: (2025)
How Well Can LLMs Echo Us? Evaluating AI Chatbots' Role-Play Ability with ECHO
von: Ng, Man Tik, et al.
Veröffentlicht: (2024)
von: Ng, Man Tik, et al.
Veröffentlicht: (2024)
When to Retrieve: Teaching LLMs to Utilize Information Retrieval Effectively
von: Labruna, Tiziano, et al.
Veröffentlicht: (2024)
von: Labruna, Tiziano, et al.
Veröffentlicht: (2024)
LLMs Can Compensate for Deficiencies in Visual Representations
von: Takishita, Sho, et al.
Veröffentlicht: (2025)
von: Takishita, Sho, et al.
Veröffentlicht: (2025)
Lexicon-Enriched Graph Modeling for Arabic Document Readability Prediction
von: Elchafei, Passant, et al.
Veröffentlicht: (2025)
von: Elchafei, Passant, et al.
Veröffentlicht: (2025)
Generative Dense Retrieval: Memory Can Be a Burden
von: Yuan, Peiwen, et al.
Veröffentlicht: (2024)
von: Yuan, Peiwen, et al.
Veröffentlicht: (2024)
Can Multimodal LLMs do Visual Temporal Understanding and Reasoning? The answer is No!
von: Imam, Mohamed Fazli, et al.
Veröffentlicht: (2025)
von: Imam, Mohamed Fazli, et al.
Veröffentlicht: (2025)
From Reviews to Requirements: Can LLMs Generate Human-Like User Stories?
von: Sakib, Shadman, et al.
Veröffentlicht: (2026)
von: Sakib, Shadman, et al.
Veröffentlicht: (2026)
Data Checklist: On Unit-Testing Datasets with Usable Information
von: Zhang, Heidi C., et al.
Veröffentlicht: (2024)
von: Zhang, Heidi C., et al.
Veröffentlicht: (2024)
LLMs Can Generate a Better Answer by Aggregating Their Own Responses
von: Li, Zichong, et al.
Veröffentlicht: (2025)
von: Li, Zichong, et al.
Veröffentlicht: (2025)
Knowledge Graph Analysis of Legal Understanding and Violations in LLMs
von: Jha, Abha, et al.
Veröffentlicht: (2025)
von: Jha, Abha, et al.
Veröffentlicht: (2025)
Can We Further Elicit Reasoning in LLMs? Critic-Guided Planning with Retrieval-Augmentation for Solving Challenging Tasks
von: Li, Xingxuan, et al.
Veröffentlicht: (2024)
von: Li, Xingxuan, et al.
Veröffentlicht: (2024)
Automating Legal Interpretation with LLMs: Retrieval, Generation, and Evaluation
von: Luo, Kangcheng, et al.
Veröffentlicht: (2025)
von: Luo, Kangcheng, et al.
Veröffentlicht: (2025)
Extracting Unlearned Information from LLMs with Activation Steering
von: Seyitoğlu, Atakan, et al.
Veröffentlicht: (2024)
von: Seyitoğlu, Atakan, et al.
Veröffentlicht: (2024)
Multi Class Depression Detection Through Tweets using Artificial Intelligence
von: Nusrat, Muhammad Osama, et al.
Veröffentlicht: (2024)
von: Nusrat, Muhammad Osama, et al.
Veröffentlicht: (2024)
The AI Co-Ethnographer: How Far Can Automation Take Qualitative Research?
von: Retkowski, Fabian, et al.
Veröffentlicht: (2025)
von: Retkowski, Fabian, et al.
Veröffentlicht: (2025)
PhageBench: Can LLMs Understand Raw Bacteriophage Genomes?
von: Hou, Yusen, et al.
Veröffentlicht: (2026)
von: Hou, Yusen, et al.
Veröffentlicht: (2026)
Let LLMs Take on the Latest Challenges! A Chinese Dynamic Question Answering Benchmark
von: Xu, Zhikun, et al.
Veröffentlicht: (2024)
von: Xu, Zhikun, et al.
Veröffentlicht: (2024)
Do Methods to Jailbreak and Defend LLMs Generalize Across Languages?
von: Atil, Berk, et al.
Veröffentlicht: (2025)
von: Atil, Berk, et al.
Veröffentlicht: (2025)
UniToMBench: Integrating Perspective-Taking to Improve Theory of Mind in LLMs
von: Thiyagarajan, Prameshwar, et al.
Veröffentlicht: (2025)
von: Thiyagarajan, Prameshwar, et al.
Veröffentlicht: (2025)
Puzzled by Puzzles: When Vision-Language Models Can't Take a Hint
von: Lee, Heekyung, et al.
Veröffentlicht: (2025)
von: Lee, Heekyung, et al.
Veröffentlicht: (2025)
Funny or Persuasive, but Not Both: Evaluating Fine-Grained Multi-Concept Control in LLMs
von: Labroo, Arya, et al.
Veröffentlicht: (2026)
von: Labroo, Arya, et al.
Veröffentlicht: (2026)
When Do LLMs Need Retrieval Augmentation? Mitigating LLMs' Overconfidence Helps Retrieval Augmentation
von: Ni, Shiyu, et al.
Veröffentlicht: (2024)
von: Ni, Shiyu, et al.
Veröffentlicht: (2024)
Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet
von: Atil, Berk, et al.
Veröffentlicht: (2025)
von: Atil, Berk, et al.
Veröffentlicht: (2025)
Can Language Models Take A Hint? Prompting for Controllable Contextualized Commonsense Inference
von: Colon-Hernandez, Pedro, et al.
Veröffentlicht: (2024)
von: Colon-Hernandez, Pedro, et al.
Veröffentlicht: (2024)
It Helps to Take a Second Opinion: Teaching Smaller LLMs to Deliberate Mutually via Selective Rationale Optimisation
von: Patnaik, Sohan, et al.
Veröffentlicht: (2025)
von: Patnaik, Sohan, et al.
Veröffentlicht: (2025)
Can LLMs Detect Their Own Hallucinations?
von: Kadotani, Sora, et al.
Veröffentlicht: (2025)
von: Kadotani, Sora, et al.
Veröffentlicht: (2025)
Can Editing LLMs Inject Harm?
von: Chen, Canyu, et al.
Veröffentlicht: (2024)
von: Chen, Canyu, et al.
Veröffentlicht: (2024)
Can LLMs Reason in the Wild with Programs?
von: Yang, Yuan, et al.
Veröffentlicht: (2024)
von: Yang, Yuan, et al.
Veröffentlicht: (2024)
BeliN: A Novel Corpus for Bengali Religious News Headline Generation using Contextual Feature Fusion
von: Osama, Md, et al.
Veröffentlicht: (2025)
von: Osama, Md, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Tree-Averaging Algorithms for Ensemble-Based Unsupervised Discontinuous Constituency Parsing
von: Shayegh, Behzad, et al.
Veröffentlicht: (2024) -
Ensemble Distillation for Unsupervised Constituency Parsing
von: Shayegh, Behzad, et al.
Veröffentlicht: (2023) -
Error Diversity Matters: An Error-Resistant Ensemble Method for Unsupervised Dependency Parsing
von: Shayegh, Behzad, et al.
Veröffentlicht: (2024) -
EBBS: An Ensemble with Bi-Level Beam Search for Zero-Shot Machine Translation
von: Wen, Yuqiao, et al.
Veröffentlicht: (2024) -
Feeding Two Birds or Favoring One? Adequacy-Fluency Tradeoffs in Evaluation and Meta-Evaluation of Machine Translation
von: Shayegh, Behzad, et al.
Veröffentlicht: (2025)