Gespeichert in:
| Hauptverfasser: | Kumar, Prince, Tamilselvam, Srikanth, Garg, Dinesh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2403.10205 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ConCodeEval: Evaluating Large Language Models for Code Constraints in Domain-Specific Languages
von: Kammakomati, Mehant, et al.
Veröffentlicht: (2024)
von: Kammakomati, Mehant, et al.
Veröffentlicht: (2024)
DocCGen: Document-based Controlled Code Generation
von: Pimparkhede, Sameer, et al.
Veröffentlicht: (2024)
von: Pimparkhede, Sameer, et al.
Veröffentlicht: (2024)
ETF: An Entity Tracing Framework for Hallucination Detection in Code Summaries
von: Maharaj, Kishan, et al.
Veröffentlicht: (2024)
von: Maharaj, Kishan, et al.
Veröffentlicht: (2024)
A Code Comprehension Benchmark for Large Language Models for Code
von: Havare, Jayant, et al.
Veröffentlicht: (2025)
von: Havare, Jayant, et al.
Veröffentlicht: (2025)
Dialog2Flow: Pre-training Soft-Contrastive Action-Driven Sentence Embeddings for Automatic Dialog Flow Extraction
von: Burdisso, Sergio, et al.
Veröffentlicht: (2024)
von: Burdisso, Sergio, et al.
Veröffentlicht: (2024)
Reading between the Lines: Can LLMs Identify Cross-Cultural Communication Gaps?
von: Saha, Sougata, et al.
Veröffentlicht: (2025)
von: Saha, Sougata, et al.
Veröffentlicht: (2025)
Reliability Analysis of Psychological Concept Extraction and Classification in User-penned Text
von: Garg, Muskan, et al.
Veröffentlicht: (2024)
von: Garg, Muskan, et al.
Veröffentlicht: (2024)
Training with Pseudo-Code for Instruction Following
von: Kumar, Prince, et al.
Veröffentlicht: (2025)
von: Kumar, Prince, et al.
Veröffentlicht: (2025)
Vocabulary embeddings organize linguistic structure early in language model training
von: Papadimitriou, Isabel, et al.
Veröffentlicht: (2025)
von: Papadimitriou, Isabel, et al.
Veröffentlicht: (2025)
Don't Believe Everything You Read: Enhancing Summarization Interpretability through Automatic Identification of Hallucinations in Large Language Models
von: Vakharia, Priyesh, et al.
Veröffentlicht: (2023)
von: Vakharia, Priyesh, et al.
Veröffentlicht: (2023)
Fill in the Blank: Exploring and Enhancing LLM Capabilities for Backward Reasoning in Math Word Problems
von: Deb, Aniruddha, et al.
Veröffentlicht: (2023)
von: Deb, Aniruddha, et al.
Veröffentlicht: (2023)
Contradiction Detection in RAG Systems: Evaluating LLMs as Context Validators for Improved Information Consistency
von: Gokul, Vignesh, et al.
Veröffentlicht: (2025)
von: Gokul, Vignesh, et al.
Veröffentlicht: (2025)
Lost in Transcription: How Speech-to-Text Errors Derail Code Understanding
von: Havare, Jayant, et al.
Veröffentlicht: (2026)
von: Havare, Jayant, et al.
Veröffentlicht: (2026)
Robustness and Reasoning Fidelity of Large Language Models in Long-Context Code Question Answering
von: Maharaj, Kishan, et al.
Veröffentlicht: (2026)
von: Maharaj, Kishan, et al.
Veröffentlicht: (2026)
Lightweight Spatial Modeling for Combinatorial Information Extraction From Documents
von: Dong, Yanfei, et al.
Veröffentlicht: (2024)
von: Dong, Yanfei, et al.
Veröffentlicht: (2024)
Functional Benchmarks for Robust Evaluation of Reasoning Performance, and the Reasoning Gap
von: Srivastava, Saurabh, et al.
Veröffentlicht: (2024)
von: Srivastava, Saurabh, et al.
Veröffentlicht: (2024)
Can LLMs Reliably Simulate Real Students' Abilities in Mathematics and Reading Comprehension?
von: Srivatsa, KV Aditya, et al.
Veröffentlicht: (2025)
von: Srivatsa, KV Aditya, et al.
Veröffentlicht: (2025)
Node-weighted Graph Convolutional Network for Depression Detection in Transcribed Clinical Interviews
von: Burdisso, Sergio, et al.
Veröffentlicht: (2023)
von: Burdisso, Sergio, et al.
Veröffentlicht: (2023)
CodeVaani: A Multilingual, Voice-Based Code Learning Assistant
von: Havare, Jayant, et al.
Veröffentlicht: (2025)
von: Havare, Jayant, et al.
Veröffentlicht: (2025)
ABEX: Data Augmentation for Low-Resource NLU via Expanding Abstract Descriptions
von: Ghosh, Sreyan, et al.
Veröffentlicht: (2024)
von: Ghosh, Sreyan, et al.
Veröffentlicht: (2024)
A Comprehensive Survey of Accelerated Generation Techniques in Large Language Models
von: Khoshnoodi, Mahsa, et al.
Veröffentlicht: (2024)
von: Khoshnoodi, Mahsa, et al.
Veröffentlicht: (2024)
Evaluating LLMs for Zeolite Synthesis Event Extraction (ZSEE): A Systematic Analysis of Prompting Strategies
von: Rathore, Charan Prakash, et al.
Veröffentlicht: (2025)
von: Rathore, Charan Prakash, et al.
Veröffentlicht: (2025)
Cyber for AI at SemEval-2025 Task 4: Forgotten but Not Lost: The Balancing Act of Selective Unlearning in Large Language Models
von: P, Dinesh Srivasthav, et al.
Veröffentlicht: (2025)
von: P, Dinesh Srivasthav, et al.
Veröffentlicht: (2025)
ToolWeave: Structured Synthesis of Complex Multi-Turn Tool-Calling Dialogues
von: Khandelwal, Dinesh, et al.
Veröffentlicht: (2026)
von: Khandelwal, Dinesh, et al.
Veröffentlicht: (2026)
Cross-Domain Content Generation with Domain-Specific Small Language Models
von: Maloo, Ankit, et al.
Veröffentlicht: (2024)
von: Maloo, Ankit, et al.
Veröffentlicht: (2024)
RexBERT: Context Specialized Bidirectional Encoders for E-commerce
von: Bajaj, Rahul, et al.
Veröffentlicht: (2026)
von: Bajaj, Rahul, et al.
Veröffentlicht: (2026)
Automatic Information Extraction From Employment Tribunal Judgements Using Large Language Models
von: de Faria, Joana Ribeiro, et al.
Veröffentlicht: (2024)
von: de Faria, Joana Ribeiro, et al.
Veröffentlicht: (2024)
LLMQuoter: Enhancing RAG Capabilities Through Efficient Quote Extraction From Large Contexts
von: Bezerra, Yuri Facanha, et al.
Veröffentlicht: (2025)
von: Bezerra, Yuri Facanha, et al.
Veröffentlicht: (2025)
Thinking About Thinking: Evaluating Reasoning in Post-Trained Language Models
von: Singla, Pratham, et al.
Veröffentlicht: (2025)
von: Singla, Pratham, et al.
Veröffentlicht: (2025)
VyAnG-Net: A Novel Multi-Modal Sarcasm Recognition Model by Uncovering Visual, Acoustic and Glossary Features
von: Pandey, Ananya, et al.
Veröffentlicht: (2024)
von: Pandey, Ananya, et al.
Veröffentlicht: (2024)
SeaAlert: Critical Information Extraction From Maritime Distress Communications with Large Language Models
von: Atia, Tomer, et al.
Veröffentlicht: (2026)
von: Atia, Tomer, et al.
Veröffentlicht: (2026)
COPAL: Continual Pruning in Large Language Generative Models
von: Malla, Srikanth, et al.
Veröffentlicht: (2024)
von: Malla, Srikanth, et al.
Veröffentlicht: (2024)
Recursive Introspection: Teaching Language Model Agents How to Self-Improve
von: Qu, Yuxiao, et al.
Veröffentlicht: (2024)
von: Qu, Yuxiao, et al.
Veröffentlicht: (2024)
ReviewEval: An Evaluation Framework for AI-Generated Reviews
von: Garg, Madhav Krishan, et al.
Veröffentlicht: (2025)
von: Garg, Madhav Krishan, et al.
Veröffentlicht: (2025)
Read Before You Think: Mitigating LLM Comprehension Failures with Step-by-Step Reading
von: Han, Feijiang, et al.
Veröffentlicht: (2025)
von: Han, Feijiang, et al.
Veröffentlicht: (2025)
ADEPT: Adaptive Dynamic Early-Exit Process for Transformers
von: Yoo, Sangmin, et al.
Veröffentlicht: (2026)
von: Yoo, Sangmin, et al.
Veröffentlicht: (2026)
Assessing the Quality of AI-Generated Clinical Notes: A Validated Evaluation of a Large Language Model Scribe
von: Palm, Erin, et al.
Veröffentlicht: (2025)
von: Palm, Erin, et al.
Veröffentlicht: (2025)
An EcoSage Assistant: Towards Building A Multimodal Plant Care Dialogue Assistant
von: Tomar, Mohit, et al.
Veröffentlicht: (2024)
von: Tomar, Mohit, et al.
Veröffentlicht: (2024)
SpecBench: Measuring Reward Hacking in Long-Horizon Coding Agents
von: Zhao, Bingchen, et al.
Veröffentlicht: (2026)
von: Zhao, Bingchen, et al.
Veröffentlicht: (2026)
Towards Event Extraction with Massive Types: LLM-based Collaborative Annotation and Partitioning Extraction
von: Liu, Wenxuan, et al.
Veröffentlicht: (2025)
von: Liu, Wenxuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
ConCodeEval: Evaluating Large Language Models for Code Constraints in Domain-Specific Languages
von: Kammakomati, Mehant, et al.
Veröffentlicht: (2024) -
DocCGen: Document-based Controlled Code Generation
von: Pimparkhede, Sameer, et al.
Veröffentlicht: (2024) -
ETF: An Entity Tracing Framework for Hallucination Detection in Code Summaries
von: Maharaj, Kishan, et al.
Veröffentlicht: (2024) -
A Code Comprehension Benchmark for Large Language Models for Code
von: Havare, Jayant, et al.
Veröffentlicht: (2025) -
Dialog2Flow: Pre-training Soft-Contrastive Action-Driven Sentence Embeddings for Automatic Dialog Flow Extraction
von: Burdisso, Sergio, et al.
Veröffentlicht: (2024)