Coordinates of Capability: A Unified MTMM-Geometric Framework for LLM Evaluation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sakhawat, Adib, Islam, Tahsin, Farhin, Takia, Raiyan, Syed Rifat, Mahmud, Hasan, Hasan, Md Kamrul |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Political Alignment in Large Language Models: A Multidimensional Audit of Psychometric Identity and Behavioral Bias
von: Sakhawat, Adib, et al.
Veröffentlicht: (2026)
von: Sakhawat, Adib, et al.
Veröffentlicht: (2026)
Beyond Symbolic Solving: Multi Chain-of-Thought Voting for Geometric Reasoning in Large Language Models
von: Siddique, Md. Abu Bakor, et al.
Veröffentlicht: (2026)
von: Siddique, Md. Abu Bakor, et al.
Veröffentlicht: (2026)
BanglaNirTox: A Large-scale Parallel Corpus for Explainable AI in Bengali Text Detoxification
von: Mohsin, Ayesha Afroza, et al.
Veröffentlicht: (2025)
von: Mohsin, Ayesha Afroza, et al.
Veröffentlicht: (2025)
Courtroom-Style Multi-Agent Debate with Progressive RAG and Role-Switching for Controversial Claim Verification
von: Chowdhury, Masnun Nuha, et al.
Veröffentlicht: (2026)
von: Chowdhury, Masnun Nuha, et al.
Veröffentlicht: (2026)
PhysicsEval: Inference-Time Techniques to Improve the Reasoning Proficiency of Large Language Models on Physics Problems
von: Siddique, Oshayer, et al.
Veröffentlicht: (2025)
von: Siddique, Oshayer, et al.
Veröffentlicht: (2025)
BanglaLorica: Design and Evaluation of a Robust Watermarking Algorithm for Large Language Models in Bangla Text Generation
von: Tariqul, Amit Bin, et al.
Veröffentlicht: (2026)
von: Tariqul, Amit Bin, et al.
Veröffentlicht: (2026)
Narrative over Numbers: The Identifiable Victim Effect and its Amplification Under Alignment and Reasoning in Large Language Models
von: Raiyan, Syed Rifat
Veröffentlicht: (2026)
von: Raiyan, Syed Rifat
Veröffentlicht: (2026)
Bangla Sign Language Translation: Dataset Creation Challenges, Benchmarking and Prospects
von: Rubaiyeat, Husne Ara, et al.
Veröffentlicht: (2025)
von: Rubaiyeat, Husne Ara, et al.
Veröffentlicht: (2025)
CircuitLM: A Multi-Agent LLM-Aided Design Framework for Generating Circuit Schematics from Natural Language Prompts
von: Hasan, Khandakar Shakib Al, et al.
Veröffentlicht: (2026)
von: Hasan, Khandakar Shakib Al, et al.
Veröffentlicht: (2026)
BDA: Bangla Text Data Augmentation Framework
von: Tariquzzaman, Md., et al.
Veröffentlicht: (2024)
von: Tariquzzaman, Md., et al.
Veröffentlicht: (2024)
AREG: Adversarial Resource Extraction Game for Evaluating Persuasion and Resistance in Large Language Models
von: Sakhawat, Adib, et al.
Veröffentlicht: (2026)
von: Sakhawat, Adib, et al.
Veröffentlicht: (2026)
CogniAlign: Survivability-Grounded Multi-Agent Moral Reasoning for Safe and Transparent AI
von: Ali, Hasin Jawad, et al.
Veröffentlicht: (2025)
von: Ali, Hasin Jawad, et al.
Veröffentlicht: (2025)
AIDG: A Formal Decomposition of Information Extraction and Containment Asymmetries in Multi-Turn LLM Dialogue
von: Sakhawat, Adib, et al.
Veröffentlicht: (2026)
von: Sakhawat, Adib, et al.
Veröffentlicht: (2026)
When Words Don't Mean What They Say: Figurative Understanding in Bengali Idioms
von: Sakhawat, Adib, et al.
Veröffentlicht: (2026)
von: Sakhawat, Adib, et al.
Veröffentlicht: (2026)
Transformer-Driven Triple Fusion Framework for Enhanced Multimodal Author Intent Classification in Low-Resource Bangla
von: Islam, Ariful, et al.
Veröffentlicht: (2025)
von: Islam, Ariful, et al.
Veröffentlicht: (2025)
BanglaSentNet: An Explainable Hybrid Deep Learning Framework for Multi-Aspect Sentiment Analysis with Cross-Domain Transfer Learning
von: Islam, Ariful, et al.
Veröffentlicht: (2025)
von: Islam, Ariful, et al.
Veröffentlicht: (2025)
A Deep Learning-based Multimodal Depth-Aware Dynamic Hand Gesture Recognition System
von: Mahmud, Hasan, et al.
Veröffentlicht: (2021)
von: Mahmud, Hasan, et al.
Veröffentlicht: (2021)
FrugalPrompt: Reducing Contextual Overhead in Large Language Models via Token Attribution
von: Raiyan, Syed Rifat, et al.
Veröffentlicht: (2025)
von: Raiyan, Syed Rifat, et al.
Veröffentlicht: (2025)
Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis
von: Bhowmik, Shimanto, et al.
Veröffentlicht: (2025)
von: Bhowmik, Shimanto, et al.
Veröffentlicht: (2025)
Are ASR foundation models generalized enough to capture features of regional dialects for low-resource languages?
von: Dipto, Tawsif Tashwar, et al.
Veröffentlicht: (2025)
von: Dipto, Tawsif Tashwar, et al.
Veröffentlicht: (2025)
Gradient Masters at BLP-2025 Task 1: Advancing Low-Resource NLP for Bengali using Ensemble-Based Adversarial Training for Hate Speech Detection
von: Hoque, Syed Mohaiminul, et al.
Veröffentlicht: (2025)
von: Hoque, Syed Mohaiminul, et al.
Veröffentlicht: (2025)
BdSLW60: A Word-Level Bangla Sign Language Dataset
von: Rubaiyeat, Husne Ara, et al.
Veröffentlicht: (2024)
von: Rubaiyeat, Husne Ara, et al.
Veröffentlicht: (2024)
BdSLW401: Transformer-Based Word-Level Bangla Sign Language Recognition Using Relative Quantization Encoding (RQE)
von: Rubaiyeat, Husne Ara, et al.
Veröffentlicht: (2025)
von: Rubaiyeat, Husne Ara, et al.
Veröffentlicht: (2025)
Fine-Tuning Video Transformers for Word-Level Bangla Sign Language: A Comparative Analysis for Classification Tasks
von: Shawon, Jubayer Ahmed Bhuiyan, et al.
Veröffentlicht: (2025)
von: Shawon, Jubayer Ahmed Bhuiyan, et al.
Veröffentlicht: (2025)
SCReedSolo: A Secure and Robust LSB Image Steganography Framework with Randomized Symmetric Encryption and Reed-Solomon Coding
von: Raiyan, Syed Rifat, et al.
Veröffentlicht: (2025)
von: Raiyan, Syed Rifat, et al.
Veröffentlicht: (2025)
Prompting with Sign Parameters for Low-resource Sign Language Instruction Generation
von: Tariquzzaman, Md, et al.
Veröffentlicht: (2025)
von: Tariquzzaman, Md, et al.
Veröffentlicht: (2025)
TabSQLify: Enhancing Reasoning Capabilities of LLMs Through Table Decomposition
von: Nahid, Md Mahadi Hasan, et al.
Veröffentlicht: (2024)
von: Nahid, Md Mahadi Hasan, et al.
Veröffentlicht: (2024)
Pruning for Protection: Increasing Jailbreak Resistance in Aligned LLMs Without Fine-Tuning
von: Hasan, Adib, et al.
Veröffentlicht: (2024)
von: Hasan, Adib, et al.
Veröffentlicht: (2024)
Pitfalls of Evaluating Language Models with Open Benchmarks
von: Hasan, Md. Najib, et al.
Veröffentlicht: (2025)
von: Hasan, Md. Najib, et al.
Veröffentlicht: (2025)
BanglaSocialBench: A Benchmark for Evaluating Sociopragmatic and Cultural Alignment of LLMs in Bangladeshi Social Interaction
von: Sijan, Tanvir Ahmed, et al.
Veröffentlicht: (2026)
von: Sijan, Tanvir Ahmed, et al.
Veröffentlicht: (2026)
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation
von: Sakib, Tanjil Hasan, et al.
Veröffentlicht: (2025)
von: Sakib, Tanjil Hasan, et al.
Veröffentlicht: (2025)
Auxilio and Beyond: Comparative Evaluation, Usability, and Design Guidelines for Head Movement-based Assistive Mouse Controllers
von: Kabir, Mohammad Ridwan, et al.
Veröffentlicht: (2022)
von: Kabir, Mohammad Ridwan, et al.
Veröffentlicht: (2022)
Are LLMs Ready to Replace Bangla Annotators?
von: Hasan, Md. Najib, et al.
Veröffentlicht: (2026)
von: Hasan, Md. Najib, et al.
Veröffentlicht: (2026)
BanglaASTE: A Novel Framework for Aspect-Sentiment-Opinion Extraction in Bangla E-commerce Reviews Using Ensemble Deep Learning
von: Islam, Ariful, et al.
Veröffentlicht: (2025)
von: Islam, Ariful, et al.
Veröffentlicht: (2025)
Bounded Behavioral Indistinguishability for Black-Box LLM Distillation
von: Hasan, Munawar
Veröffentlicht: (2026)
von: Hasan, Munawar
Veröffentlicht: (2026)
Privacy Preserving Topic-wise Sentiment Analysis of the Iran Israel USA Conflict Using Federated Transformer Models
von: Islam, Md Saiful, et al.
Veröffentlicht: (2026)
von: Islam, Md Saiful, et al.
Veröffentlicht: (2026)
DPCSpell: A Transformer-based Detector-Purificator-Corrector Framework for Spelling Error Correction of Bangla and Resource Scarce Indic Languages
von: Bijoy, Mehedi Hasan, et al.
Veröffentlicht: (2022)
von: Bijoy, Mehedi Hasan, et al.
Veröffentlicht: (2022)
SecBPMN-GPT: A Hybrid LLM & Rule-Based Framework for Automating Security Annotation in Business Process Models
von: Islam, Md. Kamrul, et al.
Veröffentlicht: (2025)
von: Islam, Md. Kamrul, et al.
Veröffentlicht: (2025)
An Efficient Approach for Solving Expensive Constrained Multiobjective Optimization Problems
von: Rahi, Kamrul Hasan
Veröffentlicht: (2024)
von: Rahi, Kamrul Hasan
Veröffentlicht: (2024)
Rep3Net: An Approach Exploiting Multimodal Representation for Molecular Bioactivity Prediction
von: Islam, Sabrina, et al.
Veröffentlicht: (2025)
von: Islam, Sabrina, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Political Alignment in Large Language Models: A Multidimensional Audit of Psychometric Identity and Behavioral Bias
von: Sakhawat, Adib, et al.
Veröffentlicht: (2026) -
Beyond Symbolic Solving: Multi Chain-of-Thought Voting for Geometric Reasoning in Large Language Models
von: Siddique, Md. Abu Bakor, et al.
Veröffentlicht: (2026) -
BanglaNirTox: A Large-scale Parallel Corpus for Explainable AI in Bengali Text Detoxification
von: Mohsin, Ayesha Afroza, et al.
Veröffentlicht: (2025) -
Courtroom-Style Multi-Agent Debate with Progressive RAG and Role-Switching for Controversial Claim Verification
von: Chowdhury, Masnun Nuha, et al.
Veröffentlicht: (2026) -
PhysicsEval: Inference-Time Techniques to Improve the Reasoning Proficiency of Large Language Models on Physics Problems
von: Siddique, Oshayer, et al.
Veröffentlicht: (2025)