Unification of Balti and trans-border sister dialects in the essence of LLMs and AI Technology
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sharif, Muhammad, Yi, Jiangyan, Shoaib, Muhammad |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multi-modal Generative AI: Multi-modal LLMs, Diffusions, and the Unification
von: Wang, Xin, et al.
Veröffentlicht: (2024)
von: Wang, Xin, et al.
Veröffentlicht: (2024)
VIST-GPT: Ushering in the Era of Visual Storytelling with LLMs?
von: Gado, Mohamed, et al.
Veröffentlicht: (2025)
von: Gado, Mohamed, et al.
Veröffentlicht: (2025)
DefAn: Definitive Answer Dataset for LLMs Hallucination Evaluation
von: Rahman, A B M Ashikur, et al.
Veröffentlicht: (2024)
von: Rahman, A B M Ashikur, et al.
Veröffentlicht: (2024)
Learning Human-Perceived Fakeness in AI-Generated Videos via Multimodal LLMs
von: Fu, Xingyu, et al.
Veröffentlicht: (2025)
von: Fu, Xingyu, et al.
Veröffentlicht: (2025)
MedObvious: Exposing the Medical Moravec's Paradox in VLMs via Clinical Triage
von: Khan, Ufaq, et al.
Veröffentlicht: (2026)
von: Khan, Ufaq, et al.
Veröffentlicht: (2026)
Unlearning Sensitive Information in Multimodal LLMs: Benchmark and Attack-Defense Evaluation
von: Patil, Vaidehi, et al.
Veröffentlicht: (2025)
von: Patil, Vaidehi, et al.
Veröffentlicht: (2025)
Qalam : A Multimodal LLM for Arabic Optical Character and Handwriting Recognition
von: Bhatia, Gagan, et al.
Veröffentlicht: (2024)
von: Bhatia, Gagan, et al.
Veröffentlicht: (2024)
LLMs can Compress LLMs: Adaptive Pruning by Agents
von: Kodathala, Sai Varun, et al.
Veröffentlicht: (2026)
von: Kodathala, Sai Varun, et al.
Veröffentlicht: (2026)
MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs
von: Fu, Chaoyou, et al.
Veröffentlicht: (2024)
von: Fu, Chaoyou, et al.
Veröffentlicht: (2024)
Vision Enhancing LLMs: Empowering Multimodal Knowledge Storage and Sharing in LLMs
von: Li, Yunxin, et al.
Veröffentlicht: (2023)
von: Li, Yunxin, et al.
Veröffentlicht: (2023)
Multimodal Generative AI with Autoregressive LLMs for Human Motion Understanding and Generation: A Way Forward
von: Islam, Muhammad, et al.
Veröffentlicht: (2025)
von: Islam, Muhammad, et al.
Veröffentlicht: (2025)
Scientific Reasoning: Assessment of Multimodal Generative LLMs
von: Dreyer, Florian, et al.
Veröffentlicht: (2025)
von: Dreyer, Florian, et al.
Veröffentlicht: (2025)
LLMs Can Compensate for Deficiencies in Visual Representations
von: Takishita, Sho, et al.
Veröffentlicht: (2025)
von: Takishita, Sho, et al.
Veröffentlicht: (2025)
AI for Service: Proactive Assistance with AI Glasses
von: Wen, Zichen, et al.
Veröffentlicht: (2025)
von: Wen, Zichen, et al.
Veröffentlicht: (2025)
Leveraging LLMs for On-the-Fly Instruction Guided Image Editing
von: Santos, Rodrigo, et al.
Veröffentlicht: (2024)
von: Santos, Rodrigo, et al.
Veröffentlicht: (2024)
Decoder-Only LLMs are Better Controllers for Diffusion Models
von: Dong, Ziyi, et al.
Veröffentlicht: (2025)
von: Dong, Ziyi, et al.
Veröffentlicht: (2025)
Bayesian Optimization for Controlled Image Editing via LLMs
von: Cai, Chengkun, et al.
Veröffentlicht: (2025)
von: Cai, Chengkun, et al.
Veröffentlicht: (2025)
Analyzing Finetuning Representation Shift for Multimodal LLMs Steering
von: Khayatan, Pegah, et al.
Veröffentlicht: (2025)
von: Khayatan, Pegah, et al.
Veröffentlicht: (2025)
Flattery in Motion: Benchmarking and Analyzing Sycophancy in Video-LLMs
von: Zhou, Wenrui, et al.
Veröffentlicht: (2025)
von: Zhou, Wenrui, et al.
Veröffentlicht: (2025)
Thinking Like a Botanist: Challenging Multimodal Language Models with Intent-Driven Chain-of-Inquiry
von: Sakib, Syed Nazmus, et al.
Veröffentlicht: (2026)
von: Sakib, Syed Nazmus, et al.
Veröffentlicht: (2026)
SCOPE: Sign Language Contextual Processing with Embedding from LLMs
von: Liu, Yuqi, et al.
Veröffentlicht: (2024)
von: Liu, Yuqi, et al.
Veröffentlicht: (2024)
Mitigating Adversarial Attacks in LLMs through Defensive Suffix Generation
von: Kim, Minkyoung, et al.
Veröffentlicht: (2024)
von: Kim, Minkyoung, et al.
Veröffentlicht: (2024)
Moment Sampling in Video LLMs for Long-Form Video QA
von: Chasmai, Mustafa, et al.
Veröffentlicht: (2025)
von: Chasmai, Mustafa, et al.
Veröffentlicht: (2025)
EasyGen: Easing Multimodal Generation with BiDiffuser and LLMs
von: Zhao, Xiangyu, et al.
Veröffentlicht: (2023)
von: Zhao, Xiangyu, et al.
Veröffentlicht: (2023)
Lost in Time: Clock and Calendar Understanding Challenges in Multimodal LLMs
von: Saxena, Rohit, et al.
Veröffentlicht: (2025)
von: Saxena, Rohit, et al.
Veröffentlicht: (2025)
Calibrating Uncertainty Quantification of Multi-Modal LLMs using Grounding
von: Padhi, Trilok, et al.
Veröffentlicht: (2025)
von: Padhi, Trilok, et al.
Veröffentlicht: (2025)
Multimodal LLMs as Customized Reward Models for Text-to-Image Generation
von: Zhou, Shijie, et al.
Veröffentlicht: (2025)
von: Zhou, Shijie, et al.
Veröffentlicht: (2025)
MLLM-CompBench: A Comparative Reasoning Benchmark for Multimodal LLMs
von: Kil, Jihyung, et al.
Veröffentlicht: (2024)
von: Kil, Jihyung, et al.
Veröffentlicht: (2024)
Euclid: Supercharging Multimodal LLMs with Synthetic High-Fidelity Visual Descriptions
von: Zhang, Jiarui, et al.
Veröffentlicht: (2024)
von: Zhang, Jiarui, et al.
Veröffentlicht: (2024)
TP-Eval: Tap Multimodal LLMs' Potential in Evaluation by Customizing Prompts
von: Xie, Yuxuan, et al.
Veröffentlicht: (2024)
von: Xie, Yuxuan, et al.
Veröffentlicht: (2024)
CHARTOM: A Visual Theory-of-Mind Benchmark for LLMs on Misleading Charts
von: Bharti, Shubham, et al.
Veröffentlicht: (2024)
von: Bharti, Shubham, et al.
Veröffentlicht: (2024)
HawkEye: Training Video-Text LLMs for Grounding Text in Videos
von: Wang, Yueqian, et al.
Veröffentlicht: (2024)
von: Wang, Yueqian, et al.
Veröffentlicht: (2024)
VF-Eval: Evaluating Multimodal LLMs for Generating Feedback on AIGC Videos
von: Song, Tingyu, et al.
Veröffentlicht: (2025)
von: Song, Tingyu, et al.
Veröffentlicht: (2025)
Image Captioning Evaluation in the Age of Multimodal LLMs: Challenges and Future Perspectives
von: Sarto, Sara, et al.
Veröffentlicht: (2025)
von: Sarto, Sara, et al.
Veröffentlicht: (2025)
Leveraging Multimodal-LLMs Assisted by Instance Segmentation for Intelligent Traffic Monitoring
von: Onsu, Murat Arda, et al.
Veröffentlicht: (2025)
von: Onsu, Murat Arda, et al.
Veröffentlicht: (2025)
Multimodal LLMs Struggle with Basic Visual Network Analysis: a VNA Benchmark
von: Williams, Evan M., et al.
Veröffentlicht: (2024)
von: Williams, Evan M., et al.
Veröffentlicht: (2024)
Weak-eval-Strong: Evaluating and Eliciting Lateral Thinking of LLMs with Situation Puzzles
von: Chen, Qi, et al.
Veröffentlicht: (2024)
von: Chen, Qi, et al.
Veröffentlicht: (2024)
AIM: Adaptive Inference of Multi-Modal LLMs via Token Merging and Pruning
von: Zhong, Yiwu, et al.
Veröffentlicht: (2024)
von: Zhong, Yiwu, et al.
Veröffentlicht: (2024)
GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs
von: Nguyen, Duy, et al.
Veröffentlicht: (2025)
von: Nguyen, Duy, et al.
Veröffentlicht: (2025)
Grasp Any Region: Towards Precise, Contextual Pixel Understanding for Multimodal LLMs
von: Wang, Haochen, et al.
Veröffentlicht: (2025)
von: Wang, Haochen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Multi-modal Generative AI: Multi-modal LLMs, Diffusions, and the Unification
von: Wang, Xin, et al.
Veröffentlicht: (2024) -
VIST-GPT: Ushering in the Era of Visual Storytelling with LLMs?
von: Gado, Mohamed, et al.
Veröffentlicht: (2025) -
DefAn: Definitive Answer Dataset for LLMs Hallucination Evaluation
von: Rahman, A B M Ashikur, et al.
Veröffentlicht: (2024) -
Learning Human-Perceived Fakeness in AI-Generated Videos via Multimodal LLMs
von: Fu, Xingyu, et al.
Veröffentlicht: (2025) -
MedObvious: Exposing the Medical Moravec's Paradox in VLMs via Clinical Triage
von: Khan, Ufaq, et al.
Veröffentlicht: (2026)