Project Details
Abstract
Develops a framework for image understanding in Multimodal Large Language Models (MLLMs) supporting Modern Standard Arabic (MSA) and Arabic dialects. Enables vision-language tasks including image captioning, image generation, and visual question answering.
Submitting Institute Name
Hamad Bin Khalifa University (HBKU)
| Sponsor's Award Number | CHSS-IG-C3-2026-003 |
|---|---|
| Proposal ID | CHSS-CORE-000016 |
| Status | Active |
| Effective start/end date | 5/01/26 → 31/12/26 |
Primary Theme
- Artificial Intelligence
Primary Subtheme
- AI - Smart Society
Secondary Theme
- None
Secondary Subtheme
- None
Keywords
- Multimodal LLMs, Arabic NLP, Image captioning, Visual question answering, Arabic dialects, Computer vision
- None
Fingerprint
Explore the research topics touched on by this project. These labels are generated based on the underlying awards/grants. Together they form a unique fingerprint.