Vision-language model for generating Arabic image captions using Bidirectional Transformers (BiT) and advanced feature fusion.
-
Updated
Jul 17, 2025 - Jupyter Notebook
Vision-language model for generating Arabic image captions using Bidirectional Transformers (BiT) and advanced feature fusion.
To associate your repository with the bertforimagecaptioning topic, visit your repo's landing page and select "manage topics."