Sharma, A., & Aggarwal, D. M. (2026). ConvNeXt-Driven Image Captioning in Hindi with Adaptive Attention and Transformer Decoder. Journal of Graphic Era University, 14(02), 607–624. https://doi.org/10.13052/jgeu0975-1416.14210