About this document
Enhancing Image Captioning with CLIP by vangipurapuvishnuvardhan is a document available to read on EtoBox.
The document discusses a project at Geethanjali Institute of Science and Technology focused on enhancing image captioning using CLIP (Contrastive Language-Image Pretraining) as a prefix model. It highlights the challenges of traditional image captioning methods and demonstrates that integrating CLIP can improve caption diversity, coherence, and relevance. The document also outlines system requirements and existing systems in the field of image captioning.
- Author
- vangipurapuvishnuvardhan
- Language
- EN