News-to-Meme Generation: Performance Analysis of GPT-3, BLIP and CLIP+GPT Using BBC News Headlines
DOI:
https://doi.org/10.61453/INTIj.20260330Keywords:
Meme Generation, Multimodal AI, GPT-3, BLIP, CLIP, Text-to-ImageAbstract
Making memes has become a powerful way to share cultural and funny commentary. This study assesses the efficacy of AI models in autonomously creating memes from real-world data. The comparative study of three prominent models—GPT-3, BLIP, and CLIP+GPT—utilized 500 BBC news headlines as input prompts. Using a consistent GPT-based evaluation framework, generated memes are rated on five scales: Humor, Relevance, Engagement, Creativity, and Clarity. The findings indicate that GPT-3 excels in Humor and Relevance, CLIP+GPT is superior in Engagement, and BLIP offers balanced results
References
Chen, Y., Yan, S., Zhu, Z., Li, Z., & Xiao, Y. (2024). XMeCap: Meme caption generation with sub-image adaptability. arXiv. https://doi.org/10.48550/arXiv.2407.17152
Kim, S., & Chilton, L. B. (2025). AI humor generation: Cognitive, social and creative skills for effective humor. arXiv. https://doi.org/10.48550/arXiv.2502.07981
Li, J., Li, D., Savarese, S., & Hoi, S. C. H. (2022). BLIP: Bootstrapping language-image pre-training for unified vision-language understanding and generation. Proceedings of the 39th International Conference on Machine Learning. https://doi.org/10.48550/arXiv.2201.12086
Radford, A., Kim, J. W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., Krueger, G., & Sutskever, I. (2021). Learning transferable visual models from natural language supervision. Proceedings of the 38th International Conference on Machine Learning, *139*, 8748–8763. https://doi.org/10.48550/arXiv.2103.00020
Sadasivam, A., Gunasekar, K., Davulcu, H., & Yang, Y. (2020). memeBot: Towards automatic image meme generation. arXiv. https://doi.org/10.48550/arXiv.2004.14571
Vyalla, S. R., Balakrishnan, A., Subramanian, A., & Das, A. (2019). Memeify: A large-scale meme generation system. arXiv. https://doi.org/10.48550/arXiv.1910.12279
Wang, H., & Lee, R. K.-W. (2024). MemeCraft: Contextual and stance-driven multimodal meme generation. arXiv. https://doi.org/10.48550/arXiv.2403.14652
Downloads
Published
How to Cite
Issue
Section
License
Copyright (c) 2026 INTI Journal

This work is licensed under a Creative Commons Attribution 4.0 International License.