Prompt Refinement from Text to Image Generation using Multi model Approach


Date Published : 14 September 2026

Contributors

MANJULA R

SRM Institute of Science and Technology
Author

S.Hemalatha

Panimalar Engineering College, Chennai;
Author

Keywords

Vision Language models; CLIP; BLIP; Prompt learning; Text-to-image Generation

Proceeding

Track

General Track

License

Copyright (c) 2026 Sustainable Global Societies Initiative

Creative Commons License

This work is licensed under a Creative Commons Attribution-NonCommercial-NoDerivatives 4.0 International License.

Abstract

The rapid expansion of digital content across textual, visual, and multimodal platforms has significantly increased the demand for intelligent systems. Text to image models, which produce realistic visuals from prompts, have been made possible by a recent discovery in generative artificial intelligence. An iterative prompt refining framework that automatically enhances the prompt for better image production is proposed in this work. The suggested approach combines several AI models, such as Large Language Models for quick refining, BLIP for picture captioning, and Stable Diffusion for image production. The suggested pipeline uses the user's prompt to create an initial image, which is then examined using BLIP to extract a textual description. This process is repeated across multiple iterations to enhance the prompt quality and generated image results. To evaluate the effectiveness of the proposed approach, the refined prompts are compared with the baseline keyword-based prompt refinement method. The images are evaluated using two metrics: CLIP score and aesthetic score, which measures prompt alignment and visual quality. Experimental results show that the LLM-based refinement approach performs better compared to the baseline refinement method as well as improves the alignment of the generated image and prompts

References

No References

How to Cite

Rajagopal, M., & S, H. (2026). Prompt Refinement from Text to Image Generation using Multi model Approach. Sustainable Global Societies Initiative, 1(11). https://vectmag.com/sgsi/paper/view/1225