Packaging Virtual Image Auxiliary Generation Algorithm based on Large Language Model (LLM)

Yang Zhou, Fan Zhang · 2023

When applying the Large Language Model (LLM) to image processing, it is crucial to control the training size and the accuracy of the algorithm. This research study proposes a novel LLM-based algorithm for the generation of auxiliary virtual images. The proposed approach is based on a two-step strategy, namely the optimized LLM and the joint pix2pix model, which integrates the neural structure into the traditional processing pipelines. For the designed LLM, this study uses the Transformer's global interactive ability that combines with the local characteristics of CNN to enrich the feature diversity, then the input feature maps are divided into multiple groups and further, then fuse with the updated regulation to achieve the initial generation task. For the joint pix2pix mode, the original image is generated by the generator to generate a new image, the new image and the original image are fused together as fake data and sent to the discriminator for training. The experimental results on the small and large datasets show that the proposed approach outperforms.

Read the paper · More papers on PaperTik