← Back to archive
Artificial Intelligence

Text-to-Image Generation Results with Baidu Wenxin

It is clear that the generation still has quite a few issues: the character's head remains less than ideal. The background is still quite good and the clothing is fine, but the face falls short. What especially needs noting is that after switching to oil-painting style, the generated results — especially faces — remain hard to control; this is a common flaw of diffusion models. Sometimes it does produce acceptable images, and overall, cartoon-style faces are relatively acceptable. So among the keywords, I tried covering the face, but unfortunately Wenxin doesn't seem to recognize such instructions — though the results turned out surprisingly decent. Currently Baidu Wenxin's text-to-image only offers watercolor, oil painting, chalk drawing, cartoon, crayon drawing, and children's drawing styles. Not yet able to try more interesting image generation — still pretty solid.

Written by Master Sanfu on September 1, 2022. Please credit the source if you share.

Translation Notice: This English version was translated with AI assistance. Specialized, historical, religious, or culturally sensitive terms may contain nuances, inaccuracies, or debatable wording. In case of ambiguity or discrepancy, the original Chinese text shall prevail.