글로 표현한 내용을 그림으로 만들어주는 인공 지능 이미지 생성 서비스인 Stable Diffusion(https://stability.ai/)을 이용해 몇 가지 모델로 남녀 인물을 묘사했을 때 어떤 차이가 있는지 간략하게 살펴보았다.
시험한 모델은 아래와 같다.
V2.1 EMA
AbyssOrangeMix2
Basil Mix
Realistic Vision V1.3
각 모델에 생성 조건은 모두 동일하게 적용했다.
Positive prompt: analog style, beautiful woman/handsome man, age 50, (high detailed skin:1.2), pale skin, modern architecture, Intricate, High Detail, Sharp focus, dramatic, proportional body, a highly detailed face by Greg Rutkowski and Magali Villeneuve, beautiful/handsome eyes, toned body, cinematic lighting, depth of field, photography
Negative prompt: (deformed iris, deformed pupils, semi-realistic, cgi, 3d, render, sketch, cartoon, drawing, anime:1.4), text, close up, cropped, out of frame, worst quality, low quality, jpeg artifacts, ugly, duplicate, morbid, mutilated, extra fingers, mutated hands, poorly drawn hands, poorly drawn face, mutation, deformed, blurry, dehydrated, bad anatomy, bad proportions, extra limbs, cloned face, disfigured, gross proportions, malformed limbs, missing arms, missing legs, extra arms, extra legs, fused fingers, too many fingers, long neck
Sampler: DPM++ SDE Karras
Steps: 25
CFG scale: 7
이런 조건에 따라 남녀, 나이(기본값, 20세, 50세), 국가(기본값, 한국, 일본 중국)별로 이미지를 만들었다.
같은 조건을 같은 모델에 적용해도 때에 따라 다른 결과가 나오기 때문에 분위기가 전체적인 스타일들만 확인하면 된다.