神秘的程序员们

Midjourney V5 比 V4 更好吗?Prompt 全公开(下篇)

  1. V5 的惊艳之处:photograpy,CG rendering, HD film style 类生成
  2.  Prompt 控制准确度的基本测试
  3. V5 比 V4 的更好的地方:
    在 V5 里可以准确控制镜头语言,光影也更真实自然
  4. V5 相比 V4 倒退的地方:
    V5 会让构图更碎片化、产生更多不必要的细节,成像的锐利和清晰程度下降

3. V5 相比 V4 的优越之处

V5 在镜头语言的控制上,相比 V4 实现了非常明显的进步,光影的渲染也更写实、自然。AI 感已经变得很弱了,几乎肉眼难辨。
POV 第一视角

Image

V4在上,V5在下

Imagepov shot of 3 cats watching you
俯拍镜头 overhead shot

Image

V4在上,V5在下

Image

overhead shot of 3 cats watching you
低角度镜头
v5有一张做到了非常标准。V4 基本是胡来

ImageV4在上,V5在下Image

super low angle shot of 3 cats watching you
高角度镜头

ImageV4在上,V5在下Imagesuper high angle shot of 3 cats watching you

浅景深,V5 比 V4 自然得多

Image

V4在上,V5在下

Image
Shallow Focus Shot of 3 cats watching you
深景深

Image

V4在上,V5在下

Image

deep Focus Shot of 3 cats watching you

V4 在生成 bird eye view 的同时还生成了 bird 和 eye Image

Image

V4在上,V5在下

Image

eye bird view shot of a white sand beach, ocean wave foam

全身像。大部分时候,用 V4 生成 full body 都不是真正的全身像(没有脚部或者膝盖以下),V5 里表现要好很多

Image

V4在上,V5在下

Imagefull body portrait of a Zombie Bride

半身像。V4 一个被诟病的问题是,每组 4v1 生成的结构,构图都过于接近,而且人像太容易出现中心对称构图。V5 应该是增加了每批次 4 个种子的随机变量,每批结果的构图会更多样。下面的对比可以观察到这个结果。

Image

V4在上,V5在下

Imagehalf body portrait of a Zombie Bride

侧面像 + knee shot

Image

V4在上,V5在下

Imageside view portrait of a Zombie Bride, knee shot

广角,场景和构图更多样

Image

V4在上,V5在下

Image

a cowboy riding a running horse, full body horse, wide shot


4. V5 相比 V4 的倒退之处
1. 虽然摄影类风格的生成更写实和自然,但比较下面放大的僵尸新娘和牛仔骑马场景,可以发现 v5 的生成都像打了柔光,都笼罩上了一层影楼滤镜或电影滤镜。相比 V4,虽然 AI 感降低了,但也一定程度上牺牲掉了成像的细节,清晰和锐利程度都明显下降。

Image

V4在上,V5在下

Image

ImageV4在上,V5在下Image

2. V5 倾向于照片化一切生成结果,而且有一种 “糖水感”。

Image

V4在上,V5在下

Image

a stunning futuristic cabin, floating on the sea level, the tumultuous sea, masterpiece, inspired by lawren Harris

Image

V4在上,V5在下

Image
by Tony Cragg , character, ink art, side view 

https://lib.kalos.art/artist/63ec1157-6cf8-4f04-871e-2ed96225db1e?model=1
3. 下面两组都是艺术媒介测试,铅笔素描和版画风格的弗兰肯斯坦,V5 会过度添加细节,也基本丢失了艺术媒介的特征。所以想用 MJ 生成 fine-art 类作品的 (除了水彩),还是退回 V4 版本吧

Image

V4在上,V5在下

Image

pencil drawimatchng of portrait of Frankenstein, artistic, detailed

Image

V4在上,V5在下

Image
fine-art woodcut pringmaking of portrait of Frankenstein, artistic, masterpiece, detailed

4. V5 生成构图更碎片化,同时也有明显的锐度丢失的倾向

Image

V4在上,V5在下

by Tomek Setowski , city landscape
https://lib.kalos.art/artist/0ce5871e-630d-4bdb-bab6-ec20357a3937?model=1

Image

V4在上,V5在下

M C Escher style Stairway to Hell 

https://lib.kalos.art/artist/f0d8bff7-db5c-4a53-a3d4-b9a201306005?model=1
5. V5 会倾向于生成过多不必要的细节,对画面主题的美感和结构都有很负面的影响

Image

V4在上,V5在下

Image


Cat Goddess by H.R. Giger, half body, super-detailed, white and Pearlescent :: 2, melting vintage gold Fragments :: 1 
https://lib.kalos.art/artist/0363b74e-fd51-44c3-bb02-5bedd13decc3?model=1
再次生成时,我去掉了 Prompt 里的 “super-detailed”,情况并没有得到改善。
ImageCat Goddess by H.R. Giger, half body, white and Pearlescent :: 2, melting vintage gold Fragments :: 1

Image

V4在上,V5在下

Image

highly detailed beautiful organic molding, white smooth shiny polished marble, art nouveau, sharp focus, dynamic lighting, elegant harmony, beauty, masterpiece, only oni mask, cyberpunk,

Image

V4在上,V5在下

Image


Demon’s Crown, by Camille Claudel 
https://lib.kalos.art/artist/f669d8ff-54d0-4160-b1c6-24e0a27c55f5?model=1
平面插画类的生成,也出现过于繁复的笔触和构图

ImageV4在上,V5在下

by Alphonso Mucha , stunning natural landscape, church https://lib.kalos.art/artist/e0cab3c5-5b91-4834-a751-0e27d6179383?model=1

Image

V4在上,V5在下

Image


feline animal painting by Utagawa Kuniyoshi 
https://lib.kalos.art/artist/67afd69d-11f5-408e-bd48-b7f2535d52eb?model=1

Image

V4在上,V5在下

Image

by Amanda Sage 
https://lib.kalos.art/artist/501c2b6a-9284-4104-b71b-9dac30b43ac2?model=1
以上对比评测都是用同样 Prompt 在两个版本里首次生成的结果,尽量避免了人为的 cherry picking。如果结果太意外,我会多次生成以确认。
个人评价意见仅供参考。在生成不同主题和风格的作品时,该选择 V4 还是 V5?希望这个对比评测能对你有所帮助。
ImagePraying Hands, young lady’s hands, shining smooth skin, realistic , close-up view, clear background, Minimalism, artistic, atmospheric, masterpiece, sharp focus, hyper-detailed, 500px
BTW, 传说中 V5 解决了的手指问题,好像并没有哦~~ ImageImageImage

我刚刚发布了 AIGC 艺术家样式库 lib.KALOS.art 。4人小团队前后忙了4周。

- 目前全球规模最大,1300+艺术家共3万余张 4v1 样式图片,
- 覆盖三个主流图像生成模型
- 为每个艺术家都生成了8~11种常见主题,如 人像、风景、科幻、街景、动物、花卉等主题

Image

Image

Image

Image

艺术家和多种主题的结合,会带来很多意想不到的结果

后现代舞台设计师去画废土科幻场景?or 立体主义雕塑家去画一张猫咪?

按人类惯有思维,用肖像画家去生成肖像,用风景画家去生成风景,其实限制了AI模型的创作力和可能性。希望 lib.kalos.art能帮你发掘AIGC的潜力,得到更多创作灵感

Image

点击阅读原文,访问最新最全的 AIGC 艺术样式数据库