為什麼突然測動漫模型
之前寫過一篇六家泛用模型的動漫風格對決,測的是「通才模型畫動漫能到什麼程度」。這次反過來:專門為動漫而生的模型同場對決,再留一個泛用模型當對照組。
起因是 Hugging Face 上一個叫 Anima-2.9B 的模型。它不是官方新版,而是社群作者把 CircleStone Labs 的動漫模型 Anima(2B、28 層)加深到 40 層、再餵 170 萬張動漫圖續訓出來的加深版,最大賣點是知識新到 2026 年 7 月——對動漫模型來說,「認得多新的角色」本身就是戰力。我隨手試了幾張,風格跟平常用的 Krea2 完全是兩個世界,於是決定認真測一輪。
一次把陣容拉滿,8 個模型同場:
- Anima 全家:Base(原始版)、Aesthetic(美感微調版)、Turbo(蒸餾加速版)、2.9B(社群加深版),四版同場
- NoobAI-XL:SDXL 世代的動漫王者血脈,社群資源最多的路線
- NetaYume Lumina v4:新架構(Lumina-Image-2.0)的動漫專精代表
- Z-Image-Anime:Z-Image Turbo 的動漫微調版
- Krea2:泛用模型對照組——動漫題到底需不需要專精模型,讓它來回答
快速結論
- 整體風格我最喜歡的三個:Anima-Aesthetic、Anima-Turbo、NetaYume v4。Aesthetic 是 Anima 全家完成度最高的版本;Turbo 快很多而且質感不打折;NetaYume 是全場最「標準漂亮動漫插畫」的一位
- Anima-2.9B 的圖跟 Anima-Base 很像:加深帶來的是更新的角色知識,不是畫質躍升。它還在 preview 階段,畫質要等正式版
- NoobAI-XL 的美感不如 NetaYume:構圖和表情夠兇夠有戲,但整體美感輸給新架構的同類
- Z-Image-Anime 會偏 2.5D、3D 質感:立體感和光影很扎實,但就不是傳統平面動漫的味道,喜不喜歡看個人
- Krea2 幾乎全崩:動作題手指解剖崩壞、畫面噴滿亂字,泛用模型在動漫題的極限非常明顯——動漫題請用動漫專精模型
- 深層原因是「語言不通」:動漫專精模型吃的是一套叫 Danbooru 標籤的暗語,連漫畫的視覺語法(速度線、衝擊格)都有專屬詞彙;泛用模型聽不懂這套語言,再會畫也使不上力
測試方法
- 規格:17 組 prompt、每組固定 seed,8 個模型同 seed 對照,統一 832×1216 直式;Anima 全家需要 ComfyUI 0.33.1 以上
- 題目設計:偏重一般評測很少碰的面向——打鬥動作誇飾(巨拳透視、飛踢、能量波)、全身完整度(五指張開的手掌、朝鏡頭奔跑)、風格多樣性(黑白網點漫畫、90 年代賽璐璐、極簡平塗)、鏡頭語言(魚眼低角度、雙人構圖)
- 採樣設定:各模型都用官方建議——Anima 系 32 步(Turbo 10 步)、NoobAI 28 步、NetaYume 40 步、Z-Image-Anime 9 步、Krea2 28 步;蒸餾版一律 CFG 1
- 欄位順序:對照圖從左到右固定為 Anima-2.9B、Anima-Base、Anima-Aesthetic、Anima-Turbo、NoobAI-XL、NetaYume v4、Z-Image-Anime、Krea2
兩個要誠實揭露的地方。第一,prompt 有兩套版本:前六個模型吃標籤式寫法,Z-Image-Anime 和 Krea2 吃自然語言描述——兩套講的是同一個畫面,但用各自最擅長的說法。所以這不是「一字不差的同文對決」,而是每個模型都拿到對自己最有利的說法之後的對決;這個設計對泛用模型反而是偏袒的,它還是崩了,這才是重點。
第二,審美判斷以我本人為準,你看完文末 17 組全圖後完全可以不同意。
先搞懂:什麼是 Danbooru 標籤
這次評測會一直提到「標籤」,值得先花一分鐘講清楚,因為它就是動漫專精模型的核心秘密。
Danbooru 是一個經營了二十年的動漫圖庫網站,站上近千萬張圖都被社群用一套標準化詞彙細細標註過:髮色、瞳色、服裝、姿勢、構圖、鏡頭角度,全部有固定寫法。連漫畫特有的視覺語法都有專屬標籤——速度線叫 speed lines、強調線叫 emphasis lines、定格衝擊畫面叫 impact frame、朝鏡頭揮來的拳頭叫 incoming punch。
動漫專精模型就是拿這些「圖+標籤」配對訓練出來的,所以對它們下指令不是寫作文,而是排列標籤。你寫 cat ears, sun hat, hat ribbon 的命中率,遠高於一段優美的英文描述。這套暗語泛用模型沒學過——這就是本次對決最大的分水嶺。
打鬥誇飾題:會不會說「漫畫語」一翻兩瞪眼
先看全場最有戲的一組。題目要求:朝鏡頭揮來的巨拳、透視誇張到爆框、衝擊定格、強調線、玻璃碎片。

Prompt(節錄):
1girl, red hair, torn clothes, school uniform, clenched fist, punching, incoming punch, foreshortening, giant fist, emphasis lines, speed lines, impact frame, fierce expression, screaming, shattered glass, debris, dutch angle, action
Anima 全家四張全部命中:拳頭誇張放大、直接衝出畫框,強調線、碎片、吶喊的表情全部到位——這就是「漫畫語」的力量,impact frame 三個標籤字換來一整套漫畫分鏡語法。NoobAI 選擇了特寫詮釋,表情是全場最兇的,但誇飾幅度收斂。NetaYume 和 Z-Image-Anime 都畫出了朝鏡頭的拳頭,比例比較保守,像「認真打人」而不是「漫畫誇飾」。
然後看最右邊的 Krea2:意圖有到——拳頭前伸、速度線都有——但手指解剖崩壞,畫面還噴滿了模仿日文擬聲詞的亂字。它「知道」這該是一張什麼樣的圖,但沒有語彙把它正確畫出來。整輪 17 題看下來,動作題的 Krea2 幾乎張張如此。
手部魔王關:五指張開朝向鏡頭
手是所有生圖模型的罩門,我們直接出最難的題型:手掌朝鏡頭伸過來、五指張開、帶透視。

Prompt(節錄):
1girl, silver hair, school uniform, reaching towards viewer, outstretched arm, open hand, spread fingers, palm, foreshortening, smile, looking at viewer, full body
這組也是本文封面的出處——封面三張就是 Anima-2.9B、Anima-Turbo、NetaYume 的作品。動漫專精模型在這題的整體表現比我預期好,多數把五指的結構撐住了;表情和動感的組合則是 Turbo 和 NetaYume 最討喜,兩位也是整輪評測裡最容易給出這種「近身特寫」構圖的模型——題目要全身它們常常自作主張拉近,扣分項,但畫面反而更好看,算是個性。
風格多樣性:動漫模型的核心價值
動漫專精模型不只會畫「一種動漫」。黑白漫畫這組是很好的試金石——要求純墨線、網點、分格感:

Prompt(節錄):
monochrome, greyscale, manga, screentone, 1girl, sailor uniform, action panel, speed lines, speech bubble, ink (medium), high contrast, dramatic shadow, surprised
screentone(網點)這個標籤讓前排模型直接切換成印刷漫畫質感。反過來,極簡平塗這組走完全相反的方向——大量留白、平塗色塊、幾乎不打陰影:

Prompt(節錄):
1girl, flat color, ligne claire, pastel colors, minimalist, bob cut, yellow raincoat, umbrella, rain, reflection, puddle, simple background, wide shot
這組讓我印象最深的反而是 Anima-Base:把人物縮成畫面一角、用大片留白撐構圖,是全場最敢的詮釋。同一個模型家族能在網點漫畫和極簡平塗之間自由切換,這種風格頻寬正是動漫專精模型的價值所在。
Anima 家族內戰:2.9B 值得換嗎
回到起點的問題。同場四個版本直接對照,先看開場的海邊立繪:

Prompt(節錄):
1girl, cat girl, cat ears, cat tail, light blue hair, sun hat, hat ribbon, sleeveless dress, sailor collar, standing, hand on headwear, looking away, ocean, sunlight, caustics, floating hair, wind, from side
我的結論:2.9B 的圖跟 Base 很像。同 seed 下兩者的構圖與用色傾向高度接近,2.9B 略偏高彩度、對比強一點的曬圖風。它的加深訓練換來的是更新的角色與題材知識,不是肉眼可見的畫質躍升——考慮到它還在 preview 階段(作者還在跑更大規模的訓練),這個結果不意外。如果你想生最新番的角色,2.9B 值得裝;如果只看畫質,Aesthetic 仍然是 Anima 全家最好的版本,光影最精緻、自然語言理解也最好,而且官方設計成不需要堆品質標籤。
Turbo 則是全家的驚喜:10 步出圖,質感卻沒有明顯打折,而且它的特寫構圖經常比 Base 更有魅力。日常快速出圖我會直接用它。
無人場景題:NoobAI 的小失手
一組無人的蒸汽龐克街景,看場景理解和氛圍:

Prompt(節錄):
no humans, scenery, steampunk, cityscape, watercolor (medium), muted color, steam locomotive, iron bridge, gothic architecture, cherry blossoms, cobblestone street, atmospheric haze, from below, dutch angle
Anima 系的水彩淡彩氛圍很完整;NetaYume 和 Z-Image-Anime 的火車頭細節最扎實。NoobAI 這組把火車弄丟了——只畫了街景和櫻花。角色圖是它的主場,無人場景相對是弱項。
該說的還是要說:訓練資料的爭議
動漫專精模型的能力來自 Danbooru 圖庫,而這件事本身有爭議:站上的圖多數是未經畫師授權的轉載,2022 年起「拿它訓練 AI」在日本繪師圈引發過大規模抗議,最痛的點是畫師名標籤讓「一鍵模仿特定畫師的風格」變成可能。法律上各地未定:日本著作權法原則允許機器學習利用,美國的相關訴訟還在進行中。
我的做法是:這次評測的 prompt 全部使用原創角色描述、不放畫師名標籤(測試中途我把唯一一組帶畫師標籤的題目改掉重跑了)、不生成版權角色。這些模型的能力真實存在,爭議也真實存在,兩件事同時成立,用的時候心裡有數。
情境推薦
| 你的情況 | 推薦 |
|---|---|
| 動漫插畫日常主力 | Anima-Aesthetic(全家畫質最佳、免品質標籤)或 NetaYume v4(最標準的漂亮動漫插畫) |
| 快速出圖、迭代風格 | Anima-Turbo(10 步出圖,質感不打折) |
| 想生最新動漫角色 | Anima-2.9B(知識新到 2026 年 7 月,畫質等正式版) |
| 立體感、厚塗質感的動漫 | Z-Image-Anime(偏 2.5D 的光影詮釋) |
| 需要海量社群資源(LoRA、ControlNet) | NoobAI-XL(SDXL 生態無可取代,但直出美感輸新架構) |
| 動漫題還想用泛用模型 | 不推薦——本次 Krea2 的表現就是答案 |
完整 17 組對照圖
前面展示過 6 組,其餘 11 組全部放在這裡,欄位順序同上(Anima-2.9B → Anima-Base → Anima-Aesthetic → Anima-Turbo → NoobAI-XL → NetaYume v4 → Z-Image-Anime → Krea2),歡迎自行下判斷:











完整 17 組 Prompt(Danbooru 標籤版)
想自己重跑的話,這裡是全部 17 組的標籤版 prompt——Anima 全家、NoobAI、NetaYume 都吃這套寫法。使用時按各模型慣例加品質前綴(例如 masterpiece, best quality, absurdres;Anima-Aesthetic 官方說不需要),負面提示詞共用這組即可:
worst quality, low quality, lowres, bad anatomy, bad hands, extra digits, fewer digits, missing fingers, jpeg artifacts, watermark, signature, blurry
01 海邊貓娘
1girl, cat girl, cat ears, cat tail, long hair, light blue hair, wavy hair, dark blue eyes, sun hat, white headwear, hat ribbon, blue ribbon, sleeveless dress, long dress, sailor collar, standing, hand on headwear, looking away, blush, ocean, island, bird, sunlight, caustics, floating ribbons, floating hair, wind, from side
02 蒸汽龐克無人街景(負面另加 1girl, 1boy, people, moe)
no humans, scenery, steampunk, cityscape, watercolor (medium), flat color, muted color, limited palette, steam locomotive, black train, iron bridge, gothic architecture, industrial pipes, cherry blossoms, falling petals, cobblestone street, wrought-iron railing, stone stairs, potted plant, wooden bench, atmospheric haze, from below, dutch angle, cinematic composition
03 天使女武神
1girl, warrior, angel, valkyrie, golden armor, breastplate, shoulder armor, gauntlets, white tunic, flowing clothes, long hair, white hair, large wings, white wings, gold trim, holding polearm, spear, glowing weapon, standing, divine light, radiant aura, dark blue background, starry background, sparkle, celestial, majestic, full body
04 巨拳衝擊
1girl, red hair, short hair, torn clothes, school uniform, clenched fist, punching, incoming punch, foreshortening, oversized forearm, giant fist, close-up fist, emphasis lines, speed lines, motion lines, impact frame, fierce expression, screaming, open mouth, shattered glass, debris, dutch angle, dramatic lighting, action
05 魔法少女變身
1girl, magical girl, transformation, pink hair, twintails, glowing, light particles, sparkle, magic circle, ribbon, floating hair, floating clothes, dress, energy, spinning, outstretched arms, closed eyes, smile, full body, colorful background, lens flare, backlighting
06 空中飛踢
1girl, flying kick, jumping, midair, full body, from below, dutch angle, motion blur, speed lines, twin braids, brown hair, sportswear, sneakers, outstretched leg, determined expression, clenched teeth, rooftop, city, sunset, orange sky, action
07 雙刀對決
2girls, duel, sword clash, katana, sparks, crossed swords, face-to-face, glaring, gritted teeth, long hair, black hair, white hair, motion blur, dynamic angle, falling petals, night, moonlight, blue theme, action, full body
08 能量波發射
1boy, muscular, spiky hair, torn clothes, energy beam, firing beam, energy, aura, glowing hands, electricity, screaming, wide-eyed, crater, rubble, floating rocks, dust, emphasis lines, from front, perspective, dramatic lighting, action
09 英雄式落地
1girl, superhero landing, on one knee, fist on ground, cracked ground, dust cloud, cape, short hair, glowing eyes, looking up, from below, low angle, full body, night, city lights, backlighting, debris
10 手掌前伸
1girl, silver hair, long hair, school uniform, reaching towards viewer, outstretched arm, open hand, spread fingers, palm, foreshortening, smile, looking at viewer, full body, wind, petals, blue sky, depth of field
11 吐司少女奔跑
1girl, running towards viewer, full body, toast in mouth, food in mouth, school uniform, school bag, brown hair, ahoge, panicked, tearing up, motion blur, morning, sunlight, residential street, dynamic pose
12 90 年代賽璐璐
1girl, retro artstyle, 1990s (style), cel shading, film grain, long hair, green hair, pilot suit, cockpit, looking back, serious expression, lens flare, muted color, anime screencap
13 黑白網點漫畫
monochrome, greyscale, manga, screentone, 1girl, short hair, sailor uniform, action panel, speed lines, speech bubble, ink (medium), high contrast, dramatic shadow, wide-eyed, surprised
14 誇張情緒 chibi
1girl, chibi, comically exaggerated, anger vein, clenched teeth,
>_<, tears, trembling, clenched fists, steam from head, emphasis lines, simple background, white background, full body
15 魚眼低角度
1girl, fisheye, from below, perspective distortion, standing on edge, skyscraper, rooftop, looking down at viewer, smirk, twintails, black dress, wind, night, neon lights, city below, full body
16 極簡平塗
1girl, flat color, ligne claire, pastel colors, minimalist, bob cut, black hair, yellow raincoat, umbrella, rain, reflection, puddle, simple background, wide shot
17 背靠背戰鬥
2girls, back-to-back, battle stance, fighting stance, contrast, red theme, blue theme, serious expression, smirk, long hair, ponytail, holding weapon, sword, gauntlets, full body, wind, embers, dramatic lighting, surrounded
模型下載
| 模型 | 大小 | 下載 | 放置位置 |
|---|---|---|---|
| Anima-2.9B preview | 5.4 GB | Hugging Face | models/diffusion_models/ |
| Anima-Base / Aesthetic / Turbo | 各 3.9 GB | Hugging Face | models/diffusion_models/ |
| Qwen3-0.6B text encoder(Anima 全家共用) | 1.1 GB | 同上 repo split_files/text_encoders/ | models/text_encoders/ |
| Qwen-Image VAE(Anima 全家共用) | 0.25 GB | 同上 repo split_files/vae/ | models/vae/ |
| NoobAI-XL v1.1 | 6.6 GB | Hugging Face | models/checkpoints/ |
| NetaYume Lumina v4(含文字編碼器與 VAE 的整合檔) | 9.9 GB | Hugging Face | models/checkpoints/ |
| Z-Image-Turbo-Anime AIO | 19 GB | Civitai | models/checkpoints/ |
Anima 全家需要 ComfyUI 0.33.1 以上原生支援;已經在用 Qwen-Image 系模型的話 VAE 不用重複下載。授權注意:Anima 全家(含 2.9B)採非商用授權——生成的圖片可以商用,但模型本身不能拿去做付費服務或商業部署。
相關文章
- 動漫風格大對決:六家開源模型 30 題實測,從賽璐璐、恐怖漫畫到水墨動畫 — 本文的前篇:泛用模型畫動漫的極限測試
- 微軟 Mage-Flow 實測:4B 小模型四變體同場對決,順便挑戰 Krea2 — 同樣的同 seed 評測方法論
- Krea2 微調模型大亂鬥:官方 Turbo 對決四個社群微調,35 組同 seed 實測 — 本文對照組 Krea2 的主場表現
- KREA 2 實測:中文 Prompt 能用嗎?四種官方 LoRA 風格全比較




