# B｜国外链路：YouTube Shorts / TikTok / Instagram Reels 上的「东方仙宫 + 珠光宝气古风美人」AI 视频实操调研

> **调研时间点：2026-09-12**（CST）。本文所有价格、模型版本均以该时点公开信息为准，AI 视频领域版本迭代以「周」计，**使用前请复核原始链接**。
> **数据性质**：工具能力/价格来自厂商页面或第三方实测；账号案例来自媒体报道与创作者自述，凡属推测的部分均标注「（未证实）」。
> **一句话前提校正**：截至 2026-09，**Sora 2 已经停服**（App/网页 2026-04-26 关停，API 2026-09-24 下线），**Veo 4 至今未发布**（2026 Google I/O 只发了 Gemini Omni）。任何还在按「Sora 2 + Veo 3.1 双主力」写方案的资料都已过期。

---

## 0. 国外链路现状速览（先看这张表）

| 环节 | 2026-09 的国外实际主力 | 备注 |
|---|---|---|
| 出图 | Midjourney V8.1/V8.2（alpha）、FLUX 3 / FLUX.1 Krea、Ideogram 4.0、Nano Banana Pro、Seedream 5.0、SDXL 生态（Pony / Illustrious） | 古风仙侠的**审美天花板仍在 MJ**；批量可控性在 FLUX/SDXL |
| 出视频 | Veo 3.1、Kling 3.0 / 3.0 Omni、Runway Gen-4.5、Seedance 2.5、Hailuo 3（MiniMax H3）、Grok Imagine Video 1.5、Higgsfield（聚合层） | 详见第 2 节「各自擅长什么」 |
| 一致性 | Kling Elements、Runway Gen-4 References、Veo Flow Ingredients、Seedance 多参考、Higgsfield Soul ID | 详见第 3 节 |
| 音频 | 模型原生音频（Kling 3.0 Omni / Veo 3.1 / Hailuo 3）、ElevenLabs（人声/旁白）、Suno v6（国风配乐） | Kling 3.0 与 Seedance 已把「对白+音效+配乐」并进一次生成 |
| 后期 | Topaz Video AI（Starlight/Astra）、DaVinci Resolve | Topaz 2025-10 起改为订阅制 |
| 分发 | 竖屏 9:16、8–15s、前 1–2 秒定生死 | Shorts 主信号是 **swipe-away rate** |

**核心判断（给制作方的三条）**

1. **仙宫题材不吃「角色一致性」，吃「空间一致性」。** 爆款（如「中式天庭」）是纯空镜 + 大景小人，没有正脸、没有台词，因此绕开了 AI 最弱的两项（人脸/口型），只压最擅长的两项（建筑/云雾/体积光）。这是该题材能低成本跑通的根本原因。
2. **古风美人题材恰好相反**，必须有稳定的脸与服装，必须走参考图/训练身份路线，成本大约是仙宫空镜的 3–5 倍。
3. **两个题材的共同最优解是「静态图 → 图生视频」**，而不是文生视频。先出 20–50 张图选 1 张，再让视频模型做微动（云飘、衣摆、鹤飞、光移），比 T2V 一次成片成功率高一个量级。

---

## 1. 图像侧：古风 / 仙侠风格到底怎么调

### 1.1 各引擎的定位与调法

**Midjourney V8.1 / V8.2（alpha.midjourney.com）**
- V8.1（2026-04-14）官方更新要点：**HD 模式默认开启、速度快 3 倍便宜 3 倍**；标准分辨率比 V7 草稿模式还快；**moodboard 与 `--sref` 现在「超级稳定」**；恢复了 image prompt 与 image weight。V8.2 于 2026-07-24 发布，2026-08-27 推出 V8 Edit 模型。
- 古风调参四件套：
  - `--sref <图片URL或风格码>` 锁整体画风（云雾质感、金色饱和度、笔触）；
  - `--sw`（style weight，0–1000）控制风格强度，古风场景建议 **`--sw 150–350`**，太高会把建筑细节糊成插画；
  - `--oref <图片URL>`（Omni Reference，V7 起）+ **`--ow`（omni weight，默认 100，0–1000）** 锁人物/器物。**古风美人正式图建议 `--ow 400–700`**，低于 200 脸会漂，高于 800 会僵化、姿态锁死。
  - **Moodboard**（个人风格板）是国外创作者做「账号统一画风」的关键：把 10–20 张自家仙宫图建板，后续所有出图保持同一套审美。这是「账号辨识度」最省力的做法。
- 竖向出图直接 `--ar 9:16`，但**仙宫大景建议先 16:9 出，再 outpaint 到 9:16**，否则广角透视容易被裁掉。

**FLUX 3 / FLUX.1 Krea dev / FLUX Kontext**
- FLUX 3（Black Forest Labs）已把能力扩展到「视频+音频同时生成」，并在官网以 `flux-3/video-edit` 形态出现在聚合平台。
- 本地/可控路线仍是 **FLUX.1 dev + 古风 LoRA**：优势是**材质词响应极准**（金属鎏金、丝绸高光、玉石次表面散射），弱项是长句审美不如 MJ。Krea 版偏「去 AI 味」，适合做「看着像实拍汉服大片」的质感。
- 提示词写法差异：FLUX 吃**名词堆叠 + 光线物理描述**，不吃 MJ 那种诗意修辞。

**Ideogram 4.0**
- Ideogram 于 2026 年发布 4.0，技术博客自述为「开放模型 / 设计前沿」（open model）。强项是**文字渲染与平面设计感**——如果你的仙宫内容要做出「卷轴式标题、篆书印章、海报式封面」，Ideogram 是国外链路里最合适的；弱项是人物写实度与电影感不如 MJ/FLUX。

**Stable Diffusion / SDXL 生态（Pony、Illustrious、NoobAI）**
- 古风仙侠在 Civitai 上仍是 LoRA 最密集的品类之一（例：Pony SDXL 的 `VoidKingdom` 风格 LoRA、LiblibAI 上的「2.5D 修仙 LoRA｜女仙」「Flux 碧影仙踪_lightGreenHanfu」等）。
- 现实取舍：**LoRA 路线适合「同一角色/同一画风批量出图」**（一次训练，长期复用），但**需要 GPU 与调参时间**。国外创作者做账号矩阵时，常见组合是「MJ 定风格 → 训练低成本 SDXL LoRA → ComfyUI 批量跑分镜」。

**NijiJourney（Niji 7）**
- Niji 7 发布时官方宣传点：眼睛高光更细腻、提示词理解提升、**`--sref` 风格迁移大幅升级**。
- 适用场景：**Q 版/2.5D 仙侠短剧、花神变装、二次元仙宫**。写实仙宫不要用 Niji，会偏插画。

### 1.2 可抄用的英文提示词模板（图像侧 8 条）

> 用法：MJ 直接把 `--ar / --sref / --oref / --ow / --style raw / --v 7` 拼在末尾；FLUX/SDXL 去掉参数只留正文。所有模板均已含**镜头 + 光照 + 材质 + 画质**四类词。

**A. 仙宫场景 4 条**

**① 九重天宫城 建立镜头（16:9，最万能）**
```
Celestial capital above a sea of clouds, nine tiers of floating palaces cascading into the distance, vermilion lacquered eaves and white jade terraces, gold filigree ridge ornaments, colossal bronze astronomical instrument as the visual anchor, low-angle 28mm wide establishing shot, deep focus, all planes tack sharp, volumetric god rays breaking through the cloud deck, warm side-backlight at dusk, transparent cyan shadows, atmospheric perspective keeping color in the far tiers, 2% human figures as scale reference only, cinematic fantasy film still, monumental Eastern production design, physically plausible light, shot on ARRI Alexa 65, 35mm, subtle film grain, no railing, no text, 8k --ar 16:9 --style raw --sref <你的仙宫风格板> --sw 250
```

**② 南天门 / 白玉长廊 竖屏沉浸（9:16 首图）**
```
9:16 vertical, immense jade stairway rising from foreground to a vermilion sky-gate that extends beyond the top of frame, single-point vertical perspective, low camera, 24mm equivalent, gate opening roughly thirty adult-heights tall, two white-robed immortals walking away from camera at 3% frame height, three-layer trailing silk hems all flowing left in the same wind, transparent shallow water mirror on unfenced white jade platform, rising morning mist, warm ivory highlights, cool blue-grey shadows, one restrained starburst lens flare at the gate edge, ethereal and quiet, cinematic fantasy film still, ultra-detailed material rendering, subsurface scattering on jade, shot on 35mm anamorphic, film grain, no people facing camera, no text --ar 9:16 --style raw
```

**③ 殿宇内部 / 藏书阁 体积光（仙宫「内景」稀缺，做差异化用）**
```
Interior of a celestial scripture hall suspended in clouds, enormous red-lacquer columns and open lattice screens, thousands of floating scrolls slowly orbiting in mid-air, one shaft of volumetric god rays entering from a high window, dust motes and incense smoke visible in the beam, polished black stone floor with faint reflections, gold filigree lotus capitals, selective high saturation — jade green and peacock blue kept mid-low, amber gold only on lit edges, wide 32mm medium shot, symmetrical composition, deep focus, cinematic chiaroscuro, physically based rendering, octane-quality material detail, shot on ARRI Alexa, 40mm, fine film grain, 8k --ar 3:2 --style raw
```

**④ 大景小人 + 仙鹤 + 瀑布（中式「意境」标准构图）**
```
Astronomical scale celestial realm, no visible horizon, a giant vermilion main hall cut off by the top of frame, white jade causeway crossing empty air toward it, a transparent waterfall spilling off a floating island into the lower cloud sea, a flock of red-crowned cranes gliding across the middle ground, one lone immortal in structured flowing robes on the right third, back turned, about 3% of frame height, 45% of the frame left as breathing air holding cloud, mist and distant pavilions, low horizon, asymmetric framing with an ancient pine as foreground frame, ultra-wide 21mm, deep depth of field, high-key jade white and pale blue-grey, restrained vermilion accents, cinematic volumetric light, hyper-detailed architecture, immersive, no railing, no text, no watermark, 8k --ar 21:9 --style raw --sw 200
```

**B. 古风美人 4 条**

**⑤ 珠光宝气皇后 特写（面部 + 首饰材质是重点）**
```
Extreme close-up beauty portrait of an ancient Chinese empress, three-quarter profile, porcelain skin with visible subsurface scattering and fine peach-fuzz, jade-and-gold filigree phoenix crown with hanging pearl tassels, gold filigree hairpins, mother-of-pearl inlay, silk collar with iridescent sheen, deep red lips, kohl-lined almond eyes looking slightly off-camera, 85mm f/1.4 portrait lens, razor-thin depth of field with creamy bokeh, warm key light from camera left with a soft silver reflector, cool rim light separating the hair from a dark background, catchlights in the eyes, shot on ARRI Alexa with vintage Cooke optics, cinematic color grade, fine 35mm film grain, hyper-detailed jewelry micro-texture, 8k --ar 9:16 --style raw --oref <角色参考图> --ow 550
```

**⑥ 汉服仙女 全身立绘（服装结构 + 比例）**
```
Full-body vertical shot of a Chinese goddess in layered hanfu, eight-head-tall elegant proportions, structured silk outer robe with bone, sheer gauze sleeves with air, restrained ornament, one dominant color scheme, wide iridescent silk sash flowing in a single consistent wind direction, gold filigree hair ornament, pearl earrings, standing on a white jade terrace above clouds, back three-quarter view then turning, 50mm lens at eye level, full-length composition with negative space above, soft overcast skylight plus a warm rim from behind, fabric physics rendered with realistic weight and drape, subsurface scattering on skin, cinematic still, fashion-editorial lighting, shot on ARRI Alexa, 50mm, subtle film grain, ultra-detailed embroidery, no extra fingers, no text --ar 9:16 --style raw
```

**⑦ 神女下凡 / 神性光（做「封面帧」最好用）**
```
Medium shot of a celestial goddess descending, ethereal halo of soft golden light behind her head, translucent layered gauze robes catching volumetric god rays, iridescent silk shifting from pearl white to pale aquamarine, gold filigree diadem and pearl drops, long black hair lifting in the updraft, calm downward gaze, arms slightly open, cloud sea and distant palace spires far below in soft bokeh, low camera looking up, 35mm, slight low-angle hero framing, dramatic backlight with lens bloom, motivated practical glow from a floating lantern, cinematic fantasy key art, hyper-real skin texture, shot on ARRI Alexa 35, 35mm, film grain, 8k --ar 2:3 --style raw
```

**⑧ 花神变装 / 回眸动态帧（用于变装类竖屏）**
```
Half-body shot of a flower goddess turning to look back over her shoulder, hair and sheer veil caught mid-motion with believable inertia, gold thread embroidery on scarlet silk, jade and pearl hairpins, petals and pollen drifting through the air in the light, shallow depth of field, 85mm, motion energy frozen at 1/250s, warm backlight through translucent fabric showing weave and translucency, cool ambient fill, cinematic and luxurious, high-end beauty campaign look, detailed skin pores and fine facial hair, shot on ARRI Alexa with Cooke S4, subtle grain, 8k --ar 9:16 --style raw --oref <同一角色> --ow 600
```

> **镜头/光照/材质/画质词库（直接往里塞）**
> - 镜头：`low-angle 28mm establishing shot`、`single-point vertical perspective`、`85mm f/1.4 portrait`、`anamorphic 2.39:1`、`deep focus`、`razor-thin DOF`
> - 光照：`volumetric god rays`、`warm side-backlight`、`transparent cyan shadows`、`motivated practical light`、`rim light`、`atmospheric perspective`
> - 材质：`gold filigree`、`iridescent silk`、`mother-of-pearl inlay`、`subsurface scattering`、`translucent gauze`、`polished jade`
> - 画质：`shot on ARRI Alexa`、`35mm film grain`、`physically based rendering`、`cinematic color grade`、`8k`

---

## 2. 视频侧：各模型擅长什么，仙宫题材上差异在哪

| 模型 | 单次时长 | 最强项 | 仙宫题材实际表现 |
|---|---|---|---|
| **Veo 3.1**（Google，2025-10/2026-01 起） | 4/6/8s，可续接 | 原生音频 + 提示词遵循 + 4K 级质感 + Ingredients 参考 | **云海/建筑质感最"贵"**，光很"电影"；缺点是贵、单段短 |
| **Kling 3.0 / 3.0 Omni**（快手，2026-02-04） | 3–15s，可续到 3 分钟 | **一次生成 6 镜头故事板 + 原生 4K 60fps + Omni Audio（对白/音效/配乐同帧生成）+ Elements 角色锁定** | 仙宫「巡游式多镜头」最强；多角色同场也能锁 |
| **Runway Gen-4.5** | 2–10s，原生 720p + 4K 放大 | **Gen-4 References（最多 3 张参考）+ Motion Brush 逐区域控制 + Director Mode + 时间线剪辑** | 适合「先定角色板再动」的精修流程；纯仙宫空镜性价比一般 |
| **Seedance 2.5**（字节） | **4–30s 任意整秒** | 首尾帧关键帧 + 最多 30 图/10 视频/10 音频参考 + 编辑式局部改 | 720p 上限，但**长镜头一次成**最省事；竖屏短剧首选 |
| **Hailuo 3 / MiniMax H3** | 5–15s | **4K 原生 + 便宜**（$0.06/s @768p、$0.16/s @4K）+ 32kHz 立体声音频；H3-Base 已开源 | 仙宫大景「清晰又便宜」，但该端点**不吃关键帧/参考图** |
| **Grok Imagine Video 1.5** | 1–15s | 极快、便宜（$0.08/s @480p），适合**草稿迭代** | 用于快速试运镜与构图，成片仍需换模型 |
| **Higgsfield** | 聚合层 | **Soul ID**（20+ 张图训练持久角色身份）+ Cinema Studio 运镜 + LipSync | 做「同一仙女跨几十条视频」的账号最省事的商业化方案 |
| **Pika / Luma Dream Machine** | 短 | 社交特效/风格化、易用 | 仙宫题材偏弱，更多用于「变装/转场」类效果 |
| ~~Sora 2~~ | — | **已停服** | 不要再纳入方案 |

**差异小结（仙宫）：** 建筑与云雾的「体量感」→ Veo 3.1 / Kling 3.0；长镜头一次成 → Seedance 2.5；4K 且便宜 → Hailuo 3；多镜头叙事 + 音频一体 → Kling 3.0 Omni；角色一致性 → Higgsfield Soul ID / Runway References。

### 2.1 可直接用的英文视频提示词（6 条）

**V1 — Veo 3.1：仙宫航拍推进（含环境音）**
```
Slow forward dolly-in through a sea of clouds toward a colossal vermilion sky-gate, camera rises slightly as the gate fills the frame, god rays sweep across the jade causeway, red-crowned cranes cross the frame left to right, silk banners lift in a steady wind. Ambient audio only: high-altitude wind, distant bronze bell, faint crane calls, no music.
```

**V2 — Kling 3.0：6 镜头仙宫巡游（原生音频）**
```
Multi-shot, 9:16 vertical.
Shot 1 (3s, wide, slow crane up): cloud sea at dawn, distant floating palaces resolve out of mist. Audio: wind, distant chime.
Shot 2 (2.5s, medium, dolly-in): white jade corridor, one immortal in white robes walking away. Audio: footsteps on stone, robe rustle.
Shot 3 (2s, macro CU, rack focus): gold filigree bracket detail, pearls on a tassel swinging. Audio: faint metallic shimmer.
Shot 4 (3s, wide, slow pan right): waterfall pouring off a floating island into the cloud deck. Audio: deep rushing water.
Shot 5 (2.5s, low-angle, push in): the main hall, cut off by the top of frame. Audio: low bronze bell swell.
Shot 6 (2s, wide, static): cranes gliding across the moon. Music: sparse guqin notes resolving on the final frame.
```

**V3 — Seedance 2.5：首尾帧 30s 长镜头**
```
First frame: wide shot of an empty jade terrace above clouds at dusk. Last frame: the same terrace, now lit by hundreds of floating lanterns, one goddess standing at the edge, back to camera. Camera: very slow continuous dolly-in with a slight upward tilt across 30 seconds, no cuts. Continuous motion: clouds drift right to left, lanterns rise one by one, robe hems lift in the same wind. Ambient audio: wind, water, distant bells.
```

**V4 — Hailuo 3：4K 竖屏 仙鹤掠过（性价比镜头）**
```
9:16 vertical, 10 seconds, ultra-detailed: a flock of red-crowned cranes glides past the camera in front of an enormous vermilion palace gate suspended above a cloud sea, slow lateral tracking shot, feathers backlit by warm sunset light, volumetric haze between layers, far palaces fading in atmospheric perspective. Ambience: high wind, crane calls.
```

**V5 — Runway Gen-4.5 + References：仙女回眸（角色一致性镜头）**
```
Reference: character plate provided. Medium close-up, a Chinese goddess in scarlet and gold hanfu turns her head over her shoulder toward camera, hair and sheer veil settle with believable inertia, gold filigree hairpins catch a warm rim light, background: cloud sea and distant palace spires in soft bokeh. Camera: locked-off 85mm, subtle handheld breathing. Keep face, hairline, jewelry, and robe exactly as in the reference. Light: warm backlight plus soft frontal fill.
```

**V6 — Grok Imagine 1.5：480p 草稿试拍（构图与运镜试验）**
```
Draft, 480p, 8s, 9:16. Fast low-angle push-in on a jade stairway leading to a giant sky-gate, god rays sweeping, mist rolling down the steps, two tiny robed figures ascending. Text-to-video, no dialogue, ambient wind.
```

**V7 — Higgsfield（Soul ID 锁定 + Cinema Studio 运镜）**
```
Soul ID: <仙女身份>. Shot: crane-up revealing a celestial terrace, then slow orbital move around the subject as her gauze sleeves lift in the wind; camera executes a 30-degree orbit at constant speed, no cut. Lighting: golden hour rim light, cool ambient fill. Keep identity, costume silhouette and jewelry consistent with Soul ID across all shots.
```

---

## 3. 一致性方案（国外链路的四套主流做法）

1. **参考图锁脸（Reference / Ingredients / Elements）**
   - **Kling Elements**：角色用 **2–4 张不同角度图**；3.0 Omni 还能上传 **3–8 秒视频**同时锁「脸 + 声音」。官方建议 5–10 张、含正面/3-4/侧面/不同光照/一张特写。
   - **Runway Gen-4 Image References**：单次最多 **3 张**生效，官方建议用**中性均匀打光**的素材；先出「character plate」审批，再进 Gen-4.5 出视频。
   - **Veo 3.1 Flow Ingredients / Vertex subject reference**：1–3 张参考图 → 8 秒片段；Flow 的 Characters 支持 1–2 张外观图并绑定语音。注意：**Quality 档不支持 Ingredients，只有 Lite/Fast 支持**。
   - **Seedance 2.5**：参考预算最大（**30 图 + 10 视频 + 10 音频**），适合把「身份/服装/场景/运镜/音色」拆成不同参考各司其职。
2. **角色表（character sheet）**
   标准做法：一次性出「正面全身 + 3/4 侧 + 侧面 + 特写 + 服装平铺」四联图，作为**项目固定资产**。国外短剧团队普遍把它当「选角照」用，任何镜头都从这张表出发，而不是从文字描述出发。
3. **首尾帧插值（first/last frame）**
   - Seedance 2.5、Kling（首尾帧控制）、Veo 3.1（首帧/尾帧）均支持。做仙宫最实用的组合是：**同一张图的「未加光」与「加光」两个版本做首尾帧**，让光在镜头里生长。
4. **身份训练（Soul ID 类）**
   - Higgsfield Soul ID：上传 **20+ 张**参考图训练一个持久身份，之后在平台内**所有模型**（Seedance 2.0 / Kling 3.0 / WAN 2.6 / Veo 3.1 / Gemini Omni Flash）复用，不必每次重传参考。这是目前国外「AI 古风美人账号」做规模化最省事的一条路，属于商业化的关键差异点。

**多镜头同一角色的实操纪律（来自 2026-09 的第三方对照测试）**：
- 保持「参考包固定、只改变化项」——把身份/服装/场景写在固定段，把动作/景别/机位写在变化段。
- 每条镜头只用**一个**变化点（换机位就不换服装）。
- 逐镜头评分四维：**身份 / 服装道具 / 场景连续性 / 动作连续性**；耳环、发际线、手、眼珠颜色是最常崩的细节。

---

## 4. 音画

- **原生音频（省 2–3 个工具步骤）**：Kling 3.0 的 Omni Audio 与视频**同一前向过程**生成，音素在扩散阶段对齐口型；Hailuo 3 原生 32kHz 立体声；Veo 3.1 原生同步音频。对「仙宫空镜 + 编钟/古筝/风声/鹤唳」这类环境音，**直接写进提示词的 audio block 即可**，不需要单独配音。
- **人声/旁白**：ElevenLabs。2026 年价目（早期口径）：Free 10k credits（**无商用权**）、Starter **$5/月**（有商用权 + 即时克隆）、Creator **$22/月**（专业克隆、192kbps）、Pro **$99/月**、Scale $330、Business $1,320。做付费内容**至少 Starter**。
- **国风配乐**：**Suno v6**（2026-09-11 发布）家族 v6 / v6-wild / v6-mini，是首个**用正版授权音乐训练**的版本（华纳音乐、BMG、Believe/TuneCore），支持段落级编辑与混音；**免费版可用 v6-mini，Pro $8/月，Premier $24/月**。注意：索尼、环球的版权诉讼仍在进行，授权尚未全覆盖。
- **国风音乐提示词关键词**：`guqin, guzheng, dizi bamboo flute, pipa, bianzhong bronze bells, taiko-like drums, pentatonic, guofeng, cinematic Chinese fantasy score, sparse, 60 BPM, no vocals`。
- **环境音清单（仙宫）**：`high-altitude wind`, `distant bronze bell`, `crane calls`, `waterfall`, `temple chime`, `fabric rustle`, `stone footsteps`。**「中式天庭」类爆款的共同点是：无台词、无字幕，只靠画面 + 环境音**，这恰恰规避了 AI 口型与配音的短板。

---

## 5. 后期与分发

**后期**
- **Topaz Video AI**：2025-10 起取消永久授权，改为订阅：**Personal $299/年、Pro $699/年**（Pro 增加 Starlight / Iris / Astra / Dione 等专用模型、批量/API、优先处理）。注意该价格来自竞品博客，**建议以 Topaz 官网复核**。AI 生成视频的典型处理链：**Starlight/Astra 生成式放大 → Proteus 细修 → Apollo/RIFE 补帧 → 稳定**。
- **DaVinci Resolve**：调色重点在**降饱和 + 分离色相**——AI 出图常见「全局橙金滤镜」和「灰雾吞色」，需要把白玉/云层/皮肤拉回中性，只让朱红与暖金保持高纯。
- **剪辑节奏**：竖屏 8–15s 一条，**仙宫大景单镜头 2–3s**，一条片子 4–6 个镜头；变装类用「同机位硬切」制造反差。**不要给仙宫加运镜特效**，AI 生成的慢推/慢摇已经足够。

**分发与算法偏好（2026）**
- **竖屏 9:16 是硬要求**。横屏重构图、黑边、静态中心裁剪会显著拉低完播。
- **YouTube Shorts 的首要信号是 swipe-away rate（划走率）的倒数**，**前 1–2 秒**决定初始分发；其次是**循环完成率（loop completion）**、评论率、分享率（分享权重最高）。新片先推给 **200–500 人**测试池，**前 2–4 小时**基本决定天花板。
- **封面/标题套路**（国外仙宫类常见）：
  - 标题：`Chinese Heaven` / `What Chinese people imagine heaven looks like` / `Celestial Palace — AI` / `Xianxia in 4K`；
  - 封面：**用最贵的那一帧**（大门 + 云海 + 体积光），不要在封面上放小字；
  - 钩子前 2 秒：**直接给最大景别**（巨门压顶 / 天阶无人尽头），不要开场黑屏、不要慢起、不要 of 标题卡；仙宫类做「无人称、无对白、无字幕」反而更利于跨国传播。
- **标签/话题**：`#ChineseHeaven`、`#Chinamaxxing`、`#BecomingChinese`、`#hanfu`、`#xianxia`、`#AIart`。2026 年「Chinamaxxing」在海外累计浏览量已破 **40 亿次**，是中国风内容的现成流量池。

---

## 6. 成本：订阅价格与单条视频折算（信息时间点 2026-08 ~ 2026-09）

**订阅/额度**

| 平台 | 价格（时点） | 额度说明 |
|---|---|---|
| Midjourney | Basic $10/月（年付 $8）、Standard $30（$24）、Pro $60（$48）、Mega $120（$96） | 按 **GPU 时间**计费：Basic ~3.3h 快速、Standard 15h + Relax 无限、Pro 30h + Stealth、Mega 60h；额外快速时长约 **$4/小时**（2026-04 口径） |
| Kling AI | 首月 Standard **$6.99**→续费 **$8.80**（660 credits）；Pro $25.99→$32.56（3,000）；Premier $64.99→$80.96（8,000）；Ultra $127.99→$159.99（26,000） | 消耗：720p 无音频 6 credits/s、1080p 无音频 8、**1080p + 原生音频 12**、4K 直接 **30 credits/s**。月 credits **不滚存** |
| Kling API | 1 unit=$0.14；Kling 3.0：720p $0.084/s、1080p $0.112/s、**1080p+音频 $0.168/s**、4K $0.42/s | API 失败任务**不扣** unit（订阅侧失败仍扣） |
| Runway | Standard ~$15/月（625 credits）、Pro ~$35（2,250）、Max ~$95（9,500） | Gen-4.5 约 **12 credits/秒**（10s 1080p ≈ 100–150 credits） |
| Veo 3.1（Vertex） | 720p 无音频：Lite $0.03/s、Fast $0.08/s、Standard $0.20/s；带音频 $0.05 / $0.10 / **$0.40** | Flow 免费档每天 **50 credits**（Lite 10 credits/次 → 每天最多 5 次 Lite） |
| MiniMax H3（Hailuo 3） | 768p **$0.08/s**、2K **$0.13/s**；fal 口径 480p $0.05 / 768p $0.06 / 2K $0.13 / 4K $0.16 | 参考图前 5 张免费，之后 $0.04/张 |
| Grok Imagine Video 1.5 | 480p **$0.08/s**、720p $0.14/s、1080p **$0.25/s**；图片输入 $0.01/张 | 1–15s，默认 480p |
| Higgsfield | Basic $9/月（120 credits）、Plus $49（1,000）、Ultra $129（3,000）；也可买 credit pack（80 credits $5 起，需有效订阅） | 共享模型单价：Kling 3.0 约 $1/条（720p 10s）、Seedance 2.0 约 $3/条、Veo 3.1 8s 约 $2.9 |
| ElevenLabs | Starter $5、Creator $22、Pro $99、Scale $330、Business $1,320 | Free 无商用权 |
| Suno | Pro **$8/月**、Premier **$24/月**（免费档 v6-mini） | 2026-09-11 v6 发布 |
| Topaz Video AI | Personal **$299/年**、Pro **$699/年** | 2025-10 起订阅制 |

**单条竖屏视频折算成本（15 秒、5 镜头、含 4 次重试）**
- **纯仙宫空镜路线**：MJ（Standard $30/月摊到 60 条 ≈ $0.5）+ Seedance/Kling 生成 15s ≈ 720p 无音频 Kling 消耗 90 credits ≈ **$1.0**，含重试 ×3 ≈ **$3**；加原生音频 +50% ≈ **$4.5/条**。用 Hailuo 3 @768p：15s × $0.08 ≈ **$1.2**，含重试 ≈ **$3.6**（最便宜）。
- **古风美人路线**：多一段「角色表 + 参考图锁定」，生成量翻倍，且需要更好分辨率 → 实际 **$6–$15/条**；若用 Higgsfield Soul ID + Kling 3.0 锁脸，约 **$5–$12/条**（含重试）。
- **国内创作者自述对照**：「中式天庭」作者陈爱军称，**低画质一条 50–60 元，4K 一条 100–300 元**，单条耗时 4–5 小时（视频侧只做微动，成本主要压在 MJ 出图与筛选上）。这条自述是「国外链路 + 国内价格」的最好锚点。

> ⚠️ 成本口径提醒：不同来源的 Veo 3.1「每有效秒」差异极大（$0.17 ~ $5.5/秒），差异来自**是否把重试、是否用 Vertex 计费、Resolution/音频档位**混在一起。**报价前务必用「有效秒 × 重试系数」自算，不要直接引用博客的单秒数字。**

---

## 7. 真实账号 / 案例（3 个）

### 案例 1｜「栖光」（陈爱军，成都眼镜店老板）——「中式天庭」，海外 500 万+ 播放
- **事实**：34 秒、无台词、无字幕的「中式天庭」短片，被海外网友**自发搬运**到 TikTok/YouTube，**不到 4 天破 500 万播放**，外交部发言人毛宁在海外账号转发，话题 `#ChineseHeaven` 刷屏；本人**没有海外账号**。央视新闻、川观新闻、辽宁日报均有报道。
- **画风**：白玉长廊、南天门、云海、飞檐琼宇、仙鹤；人物极小（背影/远景），高度靠建筑尺度而非人脸。
- **可疑工作流（本人自述）**：**Midjourney 出图 → 可灵 / 即梦做图生视频**；从上百张生成图里选最贴意境的一张；单条 4–5 小时，状态好时 2–3 小时；几乎日更；累计 200+ 条。
- **可复制的点**：①「景大人小」是中式意境的核心反差，也是规避 AI 人脸崩坏的工程选择；②纯视觉、无对白 = 无需翻译，跨语言传播零损耗。
- 来源：[川观新闻](https://cbgc.scol.com.cn/news/7847149)、[出海网](https://www.chwang.com/article/208781680464)、[辽宁日报](http://epaper.lnd.cn/lnrbepaper/pc/con/202608/20/content_336007.html)

### 案例 2｜@灵筠仙逸（完美世界，抖音，广州）——国风 AI 美人账号
- **事实**：35 岁 AI 人像创作者，**1.6 万粉丝、11.8 万点赞**；内容形态是**静态写真轮播短视频**（几乎不剪辑、不动画）。
- **画风**：定位「丰腴有韵、温婉藏柔」的东方美人，低饱和暖调、天光/海边/新中式庭院实景感，刻意避开白幼瘦模板与擦边。
- **可疑工作流（未证实）**：MJ 或 FLUX 系出图 + 静态多图轮播；同脸同风格靠固定提示词模板 + 参考图。
- **可复制的点**：**「审美定位」比技术更重要**——同质化的 AI 美女赛道里，靠一套差异化审美（体态、色调、场景）建立辨识度；且静态轮播的**算力成本几乎为零**，是冷启动账号验证审美的低成本做法。
- 来源：[网易号报道](https://www.163.com/dy/article/L15JDBAL0553KEE1.html)

### 案例 3｜Hanfu AI 变装（TikTok，2026-01 起全球扩散，印尼/东南亚媒体跟进教学）
- **事实**：TikTok 上「用 AI 把自己 P 成中国古装剧汉服美人」的玩法在 2026 年初病毒式传播，印尼 Kompas、Kapwing 等媒体/工具站专门出教程与「Chinamaxxing 趋势滤镜」，说明该趋势已从中文圈扩散到非中文创作者。
- **画风**：真实人脸 + AI 汉服/凤冠/珠翠合成，多为**前后对比 / 变装硬切**，时长 8–15s。
- **可疑工作流（未证实）**：MJ/FLUX 出汉服角色 → 换脸或 Image Reference → 图生视频做衣摆微动 → 卡点剪切；部分内容直接用滤镜类工具一键生成。
- **可复制的点**：这是「珠光宝气古风美人」在国外最**低门槛、最高传播效率**的形态——因为观众要看的是「变装反差」，而不是「AI 演技」。
- 来源：[Kompas 教程](https://www.kompas.com/tren/copy/2026/01/26/120000765/cara-edit-foto-ai-pakai-hanfu-ala-drama-china-yang-viral-di-tiktok)、[Kapwing Chinamaxxing trend](https://www.kapwing.com/kai/create/chinamaxxing-trend-filter)

**另一个可参考的邻近案例（非仙宫题材，但方法同构）**：日本一位创作者用 AI 做「AI 女神」账号，**7 天冲到 120 万点阅**，港媒报道其成本约 **$156**、月收入达 10 万（报道未说明币种，且**未经独立核实**）。来源：[ezone 报道](https://ezone.hk/article/20098103/)（未证实）

---

## 8. 落地建议：一条最小可行的国外链路 SOP

1. **定审美**：MJ V8.1（alpha）跑 30–50 张仙宫图，建 **moodboard**，固定 `--sref --sw`；确定「一个主色 + 一个点缀色」（例：月白 + 朱红鎏金）。
2. **建资产**：仙宫类做「5 个固定机位模板」；美人类做「角色四联表 + 2–4 张角度图」。
3. **出图**：MJ 定稿 →（可选）训练轻量 SDXL/FLUX LoRA → ComfyUI 批量出分镜。
4. **出视频**：仙宫空镜用 **Hailuo 3 / Seedance 2.5**（便宜 + 长镜头）；多镜头叙事用 **Kling 3.0 Omni**；角色一致性镜头用 **Higgsfield Soul ID / Runway Gen-4.5 + References**；草稿用 **Grok Imagine 480p**。
5. **音画**：优先把 audio block 写进视频提示词；需要旁白用 ElevenLabs（≥Starter）；配乐用 Suno v6（`guqin/guzheng/bianzhong/pentatonic`）。
6. **后期**：Topaz 放大 → DaVinci 降全局暖调、分离色相 → 9:16 裁切 → 8–15s、4–6 镜头、前 2 秒给最大景别。
7. **分发**：先发 Reels/TikTok 测 hook，再把高完播的搬运到 Shorts；标题用「提问式」+ `#ChineseHeaven` / `#Chinamaxxing`。

---

## 9. 不确定项与注意事项

- **Sora 2 已停服**：App/网页 2026-04-26 关停，API 2026-09-24 下线（OpenAI 帮助中心有专门说明页）。Sora 仅作为 ChatGPT Plus/Pro 内的功能残留，无 API、无批量。**任何仍以 Sora 2 为主力的方案需重写。**
- **Veo 4 未发布**：截至 2026-05 的 Google I/O，官方 Veo 系列仍是 **Veo 3.1**；Google 转向了 **Gemini Omni**。关于 Veo 3.1 的发布时间，不同来源分别写 2025-10-15 与「2026-01」，**存在冲突（未证实）**。
- **价格冲突**：Veo 3.1「每有效秒」在不同博客相差一个数量级（$0.17 vs $2.17 vs $4.67）；Topaz 价格来自竞品站点。**均需以厂商官网复核**。
- **账号与工作流**：「栖光」的 Midjourney + 可灵/即梦流程为本人受访自述（可信度较高）；案例 2、3 的具体工具链为画面特征推断，**均标注未证实**。
- **版权**：Suno v6 虽为正版授权训练，但索尼/环球诉讼未结；Kling/Hailuo 免费档产物通常**无商用权**；Midjourney 商用前请确认订阅层级与素材来源权利（人物肖像、音乐、商标）。

## 主要来源

- Midjourney 官方更新：[V8.1 Alpha](https://updates.midjourney.com/v8-1-alpha/)、[V8.2](https://updates.midjourney.com/version-8-2/)、[V8 Edit](https://updates.midjourney.com/edit-model-for-v8/)、[计费模型解析](https://dodopayments.com/blogs/midjourney-billing-model)
- Omni Reference / `--ow` 实操：[CSDN 实战文](https://blog.csdn.net/arduino9maker/article/details/153298773)；Niji 7：[chinaz](https://www.chinaz.com/ainews/24495.shtml)；Ideogram 4.0：[官方新闻稿](https://ideogram.ai/news/ideogram-4.0/)
- 视频模型对比：[Kling 3.0 官方能力页](https://pre.vivago.ai/models/kling-video.html)、[MiniMax H3 发布](https://llm-stats.com/blog/research/minimax-h3-launch)、[Hailuo 3 vs Seedance 2.5](https://www.trezalabs.com/blog/hailuo-3-vs-seedance-2-5-comparison)、[角色一致性 2026 对照测试](https://www.seeddance.io/blog/best-ai-video-generators-character-consistency-2026)、[Higgsfield vs Runway](https://higgsfield.ai/blog/higgsfield-vs-runway-2026)
- 价格：[Kling 定价与 API](https://www.modellix.ai/blog/kling-ai-pricing-per-month/)、[Grok Imagine 计费](https://wavespeed.ai/blog/cost-and-billing/grok-imagine-video-pricing/)、[Veo 3.1 成本](https://www.versely.studio/blog/veo-3-1-pricing-cost-breakdown-2026)、[Runway 定价](https://www.versely.studio/blog/runway-gen-4-pricing-vs-alternatives-2026)、[Sora 2 停服与历史价格](https://www.versely.studio/blog/sora-2-pricing-cost-breakdown-2026)、[ElevenLabs 定价](https://bigvu.tv/blog/elevenlabs-pricing-2026-plans-credits-commercial-rights-api-costs/)、[Suno v6 发布](https://www.pingwest.com/w/317314)、[Topaz 定价（竞品口径）](https://www.contenta-software.com/aivideoenhancer/blog/topaz-video-ai-price-2026.php)
- 分发算法：[YouTube Shorts 算法 2026](https://dev.to/kyle_clipspeedai/youtube-shorts-algorithm-in-2026-what-actually-determines-whether-your-clip-gets-pushed)
- 仙宫提示词工程参考：[xianxia-visual-director prompt-examples](https://github.com/liyue-aigc/xianxia-visual-director/blob/main/xianxia-visual-director/references/prompt-examples.md)
