即梦 AI(字节跳动)AI 漫剧平台研究


1. 介绍

即梦 AI 是字节跳动旗下剪映团队推出的 AIGC 创作平台,原名"剪映 Dreamina",2024 年 5 月正式定名"即梦"。在 AI 漫剧领域,即梦的独特位置在于它是唯一同时握有模型、创作工具与分发渠道三者的平台:模型侧接入字节 Seed 团队的 Seedance 系列视频模型,工具侧提供智能画布与"AI 片场",渠道侧直连抖音与红果短剧。

1.1. 开发商与版本沿革

时间事件
2024 年 3 月开启内测
2024 年 5 月由"剪映 Dreamina"正式定名"即梦"
2024 年 7 月底 / 次月推出安卓版;上线 iOS 移动端并引入会员体系
2024 年 11 月 / 12 月发布视频生成模型 Seaweed 标准版;发布图片模型 2.1
2025 年 2 月考虑接入 DeepSeek;发布多模态视频模型 OmniHuman
2025 年 3 月强化动作模仿能力
2025 年 7 月 / 8 月入选全球百大 AI 应用(内容创作类);联合发起"未来影像计划"
2025 年 9 月通过火山引擎开放企业级 API;上线文生图 3.0/3.1、图生图 3.0、OmniHuman 1.5
2025 年 12 月启动"次元折叠"微短剧计划;网页端升级为"AI 片场";推出视频 3.5 Pro 与智能多帧 2.0
2026 年 2 月接入 Seedance 2.0(图像/视频/音频/文本四模态混合输入、15 秒视频、音画同步、多镜头叙事),同步上线 Seedream 5.0 Lite;起限制真人素材使用,引入数字人分身认证机制
2026 年 4 月 9 日推出首个协作型 AI 叙事创作工具"小章鱼"Octo
2026 年 4 月 28 日因未有效落实人工智能生成合成内容标识规定要求,即梦 AI 网站被网信部门依法查处
2026 年 8 月 5 日接入 Seedance 2.5 模型

1.2. 定位

即梦的产品定位是"面向大众和专业创作者的一站式 AI 内容创作平台",在 AI 漫剧链路中覆盖从剧本意象、角色设定、分镜画面到成片生成的全段,但不覆盖投放与分账(由抖音与红果承接)。

与同组其他平台相比,即梦的定位有两个显著特征:

  • 渠道内置:抖音内置 AI 玩法模版、AI 形象与合拍,另有小云雀 AI 短剧 Agent 作为独立的短剧生产入口。生成内容可直接进入抖音分发体系,这是创业公司平台无法复制的优势。
  • 大众化优先于可编程性:即梦主推网页端与移动端交互,开发者能力通过火山引擎方舟 API 外溢,而非在即梦自身提供 CLI / 工作流编排。这决定了它更适合"人操作"而非"Agent 操作"。

1.3. 定价

档位月费月度积分备注
基础会员69 元/月725 积分/月二手研报口径
标准会员199 元/月2,210 积分/月二手研报口径
高级会员499 元/月6,160 积分/月二手研报口径
平台基础会员(另一口径)79 元/月1,200 积分/月与上方口径并存
灵活额度0.1 元/积分按量补购
连续包年首年 949 元,次年 1899 元会员续费政策调整

注意:以上数字来自中信证券研究报告转引,未能从即梦官网直接核验,属 口径。另有二手来源称即梦高级会员 499 元/月约合 0.41 元/秒,并报道"4 月连续涨价,15 秒视频成本从 5 元涨至 40 元(C 端)",同为 。本组平台定价 2026 年内已多次调整,引用时须标注"截至 2026 年 9 月"。

1.4. 开放形态

形态说明
网页端 / 移动端 App主交互入口;网页端自 2025 年 12 月起升级为"AI 片场"项目级工作区
抖音内置玩法AI 玩法模版、AI 形象与合拍
小云雀 AI 短剧 Agent面向短剧的独立 Agent 入口
火山引擎企业级 API2025 年 9 月开放;Seedance 2.0 系列 API 于 2026 年 4 月 14 日在火山方舟上线
大客户专属方案博纳影业、蓝色光标等享专属推理算力与定制化知识库

2. 名词解释

术语英文 / 缩写释义
AI 漫剧AI Comic Drama介于静态漫画与真人短剧之间的内容形态,以漫画分镜加动态视听语言构成
动态漫Motion Comic以静态漫画素材为基础,通过运镜、缩放、局部动效与配音形成的轻微动态视频形态
分镜 / 分镜脚本Storyboard将文字剧本转化为画面草图,标注每个镜头的构图、动作、时长
角色一致性Character Consistency同一角色在跨镜头、跨集、跨次生成中保持五官、服装、体型、气质稳定的能力
关键帧Keyframe定义动画或运镜变化关键状态的帧(起点与终点),对应二维动画中的"原画"
中间帧 / 过渡帧In-between / Tween关键帧之间通过插值算法自动生成的过渡帧
首尾帧First-Last Frame上传首帧与尾帧,由模型补全中间运动轨迹的图生视频控制法
图生视频Image-to-Video(I2V)输入一张静态图片,由模型生成数秒动画
口型同步 / 唇形同步Lip Sync把音频叠加到生成角色上并驱动嘴部动作匹配发音
镜头语言Camera Language通过景别、角度、运动、构图与剪辑节奏传递叙事信息的视听表达体系
智能多帧Smart Multi-Frame即梦的图像/视频多帧参考能力,2025 年 12 月升级至 2.0,用于多图条件输入
AI 片场AI Studio即梦网页端于 2025 年 12 月升级后的项目级工作区形态,承载微短剧计划"次元折叠"
小章鱼 OctoOcto即梦 2026 年 4 月 9 日推出的协作型 AI 叙事创作工具
数字分身 / 真人校验Digital Avatar Verification用户需录制本人形象与声音完成校验后才能制作本人 AI 形象出镜
显式标识 / 隐式标识Explicit / Implicit LabelAI 生成合成内容的两类法定标识:显式为用户可感知提示;隐式嵌入文件元数据
AIGC 元数据字段AIGC Metadata Field强制性国标 GB 45438—2025 规定的元数据隐式标识字段,含 Label / ContentProducer / ProduceID / ContentPropagator / PropagateID
智能画布Smart Canvas即梦的图像编辑工作台,支持多图融合、局部重绘、一键扩图、图像消除、抠图与多图层编辑
类型系数Type Coefficient抖音/红果平台按漫剧品类设定的分账系数,直接决定单部作品分账天花板
SeaweedSeaweed即梦 2024 年 11 月发布的视频生成模型标准版
OmniHumanOmniHuman即梦 2025 年 2 月发布的多模态视频模型,2025 年 9 月升级至 1.5

3. 功能说明

3.1. 视频生成

能力说明来源
文生视频 / 图生视频核心生成方式即梦官网
首帧 + 尾帧控制指定运动起点与终点,模型补全中间轨迹即梦官网
四模态混合输入图像、视频、音频、文本混合输入(最多 9 图 + 3 视频 + 3 音频 + 文本)百度百科《Seedance 2.0》
单次时长15 秒百度百科《即梦AI》
分辨率480P / 720P / 1080P / 4K(API 侧)火山引擎方舟控制台
帧率24fps火山引擎方舟控制台
音画同步与 Seedance 2.0 同步上线的原生能力百度百科《即梦AI》
多镜头叙事单次指令内自动规划并切换多个连贯镜头火山引擎方舟控制台

2026 年 8 月 5 日接入 Seedance 2.5 后,即梦的视频生成能力在时长、分辨率与一致性维度上进一步对齐字节 Seed 团队的最新模型,但具体参数提升幅度未见官方量化披露,标 [待填写]

3.2. 图像与智能画布

智能画布是即梦在 AI 漫剧链路中最具工程价值的模块,它把"生成后修图"从外部流程内化为平台能力:

功能在漫剧生产中的作用
多图融合把多个角色或角色与场景合成到同一画面
局部重绘修正角色局部(发饰、服装、手部)而不重生成整帧
一键扩图把特写构图扩展为全景,保持画风连续
图像消除去除穿帮元素或多余人物
抠图生成角色透明图层,供后续合成复用
多图层编辑以图层思路管理角色、背景、前景,是天然的资产组织方式

文生图与图生图模型于 2025 年 9 月升级至 3.0/3.1,图片模型 2.1 于 2024 年 12 月发布。

3.3. 叙事与协作工具

  • AI 片场:2025 年 12 月网页端升级形态,配合"次元折叠"微短剧计划,把零散的单次生成组织为项目级工作区。
  • 智能多帧 2.0:与 AI 片场同期推出,强化多图条件输入。
  • 小章鱼 Octo(2026 年 4 月 9 日):即梦首个协作型 AI 叙事创作工具,把"叙事"从单次生成提升为可协作的对象。
  • 小云雀 AI 短剧 Agent:面向短剧的独立 Agent 入口,承担剧本到分镜的自动化。
  • OmniHuman / 动作模仿:2025 年 2 月发布 OmniHuman 多模态视频模型,3 月强化动作模仿,9 月升级至 1.5,用于角色表演与口播类内容。

4. 平台架构

图 4-1|即梦 AI 四层平台架构:从分发与生态到治理

即梦 AI 四层平台架构(大厂垂直一体化) 信息截止 2026-09 · 示意:基于本文分析绘制 分发与生态层(本图重点) 抖音 内置 AI 玩法/合拍 红果短剧 短剧分发与分账 小云雀 短剧 Agent 入口 火山方舟 API 企业级模型调用 大客户方案 专属算力/知识库 承接创作产出 工具与工作区层 AI 片场 项目级工作区 智能画布 图像编辑与修图 智能多帧 2.0 多图条件输入 小章鱼 Octo 协作式叙事创作 调用模型能力 模型层 Seedance 2.5/2.0 视频生成 Seedream 5.0 Lite 图像生成 OmniHuman 多模态视频 视频 3.5 Pro 视频生成 图片 3.0/3.1 图像生成 Seaweed 早期视频模型 受合规约束 治理层 数字人分身认证 真人校验后方可出镜 真人素材限制 禁用真人人脸作参考 AI 生成合成内容标识 显式标识 + 元数据隐式 结构解读:即梦护城河 = 唯一内置渠道(抖音/红果)× 最快模型迭代;但各层能力边界由字节内部产品划分决定,外部无法替换。

数据来源:基于本文分析绘制的示意图。

4.1. 总体架构

┌─────────────────────────────────────────────────────────────┐
│  分发与生态层  抖音 / 红果 / 小云雀 / 火山方舟 API / 大客户方案  │
├─────────────────────────────────────────────────────────────┤
│  工具与工作区层  AI 片场 · 智能画布 · 小章鱼 Octo · 智能多帧    │
├─────────────────────────────────────────────────────────────┤
│  模型层  Seedance 2.5 / 2.0 · Seedream 5.0 Lite · OmniHuman  │
│          视频 3.5 Pro · 图片 3.0/3.1 · Seaweed                │
├─────────────────────────────────────────────────────────────┤
│  治理层  数字人分身认证 · 真人素材限制 · AI 生成合成内容标识      │
└─────────────────────────────────────────────────────────────┘

即梦的架构是典型的"大厂垂直一体化":模型层由字节 Seed 团队与剪映团队共同供给,工具层由剪映团队自建,分发层由字节内容生态承接。这种架构的优势是链路短、迭代快;代价是每一层的能力边界由字节内部的产品划分决定,外部使用者无法替换其中任何一层。

4.2. 模型层

模型类型上线时间说明
Seaweed 标准版视频生成2024 年 11 月即梦早期自研视频模型
图片模型 2.1图像生成2024 年 12 月
OmniHuman / 1.5多模态视频2025 年 2 月 / 2025 年 9 月多模态视频模型,1.5 于 9 月上线
文生图 3.0/3.1、图生图 3.0图像生成2025 年 9 月
视频 3.5 Pro视频生成2025 年 12 月与 AI 片场同期推出
Seedance 2.0视频生成2026 年 2 月四模态混合输入、15 秒、音画同步、多镜头叙事
Seedream 5.0 Lite图像生成2026 年 2 月与 Seedance 2.0 同步上线
Seedance 2.5视频生成2026 年 8 月 5 日接入最新版本,参数提升幅度未见官方量化披露

4.3. 工具与工作区层

工具与工作区层是即梦相对于纯模型服务(如 Seedance API)的增量价值所在,也是它在 L2/L3 上的主要承载。核心构件为:AI 片场(项目工作区)、智能画布(图像编辑)、智能多帧(多图条件)、小章鱼 Octo(叙事协作)。

4.4. 分发与生态层

  • 抖音内置:AI 玩法模版、AI 形象与合拍。
  • 小云雀 AI 短剧 Agent:短剧生产独立入口。
  • 影视机构合作:华策影视、柠萌影视官宣合作探索短剧创作。
  • 联合出品:与抖音、博纳影业推出 AIGC 科幻短剧集《三星堆:未来启示录》。
  • 行业活动:与釜山国际电影节合作;联合火山引擎、上海电影发起 InnoAsia AI 电影国际峰会;与 ADC 年度设计大奖合作推出 AI 视觉设计专项奖。
  • 大客户方案:博纳影业、蓝色光标等享专属推理算力与定制化知识库。

5. Harness 设计

5.1. L1 上下文工程层

即梦的上下文工程能力体现在四模态混合输入的容量与组织方式上。Seedance 2.0 支持图像、视频、音频、文本四种模态同时输入,单次最多 9 张图、3 段视频、3 条音轨加文本提示词。对 AI 漫剧而言,这意味着一次生成可以同时锚定:角色参考图(谁)、场景参考图(在哪)、运镜参考视频(怎么拍)、音色参考音频(什么声音)、文本分镜(做什么)。

配套的图像模型 Seedream 5.0 Lite 具备联网实时检索与智能逻辑推理能力,把上下文来源从"用户上传"扩展到"实时检索"。

缺口:即梦未公开上下文压缩、优先级排序或缓存复用机制。相比海螺 AI 的 H3-Context-IR(约 100,000 token 压缩至约 4,000 token)这类显式压缩方案,即梦走的是"扩大输入容量"路线而非"压缩上下文"路线。当参考素材量超过 9 图 + 3 视频 + 3 音频上限时,使用者需要自行裁剪,且无公开的分层策略。

5.2. L2 工具与执行层

工具类型可编程性
智能画布图像编辑(多图融合、局部重绘、扩图、消除、抠图、多图层)交互为主
AI 片场项目工作区交互为主
智能多帧多图条件输入交互为主
火山方舟 API模型调用可编程(/v3/contents/generations

即梦的 L2 特征是"工具丰富但封闭"。智能画布覆盖了漫剧生产中最高频的修图需求,这是创业公司平台普遍缺失的一层;但这些工具以 GUI 交互暴露,没有对应的程序化接口,无法被外部 Agent 直接调度。开发者若要编程调用,只能绕道火山方舟 API,而该 API 只覆盖模型层,不覆盖智能画布能力。

5.3. L3 编排与控制层

即梦在 L3 上的投入是含蓄但明确的

  • 多镜头叙事:Seedance 2.0 自带规划,可在单次生成内完成全景至中景至特写的镜头切换。这是把编排能力下沉进模型的做法——使用者无需显式编排,但也无法显式干预。
  • 小章鱼 Octo:2026 年 4 月推出的协作型 AI 叙事创作工具,是即梦在"叙事级编排"上的第一次明确产品化尝试。
  • 小云雀 AI 短剧 Agent:把剧本到分镜的链路做成 Agent 化流程。

判断:即梦的编排以"模型内隐式规划 + 上层 Agent 封装"为主,缺少显式的、可编辑的中间表示(如可导出的分镜表、可编辑的节点图)。这与 PixVerse Canvas、ComfyUI 节点图形成对比。对使用者而言,好处是上手快,代价是难以做精细的版本管理与回归。

5.4. L4 记忆与状态层

这是即梦相对薄弱的一层,也是本组核心论断的直接注脚。

即梦公开的 L4 相关能力包括:

  • 角色特征稳定保持(模型侧能力,火山引擎 Seedance 2.0 模型描述中提及)。
  • 资产库/素材库(AI 片场作为项目级工作区,可承载素材组织)。
  • 智能画布的多图层编辑提供了事实上的图层级状态。

但存在三个明确缺口

  1. 资产库口径未公开。与 Vidu 的"主体库(角色/道具/场景,最多 7 张参考图)"相比,即梦未公开其资产库的容量、持久化范围与跨项目复用规则,本组标注为 [待填写]
  2. 无镜头状态机。跨镜头的剧情状态(角色伤势、人物关系、时间线推进)没有公开的持久化机制。
  3. 跨会话续写依赖人工。AI 片场是项目工作区,但不是状态机;一百集的项目在即梦上仍需要使用者自行维护剧情状态表。

对 AI 漫剧这种 L4 压力最大的场景,即梦给出了"渠道与工具"的强答案,但把"状态管理"这道题留给了使用者。

5.5. L5 评估与观测层

即梦平台侧未见公开的评估与观测产品(如轨迹追踪、回归集、可用率统计面板),本组标注为 [待填写]。目前可获得的评估信号全部来自平台外部:

信号数值性质
视频可用率行业通用指标,即梦侧未见官方数值内部指标,不公开
爆款率0.47%(全行业,非即梦专属)外部市场信号
回本率不足 1.3%(全行业)外部市场信号
万播收益5~10 元(全行业)外部市场信号

这说明即梦的 L5 是"由市场而非平台承担的评估":作品好不好,不看平台指标,看抖音播放量与分账金额。对高频连续生产而言,这是明显的工程缺口。

5.6. L6 治理与安全层

即梦是本组治理样本最特殊的平台——它既有最严格的真人形象管控,也发生过最严重的合规事故。

正向机制

  • 数字人分身认证:2026 年 2 月起限制真人素材使用,引入数字人分身认证机制。在即梦 App 与豆包 App 中,用户若希望真人形象出镜,须录制本人形象与声音完成真人校验后方可制作数字分身。
  • 真人人脸禁用:在即梦 Web 端、小云雀等平台使用 Seedance 2.0 时,系统明确提示暂不支持真人人脸作为参考素材。
  • 法规约束:受《生成式人工智能服务管理暂行办法》(2023 年 8 月生效)约束,须完成算法备案、维护核心价值观、对合成媒体加标识。

负向事件

2026 年 4 月 28 日,因未有效落实人工智能生成合成内容标识规定要求,即梦 AI 网站被网信部门依法查处。

这起事件是理解 AI 漫剧场景 L6 重要性的关键证据:连拥有完整法务与合规团队的一线大厂平台,其标识管线都可能断裂。它证明《人工智能生成合成内容标识办法》与 GB 45438—2025 的执行不是"上线时配一次"的工作,而是需要持续验证的管线能力。

按 GB 45438—2025,视频内容显式标识须含"人工智能(或 AI)"要素加"生成/合成"要素,位于起始画面(可含末尾、中间)的边或角,文字高度不低于画面最短边长度的 5%,正常播放速度下持续时间不少于 2 秒;元数据隐式标识须写入含 Label / ContentProducer / ProduceID / ContentPropagator / PropagateID 等字段的 AIGC 结构。

5.7. 六层能力矩阵

即梦的实现成熟度主要缺口
L1 上下文工程四模态混合输入(9 图 + 3 视频 + 3 音频 + 文本);Seedream 5.0 Lite 联网检索无公开压缩与优先级机制
L2 工具与执行智能画布(6 类图像编辑)、AI 片场、智能多帧、火山方舟 API工具以 GUI 暴露,无程序化接口
L3 编排与控制多镜头叙事(模型内)、小章鱼 Octo、小云雀 Agent中强无显式可编辑的中间表示
L4 记忆与状态角色特征稳定保持、AI 片场项目工作区资产库口径未公开;无镜头状态机
L5 评估与观测未见公开评估产品,依赖市场信号缺轨迹追踪、回归集、可用率面板
L6 治理与安全数字人分身认证、真人人脸禁用;2026-04-28 被查处后整改最强样本标识管线曾断裂,需持续验证

6. 实际案例

6.1. 案例一:AIGC 科幻短剧集《三星堆:未来启示录》

  • 背景:影视级 AIGC 内容需要兼顾视觉奇观与叙事连贯,是检验平台连续生产能力的典型场景。
  • 方案:即梦联合抖音、博纳影业推出 AIGC 科幻短剧集《三星堆:未来启示录》。
  • 效果:作品成功进入影视级 AIGC 序列;具体的制作周期、成本与播放数据未见公开披露,标 [待填写]

6.2. 案例二:2026 年总台春晚视觉效果生成

  • 背景:春晚级别的内容对画质、稳定性与交付时限的要求极高,不能用"反复试错"的方式生产。
  • 方案:即梦与央视合作,为 2026 年总台春晚节目生成视觉效果。
  • 效果:验证了即梦在高规格交付场景下的可用性;具体生成镜头数与制作周期未见公开披露,标 [待填写]

6.3. 案例三:机构客户的专属推理方案

  • 背景:大型影视与营销机构的诉求与 C 端用户不同——需要稳定的算力保障与私有的知识/素材沉淀。
  • 方案:博纳影业、蓝色光标等大客户享专属推理算力与定制化知识库。
  • 效果:这是即梦在 L4(定制化知识库)与 L6(专属资源隔离)上少见的企业级交付;具体 SLA 与隔离机制未见公开披露,标 [待填写]

6.4. 案例四:合规事故与整改

  • 背景:《人工智能生成合成内容标识办法》于 2025 年 9 月 1 日施行,配套强制性国标 GB 45438—2025 对标识提出量化要求。
  • 方案:2026 年 4 月 28 日,即梦 AI 网站因未有效落实人工智能生成合成内容标识规定要求,被网信部门依法查处,随后整改。
  • 效果:该事件成为本组唯一的公开处罚案例,直接影响了对各平台 L6 成熟度的评估基线——即梦以"最强治理样本"与"唯一被查处平台"双重身份出现在同一张矩阵里

7. 总结

7.1. 优势

  1. 渠道闭环唯一:模型 + 创作工具 + 抖音/红果分发三者一体,是创业者平台无法复制的结构性优势。
  2. 工具层最贴合漫剧修图需求:智能画布的多图融合、局部重绘、扩图、消除、抠图、多图层编辑,精确对应 AI 漫剧最频繁的手工补救场景。
  3. 模型迭代节奏最快:2024-11 Seaweed 至 2026-08 Seedance 2.5,两年内视频模型迭代至少四代。
  4. 真人形象管控最严:数字人分身认证 + 真人人脸禁用,是行业内最早、最明确的真人形象治理。

7.2. 局限

  1. L4 是明显短板:无公开的资产库容量与镜头状态机,连续生产的状态管理完全依赖使用者自建。
  2. 工具不可编程:智能画布等核心工具无程序化接口,无法接入自有 Agent 或 CI。
  3. L5 缺位:无公开的评估与观测产品,质量判定完全交给市场。
  4. 定价不透明:会员定价存在两套二手口径,且 2026 年内多次调整,成本核算困难。
  5. 有合规事故记录:2026 年 4 月被查处,说明标识管线可靠性需要使用者额外复核。

7.3. 适用边界

适合不适合
追求抖音/红果分发的短剧与漫剧团队需要把生成能力接入自有 Agent / CI 的工程团队
以人工操作为主、需要高频修图的创作流程需要严格成本可预测与大批量自动化的生产
单条广告片、概念片、MV100 集以上的连续剧集(L4 支撑不足)
影视/营销机构的品牌内容强合规审计要求且无法自建标识复核的场景

7.4. 选型建议

  • 选即梦的核心理由是渠道与工具,不是 Harness 完整度。若你的分发主阵地在抖音/红果,即梦的分发优势大于其 L4 短板。
  • 必须在即梦之上自建项目级状态层:分镜表(结构化 JSON)、角色卡(含参考图与文本锚点)、场景卡、剧情状态机。不要指望 AI 片场替你管理剧情状态。
  • 必须自建合规复核:在导出后校验显式标识(起始画面、边角位置、字高不低于最短边 5%、持续不少于 2 秒)与 AIGC 元数据隐式标识字段,不要默认平台已完整落实。
  • 成本核算按最保守口径:采用 499 元/月高级会员档并预留 30% 以上返工冗余;不要以"1000 元/部"的行业乐观测算做预算。

信息缺口声明

  1. 即梦 AI 官方定价页的实时数字:69/199/499 元/月 与 79 元/月 两套口径均来自二手研报转引,未能从即梦官网直接核验。
  2. 即梦资产库的具体形态:容量上限、持久化范围、跨项目复用规则、是否支持跨集引用,均未检索到公开说明。
  3. Seedance 2.5 接入即梦后的参数变化:时长、分辨率、一致性指标的具体提升幅度未见官方量化披露。
  4. 即梦平台侧的评估与观测能力:是否存在可用率统计、生成历史追踪、回归集等机制,无公开结果。
  5. 即梦是否内置符合 GB 45438—2025 的元数据隐式标识:仅有被查处这一负面证据,未见正面合规说明。
  6. 机构客户方案的技术细节:专属推理算力的隔离方式、定制化知识库的容量与检索机制、SLA 条款,均无公开结果。
  7. 小章鱼 Octo 的功能边界:作为"协作型 AI 叙事创作工具",其协作模型、导出格式与是否支持外部集成,未见详细说明。

8. 参考资料

  1. 百度百科《即梦AI》 — https://baike.baidu.com/item/%E5%8D%B3%E6%A2%A6/64611350
  2. 百度百科《Seedance 2.0》 — https://baike.baidu.com/item/Seedance%202/67291551
  3. 百度百科《AI漫剧》 — https://baike.baidu.com/item/AI%E6%BC%AB%E5%89%A7/68788906
  4. 百度百科《关键帧动画》 — https://baike.baidu.com/item/%E5%85%B3%E9%94%AE%E5%B8%A7%E5%8A%A8%E7%94%BB/10223838
  5. 即梦 AI 官网 — https://jimeng.jianying.com/
  6. 火山引擎方舟控制台 Doubao-Seedance-2.0 系列 — https://console.volcengine.com/ark/region:ark+cn-beijing/model/detail?Id=doubao-seedance-2-0
  7. 国家网信办等四部门《人工智能生成合成内容标识办法》(国信办通字〔2025〕2 号) — https://www.cac.gov.cn/2025-03/14/c_1743654684782215.htm
  8. 强制性国家标准 GB 45438—2025《网络安全技术 人工智能生成合成内容标识方法》 — https://www.tc260.org.cn/upload/2025-03-15/1742009439794081593.pdf
  9. 央视网新闻《人工智能生成合成内容标识、传播、审核等如何发展?》 — https://news.cctv.cn/2025/03/15/ARTI3wMX1ohsE7LsLT4qBVrv250315.shtml
  10. 21 财经《1 元 1 秒!字节 Seedance2.0 定价出炉》 — https://m.21jingji.com/article/20260305/herald/9065ec418e98b284842d12cab6aa23fb.html
  11. 今日头条《AI 漫剧的变现逻辑,可以总结为"一基三翼"》 — https://www.toutiao.com/article/7678962433607664163
  12. 澎湃新闻《1914 元制作 1 集?漫剧仍困在隐性成本中》 — https://www.thepaper.cn/newsDetail_forward_33993493
  13. 腾讯新闻《抖音登顶、红果凶猛:字节的流量生意与下一道考题》 — https://new.qq.com/rain/a/20260909A0EWUR00
  14. KOCPC(英文)DataEye 2026 H1 AI 短剧报告转述 — https://en.kocpc.com.tw/archives/25203
  15. 一品威客《AI 漫剧分镜设计指南》 — https://gonglue.epwk.com/322844.html
  16. AI Wiki《Doubao》 — https://aiwiki.ai/wiki/doubao

Jimeng AI (ByteDance) AI Comic Drama Platform Research

1. Introduction

Jimeng AI is an AIGC creation platform launched by ByteDance's CapCut (Jianying) team, formerly named "CapCut Dreamina" and officially renamed "Jimeng" in May 2024. In the AI comic drama space, Jimeng's distinctive position is that it is the only platform that simultaneously holds models, creation tools, and distribution channels: on the model side, it integrates the Seedance series of video models from ByteDance's Seed team; on the tool side, it offers Smart Canvas and the "AI Studio"; on the channel side, it connects directly to Douyin and Hongguo short dramas.

1.1. Developer and Version History

DateEvent
March 2024Opened closed beta
May 2024Officially renamed from "CapCut Dreamina" to "Jimeng"
End of July / following month, 2024Launched Android version; released iOS mobile app and introduced a membership system
November / December 2024Released video generation model Seaweed standard edition; released image model 2.1
February 2025Considered integrating DeepSeek; released multimodal video model OmniHuman
March 2025Strengthened motion imitation capabilities
July / August 2025Selected among the world's top 100 AI apps (content creation category); co-launched the "Future Imaging Initiative"
September 2025Opened enterprise-grade API via Volcano Engine; launched text-to-image 3.0/3.1, image-to-image 3.0, and OmniHuman 1.5
December 2025Launched the "Dimension Fold" micro-drama initiative; upgraded the web client to "AI Studio"; released Video 3.5 Pro and Smart Multi-Frame 2.0
February 2026Integrated Seedance 2.0 (four-modality mixed input of image/video/audio/text, 15-second videos, audio-visual sync, multi-shot storytelling), launched Seedream 5.0 Lite in tandem; began restricting the use of real-person material and introduced a digital avatar verification mechanism
April 9, 2026Launched the first collaborative AI narrative creation tool, "Little Octopus" Octo
April 28, 2026Jimeng AI's website was lawfully penalized by cyberspace authorities for failing to effectively implement the labeling requirements for AI-generated synthetic content
August 5, 2026Integrated the Seedance 2.5 model

1.2. Positioning

Jimeng's product positioning is "a one-stop AI content creation platform for the general public and professional creators." In the AI comic drama pipeline, it covers the full span from script imagery, character design, and storyboard frames to final video generation, but does not cover distribution or revenue sharing (handled by Douyin and Hongguo).

Compared with other platforms in the same group, Jimeng's positioning has two notable characteristics:

  • Built-in channel: Douyin includes built-in AI gameplay templates, AI avatars, and duet features, with Xiao Yun Que AI Short Drama Agent as a separate short-drama production entry point. Generated content can flow directly into the Douyin distribution system, an advantage that startup platforms cannot replicate.
  • Consumer-friendliness over programmability: Jimeng prioritizes web and mobile interaction, with developer capabilities spilling out through the Volcano Engine Ark API rather than being offered as CLI / workflow orchestration within Jimeng itself. This makes it better suited for "human operation" than "Agent operation".

1.3. Pricing

TierMonthly FeeMonthly CreditsNotes
Basic Member69 RMB/month725 credits/monthPer secondary research report
Standard Member199 RMB/month2,210 credits/monthPer secondary research report
Premium Member499 RMB/month6,160 credits/monthPer secondary research report
Platform Basic Member (alternate figure)79 RMB/month1,200 credits/monthCoexists with the figure above
Flexible credits0.1 RMB/creditPay-as-you-go top-up
Consecutive yearly plan949 RMB first year, 1,899 RMB following yearsMembership renewal policy adjustment

Note: the figures above are quoted from a CITIC Securities research report and could not be directly verified on Jimeng's official website, so they are [To be verified]. Another secondary source claims Jimeng Premium membership at 499 RMB/month equates to roughly 0.41 RMB/second, and reports "successive price increases in April, with 15-second video cost rising from 5 RMB to 40 RMB (C-end)", also. Platform pricing in this group was adjusted multiple times during 2026, so citations should note "as of September 2026".

1.4. Open Forms

FormDescription
Web client / Mobile AppPrimary interaction entry; the web client has been upgraded into an "AI Studio" project-level workspace since December 2025
Built-in Douyin gameplayAI gameplay templates, AI avatars, and duet features
Xiao Yun Que AI Short Drama AgentIndependent Agent entry point for short dramas
Volcano Engine enterprise-grade APIOpened September 2025; the Seedance 2.0 series API went live on Volcano Ark on April 14, 2026
Dedicated large-customer solutionsBona Film Group, BlueFocus, and others enjoy dedicated inference compute and customized knowledge bases

2. Glossary

TermEnglish / AbbreviationDefinition
AI 漫剧AI Comic DramaA content format between static comics and real-person short dramas, composed of comic storyboards plus dynamic audiovisual language
动态漫Motion ComicA lightly dynamic video format based on static comic material, formed through camera movement, zooming, localized motion effects, and dubbing
分镜 / 分镜脚本StoryboardConverting a written script into visual sketches, annotating each shot's composition, action, and duration
角色一致性Character ConsistencyThe ability to keep the same character's facial features, clothing, physique, and temperament stable across shots, episodes, and generations
关键帧KeyframeFrames that define key states of animation or camera movement change (start and end points), corresponding to "originals" in 2D animation
中间帧 / 过渡帧In-between / TweenTransition frames automatically generated between keyframes via interpolation algorithms
首尾帧First-Last FrameAn image-to-video control method in which the user uploads the first and last frames and the model completes the intermediate motion trajectory
图生视频Image-to-Video(I2V)Inputting a single static image for the model to generate several seconds of animation
口型同步 / 唇形同步Lip SyncOverlaying audio onto a generated character and driving mouth movements to match the pronunciation
镜头语言Camera LanguageAn audiovisual expression system that conveys narrative information through shot size, angle, movement, composition, and editing rhythm
智能多帧Smart Multi-FrameJimeng's image/video multi-frame reference capability, upgraded to 2.0 in December 2025, used for multi-image conditional input
AI 片场AI StudioThe project-level workspace form reached by Jimeng's web client after its December 2025 upgrade, hosting the "Dimension Fold" micro-drama initiative
小章鱼 OctoOctoThe collaborative AI narrative creation tool Jimeng launched on April 9, 2026
数字分身 / 真人校验Digital Avatar VerificationUsers must record their own appearance and voice to complete verification before they can create their own AI avatar for on-screen use
显式标识 / 隐式标识Explicit / Implicit LabelTwo types of statutory labeling for AI-generated synthetic content: explicit labels are user-perceivable cues; implicit labels are embedded in file metadata
AIGC 元数据字段AIGC Metadata FieldThe metadata implicit-label fields required by mandatory national standard GB 45438—2025, including Label / ContentProducer / ProduceID / ContentPropagator / PropagateID
智能画布Smart CanvasJimeng's image editing workbench, supporting multi-image fusion, localized repainting, one-click expansion, image removal, background removal, and multi-layer editing
类型系数Type CoefficientThe revenue-sharing coefficient set by Douyin/Hongguo per comic drama category, which directly determines the revenue-sharing ceiling for a single work
SeaweedSeaweedThe standard edition of Jimeng's video generation model released in November 2024
OmniHumanOmniHumanA multimodal video model Jimeng released in February 2025, upgraded to 1.5 in September 2025

3. Feature Description

3.1. Video Generation

CapabilityDescriptionSource
Text-to-video / Image-to-videoCore generation methodJimeng official website
First-frame + last-frame controlSpecify the motion start and end points; the model completes the intermediate trajectoryJimeng official website
Four-modality mixed inputMixed input of image, video, audio, and text (up to 9 images + 3 videos + 3 audio clips + text)Baidu Baike 《Seedance 2.0》
Duration per generation15 secondsBaidu Baike 《Jimeng AI》
Resolution480P / 720P / 1080P / 4K (API side)Volcano Engine Ark console
Frame rate24fpsVolcano Engine Ark console
Audio-visual syncNative capability launched in tandem with Seedance 2.0Baidu Baike 《Jimeng AI》
Multi-shot storytellingAutomatically plans and switches between multiple coherent shots within a single instructionVolcano Engine Ark console

After integrating Seedance 2.5 on August 5, 2026, Jimeng's video generation capability has further aligned with ByteDance's Seed team's latest models in terms of duration, resolution, and consistency, but the specific magnitude of parameter improvements has not been publicly quantified, marked [To be filled].

3.2. Image and Smart Canvas

Smart Canvas is the most engineering-valuable module in Jimeng's AI comic drama pipeline; it internalizes "post-generation image retouching" from an external process into an in-platform capability:

FunctionRole in comic drama production
Multi-image fusionComposes multiple characters, or characters and scenes, into the same frame
Localized repaintingCorrects parts of a character (hair ornaments, clothing, hands) without regenerating the whole frame
One-click expansionExtends a close-up composition into a wide shot while keeping the art style consistent
Image removalRemoves continuity-breaking elements or superfluous people
Background removalGenerates a transparent layer of a character for reuse in later compositing
Multi-layer editingManages characters, backgrounds, and foregrounds with a layer approach, a natural way to organize assets

Text-to-image and image-to-image models were upgraded to 3.0/3.1 in September 2025, and image model 2.1 was released in December 2024.

3.3. Narrative and Collaboration Tools

  • AI Studio: the upgraded form of the web client since December 2025, which, together with the "Dimension Fold" micro-drama initiative, organizes scattered one-off generations into project-level workspaces.
  • Smart Multi-Frame 2.0: launched at the same time as AI Studio, strengthening multi-image conditional input.
  • Little Octopus Octo (April 9, 2026): Jimeng's first collaborative AI narrative creation tool, elevating "narrative" from a one-off generation into a collaborable object.
  • Xiao Yun Que AI Short Drama Agent: an independent Agent entry point for short dramas, handling the automation from script to storyboard.
  • OmniHuman / Motion imitation: released the OmniHuman multimodal video model in February 2025, strengthened motion imitation in March, and upgraded to 1.5 in September, used for character performance and talking-head content.

4. Platform Architecture

图 4-1|即梦 AI 四层平台架构:从分发与生态到治理

即梦 AI 四层平台架构(大厂垂直一体化) 信息截止 2026-09 · 示意:基于本文分析绘制 分发与生态层(本图重点) 抖音 内置 AI 玩法/合拍 红果短剧 短剧分发与分账 小云雀 短剧 Agent 入口 火山方舟 API 企业级模型调用 大客户方案 专属算力/知识库 承接创作产出 工具与工作区层 AI 片场 项目级工作区 智能画布 图像编辑与修图 智能多帧 2.0 多图条件输入 小章鱼 Octo 协作式叙事创作 调用模型能力 模型层 Seedance 2.5/2.0 视频生成 Seedream 5.0 Lite 图像生成 OmniHuman 多模态视频 视频 3.5 Pro 视频生成 图片 3.0/3.1 图像生成 Seaweed 早期视频模型 受合规约束 治理层 数字人分身认证 真人校验后方可出镜 真人素材限制 禁用真人人脸作参考 AI 生成合成内容标识 显式标识 + 元数据隐式 结构解读:即梦护城河 = 唯一内置渠道(抖音/红果)× 最快模型迭代;但各层能力边界由字节内部产品划分决定,外部无法替换。

数据来源:基于本文分析绘制的示意图。

4.1. Overall Architecture

┌─────────────────────────────────────────────────────────────┐
│  分发与生态层  抖音 / 红果 / 小云雀 / 火山方舟 API / 大客户方案  │
├─────────────────────────────────────────────────────────────┤
│  工具与工作区层  AI 片场 · 智能画布 · 小章鱼 Octo · 智能多帧    │
├─────────────────────────────────────────────────────────────┤
│  模型层  Seedance 2.5 / 2.0 · Seedream 5.0 Lite · OmniHuman  │
│          视频 3.5 Pro · 图片 3.0/3.1 · Seaweed                │
├─────────────────────────────────────────────────────────────┤
│  治理层  数字人分身认证 · 真人素材限制 · AI 生成合成内容标识      │
└─────────────────────────────────────────────────────────────┘

Jimeng's architecture is a typical "big-tech vertical integration": the model layer is jointly supplied by ByteDance's Seed team and the CapCut team, the tool layer is built in-house by the CapCut team, and the distribution layer is handled by ByteDance's content ecosystem. The advantage of this architecture is a short pipeline and fast iteration; the cost is that each layer's capability boundaries are determined by ByteDance's internal product divisions, and external users cannot replace any single layer.

4.2. Model Layer

ModelTypeLaunch DateDescription
Seaweed Standard EditionVideo generationNovember 2024Jimeng's early in-house video model
Image Model 2.1Image generationDecember 2024
OmniHuman / 1.5Multimodal videoFebruary 2025 / September 2025Multimodal video model; 1.5 launched in September
Text-to-image 3.0/3.1, image-to-image 3.0Image generationSeptember 2025
Video 3.5 ProVideo generationDecember 2025Launched at the same time as AI Studio
Seedance 2.0Video generationFebruary 2026Four-modality mixed input, 15 seconds, audio-visual sync, multi-shot storytelling
Seedream 5.0 LiteImage generationFebruary 2026Launched in tandem with Seedance 2.0
Seedance 2.5Video generationIntegrated on August 5, 2026Latest version; parameter improvement magnitude not publicly quantified

4.3. Tool and Workspace Layer

The tool and workspace layer is where Jimeng's incremental value lies relative to pure model services (such as the Seedance API), and it is its primary carrier at L2/L3. Core components are: AI Studio (project workspace), Smart Canvas (image editing), Smart Multi-Frame (multi-image condition), and Little Octopus Octo (narrative collaboration).

4.4. Distribution and Ecosystem Layer

  • Built into Douyin: AI gameplay templates, AI avatars, and duet features.
  • Xiao Yun Que AI Short Drama Agent: independent entry point for short-drama production.
  • Film & TV institution cooperation: Huace Film & TV and Linmon Pictures announced cooperation to explore short-drama creation.
  • Co-production: launched the AIGC sci-fi short-drama series 《Sanxingdui: The Future Revelation》 with Douyin and Bona Film Group.
  • Industry activities: cooperated with the Busan International Film Festival; co-launched the InnoAsia AI Film International Summit with Volcano Engine and Shanghai Film (Group); cooperated with the ADC Annual Design Awards to launch a special award for AI visual design.
  • Large-customer solutions: Bona Film Group, BlueFocus, and others enjoy dedicated inference compute and customized knowledge bases.

5. Harness Design

5.1. L1 Context Engineering Layer

Jimeng's context engineering capability is embodied in the capacity and organization of four-modality mixed input. Seedance 2.0 supports simultaneous input of four modalities — image, video, audio, and text — with up to 9 images, 3 videos, 3 audio tracks, plus a text prompt per generation. For AI comic drama, this means a single generation can simultaneously anchor: the character reference image (who), the scene reference image (where), the camera-movement reference video (how to shoot), the voice reference audio (what sound), and the text storyboard (what to do).

The accompanying image model Seedream 5.0 Lite has real-time web retrieval and intelligent logical reasoning capabilities, expanding the source of context from "user uploads" to "real-time retrieval".

Gap: Jimeng has not disclosed mechanisms for context compression, prioritization, or cache reuse. Compared with explicit compression schemes such as Hailuo AI's H3-Context-IR (compressing roughly 100,000 tokens to about 4,000 tokens), Jimeng follows the "expand input capacity" route rather than the "compress context" route. When reference material exceeds the 9 images + 3 videos + 3 audio clips limit, users must trim it themselves, with no public tiering strategy.

5.2. L2 Tool and Execution Layer

ToolTypeProgrammability
Smart CanvasImage editing (multi-image fusion, localized repainting, expansion, removal, background removal, multi-layer)Interaction-first
AI StudioProject workspaceInteraction-first
Smart Multi-FrameMulti-image conditional inputInteraction-first
Volcano Ark APIModel invocationProgrammable (/v3/contents/generations)

Jimeng's L2 is characterized as "rich tools but closed". Smart Canvas covers the most frequent image-retouching needs in comic drama production, a layer that startup platforms generally lack; however, these tools are exposed through GUI interaction with no corresponding programmatic interface, so they cannot be directly orchestrated by external Agents. Developers who want to programmatically invoke them can only detour through the Volcano Ark API, which covers only the model layer and not Smart Canvas capabilities.

5.3. L3 Orchestration and Control Layer

Jimeng's investment in L3 is understated but clear:

  • Multi-shot storytelling: Seedance 2.0 comes with built-in planning and can complete shot transitions from wide to medium to close-up within a single generation. This embeds orchestration capability into the model — users need not explicitly orchestrate, but also cannot explicitly intervene.
  • Little Octopus Octo: the collaborative AI narrative creation tool launched in April 2026, Jimeng's first explicit productization attempt at "narrative-level orchestration".
  • Xiao Yun Que AI Short Drama Agent: turns the script-to-storyboard pipeline into an Agent-driven process.

Assessment: Jimeng's orchestration relies primarily on "implicit in-model planning + upper-level Agent encapsulation", lacking an explicit, editable intermediate representation (such as an exportable storyboard table or an editable node graph). This contrasts with PixVerse Canvas and ComfyUI node graphs. For users, the benefit is a fast learning curve; the cost is difficulty in managing fine-grained versioning and regression.

5.4. L4 Memory and State Layer

This is Jimeng's relatively weakest layer, and a direct footnote to this group's core thesis.

The L4-related capabilities Jimeng publicly discloses include:

  • Stable preservation of character features (a model-side capability, mentioned in the Volcano Engine Seedance 2.0 model description).
  • Asset library / material library (AI Studio, as a project-level workspace, can host material organization).
  • Smart Canvas's multi-layer editing provides de facto layer-level state.

But there are three explicit gaps:

  1. Portfolio scope not disclosed. Compared with Vidu's "subject library (characters/props/scenes, up to 7 reference images)", Jimeng does not disclose the capacity, persistence scope, or cross-project reuse rules of its asset library; this group marks it as [To be filled].
  2. No shot state machine. Cross-shot plot state (character injuries, character relationships, timeline progression) has no public persistence mechanism.
  3. Cross-session continuation depends on manual work. AI Studio is a project workspace, not a state machine; a 100-episode project on Jimeng still requires users to maintain the plot-state table themselves.

For AI comic drama, the scenario that puts the most pressure on L4, Jimeng provides a strong "channel and tools" answer but leaves the "state management" problem to the user.

5.5. L5 Evaluation and Observation Layer

Jimeng's platform side shows no public evaluation and observation products (such as trajectory tracking, regression sets, or availability statistics dashboards); this group marks it as [To be filled]. All currently available evaluation signals come from outside the platform:

SignalValueNature
Video availability rateIndustry-standard metric; no official figure on Jimeng's sideInternal metric, not public
Hit rate0.47% (industry-wide, not Jimeng-specific)External market signal
Cost-recovery rateBelow 1.3% (industry-wide)External market signal
Revenue per 10K plays5~10 RMB (industry-wide)External market signal

This shows that Jimeng's L5 is "evaluation borne by the market rather than the platform": whether a work is good depends not on platform metrics but on Douyin play counts and revenue-sharing amounts. For high-frequency continuous production, this is an obvious engineering gap.

5.6. L6 Governance and Security Layer

Jimeng is the most distinctive governance sample in this group — it has both the strictest real-person appearance controls and a record of the most serious compliance incident.

Positive mechanisms:

  • Digital avatar verification: since February 2026, it has restricted the use of real-person material and introduced a digital avatar verification mechanism. In the Jimeng App and Doubao App, if users want their real-person appearance on screen, they must record their own appearance and voice to complete real-person verification before creating a digital avatar.
  • Real-person faces disabled: when using Seedance 2.0 on Jimeng's Web client, Xiao Yun Que, and other platforms, the system explicitly indicates that real-person faces are not yet supported as reference material.
  • Regulatory constraints: subject to the 《Interim Measures for the Management of Generative AI Services》(effective August 2023), it must complete algorithm filing, uphold core values, and label synthetic media.

Negative event:

On April 28, 2026, Jimeng AI's website was lawfully penalized by cyberspace authorities for failing to effectively implement the labeling requirements for AI-generated synthetic content.

This incident is key evidence for understanding the importance of L6 in the AI comic drama scenario: even a leading big-tech platform with a full legal and compliance team can see its labeling pipeline break. It proves that enforcement of the 《Measures for the Labeling of AI-Generated Synthetic Content》 and GB 45438—2025 is not a one-time "configure at launch" task but a pipeline capability that needs continuous verification.

Under GB 45438—2025, the explicit label on video content must contain an "artificial intelligence (or AI)" element plus a "generated/synthetic" element, positioned on the edge or corner of the starting frame (may also include the end or middle), with a text height no less than 5% of the shortest side of the frame, and a duration of no less than 2 seconds at normal playback speed; the metadata implicit label must be written into an AIGC structure containing fields such as Label / ContentProducer / ProduceID / ContentPropagator / PropagateID.

5.7. Six-Layer Capability Matrix

LayerJimeng's implementationMaturityPrimary gap
L1 Context EngineeringFour-modality mixed input (9 images + 3 videos + 3 audio clips + text); Seedream 5.0 Lite web retrievalStrongNo public compression or prioritization mechanism
L2 Tool and ExecutionSmart Canvas (6 types of image editing), AI Studio, Smart Multi-Frame, Volcano Ark APIStrongTools exposed via GUI, no programmatic interface
L3 Orchestration and ControlMulti-shot storytelling (in-model), Little Octopus Octo, Xiao Yun Que AgentMid-strongNo explicit editable intermediate representation
L4 Memory and StateStable preservation of character features, AI Studio project workspaceMediumPortfolio scope not disclosed; no shot state machine
L5 Evaluation and ObservationNo public evaluation products; relies on market signalsMediumLacks trajectory tracking, regression sets, availability dashboard
L6 Governance and SecurityDigital avatar verification, real-person faces disabled; rectified after being penalized on 2026-04-28Strongest sampleLabeling pipeline once broke; requires continuous verification

6. Use Cases

6.1. Case 1: AIGC Sci-Fi Short-Drama Series 《Sanxingdui: The Future Revelation》

  • Background: film-grade AIGC content must balance visual spectacle with narrative coherence, a typical scenario for testing a platform's continuous production capability.
  • Solution: Jimeng, together with Douyin and Bona Film Group, launched the AIGC sci-fi short-drama series 《Sanxingdui: The Future Revelation》.
  • Result: the work successfully entered the film-grade AIGC tier; specific production duration, cost, and playback data have not been publicly disclosed, marked [To be filled].

6.2. Case 2: Visual Effects Generation for the 2026 CCTV Spring Festival Gala

  • Background: Gala-level content imposes extremely high demands on image quality, stability, and delivery deadlines, and cannot be produced through a "trial and error" approach.
  • Solution: Jimeng cooperated with CCTV to generate visual effects for the 2026 CCTV Spring Festival Gala programs.
  • Result: it validated Jimeng's usability in high-spec delivery scenarios; the specific number of generated shots and production duration have not been publicly disclosed, marked [To be filled].

6.3. Case 3: Dedicated Inference Solutions for Institutional Clients

  • Background: the demands of large film and marketing institutions differ from those of C-end users — they need stable compute guarantees and private knowledge/material accumulation.
  • Solution: major clients such as Bona Film Group and BlueFocus enjoy dedicated inference compute and customized knowledge bases.
  • Result: this is one of Jimeng's rare enterprise-grade deliveries at L4 (customized knowledge base) and L6 (dedicated resource isolation); specific SLAs and isolation mechanisms have not been publicly disclosed, marked [To be filled].

6.4. Case 4: Compliance Incident and Rectification

  • Background: the 《Measures for the Labeling of AI-Generated Synthetic Content》 took effect on September 1, 2025, and the supporting mandatory national standard GB 45438—2025 imposes quantitative labeling requirements.
  • Solution: on April 28, 2026, Jimeng AI's website was lawfully penalized by cyberspace authorities for failing to effectively implement the labeling requirements for AI-generated synthetic content, and was subsequently rectified.
  • Result: this incident became this group's only public penalty case, directly influencing the evaluation baseline for each platform's L6 maturity — Jimeng appears in the same matrix in the dual capacity of "strongest governance sample" and "only penalized platform".

7. Summary

7.1. Strengths

  1. Unique channel closed-loop: the model + creation tools + Douyin/Hongguo distribution are one integrated whole, a structural advantage that startup platforms cannot replicate.
  2. Tool layer best fits comic drama retouching needs: Smart Canvas's multi-image fusion, localized repainting, expansion, removal, background removal, and multi-layer editing map precisely to the most frequent manual-repair scenarios in AI comic drama.
  3. Fastest model iteration cadence: from Seaweed in 2024-11 to Seedance 2.5 in 2026-08, the video model iterated at least four generations in two years.
  4. Strictest real-person appearance controls: digital avatar verification + real-person faces disabled is the earliest and clearest real-person appearance governance in the industry.

7.2. Limitations

  1. L4 is an obvious shortcoming: no public asset library capacity or shot state machine; the state management for continuous production relies entirely on users building it themselves.
  2. Tools are not programmable: core tools such as Smart Canvas have no programmatic interface and cannot be integrated into one's own Agent or CI.
  3. L5 is absent: no public evaluation and observation products; quality judgment is left entirely to the market.
  4. Opaque pricing: membership pricing has two sets of secondary figures and was adjusted multiple times during 2026, making cost accounting difficult.
  5. Has a compliance-incident record: penalized in April 2026, indicating that labeling-pipeline reliability requires additional review by users.

7.3. Applicability Boundary

Suitable forNot suitable for
Short-drama and comic-drama teams pursuing Douyin/Hongguo distributionEngineering teams that need to integrate generation capabilities into their own Agent / CI
Creation workflows that are primarily manual and need high-frequency retouchingProduction requiring strict cost predictability and large-scale automation
Single advertisement films, concept films, MVsSerial dramas of more than 100 episodes (insufficient L4 support)
Brand content for film/marketing institutionsScenarios with strict compliance-audit requirements that cannot self-build label review

7.4. Selection Recommendations

  • The core reason to choose Jimeng is channel and tools, not Harness completeness. If your main distribution front is Douyin/Hongguo, Jimeng's distribution advantage outweighs its L4 shortcoming.
  • You must build a project-level state layer on top of Jimeng: storyboard tables (structured JSON), character cards (with reference images and text anchors), scene cards, and a plot state machine. Do not expect AI Studio to manage plot state for you.
  • You must build your own compliance review: after export, verify the explicit label (starting frame, corner position, text height no less than 5% of the shortest side, duration no less than 2 seconds) and the AIGC metadata implicit-label fields; do not assume the platform has fully implemented them.
  • Do cost accounting with the most conservative figures: use the 499 RMB/month Premium membership tier and reserve more than 30% rework redundancy; do not budget using the industry's optimistic "1,000 RMB per work" estimate.

Information Gap Statement

  1. Real-time figures on Jimeng AI's official pricing page: the two sets of figures — 69/199/499 RMB/month and 79 RMB/month — are both quoted from secondary research reports and could not be directly verified on Jimeng's official website.
  2. The specific form of Jimeng's asset library: capacity limit, persistence scope, cross-project reuse rules, and whether cross-episode references are supported all have no public documentation found.
  3. Parameter changes after Seedance 2.5 was integrated into Jimeng: the specific improvement magnitude for duration, resolution, and consistency metrics has not been officially quantified.
  4. Jimeng's platform-side evaluation and observation capabilities: whether mechanisms such as availability statistics, generation-history tracking, and regression sets exist has no public result.
  5. Whether Jimeng has built-in GB 45438—2025-compliant metadata implicit labels: only the negative evidence of being penalized exists, with no positive compliance statement found.
  6. Technical details of institutional-customer solutions: the isolation method of dedicated inference compute, the capacity and retrieval mechanism of customized knowledge bases, and SLA terms all have no public result.
  7. The functional boundary of Little Octopus Octo: as a "collaborative AI narrative creation tool", its collaboration model, export format, and whether external integration is supported have not been described in detail.

8. References

  1. Baidu Baike 《Jimeng AI》 — https://baike.baidu.com/item/%E5%8D%B3%E6%A2%A6/64611350
  2. Baidu Baike 《Seedance 2.0》 — https://baike.baidu.com/item/Seedance%202/67291551
  3. Baidu Baike 《AI Comic Drama》 — https://baike.baidu.com/item/AI%E6%BC%AB%E5%89%A7/68788906
  4. Baidu Baike 《Keyframe Animation》 — https://baike.baidu.com/item/%E5%85%B3%E9%94%AE%E5%B8%A7%E5%8A%A8%E7%94%BB/10223838
  5. Jimeng AI official website — https://jimeng.jianying.com/
  6. Volcano Engine Ark console, Doubao-Seedance-2.0 series — https://console.volcengine.com/ark/region:ark+cn-beijing/model/detail?Id=doubao-seedance-2-0
  7. Four departments including the Cyberspace Administration of China, 《Measures for the Labeling of AI-Generated Synthetic Content》(National Cyberspace Administration Notice No. 2 of 2025) — https://www.cac.gov.cn/2025-03/14/c_1743654684782215.htm
  8. Mandatory national standard GB 45438—2025 《Network Security Technology — Methods for Labeling AI-Generated Synthetic Content》 — https://www.tc260.org.cn/upload/2025-03-15/1742009439794081593.pdf
  9. CCTV News, 《How will labeling, dissemination, and review of AI-generated synthetic content develop?》 — https://news.cctv.cn/2025/03/15/ARTI3wMX1ohsE7LsLT4qBVrv250315.shtml
  10. 21st Century Business Herald, 《1 RMB per second! ByteDance's Seedance 2.0 pricing released》 — https://m.21jingji.com/article/20260305/herald/9065ec418e98b284842d12cab6aa23fb.html
  11. Toutiao, 《The monetization logic of AI comic dramas can be summed up as "one foundation, three wings"》 — https://www.toutiao.com/article/7678962433607664163
  12. The Paper, 《Producing 1 episode for 1,914 RMB? Comic dramas remain trapped in hidden costs》 — https://www.thepaper.cn/newsDetail_forward_33993493
  13. Tencent News, 《Douyin tops the charts, Hongguo is fierce: ByteDance's traffic business and its next exam》 — https://new.qq.com/rain/a/20260909A0EWUR00
  14. KOCPC (English) retelling of the DataEye 2026 H1 AI short-drama report — https://en.kocpc.com.tw/archives/25203
  15. EPWK, 《AI Comic Drama Storyboard Design Guide》 — https://gonglue.epwk.com/322844.html
  16. AI Wiki 《Doubao》 — https://aiwiki.ai/wiki/doubao