正文
Maybe instead of generating video from scratch we’ll have LLMs animate scenes and then use diffusion models as more like a rendering layer
正在加载…
这段话设想用大语言模型编排场景,再用扩散模型负责渲染。
Maybe instead of generating video from scratch we’ll have LLMs animate scenes and then use diffusion models as more like a rendering layer
提出“LLMs animate scenes”的设想。
提出使用扩散模型作为“rendering layer”。
原文没有说明具体项目、产品或实施时间。
以下为辅助分析,与已报道事实分开展示;重要信息请进一步核实。
意思是:模型可能先安排视频中的场景和动作,再由另一类模型把这些安排渲染成画面。
它描述了一种将场景编排与画面生成分开的思路。
可以把它理解为先写好镜头安排,再让模型负责生成视觉效果。
正在加载事件时间线…