Sora视频模型
Sora 是 OpenAI 开发的一个扩散模型,能够根据文本指令生成长达 60 秒的视频。其核心突破在于能够生成具有高度细节的场景、复杂的摄像机运动,以及充满活力的多角色情感。Sora 不仅理解用户的 Prompt,还能在一定程度上模拟真实世界的物理属性。
2. 官方震撼示例
Section titled “2. 官方震撼示例”以下是 Sora 发布初期最具代表性的几个演示案例:
2.1. 东京街头 (Tokyo Walk)
Section titled “2.1. 东京街头 (Tokyo Walk)”[!info] 提示词 (Prompt) 中文:一位时尚女性走在充满温暖霓虹灯和动画城市标牌的东京街道上。她穿着黑色皮夹克、红色长裙和黑色靴子,拎着黑色钱包。她戴着太阳镜,涂着红色口红。她走路自信又随意。街道潮湿且反光,在彩色灯光的照射下形成镜面效果。 English: A stylish woman walks down a Tokyo street filled with warm glowing neon and animated city signage. She wears a black leather jacket, a long red dress, and black boots, and carries a black purse. She wears sunglasses and red lipstick. She walks confidently and casually. The street is damp and reflective, creating a mirror effect of the colorful lights.
2.2. 艺术画廊 (Art Gallery)
Section titled “2.2. 艺术画廊 (Art Gallery)”[!info] 提示词 (Prompt) 中文:参观艺术画廊,里面有许多不同风格的美丽艺术品。 English: Tour of an art gallery with many beautiful works of art in different styles.
2.3. 禅宗花园球体 (Zen Garden Sphere)
Section titled “2.3. 禅宗花园球体 (Zen Garden Sphere)”[!info] 提示词 (Prompt) 中文:玻璃球的特写视图,里面有一个禅宗花园。球体中有一个小矮人正在耙禅宗花园并在沙子上创造图案。 English: A close up view of a glass sphere that has a zen garden within it. There is a small dwarf in the sphere who is raking the zen garden and creating patterns in the sand.
3. 测试资格与使用流程 (历史回顾)
Section titled “3. 测试资格与使用流程 (历史回顾)”在发布初期,Sora 仅对受邀的 Red Teamers(红队测试人员)以及部分视觉艺术家、设计师和电影制作人开放,以评估风险。
3.1. 申请与使用逻辑
Section titled “3.1. 申请与使用逻辑”- Plus 账户权益:早期倾向于在 ChatGPT Plus 用户中逐步释放测试名额。
- 场景描述:在特定界面输入极致详细的 Prompt,描述场景、镜头运动、光影细节。
- 视频生成:云端利用算力集群进行扩散生成,片刻后即可获取高质量视频。
4. 技术亮点
Section titled “4. 技术亮点”- 三维空间一致性:角色在移动或被遮挡时,能保持身份和外貌的一致性。
- 物理交互模拟:如笔划过沙子留下痕迹,或者一个人吃饼干留下咬痕。
- 超长连贯性:支持生成长达一分钟的视频,远超当时其他 AI 视频工具。