Hotel Lobby AI 视频生成器

用两张照片(人或宠物)做 Hotel Lobby AI 视频:橙色影棚里隔着一支吊麦对唱,固定机位全身镜头,5–15 秒。10 秒 125 积分。

  • 每条 63 积分起
  • 用你的照片生成
  • 5s / 10s / 15s
  • 480P
  • 可带音频
模型

MiniMax H3 Max

按场景选择

全身或半身、能看清脸,每张照片一个主角。请用已同意出镜的人或你的宠物,不要用明星。 人物必须成年,不得上传儿童照片。 JPG、PNG 或 WebP,每张最多 10 MB。

你想创建什么场景?

点选场景会填写下方提示词。提示词用 Image 1、Image 2 指代你的照片,修改时请保留这些名称。

提示词0/2000
更多设置 · 9:16 · 10 秒 · 音频
画幅
时长
分辨率
音频

正在检查生成与私有存储配置。

计费与交付说明

点击生成后按当前时长与分辨率的报价扣除积分;生成失败或文件未通过检查会自动退回,成功生成但不满意的镜头仍会计费。成片保存在“我的视频”。

预计花费: 125 积分
登录后生成

样片效果

10s · MP4
两个人,一支吊着的麦克风。本站生成 · MiniMax H3 Max · 10 秒
用这几张照片生成(AI 生成的虚构人物)
  • 左边左边
  • 右边右边
套用这条配方

关于 Hotel Lobby AI

Hotel Lobby AI 把两个主角(人或宠物)放进一间橙色影棚,隔着一支吊着的麦克风你一句我一句,固定机位全身镜头。Cicadas 用 MiniMax H3 Max 根据两张照片生成一条全新的竖屏视频,5、10 或 15 秒。不包含这个趋势用的歌,也不会把脸换到原始表演上。10 秒为 125 积分。

Hotel Lobby AI 真实样片

用你的照片生成10s · 768P

两个人,一支吊着的麦克风。

MiniMax H3 Max,10 秒,768P,带声音,竖屏,用两张照片生成。固定机位全身镜头:两人轮流对着麦克风,脸和穿着都与照片一致。橙色背景上方露出了带射灯的棚顶;提示词要求不要人声,模型还是加了一段听不懂词的说唱。

用这几张照片生成(AI 生成的虚构人物)
  • 左边左边
  • 右边右边
使用这个配方 ↗
用你的照片生成10s · 480P

他和他的猫,对着麦克风。

MiniMax H3 Max,10 秒,带声音,竖屏,用一张男人的照片和一张虎斑猫的照片生成。猫坐在高脚凳上,抬头张着嘴对着麦克风,他在旁边说唱;猫的花纹与照片一致。声音是嘻哈节拍加一段听不懂词的人声。

用这几张照片生成(AI 生成的虚构人物)
  • 左边左边
  • 右边右边
使用这个配方 ↗

为什么在 Cicadas 做 Hotel Lobby AI

左右两边:两张照片,人或宠物都行

第一张照片站左边,第二张站右边,一键即可交换。宠物会坐在与麦克风等高的凳子上。实测样片里,两张脸、穿着和猫的花纹都与照片一致。

Hotel Lobby AI 的画面,不含原曲

一支吊麦、橙色影棚、固定机位全身镜头。MiniMax H3 Max 会自己生成声音:两条实测样片都是嘻哈节拍加一段听不懂词的说唱人声,尽管提示词要求不要人声。发布时换成你自己的音乐。

每条 Hotel Lobby AI 视频先看价格

480P:5 秒 63 积分,10 秒 125,15 秒 188;768P 分别为 100、200、300。生成失败会退回积分。

Hotel Lobby AI 使用步骤

  1. 01

    登录后添加两张照片,左右各一张:全身或半身,能看清脸。

  2. 02

    选一个场景(两个人、你和宠物、两只宠物),保持 10 秒,然后核对报价。

  3. 03

    勾选照片使用确认并生成,在“我的视频”下载 MP4,发布时自己配音乐。

Hotel Lobby AI 提示词示例

复制后粘贴到上方的提示词框,再按你的画面修改。

  • “Vertical 9:16, one locked-off wide shot in a seamless bright orange studio, both people in frame from head to shoes for the whole clip. The person from Image 1 stands on the left and the person from Image 2 stands on the right, facing each other across one silver microphone that hangs from the ceiling between them. They take turns leaning toward the microphone and rapping with relaxed hand gestures, then nod along together. Keep both faces, hair and clothes exactly as in the photos. The camera never moves or cuts. Even studio light. Sound: a laid-back hip-hop drum beat, no vocals. No text, no logos.”

  • “Vertical 9:16, one locked-off full-body shot in a seamless bright orange studio. The person from Image 1 stands on the left. On the right, the pet from Image 2 sits upright on a tall wooden stool, level with one silver microphone that hangs from the ceiling between them. The person leans toward the microphone and raps with relaxed hand gestures; then the pet leans in, opens its mouth and bobs its head as if rapping back. Keep the person and the pet exactly as in the photos. Even studio light. Sound: a laid-back hip-hop drum beat, no vocals. No text, no logos.”

  • “Vertical 9:16, one locked-off full-body shot in a seamless bright orange studio. The pet from Image 1 sits upright on a tall wooden stool on the left and the pet from Image 2 sits on a matching stool on the right, facing each other across one silver microphone that hangs from the ceiling between them. They take turns leaning toward the microphone, opening their mouths and bobbing their heads as if rapping. Keep both pets exactly as in the photos. Even studio light. Sound: a laid-back hip-hop drum beat, no vocals. No text, no logos.”

Hotel Lobby AI 常见问题

在 Cicadas 的实测记录 · 审核于 2026-10-05

2026-10-05 通过 fal.ai 的 MiniMax H3 Max 参考图生视频实测:一条 10 秒、768P 的双人视频(本站 200 积分),一条 10 秒、480P 的“人和猫”视频(125 积分),均为 9:16、带声音。供应商分别用 19 秒和 8 秒返回。脸、穿着和猫的花纹与照片一致。两条的声音都是嘻哈节拍加听不懂词的人声(两个自动听辨模型),尽管提示词要求不要人声。第一次 768P 双人实测只拍到大腿,之后改了提示词,那条没有展示。“两只宠物”场景、5 秒和 15 秒、合照未实测。

更新于 2026年10月5日 · 我们怎么测试