如何通过 Google Vids API 用 Gemini Omni 1.1 Flash 生成会说话的数字人视频(中文教程)

Chinese-language tutorial. This page is written in Simplified Chinese for Chinese-speaking developers. The same material in English: How to Make AI Influencer Videos with Avatars via the Google Vids API · Google Vids API docs

9 分钟阅读 • 2026 年 10 月 8 日

目录

  1. 简介
  2. 为什么选 Google Vids API
  3. Google Vids API 价格与额度
  4. 如何绑定 Google 账号
  5. 调用方式:同步、异步与 Webhook
  6. 实战:两个数字人发布一款球鞋
    1. 1. 生成 Mia 的人像
    2. 2. 生成 AeroLoop 产品图
    3. 3. 选择音色
    4. 4. 用人像创建 Mia 的数字人
    5. 5. 用文字描述创建 Leo 的数字人
    6. 6. 两个数字人和产品同框
    7. 7. 延长到 20 秒
    8. 8. 延长到 30 秒
    9. 9. 放大到 1080p
  7. 生成效果示例
  8. 常见问题
  9. 结语

简介

useapi.net 的 Google Vids API 用你自己的 Google 账号和 Google AI 套餐,通过 REST 接口调用 Google Vids 里的 Gemini Omni 1.1 Flash:生成带声音和对白的视频,延长到约 41 秒,还能创建带固定配音的数字人(avatar),让它在视频里说出你写的台词。

Google Vids 是 Google Workspace 里的视频应用,Google 在 2026 年 9 月为它加入了 Gemini Omni 1.1 Flash(Google 官方公告)。截至 2026 年 10 月,Google 没有公开 Vids 的视频生成 API,本 API 是第三方实现,生成在你自己的账号上进行,消耗的是套餐里的 Vids 月度额度,不按次付费。

本文由 useapi.net 维护者撰写,介绍的是我们自己的托管 API,示例代码开源在 GitHub。

为什么选 Google Vids API

  • 没有验证码。Google Flow 每次生成都要过 reCAPTCHA,需要接打码服务。Google Vids 没有这一步,不用打码,也没有打码费用。
  • 独立的额度。Vids 有自己的月度额度,不消耗 Google Flow 的积分。同一个 Google 账号同时绑定 Vids API 和 Google Flow API,两份额度都能用上。
  • 可以延长 Omni 视频。POST /videos/extend 每次加 3 到 10 秒,最长约 41 秒。Flow 里只有 Veo 视频能延长,Omni 不行。
  • 带配音的数字人。从 30 种 Gemini 音色里选一种,配上一张人物图片保存为数字人,之后在任何视频里用 avatarId 引用,人物用自己的声音说台词。
  • 时长 3 到 10 秒(按 1 秒递增),720p 或 1080p,横屏或竖屏,两种分辨率消耗相同。
  • 向 useapi.net 支付每月固定 15 美元,同一订阅包含 useapi.net 的所有 API。
  Google Vids API Google Flow API
额度 每月若干秒视频,与 Flow 分开 每月 Flow 积分
一条 10 秒 Omni 视频 10 秒额度,720p 或 1080p 15 积分(720p,可免费放大到 1080p)
Ultra(199 美元) 每月可生成的 10 秒 Omni 视频 约 1,000 条(10,000 秒,家庭方案共享) 约 1,666 条(25,000 积分)
Omni 时长 3 到 10 秒,按 1 秒递增 4、6、8 或 10 秒
延长 Omni 视频 ✅ 最长约 41 秒 ❌ 仅 Veo 视频
验证码 无 每次生成都要 reCAPTCHA
其他模型 — Veo 3.1、最高 4K 的 Nano Banana 图片

Google Vids API 价格与额度

官方 Gemini API 按秒计费:Omni 1.1 Flash 的 720p 视频每秒按 5,792 个 token、每百万 token 17.50 美元计算,约合每秒 0.10 美元。通过 Vids 或 Flow 生成,消耗的是 Google AI 套餐里的月度额度,下表按 Ultra $199 套餐折算单条成本(2026 年 10 月的价格):

渠道 一条 10 秒 Omni 1.1 Flash 视频 每月可生成的 10 秒视频
Google 官方 Gemini API(按量计费) 720p 约 1.01 美元 按付费量
Google Flow API,Google AI Ultra(199 美元) 15 积分,约 0.12 美元 约 1,666 条
Google Vids API,Google AI Ultra(199 美元) 10 秒额度,约 0.20 美元 约 1,000 条,家庭方案共享
同一个 Ultra 账号同时用 Flow 和 Vids 约 0.07 美元 约 2,666 条

按 Gemini API 的价格,2,666 条 10 秒视频每月约需 2,700 美元,而这里只需要一个 Ultra 套餐加 15 美元的 useapi.net 订阅。

各套餐的 Vids 额度

Google 按月计量 Vids:视频按秒,图片按张。以下为 Google 在 Gemini in Google Vids 帮助页公布的数字:

  Google AI Plus Google AI Pro Google AI Ultra 5x Google AI Ultra 20x
AI 视频 每月 6 条,与 AI 数字人合计 每月 500 秒 每月 2,500 秒 每月 10,000 秒
图片生成与编辑 每次图片操作 1 点 每月 30 张 每月 300 张 每月 1,000 张

我们在真实账号上测得(2026 年 10 月):

  • Google AI Ultra(199 美元) 就是 Google 所说的 Ultra 20x:每月 10,000 秒视频、1,000 张图片。
  • 视频秒数由 Google One 家庭方案的所有成员共享,图片按账号计算。多绑定同一家庭方案的账号,能增加图片额度和并行任务数,但不增加视频秒数。
  • 免费 Google 账号在 Vids 里是 0 秒视频、0 张图片。它仍然可以绑定,但只能创建数字人(包括根据文字描述绘制人物图片),不消耗任何额度(需用 email 指定该账号;这些数字人无法在该账号上生成视频),视频和图片任务会跳过它。
  • 一条视频消耗等于其时长的秒数,720p 和 1080p 消耗相同。延长只计新增的秒数,放大(upscale)会再计一次整条视频的秒数。
  • 被 Google 拒绝的请求(内容或额度原因)不消耗额度。
  • 额度在每月 1 日太平洋时间 00:00 重置。用完时任务以 429 失败,retryAt 为重置时间。

每个完成的任务都会在 result.quota 里返回账号剩余额度,GET /accounts/email 可以实时读取全部计数。

useapi.net 每月固定 15 美元,包含所有 API,不只是 Google Vids。每个订阅可以绑定 3 个 Google 账号,总数最多 50 个。生成在你自己的账号上进行,我们不按条收费。

开始使用

每月 15 美元的订阅即可使用 useapi.net 上的所有 API:一个 Token,所有服务。每个服务最多绑定 3 个你自己的账号,useapi.net 负责 REST 接口、负载均衡和轮询。

订阅 — 每月 15 美元

没有国际信用卡?可以用加密货币订阅(加密货币付款不退款)。价格与批量方案见订阅页面。用 Stripe 付款,首次购买 14 天内、且成功生成少于 50 次可全额退款。

如何绑定 Google 账号

  1. 注册 useapi.net 并获取 API Token。
  2. 用自动化浏览器流程绑定一个 Google AI 付费套餐(Pro 或 Ultra)的 Google 账号:输入 Token,在远程浏览器里登录 Google 即可,不需要手动复制 Cookie。远程浏览器用完即关闭删除。建议使用专用的 Gmail 账号。

⚠️ 绑定之后,不要再在任何浏览器里打开这个 Google 账号,包括 Google Vids、Gmail 和其他 Google 页面。Vids 依赖一个浏览器大约每 10 分钟续期一次的会话 Cookie,只要任何浏览器登录了同一个账号,就会抢先续期这个 Cookie,API 手里的那份几分钟内就会失效。Vids 在这一点上比 Google Flow 和 Gemini Notebook 都严格。如果 Google 让账号下线,API 会发邮件通知你,用自动化流程重新绑定大约一分钟。

所有请求也都在 Postman 集合里。

调用方式:同步、异步与 Webhook

每次生成都在我们这边作为一个任务运行。默认是同步调用:POST 会等待任务完成并返回任务记录,视频通常需要 20 到 100 秒,图片约 10 秒。超过约 100 秒仍未完成时返回 202,之后用 GET /jobs/jobid 查询结果。传 async: true 则立即返回 202,再轮询或等待 Webhook。在请求里加 replyUrl,任务完成或失败时会收到一次 POST,内容是最终的任务记录。

生成的文件存在 Google 的临时存储里,不会存进 Vids 文档或 Google Drive,所以要尽快用 GET /media/mediaId 下载:图片只保留几个小时,视频至少半天。

下面的步骤用到 curl、jq 和这些变量与辅助函数。run 以异步方式提交视频任务,每 10 秒轮询一次,把最终的任务记录存到文件:

展开变量与辅助函数
export USEAPI_TOKEN="user:12345-..."
export EMAIL="[email protected]"
API=https://api.useapi.net/v1/google-vids
AUTH="Authorization: Bearer $USEAPI_TOKEN"
enc() { jq -rn --arg v "$1" '$v|@uri'; }
post() { curl -sS -X POST "$API/$1" -H "$AUTH" -H "Content-Type: application/json" -d "$2"; }
get() { curl -sS "$API/$1" -H "$AUTH"; }
download() { curl -sSf "$API/media/$(enc "$1")" -H "$AUTH" -o "$2"; }

# run a video job async: submit it, poll GET /jobs/{jobid} every 10 s, save the final job record to $1
run() {
  local job id
  job=$(post "$2" "$(jq '. + {async: true}' <<< "$3")")
  id=$(jq -r .jobid <<< "$job")
  while [ "$(jq -r .status <<< "$job")" = processing ]; do
    sleep 10
    job=$(get "jobs/$(enc "$id")")
  done
  echo "$job" > "$1"
  jq '{status, error, duration: .result.duration, resolution: .result.resolution, videoLeft: .result.quota.video.left}' "$1"
}

各种 id 里含有 : 和 @,放进 URL 路径时要用 enc 编码。同一条视频的所有输入必须来自同一个 Google 账号,视频任务会在持有这些输入的账号上运行。如果 run 输出的 status 不是 completed,先看 error,后面的步骤都需要一条完成的视频。

实战:两个数字人发布一款球鞋

这个例子做两个虚拟网红,让他们发布一款虚构的球鞋 AeroLoop:Mia 的数字人来自用 API 生成的人像,Leo 的数字人由 Google 根据文字描述绘制。两人和产品图出现在同一条 10 秒视频里,再延长两次到 30 秒。以下每个请求都是生成本页素材时实际用的请求,只替换了 id。提示词保持英文原样,因为它们是直接发给模型的。

完整流程也做成了一个脚本(Node.js 和 Python):GitHub 上 useapi/google-vids-api 仓库的 influencer-avatars/。

1. 生成 Mia 的人像

POST /images 每次生成一张 JPEG,这里用竖幅(9:16,768×1376)和 photography 风格。耗时 9 秒,消耗 1 张图片额度。

curl — POST /images
post images "$(jq -n --arg email "$EMAIL" '{email: $email, aspectRatio: "9:16", style: "photography",
  prompt: "Full-length studio portrait of a confident 24-year-old female fashion influencer, platinum bob, oversized cream hoodie, wide-leg cargo pants, gold hoops, plain light-grey backdrop, facing camera"}')" > mia-portrait.json
MIA_IMAGE=$(jq -r .result.mediaId mia-portrait.json)
download "$MIA_IMAGE" mia-portrait.jpg

Mia:用 POST /images 生成的时尚博主全身人像,奶油色卫衣、工装裤

生成的图片只保留几个小时,要尽快用它创建数字人。

2. 生成 AeroLoop 产品图

同一个接口换成横幅(16:9,1376×768),得到视频里用作产品参考的图片。

curl — POST /images
post images "$(jq -n --arg email "$EMAIL" '{email: $email, aspectRatio: "16:9", style: "photography",
  prompt: "The AeroLoop sneaker: chunky white-and-electric-blue running shoe with a translucent air sole, on a concrete pedestal, dramatic side light, dark background"}')" > aeroloop.json
SHOE_IMAGE=$(jq -r .result.mediaId aeroloop.json)
download "$SHOE_IMAGE" aeroloop.jpg

AeroLoop 球鞋:白色与电光蓝配色、透明气垫鞋底,放在水泥台座上,用 POST /images 生成

3. 选择音色

GET /voices 列出 30 种音色,附 Google 预生成的 18 种语言试听样本。这个接口公开,不需要 Token。Mia 用 Aoede(Vids 里叫 Tova,风格 “Breezy”),Leo 用 Puck(Neo,”Upbeat”)。POST /avatars 两种名字都接受,音色不额外收费。

curl -sS "$API/voices" | jq -r '.voices[] | "\(.name)\t\(.voice)\t\(.style)\t\(.preview)"'

4. 用人像创建 Mia 的数字人

POST /avatars 的 image 传人像的 mediaId,约 2 秒即可在同一账号上保存数字人,除生成人像用掉的 1 张图片额度外,不再消耗额度。

curl — POST /avatars
post avatars "$(jq -n --arg img "$MIA_IMAGE" '{name: "Mia", voice: "Aoede", image: $img}')" > mia-avatar.json
MIA_AVATAR_ID=$(jq -r .avatarId mia-avatar.json)

数字人不能用你自己上传的照片创建。真人数字人属于 Google 的肖像(likeness)功能,需要在浏览器里录自拍视频并验证手机,所以 POST /avatars 会以 400 拒绝上传的照片。请用 API 生成的图片,或者像下一步一样用文字描述。

5. 用文字描述创建 Leo 的数字人

用 appearance 代替 image,由 Google 绘制人物,outfit、shot、expression、details 都是可选的。绘制图片消耗 1 张图片额度,约 20 秒,图片在响应的 previewMediaId 里。我们测试时,描述生成的数字人都是 1280×720(16:9),与 shot 取值无关。

curl — POST /avatars
post avatars "$(jq -n --arg email "$EMAIL" '{email: $email, name: "Leo", voice: "Puck", shot: "upper-body",
  appearance: "a 26-year-old white European male streetwear influencer with short textured hair with a fade and a confident half-smile",
  outfit: "black puffer vest over a white tee, silver chain"}')" > leo-avatar.json
LEO_AVATAR_ID=$(jq -r .avatarId leo-avatar.json)
download "$(jq -r .previewMediaId leo-avatar.json)" leo-avatar.jpg

Leo:Google 根据文字描述绘制的数字人,黑色羽绒背心、白 T 恤、银项链,1280x720

数字人保存在账号的 Vids 文档里,直到你用 DELETE /avatars/avatarId 删除。GET /avatars 随时可以列出它们和各自的 id,包括在 Vids 网页版里做的数字人。

6. 两个数字人和产品同框

POST /videos 最多接受 3 个参考,数字人(avatar_1..avatar_3)和图片(referenceImage_1..referenceImage_3)合计。提示词里用 @avatar_1、@avatar_2、@referenceImage_1 指代它们,每个数字人用自己的声音说出引号里的台词。产品图的 mediaId 可以直接传,API 会取回图片并上传到该账号。

curl — POST /videos
PROMPT=$(cat <<'EOF'
@avatar_1, holding @referenceImage_1 in both hands at chest height, and @avatar_2 stand side by side on a well-lit stage in a small, intimate theater at a sneaker launch, shown full length from head to toe in a wide, steady shot. Behind them a large LED backdrop glows with soft white-and-blue light that moves in slow, gentle waves, with the words "WEAR AEROLOOP" in large, clean, bold white letters across the middle of the screen. The two of them stand on the lower part of the stage, so the words "WEAR AEROLOOP" stay fully readable above their heads. Bright, even stage lighting on both of them. The audience sits in the dark in front of the stage, only the tops of a few heads barely visible at the bottom of the frame. @avatar_1 keeps holding @referenceImage_1 the whole time and says: "Say hi to the AeroLoop!" @avatar_2 turns to her, smiles and says: "Lighter than anything we've ever worn." Simple, uncluttered composition, the two of them clearly the focus.
EOF
)
run launch-10s.json videos "$(jq -n --arg p "$PROMPT" --arg mia "$MIA_AVATAR_ID" --arg leo "$LEO_AVATAR_ID" --arg shoe "$SHOE_IMAGE" \
  '{prompt: $p, avatar_1: $mia, avatar_2: $leo, referenceImage_1: $shoe, duration: 10, resolution: "720p", aspectRatio: "landscape"}')"
LAUNCH_10S=$(jq -r .result.mediaId launch-10s.json)
download "$LAUNCH_10S" launch-10s.mp4

10 秒,720p,生成耗时 38 秒,消耗 10 秒额度。

同一条视频用我们的 watermark remover 去掉可见的 Gemini ✦ 标记后的效果,只需一条命令:python3 vids_watermark_remover.py launch-10s.mp4。

去水印工具适用于直接生成的视频及其放大版本,720p 和 1080p 都可以。它不适用于延长或编辑过的视频:Google 会把带标记的视频交给模型重新生成,星标被画进了画面里。不可见的 SynthID 水印不受影响。

7. 延长到 20 秒

POST /videos/extend 每次增加 3 到 10 秒,返回带新 mediaId 的完整视频。延长只接收视频和提示词,不接收数字人,所以提示词只描述新的动作,并用衣着指代人物(”the man in the black puffer vest”)。这一段里,一个穿红色卫衣的粉丝冲上了舞台。

curl — POST /videos/extend
PROMPT=$(cat <<'EOF'
A young man in a red hoodie climbs onto the stage from the front right edge and plants himself to the right of the man in the black puffer vest, an arm's length away, so all three stand apart in a row with nobody overlapping. Visibly overexcited, he waves a thick stack of banknotes in his right hand above his head and screams: "Take my money!" The man in the black puffer vest laughs, raises both palms toward him and says: "Whoa, whoa, buddy, be patient!" The blonde woman in the cream hoodie laughs, holding the sneaker. Only these three people are on stage. Same stage, same LED screen with WEAR AEROLOOP, same wide steady shot.
EOF
)
run launch-20s.json videos/extend "$(jq -n --arg id "$LAUNCH_10S" --arg p "$PROMPT" '{mediaId: $id, prompt: $p, duration: 10}')"
LAUNCH_20S=$(jq -r .result.mediaId launch-20s.json)
download "$LAUNCH_20S" launch-20s.mp4

20 秒,720p,生成耗时 70 秒,只消耗新增的 10 秒额度。

8. 延长到 30 秒

第二次延长接着 20 秒的视频继续。

curl — POST /videos/extend
PROMPT=$(cat <<'EOF'
The blonde woman in the cream hoodie turns to the camera with a big smile and says: "AeroLoop drops Friday. Link in bio!" Right after that, the young man in the red hoodie throws ALL of his banknotes out over the audience in one big toss, keeping none, and the bills flutter down over their heads. From then on his hands are completely empty: he holds no money at all. Then all three cheer together, the man in the black puffer vest and the young man in the red hoodie throwing their empty fists in the air, as blue and white confetti rains down over the stage. Only these three people are on stage. Same stage, same LED screen with WEAR AEROLOOP, same wide steady shot.
EOF
)
run launch-30s.json videos/extend "$(jq -n --arg id "$LAUNCH_20S" --arg p "$PROMPT" '{mediaId: $id, prompt: $p, duration: 10}')"
LAUNCH_30S=$(jq -r .result.mediaId launch-30s.json)
download "$LAUNCH_30S" launch-30s.mp4

成片:30 秒,720p,生成耗时 85 秒,同样消耗 10 秒额度。

还可以继续延长:连续延长最多到 40 秒,例如 4 段 10 秒。超过之后,Google 只保留原视频的前约 31 秒,再接上新增的部分,所以最长约 41 秒。

9. 放大到 1080p

POST /videos/upscale 把 720p 视频放大到 1080p,时长不变,只接受 20 秒以内的视频,所以这里放大的是第 7 步的 20 秒版本。放大会再消耗一次整条视频的 20 秒额度,耗时 62 秒。

curl — POST /videos/upscale
run launch-20s-1080p.json videos/upscale "$(jq -n --arg id "$LAUNCH_20S" '{mediaId: $id}')"
download "$(jq -r .result.mediaId launch-20s-1080p.json)" launch-20s-1080p.mp4

20 秒视频放大到 1920×1080。

我们测试中,1080p 视频的表现和放大过的 720p 视频一样,生成时间约为两倍,而且每次延长或编辑 1080p 视频,人脸都会更软、更像蜡像。建议:

  • 生成、延长、编辑都用 720p。
  • 成片在 20 秒以内时,最后放大一次。
  • 成片超过 20 秒时,可只在最后一次延长时设 resolution: "1080p"。
  • 只生成单条、不再延长或编辑的视频,可以直接选 1080p,消耗与 720p 相同。

本页保留下来的所有结果(含下方三个示例)共消耗 85 秒视频额度和 3 张图片额度,明细见英文教程。

生成效果示例

同一个账号、同样的输入做出的另外三条视频。

用 POST /videos/edit 把第 6 步的 10 秒视频搬到日落时分的屋顶,消耗 10 秒额度,耗时 64 秒。编辑最多处理 10 秒,更长的视频只会返回前 10 秒的编辑结果。

curl — POST /videos/edit
run launch-rooftop.json videos/edit "$(jq -n --arg v "$LAUNCH_10S" \
  '{video: $v, prompt: "Move the whole scene to a rooftop at sunset, with the city skyline behind them. Same two people, same action, same words."}')"
download "$(jq -r .result.mediaId launch-rooftop.json)" launch-rooftop.mp4

以 startImage 作为首帧的产品展示视频,直接用 1080p 生成。生成的图片链接几小时内就会过期,所以先用 POST /assets 上传保存好的 aeroloop.jpg,再传它的 assetId。

curl — POST /assets 和 POST /videos
SHOE_ASSET=$(curl -sS -X POST "$API/assets?email=$(enc "$EMAIL")" -H "$AUTH" -H "Content-Type: image/jpeg" \
  --data-binary @aeroloop.jpg | jq -r .assetId)
run hero.json videos "$(jq -n --arg img "$SHOE_ASSET" '{startImage: $img, duration: 10, resolution: "1080p", aspectRatio: "landscape",
  prompt: "Slow 360° orbit around the sneaker, blue light rippling through the translucent sole, dust sparkling in the air, cinematic."}')"
download "$(jq -r .result.mediaId hero.json)" hero.mp4

纯文字生成的竖屏视频,1080p(1080×1920),5 秒,耗时 54 秒,消耗 5 秒额度。

curl — POST /videos
run pigeon.json videos "$(jq -n '{duration: 5, resolution: "1080p", aspectRatio: "portrait",
  prompt: "A pigeon in tiny white-and-electric-blue AeroLoop sneakers struts down a city sidewalk like a runway model, head bobbing to the beat. Pedestrians stop and film it with their phones. It pauses, does a slow-motion hair flip with its head feathers, and struts on. Funny, sunny, shot on a phone at pigeon height."}')"
download "$(jq -r .result.mediaId pigeon.json)" pigeon.mp4

同样方法去掉 ✦ 标记后的 1080p 竖屏(1080×1920)鸽子视频:python3 vids_watermark_remover.py pigeon-5s.mp4。脚本根据画面尺寸算出星标的位置和大小(星标按短边缩放,这里是 720p 星标的 1.5 倍),对全部 120 帧逆向还原混合,以 CRF 18 重新编码视频,音频原样复制。这条 5 秒视频在一台 14 核电脑上处理约 10 秒,与上面 10 秒的 720p 视频相当。需要 Python 3.8+、numpy 和 ffmpeg,详见 GitHub 上的 watermark remover。

每个 MP4 右下角都有 Google 加上的可见 Gemini ✦ 标记,以及 Google 声明的不可见 SynthID 水印(见水印说明),可见标记可以用我们 GitHub 示例里的 watermark remover 去掉。

常见问题

  • Google Vids 有官方 API 吗? 截至 2026 年 10 月,Google 没有公开 Vids 的视频生成 API。useapi.net 的 Google Vids API 是第三方 REST API,运行在你自己的 Google 账号上。Omni 1.1 Flash 在官方 Gemini API 上单独按秒计费,720p 约每秒 0.10 美元。
  • 需要验证码或打码服务吗? 不需要。Vids 生成不要求验证码,这一点和每次生成都要 reCAPTCHA 的 Google Flow API 不同。
  • 会消耗 Google Flow 的积分吗? 不会。Vids 有独立的月度额度,同一个 Google 账号可以同时绑定两个 API,两份额度都能用。
  • 免费 Google 账号能用吗? 免费账号在 Vids 里是 0 秒视频、0 张图片,只能创建数字人。生成视频和图片需要付费的 Google AI 套餐,各套餐额度见价格与额度。
  • 能用自己的照片做数字人吗? 不能。真人数字人属于 Google 的肖像功能,需要在浏览器里录自拍视频并验证手机,API 会以 400 拒绝上传的照片。请用 POST /images 生成的图片,或用 appearance 描述人物。
  • Google Vids 生成的视频最长多少秒? 单次生成 3 到 10 秒,每次延长再加 3 到 10 秒,最长约 41 秒,延长只计新增的秒数。
  • 能生成竖屏 9:16 视频吗? 可以。POST /videos 传 aspectRatio: "portrait",720p 为 720×1280,1080p 为 1080×1920,见上面的竖屏示例。
  • 同步请求为什么返回 202? 任务运行超过约 100 秒仍未完成,它会在我们这边继续运行。用 GET /jobs/jobid 查询,或者对耗时长的任务直接用 async: true。
  • Google Vids API 多少钱? 向 useapi.net 支付每月固定 15 美元,再加上你自己 Google AI 套餐里的 Vids 额度,例如 Ultra(199 美元) 每月 10,000 秒视频和 1,000 张图片。更多问答见 Google Vids API 文档(英文)。

结语

有问题欢迎加入我们的 Discord 服务器 或 Telegram 频道。

示例代码见 GitHub 仓库。