还没有视频。提交表单以生成视频。
探索不同用例和参数配置
透明定价,无隐藏费用。按需支付。
| 规则与模式 | 渠道 | 积分 | 价格(美元) | 官方/参考价格 | 每日节省 |
|---|---|---|---|---|---|
with-video-720p videoGoogleresolution: 720p | Regular | 268per run | $1.218 | - | - |
with-video-1080p videoGoogleresolution: 1080p | Regular | 270per run | $1.227 | - | - |
with-video-4k videoGoogleresolution: 4k | Regular | 400per run | $1.818 | - | - |
no-video-720p-4s videoGoogleduration: 4, resolution: 720p | Regular | 102per run | $0.464 | - | - |
no-video-720p-6s videoGoogleduration: 6, resolution: 720p | Regular | 135per run | $0.614 | - | - |
no-video-720p-8s videoGoogleduration: 8, resolution: 720p | Regular | 168per run | $0.764 | - | - |
no-video-720p-10s videoGoogleduration: 10, resolution: 720p | Regular | 200per run | $0.909 | - | - |
no-video-1080p-4s videoGoogleduration: 4, resolution: 1080p | Regular | 103per run | $0.468 | - | - |
no-video-1080p-6s videoGoogleduration: 6, resolution: 1080p | Regular | 136per run | $0.618 | - | - |
no-video-1080p-8s videoGoogleduration: 4, resolution: 1080p | Regular | 168per run | $0.764 | - | - |
no-video-1080p-10s videoGoogleduration: 10, resolution: 1080p | Regular | 202per run | $0.918 | - | - |
no-video-4k-4s videoGoogleduration: 4, resolution: 4k | Regular | 233per run | $1.059 | - | - |
no-video-4k-6s videoGoogleduration: 6, resolution: 4k | Regular | 268per run | $1.218 | - | - |
no-video-4k-8s videoGoogleduration: 8, resolution: 4k | Regular | 300per run | $1.364 | - | - |
no-video-4k-10s videoGoogleduration: 10, resolution: 4k | Regular | 335per run | $1.523 | - | - |
Gemini Omni 使用完整指南
通过单一 ApiPass 端点,使用 Google 的 Gemini Omni 模型生成具有电影感、符合物理规律的 AI 视频,并可结合可复用的角色和声音。

ApiPass 上的 Gemini Omni Video API 是 Google Gemini Omni Flash 的视频生成接口。Gemini Omni Flash 是一款高性能多模态模型,专为高速视频生成、编辑和电影级控制而设计。ApiPass 并不是把文生视频作为一次孤立调用来运行,而是将 Gemini Omni Video 作为 Omni 系列的组合层对外提供:你可以传入已在 ApiPass 上创建的可复用角色资源和声音配置,并获得一段视觉与音频同步的完整渲染视频。Gemini Omni 可同时处理文本、图像、音频和视频,让输出更连贯、更一致、更可控;ApiPass 则将这种能力与其余模型目录一起封装在同一个 API key 之后。
无需管理多个独立账号、预览权限或云端配置,ApiPass 让你使用同一个已用于其他模型目录的 API key,即可调用 Gemini Omni Video。
Gemini Omni Video 可在单个 prompt 中接受文本、图像、视频和音频的任意组合,让你能基于真实创作素材发起生成,而不只是从空白文本 prompt 开始。
该 API 基于真实世界物理模型进行渲染,能够呈现可信的反射、重力、光照和天气效果,即使在动态镜头中,画面也能保持视觉一致,而不是逐渐产生瑕疵。
你可以使用多张参考图和短视频片段引导生成,模型会保持主体、风格和场景的一致性,在编辑和不同镜头之间稳定维持角色身份。
对话式编辑让你可以通过自然语言逐步优化和修改视频,因此无需从头重写整个 prompt,就能重新调整某个场景。
Gemini Omni 将广泛的世界知识与强大的视频能力相结合——它能理解文化背景、历史设定和科学准确性,生成的不只是好看的视频,更是合乎逻辑的视频。当 prompt 涉及特定时代、地点或需要事实一致性的概念时,这一点尤其有用。
Gemini Omni API 支持自然语言视频编辑,允许用户逐步优化场景,而不是每次都重新构建完整 prompt。用户可以更改环境、调整动作、替换物体、移动镜头角度或添加视觉效果,同时保持原始场景的连贯性。这使其非常适用于 AI 视频编辑器、创作者工具,以及需要用更直观方式转换现有素材的应用。
Gemini Omni Video 将视觉创作与物理、历史、生物、文化和叙事逻辑知识连接起来。这有助于让生成视频减少随机感、更具意图性,对解说视频、电影化叙事、产品概念展示和教育内容都很有价值。
文本可定义方向,图像可引导主体或风格,视频可提供动作和场景上下文,音频或角色资源则可支持由语音驱动或具备身份感知的工作流。这有助于开发者构建从真实创作素材出发的视频工具,而不仅仅依赖空白 prompt。
Gemini Omni Video 可支持化身式场景,在这些场景中,角色存在感、表情、表达方式和环境需要自然衔接。这对于主持人片段、角色主导内容、互动媒体以及面向未来的创意视频产品都很有用。
基于 prompt 加单张参考图生成竖版短片,并在每一集中保持可复用的主角。
生成本地化广告变体,使用同一个符合品牌调性的代言人和声音,同时替换产品、信息或目标市场。
适用于教育内容、历史准确的场景,以及事实一致性与视觉质量同样重要的知识密集型叙事。
由语音驱动的化身式工作流,可让声音表现引导最终结果,适用于主持人视频、角色对话、旁白场景,以及需要让声音、表情和屏幕动作自然衔接的生成片段。
官方 Gemini Omni 访问目前分散在预览通道、多个控制台和企业级入口中。ApiPass 将其与其余模型栈一起整合到单个 API key 之后,因此集成只需更改 base-URL 和 key,而不是重新接入一个新的供应商。
官方 API 将每次生成都视为一个自包含请求。在 ApiPass 上,Gemini Omni Video 被设计为可通过 Gemini Omni Character 和 Gemini Omni Audio 能力,使用你已在平台上构建的角色和声音资源——因此你可以维护一个身份和声音库,并将它们插入到后续任何视频任务中。
制作需要可重复主角、旁白和视觉风格的系列化短视频。
发布本地化、个性化的广告变体,并在不同市场和创意组合中保持品牌一致性。
将课程脚本和教案转化为带旁白、由角色主导的教学视频,并生成具备历史或科学依据的场景。
制作电影化片段、预告片和应用内故事序列,让同一角色在多个镜头和迭代中稳定出现。
创建 ApiPass 账号,并在控制台生成 API key,即可解锁完整的 Gemini Omni 系列——视频、角色和音频——以及其余模型目录。
编写场景 prompt,并可选择收集参考素材——图像、短源视频、可复用角色资源,以及来自 Gemini Omni Audio 的声音配置——让模型具备保持身份、语气和风格一致所需的一切信息。
使用你的 prompt 和选定输入调用 Gemini Omni Video,查看生成片段,并通过对话式编辑不断优化,直到场景、动作和表达方式符合你的创作方向。
所有 API 都需要通过 Bearer Token 进行身份验证。
Authorization: Bearer
提交新的 Nano Banana 2 图像生成或编辑任务
该 API 接受具有以下结构的 JSON payload:
1{
2 "model": "string",
3 "callBackUrl": "string (optional)",
4 "channel": "auto",
5 "input": {
6 // Input parameters
7 }
8}model必填string用于生成的模型名称
"google/nano-banana-2"
callBackUrl可选string用于任务完成通知的回调 URL。如果省略,则不会发送回调。
"https://your-domain.com/api/callback"
channel可选string你可以通过 channel 参数指定 APIPASS 中对应的提供商;这些提供商负责实际的图像和视频生成任务。APIPASS 目前提供三种提供商选项:
channel 参数的默认值为 auto。启用后,APIPASS 会根据实时价格和稳定性指标,在可用提供商之间自动分配任务,以平衡最低成本和可靠性能。除非你有自定义路由需求,否则请保留默认值 auto。
可用选项:
auto
input 对象包含以下参数:
input.prompt必填string你想生成的图像的文本描述
描述主体、风格、光照和构图,以获得最佳效果
"A serene alpine lake reflecting snow-capped mountains at golden hour, photorealistic"
input.image_input可选array(URL)用于转换或作为参考的输入图像。最多支持 14 张图片。
支持的类型:image/jpeg、image/png;每张图片最大大小:30MB;最多文件数:14
["https://example.com/reference.jpg"]
input.aspect_ratio可选string生成图像的宽高比。提供 image_input 时默认为 match_input_image,否则默认为 1:1。
可用选项:
"16:9"
input.resolution可选string生成图像的分辨率。更高分辨率会产生更多细节,但生成时间更长。默认值:1K。
可用选项:
"1K"
input.output_format可选string输出图像的格式。默认值:jpg。
可用选项:
"jpg"
1curl -X POST "https://api.apipass.dev/api/v1/jobs/createTask" \
2 -H "Content-Type: application/json" \
3 -H "Authorization: Bearer YOUR_API_KEY" \
4 -d '{
5 "model": "google/nano-banana-2",
6 "callBackUrl": "https://your-domain.com/api/callback",
7 "input": {
8 "prompt": "A serene alpine lake reflecting snow-capped mountains at golden hour, photorealistic",
9 "aspect_ratio": "16:9",
10 "resolution": "1K",
11 "google_search": false,
12 "image_search": false,
13 "output_format": "jpg"
14 }
15 }'1{
2 "code": 200,
3 "message": "success",
4 "data": {
5 "taskId": "task_12345678"
6 }
7}code状态码,200 表示成功,其他表示失败
message响应消息,失败时为错误描述
data.taskId用于查询任务状态和结果的任务 ID

ByteDance
ByteDance
起价
0 积分
ByteDance
ByteDance
起价
0 积分
ByteDance
ByteDance
Turn text, images, video clips, and audio into cinematic, audio-synced videos with a single API call.
起价
0 积分
MiniMax
MiniMax
起价
0 积分
ByteDance
ByteDance
Generate cinematic 4K AI videos up to 30 seconds with Seedance 2.5's multimodal power.
起价
0 积分
Google Veo 3.1 is the latest state-of-the-art generative video model designed to transform text and image prompts into high-fidelity cinematic visuals. Building upon its predecessors, Veo 3.1 features significant upgrades in prompt adherence, visual realism, and native audio generation, creating 8-second clips with synchronized sound effects and dialogue.
起价
0 积分
wan
wan
起价
0 积分
Runway
Runway
起价
0 积分
Kling
Kling
Transform text prompts and images into high-quality videos with advanced AI. Generate cinematic content with customizable duration, aspect ratio, and audio capabilities.
起价
0 积分
MiniMax
MiniMax
Generate high-quality AI videos from text prompts or images with MiniMax's Hailuo 2.3 model. Support for both 6s and 10s durations, 768p and 1080p resolutions, with intelligent prompt optimization.
起价
0 积分
grok
grok
Grok Imagine is xAI’s AI video generation model for creating high-quality videos with realistic motion, native audio, and synchronized speech. Grok Imagine 1.5 introduces faster generation speeds, improved motion physics, and stronger visual consistency for modern AI video applications.
起价
0 积分
Luma
Luma
起价
0 积分
Kling
Kling
Kling v3 video generation model, supports text-to-video and image-to-video, up to 15 seconds, 1080p pro mode
起价
0 积分
wan
wan
Generate videos with cinematic consistency by reimagining existing footage through high-fidelity visual transformation and structural control.
起价
0 积分
grok
grok
起价
0 积分
kling
kling
起价
0 积分