Grok Imagine is xAI’s AI video generation model for creating high-quality videos with realistic motion, native audio, and synchronized speech. Grok Imagine 1.5 introduces faster generation speeds, improved motion physics, and stronger visual consistency for modern AI video applications.
还没有视频。提交表单以生成视频。
探索不同用例和参数配置
透明定价,无隐藏费用。按需支付。
| 规则与模式 | 渠道 | 积分 | 价格(美元) | 官方/参考价格 | 每日节省 |
|---|---|---|---|---|---|
grok- imagine-text-to-video-480p videogrokresolution: 480p | Starter | 0.2per second | $0.001 | - | - |
grok- imagine-text-to-video-720p videogrokresolution: 720p | Starter | 0.7per second | $0.003 | - | - |
grok-imagine 使用完整指南
通过 ApiPass 代理平台,将文本提示词转换为电影感短视频。无缝集成 xAI 的 grok-imagine 模型,生成具备连贯运动、场景连续性的视频,并可全面控制宽高比、时长、分辨率和创意模式。ApiPass 异步 API 网关专为生产级扩展而构建,可简化你的工作流程:创建任务以获取 taskId,然后通过自动回调 Webhook 或状态轮询快速获取输出结果。

Grok Imagine 文本生成视频功能可通过 ApiPass API 市场访问,能够即时将自然语言提示词转换为电影感 AI 视频片段。该 API 端点由 xAI 的先进生成式模型驱动,并通过 ApiPass 进行代理,适用于快速概念原型、自动化短内容生成以及企业级视频创作流程。开发者可以通过程序化方式定义复杂的主体运动、丰富的环境、精准的镜头运动和场景时序。
通过 Grok Imagine 模型,将详细的文本提示词即时转换为电影感 AI 视频片段。描述你的主体、环境、镜头语言和氛围——底层 Grok 流水线将负责渲染,并通过 ApiPass 交付一个可追踪、面向生产工作流优化的任务。
可从 Grok 原生的 fun、normal 或 spicy 模式中选择,以控制模型对运动和视觉能量的理解方式。fun 提供更具创意和趣味性的演绎,normal 提供均衡的运动质量,spicy 则生成更动态、更强烈的运动效果。
ApiPass 平台为 Grok 模型提供异步任务工作流。向 ApiPass 端点 /api/v1/jobs/createTask 发送请求,保存返回的 taskId,然后查询 /api/v1/jobs/recordInfo,或提供 callBackUrl 以接收自动完成通知。
通过指定动作、镜头运动和场景动态,将文本描述转换为电影感视频。详细提示词会直接引导 grok-imagine 流水线,捕捉你所需的视觉节奏、序列和时序。
根据项目调性定制运动强度。Grok Imagine API 提供三种不同的生成模式:Normal 用于稳定、均衡的输出;Fun 用于趣味且富有想象力的演绎;Spicy 用于高度动态、富有表现力的动作效果。
可创建竖屏、横屏、方形、宽屏或高竖屏格式的视频片段。可将时长设置为 6 到 30 秒,并选择 480p 或 720p 分辨率,以匹配原型、预览或生产需求。
从简单文本提示词快速创建易于分享的动态视频片段。可使用 9:16 等竖屏格式制作移动优先内容,或使用 16:9 等宽屏格式制作广告、预览和落地页视觉素材。
将生成的视频集成到开发平台或软件解决方案中,用于动态横幅、动画预览、自动化素材背景,以及可根据用户提示词以程序化方式适配的交互式创意工具。
在进入完整制作前,快速测试视觉概念、角色运动、场景转场和镜头创意。异步 API 工作流让批量自动化提示词实验更具可行性。
配置 Grok Imagine API 的原生参数,为你的生产流水线优化运动风格、画面构图和处理速度。
使用 normal 获得均衡结果,使用 fun 获得趣味且富有想象力的运动,使用 spicy 获得更强烈、更有能量或更具表现力的运动。未提供时,mode 参数默认值为 normal。
选择 2:3 用于竖屏,3:2 用于横屏,1:1 用于方形,16:9 用于宽屏,或 9:16 用于高竖屏内容。Duration 接受 6 到 30 秒之间的整数值,步长为 1。
使用 480p 获得更快的预览,在需要 Grok 输出更高分辨率结果时使用 720p。对于生产集成,请提供 callBackUrl,以便 ApiPass 代理网关在异步任务完成时通知你的服务。
新的 ApiPass 用户可以使用免费额度测试受支持的生成工作流,便于在扩展前评估提示词表现、运动质量、宽高比、模式和集成流程。
可选的 channel 参数支持 auto、starter、regular 和 official。auto 允许 ApiPass 根据实时价格和稳定性指标,在可用渠道之间分配任务。
每个请求都会创建一个生成任务并返回 taskId。该设计可确保长时间运行的生成过程保持可靠,并让开发者在不阻塞应用工作流的情况下查询进度或接收完成回调。
打开 ApiPass playground,输入详细提示词,并测试受支持的设置,例如 aspect_ratio、mode、duration 和 resolution。建议从默认值开始:aspect_ratio 2:3、mode normal、duration 6,以及 resolution 480p。
向 /api/v1/jobs/createTask 发送 POST 请求,将 model 设置为 grok-imagine/text-to-video,并传入你的 input 对象。对于需要接收自动完成通知的生产系统,请包含 callBackUrl。
当轮询可接受时,使用返回的 taskId 查询 /api/v1/jobs/recordInfo;或者在配置 callBackUrl 后依赖回调通知。任务成功后,result 中会包含生成的视频输出。
所有 API 都需要通过 Bearer Token 进行身份验证。
Authorization: Bearer
根据文本提示词创建一个异步 Grok Imagine 文本生成视频任务。
1{
2 "model": "grok-imagine/text-to-video",
3 "channel": "auto",
4 "callBackUrl": "https://your-domain.com/api/callback",
5 "input": {
6 "prompt": "A couple of doors open to the right one by one randomly and stay open, to show the inside, each is either a living room, or a kitchen, or a bedroom or an office, with little people living inside.",
7 "aspect_ratio": "2:3",
8 "mode": "normal",
9 "duration": 6,
10 "resolution": "480p"
11 }
12}model必填string模型端点名称。此 API 请使用 grok-imagine/text-to-video。
grok-imagine/text-to-video
channel可选string默认 channel 为 auto。APIPASS 会根据实时价格和稳定性指标,在可用提供商之间自动分配任务;starter 成本极低,但配额有限且稳定性较弱;regular 为标准渠道,价格远低于官方 API,稳定性适中;official 使用模型原生 API,稳定性高、任务执行速度快,并采用官方定价。
默认 channel 为 auto。APIPASS 会根据实时价格和稳定性指标,在可用提供商之间自动分配任务;starter 成本极低,但配额有限且稳定性较弱;regular 为标准渠道,价格远低于官方 API,稳定性适中;official 使用模型原生 API,稳定性高、任务执行速度快,并采用官方定价。
可用选项:
auto
callBackUrl可选string可选的回调 URL,用于接收任务完成的自动通知。生产环境建议使用回调通知,而不是轮询查询接口。
https://your-domain.com/api/callback
prompt必填string描述期望视频运动效果的文本提示词。请详细、具体地说明期望的画面、运动、动作序列、镜头语言、时序、主体、环境和运动动态。最大长度:5000 个字符。支持英文提示词。
必填。请清晰描述运动、动作、镜头移动和场景动态,以获得更好的结果。
A couple of doors open to the right one by one randomly and stay open, to show the inside, each is either a living room, or a kitchen, or a bedroom or an office, with little people living inside.
aspect_ratio可选string指定生成视频的宽高比。默认值:2:3。
2:3 为竖版,3:2 为横版,1:1 为方形,16:9 为宽屏,9:16 为高竖屏格式。
可用选项:
2:3
mode可选string影响运动风格和强度的生成模式。默认值:normal。
fun 会给出更具创意和趣味性的演绎;normal 提供均衡效果,并具备良好的运动质量;spicy 会生成更动态、更强烈的运动效果。
可用选项:
normal
duration可选number生成视频的时长,单位为秒。最小值:6,最大值:30,步长:1。
请使用 6 到 30 之间的整数值。
6
resolution可选string生成视频的分辨率。默认值:480p。
可用选项:
480p
1curl -X POST "https://api.apipass.dev/api/v1/jobs/createTask" \
2 -H "Authorization: Bearer YOUR_API_KEY" \
3 -H "Content-Type: application/json" \
4 -d '{
5 "model": "grok-imagine/text-to-video",
6 "channel": "auto",
7 "callBackUrl": "https://your-domain.com/api/callback",
8 "input": {
9 "aspect_ratio": "2:3",
10 "mode": "normal",
11 "duration": 6,
12 "resolution": "480p",
13 "prompt": "A couple of doors open to the right one by one randomly and stay open, to show the inside, each is either a living room, or a kitchen, or a bedroom or an office, with little people living inside."
14 }
15}'1{
2 "taskId": "task_44e091a72fa8424b",
3 "status": "queued"
4}taskId唯一任务标识符。使用此值查询任务状态和结果。
status创建后的当前任务状态。

ByteDance
ByteDance
起价
0 积分
ByteDance
ByteDance
起价
0 积分
ByteDance
ByteDance
Turn text, images, video clips, and audio into cinematic, audio-synced videos with a single API call.
起价
0 积分
MiniMax
MiniMax
起价
0 积分
ByteDance
ByteDance
Generate cinematic 4K AI videos up to 30 seconds with Seedance 2.5's multimodal power.
起价
0 积分
Google Veo 3.1 is the latest state-of-the-art generative video model designed to transform text and image prompts into high-fidelity cinematic visuals. Building upon its predecessors, Veo 3.1 features significant upgrades in prompt adherence, visual realism, and native audio generation, creating 8-second clips with synchronized sound effects and dialogue.
起价
0 积分
wan
wan
起价
0 积分
Runway
Runway
起价
0 积分
Kling
Kling
Transform text prompts and images into high-quality videos with advanced AI. Generate cinematic content with customizable duration, aspect ratio, and audio capabilities.
起价
0 积分
MiniMax
MiniMax
Generate high-quality AI videos from text prompts or images with MiniMax's Hailuo 2.3 model. Support for both 6s and 10s durations, 768p and 1080p resolutions, with intelligent prompt optimization.
起价
0 积分
Luma
Luma
起价
0 积分
Kling
Kling
Kling v3 video generation model, supports text-to-video and image-to-video, up to 15 seconds, 1080p pro mode
起价
0 积分
wan
wan
Generate videos with cinematic consistency by reimagining existing footage through high-fidelity visual transformation and structural control.
起价
0 积分
起价
0 积分
grok
grok
起价
0 积分
kling
kling
起价
0 积分