아직 출력이 없습니다. 폼을 제출해 콘텐츠를 생성하세요.
Gemini Omni 사용 전체 가이드
참조 이미지와 짧은 설명을 ApiPass의 모든 Gemini Omni 동영상 생성에 연결해 사용할 수 있는 재사용 가능한 AI 캐릭터로 변환하세요.

ApiPass의 Gemini Omni Character를 사용하면 짧은 설명과 참조 이미지로 재사용 가능한 캐릭터 리소스를 만들 수 있습니다. 매 호출마다 캐릭터를 다시 설명하거나 사진을 다시 업로드하는 대신, 캐릭터를 한 번 등록하고 향후 모든 Gemini Omni Video 요청에서 재사용할 수 있는 간결한 리소스를 받을 수 있습니다. 이 모든 기능은 ApiPass 모델 카탈로그의 다른 모델에 이미 사용 중인 동일한 API 키로 이용할 수 있습니다.
캐릭터를 한 번만 등록한 뒤 이후 모든 Gemini Omni Video 호출에서 참조하세요. 사진을 다시 업로드하거나 설명을 다시 입력할 필요가 없습니다.
캐릭터를 참조 이미지에 고정해 동일한 얼굴, 스타일, 정체성이 클립에서 클립으로, 장면에서 장면으로 자연스럽게 이어지도록 합니다.
캐릭터 리소스는 Gemini Omni Audio를 통해 생성한 음성 프로필과 자연스럽게 함께 사용할 수 있어, 하나의 캐릭터가 전체 시리즈나 캠페인에서 일관된 외형과 음성을 모두 유지할 수 있습니다.
생성된 각 캐릭터는 자체 데이터베이스에 쉽게 저장하고 제품 UI에 표시할 수 있는 간결한 리소스를 반환합니다.
캐릭터 생성은 Gemini Omni Video 및 Gemini Omni Audio와 동일한 ApiPass 환경 안에 있으므로, 여러 제공업체를 오갈 필요 없이 캐릭터 중심의 엔드투엔드 파이프라인을 구축할 수 있습니다.
선명한 참조 이미지 하나와 외모, 스타일, 성격을 다루는 자연어 설명을 함께 사용해 새 캐릭터를 등록하세요.
리소스를 만들 때 각 캐릭터에 이름을 지정해 자체 라이브러리나 UI 안에서 쉽게 식별할 수 있도록 하세요.
Gemini Omni Audio의 음성 프로필을 캐릭터와 연결해 처음부터 캐릭터에 음성 페르소나를 함께 부여하세요.
API는 필요할 때 캐릭터 일관성이 유지되는 생성을 구동할 수 있도록, 모든 Gemini Omni Video 요청에 첨부 가능한 재사용 가능한 캐릭터 리소스를 반환합니다.
등록된 여러 캐릭터를 하나의 Gemini Omni Video 요청에서 조합할 수 있어, 반복 등장 캐릭터가 둘 이상인 장면에서도 함께 모델 일관성을 유지할 수 있습니다.
ApiPass에서 AI 아바타를 한 번 만들고, 그 아바타가 대신 영상을 진행하게 하세요. YouTube 채널, TikTok 시리즈, 주간 팟캐스트 등 무엇을 운영하든 캐릭터는 모든 에피소드에 일관되게 등장합니다. 같은 얼굴, 같은 목소리, 같은 분위기 그대로요. 더 이상 카메라를 세팅하거나 머리 상태를 걱정할 필요가 없습니다. 스크립트만 작성하면 AI로 구현된 당신이 말하게 됩니다.
광고 변형, 제품 튜토리얼, 현지화된 시장 전반에 등장하는 승인된 단일 캐릭터로, 어디에 나타나든 브랜드 정체성을 일관되게 유지합니다.
에피소드형 숏폼 콘텐츠와 연재형 스토리 세계관을 위한 주인공, 조력자, 진행자로, 한 에피소드에서 다음 에피소드까지 모두 모델 일관성을 유지합니다.
강의, 설명 영상, 교육 모듈을 위한 일관된 교사, 내레이터, 진행자를 제공해 모든 레슨이 같은 익숙한 얼굴이 전달하는 것처럼 느껴지게 합니다.
동일한 진행자나 주인공이 모든 영상에 안정적으로 등장해야 하는 에피소드형 시리즈를 구축합니다.
모든 광고, 설명 콘텐츠, 현지화된 변형에 주연으로 등장하는 온브랜드 대변인이나 마스코트를 유지합니다.
전체 강좌 카탈로그를 이끄는 일관된 강사 또는 발표자 페르소나를 설계합니다.
시네마틱, 트레일러, 앱 내 스토리 시퀀스 전반에 등장하는 반복 캐릭터를 등록합니다.
최종 사용자에게 자신의 AI 정체성을 저장, 관리, 재사용할 수 있는 "My Characters" 라이브러리를 제공합니다.
ApiPass 계정을 만들고 대시보드에서 API 키를 생성해, 하나의 자격 증명으로 character, video, audio를 포함한 전체 Gemini Omni 제품군을 사용할 수 있습니다.
선명한 참조 이미지 하나를 선택하고, 외모와 스타일을 다루는 짧은 설명을 작성한 다음 캐릭터 이름을 선택하세요. 필요하다면 Gemini Omni Audio의 음성 프로필과 캐릭터를 연결해 완성도 높은 정체성을 만들 수 있습니다.
Gemini Omni Character를 호출해 캐릭터를 등록한 뒤, 이후 모든 Gemini Omni Video 요청에 첨부해 일관된 캐릭터 중심 영상을 대규모로 생성하세요.
모든 API는 Bearer Token을 통한 인증이 필요합니다.
Authorization: Bearer
Create a reusable Gemini Omni character resource from a character description and one public reference image. The returned characterId can be used as a character_ids value with google/gemini-omni-video.
1{
2 "model": "google/gemini-omni-character",
3 "channel": "auto",
4 "input": {
5 "descriptions": "A young female character with short silver hair, a black futuristic utility jacket, calm expression, agile posture, and a clean cyberpunk visual style.",
6 "image_urls": [
7 "https://cdn.apipass.dev/apipass/results/task_78514b0d12c44ad4_0.png"
8 ],
9 "character_name": "Jenny"
10 }
11}model필수stringModel endpoint name. Use google/gemini-omni-character for this endpoint.
google/gemini-omni-character
channel선택적stringDefault channel is auto. APIPASS automatically allocates tasks across available providers based on real-time pricing and stability metrics; starter is ultra-low-cost with limited quotas and weaker stability; regular is standard and much cheaper than official APIs with moderate stability; official uses the model native API with high stability, fast task execution, and official pricing.
Default channel is auto. APIPASS automatically allocates tasks across available providers based on real-time pricing and stability metrics. starter is ultra-low-cost with limited quotas and weaker stability. regular is standard and much cheaper than official APIs with moderate stability. official uses the model native API with high stability, fast task execution, and official pricing.
사용 가능한 옵션:
auto
callBackUrl선택적stringOptional callback URL for receiving task result updates. For this model, callback behavior depends on the adapter implementation.
https://example.com/callback
descriptions필수stringCharacter description used to define the appearance, identity, style, clothing, or personality of the character.
Use the plural field name descriptions. The official schema requires descriptions and image_urls; although some examples may show description, descriptions is the recommended field name.
A young female character with short silver hair and a futuristic utility jacket, calm, agile, and strongly cyberpunk in style.
image_urls필수array<string>Array of public character reference image URLs. Only 1 image is supported, and each image must be no larger than 20MB.
The image URL must be publicly accessible by the upstream service. Maximum items: 1.
[ "https://cdn.apipass.dev/apipass/results/task_78514b0d12c44ad4_0.png" ]
audio_ids선택적array<string>Array of audio IDs generated by the google/gemini-omni-audio endpoint. These can provide voice traits, tone, or persona guidance for the character.
audio_ids must come from the corresponding KIE gemini-omni-audio resource.
[ "audio_xxx" ]
character_name선택적stringCharacter name.
Jenny
1curl -X POST "https://api.apipass.dev/api/v1/jobs/createTask" \
2 -H "Authorization: Bearer YOUR_API_KEY" \
3 -H "Content-Type: application/json" \
4 -d '{
5 "model": "google/gemini-omni-character",
6 "channel": "auto",
7 "input": {
8 "descriptions": "A young female character with short silver hair, a black futuristic utility jacket, calm expression, agile posture, and a clean cyberpunk visual style.",
9 "image_urls": [
10 "https://cdn.apipass.dev/apipass/results/task_78514b0d12c44ad4_0.png"
11 ],
12 "character_name": "Jenny"
13 }
14}'1{
2 "code": 200,
3 "msg": "success",
4 "data": {
5 "characterId": "b09dbf56...",
6 "characterName": "Jenny",
7 "imageUrl": "https://xx.com/a.png"
8 }
9}codeResponse status code. 200 indicates a successful upstream character creation response.
msgResponse message.
dataCreated character resource information.
data.characterIdGenerated character ID. Use this value in google/gemini-omni-video input.character_ids when generating videos with this character.
data.characterNameCharacter name returned by the provider.
data.imageUrlURL of the character reference or generated character image returned by the provider.

ByteDance
ByteDance
시작가
0 크레딧
ByteDance
ByteDance
시작가
0 크레딧
ByteDance
ByteDance
Turn text, images, video clips, and audio into cinematic, audio-synced videos with a single API call.
시작가
0 크레딧
MiniMax
MiniMax
시작가
0 크레딧
ByteDance
ByteDance
Generate cinematic 4K AI videos up to 30 seconds with Seedance 2.5's multimodal power.
시작가
0 크레딧
Google Veo 3.1 is the latest state-of-the-art generative video model designed to transform text and image prompts into high-fidelity cinematic visuals. Building upon its predecessors, Veo 3.1 features significant upgrades in prompt adherence, visual realism, and native audio generation, creating 8-second clips with synchronized sound effects and dialogue.
시작가
0 크레딧
wan
wan
시작가
0 크레딧
Runway
Runway
시작가
0 크레딧
Kling
Kling
Transform text prompts and images into high-quality videos with advanced AI. Generate cinematic content with customizable duration, aspect ratio, and audio capabilities.
시작가
0 크레딧
MiniMax
MiniMax
Generate high-quality AI videos from text prompts or images with MiniMax's Hailuo 2.3 model. Support for both 6s and 10s durations, 768p and 1080p resolutions, with intelligent prompt optimization.
시작가
0 크레딧
grok
grok
Grok Imagine is xAI’s AI video generation model for creating high-quality videos with realistic motion, native audio, and synchronized speech. Grok Imagine 1.5 introduces faster generation speeds, improved motion physics, and stronger visual consistency for modern AI video applications.
시작가
0 크레딧
Luma
Luma
시작가
0 크레딧
Kling
Kling
Kling v3 video generation model, supports text-to-video and image-to-video, up to 15 seconds, 1080p pro mode
시작가
0 크레딧
wan
wan
Generate videos with cinematic consistency by reimagining existing footage through high-fidelity visual transformation and structural control.
시작가
0 크레딧
grok
grok
시작가
0 크레딧
kling
kling
시작가
0 크레딧