mirror of
https://github.com/HeyPuter/puter.git
synced 2026-09-28 16:17:05 +00:00
add google video native provider, imagen models (#2759)
Docker Image CI / build-and-push-image (push) Has been cancelled
Maintain Release Merge PR / update-release-pr (push) Has been cancelled
Notify HeyPuter / notify (push) Has been cancelled
release-please / release-please (push) Has been cancelled
test / test-backend (24.x) (push) Has been cancelled
test / API tests (node env, api-test) (24.x) (push) Has been cancelled
test / puterjs (node env, vitest) (24.x) (push) Has been cancelled
Docker Image CI / build-and-push-image (push) Has been cancelled
Maintain Release Merge PR / update-release-pr (push) Has been cancelled
Notify HeyPuter / notify (push) Has been cancelled
release-please / release-please (push) Has been cancelled
test / test-backend (24.x) (push) Has been cancelled
test / API tests (node env, api-test) (24.x) (push) Has been cancelled
test / puterjs (node env, vitest) (24.x) (push) Has been cancelled
* update gemini .chat models, add imagen, add gemini veo * add models * update documentation * add veo 3.1 lite, 1080p pricing
This commit is contained in:
@@ -49,14 +49,14 @@ For more details, see the [OpenAI API reference](https://platform.openai.com/doc
|
||||
|
||||
#### Gemini Options
|
||||
|
||||
Available when `provider: 'gemini'` or inferred from model (`gemini-2.5-flash-image-preview`, `gemini-3-pro-image-preview`):
|
||||
Available when `provider: 'gemini'` or inferred from model:
|
||||
|
||||
| Option | Type | Description |
|
||||
|--------|------|-------------|
|
||||
| `model` | `String` | Image model to use. |
|
||||
| `ratio` | `Object` | Currently only `{ w: 1024, h: 1024 }` is supported |
|
||||
| `input_image` | `String` | Base64 encoded input image for image-to-image generation |
|
||||
| `input_image_mime_type` | `String` | MIME type of the input image. Options: `'image/png'`, `'image/jpeg'`, `'image/jpg'`, `'image/webp'` |
|
||||
| `ratio` | `Object` | Aspect ratio as `{ w, h }` (e.g., `{ w: 16, h: 9 }`). |
|
||||
| `quality` | `String` | Output size tier: `'512'`, `'1K'`, `'2K'`, `'4K'` (availability varies by model) |
|
||||
| `input_images` | `Array<String>` | Base64 input images for image-to-image (Gemini models only) |
|
||||
|
||||
#### xAI (Grok) Options
|
||||
|
||||
|
||||
@@ -31,14 +31,13 @@ Additional settings for the generation request. Available options depend on the
|
||||
| Option | Type | Description |
|
||||
|--------|------|-------------|
|
||||
| `prompt` | `String` | Text description for the video generation |
|
||||
| `provider` | `String` | The AI provider to use. `'openai' (default) \| 'together'` |
|
||||
| `model` | `String` | Video model to use (provider-specific). Defaults to `'sora-2'` |
|
||||
| `seconds` | `Number` | Target clip length in seconds |
|
||||
| `test_mode` | `Boolean` | When `true`, returns a sample video without using credits |
|
||||
|
||||
#### OpenAI Options
|
||||
|
||||
Available when `provider: 'openai'` or inferred from model (`sora-2`, `sora-2-pro`):
|
||||
Available when using model `sora-2` or `sora-2-pro`:
|
||||
|
||||
| Option | Type | Description |
|
||||
|--------|------|-------------|
|
||||
@@ -49,9 +48,25 @@ Available when `provider: 'openai'` or inferred from model (`sora-2`, `sora-2-pr
|
||||
|
||||
For more details about each option, see the [OpenAI API reference](https://platform.openai.com/docs/api-reference/videos/create).
|
||||
|
||||
#### Google (Veo) Options
|
||||
|
||||
Available when using a Veo model (`veo-2.0-generate-001`, `veo-3.0-generate-001`, `veo-3.1-generate-preview`, etc.):
|
||||
|
||||
| Option | Type | Description |
|
||||
|--------|------|-------------|
|
||||
| `model` | `String` | Video model to use. Available: `'veo-2.0-generate-001'`, `'veo-3.0-generate-001'`, `'veo-3.0-fast-generate-001'`, `'veo-3.1-generate-preview'`, `'veo-3.1-fast-generate-preview'`, `'veo-3.1-lite-generate-preview'` |
|
||||
| `seconds` | `Number` | Target clip length in seconds. Veo 2.0: `5`, `6`, `8`. Veo 3.x: `4`, `6`, `8`. Note: 1080p and 4K output require `seconds: 8` |
|
||||
| `size` | `String` | Output dimensions (e.g., `'1280x720'`, `'1920x1080'`, `'3840x2160'`). `resolution` is an alias. 4K sizes only available on Veo 3.1 models |
|
||||
| `negative_prompt` | `String` | Text describing what to avoid in the video |
|
||||
| `input_reference` | `String` | Base64 image used as the first frame (image-to-video). |
|
||||
| `reference_images` | `Array<String>` | Up to 3 base64 images used as style/asset references. Supported on Veo 3.1 models only |
|
||||
| `last_frame` | `String` | Base64 image used as the last frame |
|
||||
|
||||
For more details, see the [Google Veo API reference](https://ai.google.dev/gemini-api/docs/video).
|
||||
|
||||
#### TogetherAI Options
|
||||
|
||||
Available when `provider: 'together'` or inferred from model:
|
||||
Available when using a TogetherAI model:
|
||||
|
||||
| Option | Type | Description |
|
||||
|--------|------|-------------|
|
||||
@@ -76,7 +91,7 @@ Any properties not set fall back to provider defaults.
|
||||
|
||||
A `Promise` that resolves to an `HTMLVideoElement`. The element is preloaded, has `controls` enabled, and exposes metadata via `data-mime-type` and `data-source` attributes. Append it to the DOM to display the generated clip immediately.
|
||||
|
||||
> **Note:** Real Sora renders can take a couple of minutes to complete. The returned promise resolves only when the MP4 is ready, so keep your UI responsive (for example, by showing a spinner) while you wait. Each successful generation consumes the user’s AI credits in accordance with the model, duration, and resolution you request.
|
||||
> **Note:** Video generation can take several minutes to complete. The returned promise resolves only when the video is ready, so keep your UI responsive (for example, by showing a spinner) while you wait. Each successful generation consumes the user’s AI credits in accordance with the model, duration, and resolution you request.
|
||||
|
||||
## Examples
|
||||
|
||||
|
||||
Reference in New Issue
Block a user