AI モデル

すべてのフレームに最適なモデルを。

AIVideoFlyで利用できる動画・画像モデルを比較し、適切な生成ツールですぐに制作を始められます。

Multimodal reference

Seedance

Longer cinematic clips with synchronized native audio.

テキストから動画画像から動画ネイティブ音声
長さ
4–30s
出力
480p / 720p

プロンプト例

“A fashion film of a chrome perfume bottle on black glass, slow dolly-in, rain droplets, soft bass and distant thunder.”

このモデルで作成

Controllable motion with image and native audio.

テキストから動画画像から動画ネイティブ音声
長さ
4–15s
出力
480p / 720p / 1080p

プロンプト例

“A fashion film of a chrome perfume bottle on black glass, slow dolly-in, rain droplets, soft bass and distant thunder.”

このモデルで作成

Lower-cost drafts with image and audio support.

テキストから動画画像から動画ネイティブ音声
長さ
4–15s
出力
480p / 720p / 1080p

プロンプト例

“A fashion film of a chrome perfume bottle on black glass, slow dolly-in, rain droplets, soft bass and distant thunder.”

このモデルで作成

Cinematic video with synchronized native audio.

テキストから動画画像から動画ネイティブ音声
長さ
4–12s
出力
480p / 720p / 1080p

プロンプト例

“A fashion film of a chrome perfume bottle on black glass, slow dolly-in, rain droplets, soft bass and distant thunder.”

このモデルで作成

Native audio

Veo

Fast generation with optional native audio.

テキストから動画画像から動画ネイティブ音声
長さ
8–8s
出力
720p / 1080p

プロンプト例

“A close-up of a chef plating ramen in a busy Tokyo kitchen, steam rising, handheld camera, natural kitchen sounds.”

このモデルで作成

High-fidelity generation with native audio.

テキストから動画画像から動画ネイティブ音声
長さ
8–8s
出力
720p / 1080p

プロンプト例

“A close-up of a chef plating ramen in a busy Tokyo kitchen, steam rising, handheld camera, natural kitchen sounds.”

このモデルで作成

Flexible generation

Wan

Alibaba

Flexible, high-quality generation for everyday use.

テキストから動画画像から動画
長さ
5–15s
出力
720p / 1080p

プロンプト例

“A woman in a red coat crosses a misty city bridge at dawn; wide establishing shot, then a slow tracking close-up.”

このモデルで作成

Reliable video generation at a lower cost.

テキストから動画画像から動画
長さ
5–15s
出力
720p / 1080p

プロンプト例

“A woman in a red coat crosses a misty city bridge at dawn; wide establishing shot, then a slow tracking close-up.”

このモデルで作成

Expressive motion

HappyHorse

Expressive motion and polished visual detail.

テキストから動画画像から動画
長さ
5–10s
出力
720p / 1080p

プロンプト例

“A toy robot wakes on a child’s desk, looks around, and takes its first curious steps in warm morning light.”

このモデルで作成

Stable motion generation with simple controls.

テキストから動画画像から動画
長さ
5–10s
出力
720p / 1080p

プロンプト例

“A toy robot wakes on a child’s desk, looks around, and takes its first curious steps in warm morning light.”

このモデルで作成

Fast creative drafts

Grok

Fast creative drafts for text and image prompts.

テキストから動画画像から動画
長さ
6–10s
出力
480p / 720p

プロンプト例

“A vibrant paper-cut city comes alive at sunset, cars flow through the streets, playful stop-motion movement.”

このモデルで作成

Cinematic camera

Kling

Cinematic camera motion with optional sound.

テキストから動画画像から動画ネイティブ音声
長さ
3–15s
出力
720p / 1080p

プロンプト例

“A luxury watch floats above dark water, camera orbits slowly as ripples and highlights reveal the metal texture.”

このモデルで作成

Advanced cinematic generation with precise motion.

テキストから動画画像から動画ネイティブ音声
長さ
3–15s
出力
720p / 1080p

プロンプト例

“A luxury watch floats above dark water, camera orbits slowly as ripples and highlights reveal the metal texture.”

このモデルで作成

High-detail cinematic motion

Hailuo

High-detail 768p and 2K generation for cinematic motion.

テキストから動画画像から動画
長さ
4–15s
出力
768p / 2k

プロンプト例

“A premium skincare bottle rotates in a sunlit studio, soft reflections travel across the glass, controlled camera orbit, high-detail product film.”

このモデルで作成

完成画像の生成や参照画像の編集に対応。

画像モデル

モデルの機能は本番の生成設定と同期しています。

AI動画モデル | AIVideoFly