AI 모델 카탈로그

모델 라이브러리

Virse에서 사용할 수 있는 모든 모델을 만들어 내는 결과물별로 모았습니다. 모델을 선택해 상세 페이지를 열거나, 캔버스에서 바로 제작을 시작하세요.

비디오모델 7개

Kuaishou의 Kling 3.0은 말합니다 — 5개 언어의 립싱크 대화, 캐릭터별 서로 다른 음성, 전체 시퀀스를 구성하는 스토리보드 모드를 제공합니다.

Kuaishou

동기화된 오디오, 멀티모달 레퍼런스, 정밀한 크리에이티브 제어 기능으로 Seedance 2.5에서 영화 같은 30초 비디오를 생성하세요.

ByteDance

Google DeepMind의 Veo 3.1은 동기화된 48kHz 오디오가 포함된 짧은 클립을 생성하고, 마지막 프레임을 이어받아 더 긴 시퀀스로 확장합니다.

Google

ByteDance's Seedance 2.0 produces picture and sound in one pass, from a first and last frame or from reference images, at sizes from 480P to native 4K.

ByteDance

MiniMax's Hailuo 03 model produces silent 2K video, driven either by a first and last frame or by a set of reference images.

MiniMax

The Happy Horse 1.0 AI video model needs one still image and a sentence, producing motion from the frame you already have.

Alibaba

Google's natively multimodal video model, taking text, images, video, and audio as input, offered in Virse as frame-driven and reference-driven entries.

Google

이미지모델 14개

OpenAI의 ChatGPT Images 2.5가 GPT Image 2.5 Flare와 GPT Image 2.5 Sunburst, 두 API 에디션으로 Virse에 도입되어 여섯 가지 품질 등급과 최대 4K 출력을 제공합니다.

OpenAI

FLUX 2 이전에 출시된 Black Forest Labs의 FLUX 라인 세대로, 표준 모델과 Ultra라는 두 항목으로 계속 제공됩니다.

Black Forest Labs

ByteDance의 Seedream 4.0 및 Seedream 4.5는 작성한 설명과 제공한 이미지로부터 생성하여, 하나의 원본 사진을 프로젝트에 필요한 만큼 다양한 방향으로 전개합니다.

ByteDance

ByteDance의 최신 Seedream 세대로, Seedream 5.0 Pro와 Seedream V5 Lite로 제공되며 둘 다 동일한 작성 브리프를 받습니다.

ByteDance

60억 개의 파라미터를 갖춘 모델을 8단계 추론 파이프라인으로 압축해, 생각의 흐름이 넘어가기 전에 바로 사용할 수 있는 이미지를 반환하도록 설계했습니다.

Alibaba

Reve AI's second-generation model builds an addressable layout of positioned elements before it renders, so a specified arrangement comes back arranged.

Reve

Alibaba's Qwen Image 3.0 Pro renders type as language across a dozen scripts, holding legibility down to ten pixels in dense layouts.

Alibaba

Generate studio-grade 4K images with Nano Banana Pro, blending multiple references while holding characters, products, and typography consistent.

Google

Google's Nano Banana 2 pairs flagship-level image quality with Flash-class speed, across four output sizes from rapid 0.5K prototypes to finished 4K.

Google

A text-to-image foundation model trained from scratch around layout, with bounding-box placement, hex colour conditioning, and native 2K output.

Ideogram

Google's Gemini 2.5 Flash Image — the original Nano Banana — generates and edits pictures at conversation speed, so a correction takes a sentence instead of a session.

Google

OpenAI's GPT Image 2.0 reads a hundred-word brief and attempts all of it, with independent quality and resolution controls from 1K to 4K.

OpenAI

Black Forest Labs' editing models, FLUX Kontext Pro and Kontext Max, built to apply one described change while the rest of the frame survives untouched.

Black Forest Labs

Black Forest Labs' production model reads briefs up to 32,000 tokens, so brand rules, colour values, and compositional constraints can all go in at once.

Black Forest Labs