Pustaka Model
Semua model yang tersedia di dalam Virse, dikelompokkan menurut apa yang dihasilkannya. Pilih satu untuk membuka halamannya, atau langsung berkarya dengannya di kanvas.
Video7 model
Kling 3.0 dari Kuaishou dapat berbicara — dialog tersinkronisasi bibir dalam lima bahasa, suara berbeda untuk tiap karakter, dan mode storyboard yang menyusun keseluruhan rangkaian.
KuaishouBuat video sinematik 30 detik dengan Seedance 2.5, dilengkapi audio tersinkronisasi, referensi multimodal, dan kontrol kreatif yang presisi.
ByteDanceVeo 3.1 dari Google DeepMind menghasilkan klip pendek dengan audio 48kHz yang tersinkronisasi dan memperpanjangnya menjadi rangkaian yang lebih panjang dengan meneruskan frame penutup ke depan.
GoogleByteDance's Seedance 2.0 produces picture and sound in one pass, from a first and last frame or from reference images, at sizes from 480P to native 4K.
ByteDanceMiniMax's Hailuo 03 model produces silent 2K video, driven either by a first and last frame or by a set of reference images.
MiniMaxThe Happy Horse 1.0 AI video model needs one still image and a sentence, producing motion from the frame you already have.
AlibabaGoogle's natively multimodal video model, taking text, images, video, and audio as input, offered in Virse as frame-driven and reference-driven entries.
GoogleGambar14 model
ChatGPT Images 2.5 dari OpenAI hadir di Virse dalam kedua edisi API, GPT Image 2.5 Flare dan GPT Image 2.5 Sunburst, dengan enam tingkat kualitas dan output hingga 4K.
OpenAIGenerasi lini FLUX dari Black Forest Labs yang hadir sebelum FLUX 2, tetap tersedia sebagai dua entri — model standar dan Ultra.
Black Forest LabsSeedream 4.0 dan Seedream 4.5 dari ByteDance menghasilkan gambar dari deskripsi tertulis dan dari gambar yang Anda berikan, mengubah satu gambar sumber menjadi sebanyak mungkin arah yang dibutuhkan proyek.
ByteDanceGenerasi Seedream terbaru dari ByteDance, tersedia sebagai Seedream 5.0 Pro dan Seedream V5 Lite, keduanya menggunakan brief tertulis yang sama.
ByteDanceModel 6 miliar parameter yang dipadatkan menjadi pipeline inferensi delapan langkah, dibuat untuk menghasilkan gambar yang siap pakai sebelum alur pikiran Anda beralih.
AlibabaReve AI's second-generation model builds an addressable layout of positioned elements before it renders, so a specified arrangement comes back arranged.
ReveAlibaba's Qwen Image 3.0 Pro renders type as language across a dozen scripts, holding legibility down to ten pixels in dense layouts.
AlibabaGenerate studio-grade 4K images with Nano Banana Pro, blending multiple references while holding characters, products, and typography consistent.
GoogleGoogle's Nano Banana 2 pairs flagship-level image quality with Flash-class speed, across four output sizes from rapid 0.5K prototypes to finished 4K.
GoogleA text-to-image foundation model trained from scratch around layout, with bounding-box placement, hex colour conditioning, and native 2K output.
IdeogramGoogle's Gemini 2.5 Flash Image — the original Nano Banana — generates and edits pictures at conversation speed, so a correction takes a sentence instead of a session.
GoogleOpenAI's GPT Image 2.0 reads a hundred-word brief and attempts all of it, with independent quality and resolution controls from 1K to 4K.
OpenAIBlack Forest Labs' editing models, FLUX Kontext Pro and Kontext Max, built to apply one described change while the rest of the frame survives untouched.
Black Forest LabsBlack Forest Labs' production model reads briefs up to 32,000 tokens, so brand rules, colour values, and compositional constraints can all go in at once.
Black Forest Labs