google/nano-banana-2
Google 图片 generation with conversational editing, multi-image fusion and character consistency.
Open model文本 and reference-image models for 高级 visual creation.
Google 图片 generation with conversational editing, multi-image fusion and character consistency.
Open modelOpenAI 图片 generation and editing with strong instruction following and sharp 文本 rendering.
Open modelBytedance Seedream 4.5 图片 generation with stronger spatial understanding and world knowledge.
Open modelxAI 图片 generation for expressive, high-detail creative scenes.
Open modelFast Imagen 4 generation when speed and cost matter.
Open modelAnime generation and anime-style 图片 transformation models.
创建 custom emoji-style assets from short prompts.
Restore old photos, damaged 图片 and faces with AI repair models.
创建 textured 3D assets and downloadable model 文件 from prompts or reference 图片.
Prompt-到-video and reference-video generation models.
Generate videos using xAI Grok Imagine 视频.
Open modelSeedance Lite text-到-video and image-到-video generation.
Open modelOpenAI Sora 2 Pro synced-audio 视频 generation.
Open modelA fast optimized Wan 2.2 text-到-video model.
Open modelAlibaba Happy Horse turns prompts or a reference 图片 into cinematic HD 视频 clips with duration and aspect controls.
Open model创建 a talking avatar 视频 from one portrait 图片, a script, selected 语音 and language.
Open modelSwap a source face into a target 视频 clip with a dedicated 视频 face-swap workflow.
Open modelFace-aware 图片 editing, reference and identity-preserving 图片 models.
Super-resolution and crisp 图片 enhancement models.
Full songs, instrumentals and 音频 generation models.
语音 synthesis and 语音 cloning models.
Inworld realtime text-到-speech with preset voices.
Open modelMiniMax low-latency multilingual text-到-audio with emotion control.
Open modelKokoro 82M text-到-speech based on StyleTTS2.
Open modelCoqui XTTS-v2 multilingual text-到-speech 语音 cloning.
Open model视频 understanding, transcript and caption generation models.
聊天, coding, reasoning and multimodal language models.
Google Gemini 3.1 Pro for reasoning, research and high-quality 聊天.
Open modelOpenAI GPT-5.2 for coding, agentic work, reasoning and multimodal tasks.
Open modelGoogle Gemini 3 Pro advanced reasoning with multimodal input support.
Open modelDeepSeek V3.1 hybrid thinking model for coding and technical work.
Open modelQwen3 large instruction model for writing, planning and structured answers.
Open model