Wiki articles, quality evaluations, and an AI assistant that can recommend the right model for your next creation.
Sorted by most recently published

CRAISEE · Jun 24, 2026
OpenAI's state-of-the-art image generation model, excelling at prompt adherence, crisp text rendering, and precise editing capabilities.

CRAISEE · Jun 24, 2026
A complete guide to Seedance 2.0: ByteDance's multimodal AI video model — covering architecture, core features, and practical prompts all in one place.

CRAISEE · Jun 24, 2026
Seedance 2.0 완전 가이드: ByteDance의 멀티모달 AI 영상 모델 아키텍처, 핵심 기능, 실전 프롬프트까지 한 번에 정리했다.

CRAISEE · Jun 23, 2026
Alibaba's Happy Horse 1.1 generates videos from text, animates a single image, or builds a video from multiple reference images. Supports 720p and 1080p, 3-15 second durations, and five aspect ratios.

CRAISEE · Jun 23, 2026
xAI의 grok-imagine-video-1.5는 단일 추론 패스에서 동기화된 영상과 오디오를 생성하고, 리더보드 1위를 차지하며, 경쟁사 대비 저렴한 가격을 자랑합니다. 전체 개발자 가이드를 확인하세요.

CRAISEE · Jun 23, 2026
Kuaishou의 Kling v3 (Video 3.0): 네이티브 4K, 60fps, 멀티샷 컷, 다국어 오디오. 스펙, 가격, 프롬프트, Runway Gen-4 및 Veo 3.1 비교까지 총망라한 완전 위키.

CRAISEE · May 20, 2026
OpenAI의 최첨단 이미지 생성 모델. 뛰어난 프롬프트 준수 능력, 선명한 텍스트 렌더링, 그리고 정밀한 편집 기능이 강점인 모델

CRAISEE · Apr 25, 2026
Kling v3 (Video 3.0) by Kuaishou: native 4K, 60fps, multi-shot cuts, multilingual audio. Full wiki covering specs, pricing, prompts & comparisons with Runway Gen-4 and Veo 3.1.

CRAISEE · Sep 4, 2026
Camera-aware edits for Qwen/Qwen-Image-Edit-2509 with Lightning + multi-angle LoRA

CRAISEE · Sep 4, 2026
A next-generation image generation and editing model from Alibaba's Qwen team. Supports text-to-image and image editing with strong text rendering, especially for Chinese.
CRAISEE · Sep 4, 2026
Fast video generation with text-to-video, image-to-video, and start-end-to-video modes. Up to 16 seconds at 1080p with synchronized audio.
CRAISEE · Sep 4, 2026
PixVerse's flagship video generation model. Generate cinematic videos with synchronized audio, multi-shot sequences, and precise camera control.
CRAISEE · Sep 4, 2026
High-fidelity video generation with text-to-video, image-to-video, and start-end-to-video modes. Up to 16 seconds at 1080p with synchronized audio.

CRAISEE · Sep 4, 2026
Latest video model from Pixverse with astonishing physics
CRAISEE · Sep 4, 2026
p-video-animate animates a reference image with the motion and audio of a source video. Optimized for speed and cost — 5.24s per 1s of video.
CRAISEE · Sep 4, 2026
p-video-replace swaps the person in a video with one from a reference image, keeping motion, timing, camera, and scene exactly as they were. 3.58s per 1s of video generated.
CRAISEE · Sep 4, 2026
p-video-avatar is the fastest and cheapest avatar/lipsync video model on the market.

CRAISEE · Sep 4, 2026
Fast video generation with built-in draft mode for rapid creative iteration. Text-to-video, image-to-video, and audio-to-video in a single endpoint.

CRAISEE · Sep 4, 2026
P-Image-Ideogram is Pruna AI’s text-to-image model starting from 0,003$ per generation.

CRAISEE · Sep 4, 2026
A sub 1 second 0.01$ multi-image editing model built for production use cases. For image generation, check out p-image here: https://replicate.com/prunaai/p-image

CRAISEE · Sep 4, 2026
A sub 1 second text-to-image model built for production use cases.
CRAISEE · Sep 4, 2026
A film-grade digital human model that generates realistic video from a single image, audio clip, and optional text prompt.
CRAISEE · Sep 4, 2026
Ovi: generate videos with audio from image and text inputs

CRAISEE · Sep 4, 2026
OpenAI's fast, lightweight reasoning model

CRAISEE · Sep 4, 2026
A small model alternative to o1

CRAISEE · Sep 4, 2026
Advanced reasoning model

CRAISEE · Sep 4, 2026
Google's state of the art image generation and editing model 🍌🍌

CRAISEE · Sep 4, 2026
Google's latest image editing model in Gemini 2.5

CRAISEE · Sep 4, 2026
OpenAI's first o-series reasoning model

CRAISEE · Sep 4, 2026
Reimagine any song in a different style — change voice, instruments, genre, and arrangement while keeping the original melody

CRAISEE · Sep 4, 2026
Generate full-length songs with vocals, lyrics, and rich instrumentation from a text prompt

CRAISEE · Sep 4, 2026
Generate full-length songs or instrumentals from a text prompt, with optional auto-generated lyrics

CRAISEE · Sep 4, 2026
Compose a song from a prompt or a composition plan

CRAISEE · Sep 4, 2026
FLUX Kontext max with list input for multiple images

CRAISEE · Sep 4, 2026
Create 5s 480p videos from a text prompt

CRAISEE · Sep 4, 2026
Generate 30-second music clips from text prompts or images with Lyria 3, Google's music generation model

CRAISEE · Sep 4, 2026
Modify a video with style transfer and prompt-based editing

CRAISEE · Sep 4, 2026
Generate full-length songs up to 3 minutes from text prompts or images with Lyria 3 Pro, Google's most capable music generation model

CRAISEE · Sep 4, 2026
Fast video generation with text-to-video and image-to-video, portrait and landscape support, synchronized audio, and frame interpolation. Up to 20 seconds at 1080p, and 4K resolution.

CRAISEE · Sep 4, 2026
Lyria 2 is a music generation model that produces 48kHz stereo audio through text-based prompts
CRAISEE · Sep 4, 2026
Edit and transform videos with text prompts and reference images. Style transfers, object replacement, character transformation, and more.

CRAISEE · Sep 4, 2026
High-fidelity video generation with portrait support, audio-to-video, retake, and extend. Text, image, and audio-driven creation up to 4K at 50 FPS.

CRAISEE · Sep 4, 2026
Foundation image model from Krea, tuned for expressive illustration, anime, and painterly styles. Fast and consistent across artistic directions.

CRAISEE · Sep 4, 2026
Kling Video 3.0: Generate cinematic videos up to 15 seconds with multi-shot control, native audio, and improved consistency

CRAISEE · Sep 4, 2026
Krea's flagship foundation image model. Larger and more flexible than Krea 2 Medium, with particular strength in photorealism and expressive artistic styles.

CRAISEE · Sep 4, 2026
Moonshot AI's frontier open model, built for long-horizon coding, agent swarms, and autonomous software engineering. 1 trillion parameters, 262k context window, vision and tool use.

CRAISEE · Sep 4, 2026
The highest quality Ideogram v4 model. v4 creates images with stunning realism, creative designs, and consistent styles

CRAISEE · Sep 4, 2026
Balance speed, quality and cost. Ideogram v4 creates images with stunning realism, creative designs, and consistent styles

CRAISEE · Sep 4, 2026
Alibaba's Happy Horse 1.0 generates videos from text prompts or animates a single image into video. Supports 720p and 1080p, 3-15 second durations, and five aspect ratios.

CRAISEE · Sep 4, 2026
Convert text to natural-sounding speech with xAI's Grok TTS. 5 voices, 20 languages, expressive speech tags, and high-fidelity MP3 / WAV / telephony audio output.

CRAISEE · Sep 4, 2026
Transcribe audio to text with xAI's Grok. Handles 25 languages, word-level timestamps, speaker diarization, multichannel audio, and files up to 500 MB.
CRAISEE · Sep 4, 2026
Image-to-video with synchronized audio using xAI's Grok Imagine Video 1.5 preview model

CRAISEE · Sep 4, 2026
Granite Vision 4.1 4B is a vision-language model (VLM) that delivers frontier-level performance on structured document extraction tasks — chart extraction, table extraction, and semantic key-value pair extraction — in a compact 4B parameter footprint

CRAISEE · Sep 4, 2026
xAI's higher-quality image model with sharper details, better text rendering, and 2k output

CRAISEE · Sep 4, 2026
xAI's Grok Imagine Image 2.0 — text-to-image generation and editing with a quality control and output up to 2k

CRAISEE · Sep 4, 2026
Granite Speech 4.1 2B is a compact and efficient speech-language model, specifically designed for multilingual automatic speech recognition (ASR) and bidirectional automatic speech translation (AST) for English, French, German, Spanish, Portuguese and Jap

CRAISEE · Sep 4, 2026
Granite-embedding-small-english-r2 is a 47M parameter dense biencoder embedding model from the Granite Embeddings collection that can be used to generate high quality text embeddings.

CRAISEE · Sep 4, 2026
Granite-4.2-8B is the mid-size reasoning model in the Granite 4.2 family. It delivers strong performance on reasoning-intensive tasks by leveraging built-in <think>...</think> chain-of-thought.

CRAISEE · Sep 4, 2026
Granite-4.1-8B is a 8B parameter long-context instruct model finetuned from Granite-4.1-8B-Base using a combination of open source instruction datasets with permissive license and internally collected synthetic datasets.

CRAISEE · Sep 4, 2026
OpenAI's GPT-5.6 balanced tier, tuned for everyday production work at roughly half the cost of the flagship.

CRAISEE · Sep 4, 2026
OpenAI's GPT-5.6 flagship tier, built for complex professional work, coding, and deep multi-step reasoning.

CRAISEE · Sep 4, 2026
OpenAI's GPT-5.6 cost-optimized tier, built for fast, high-volume, latency-sensitive workloads.

CRAISEE · Sep 4, 2026
Google's fast multimodal model with frontier reasoning across agents, coding, and long-context tasks

CRAISEE · Sep 4, 2026
Google's fast multimodal video generation and editing model with native audio, using the Interactions API

CRAISEE · Sep 4, 2026
Upscale videos to higher resolution with FLUX super-resolution. Precise mode sharpens and stays faithful to the source; creative mode restores and invents fine detail.

CRAISEE · Sep 4, 2026
Translate audio and video into 90+ languages while preserving each speaker's voice, emotion, and timing

CRAISEE · Sep 4, 2026
Rig any 3D bipedal character mesh

CRAISEE · Sep 4, 2026
Anthropic's most agentic Sonnet model, bringing frontier-level coding and tool use at Sonnet's speed and price

CRAISEE · Sep 4, 2026
Claude Fable 5 from Anthropic: the next generation of intelligence for the hardest knowledge work and coding problems.

CRAISEE · Sep 4, 2026
Remove backgrounds from images.

CRAISEE · Sep 4, 2026
Wan 3.0 Video Prime is Alibaba’s high-speed, all-in-one AI video generation model for creating polished clips from text, images, video, and audio references. It produces videos up to 30 seconds long with synchronized dialogue, music, and sound effects, while accelerated generation makes it ideal for rapid creative iteration and production workflows.
CRAISEE · Sep 4, 2026
Convert raster images (PNG, JPEG, WebP) into clean SVGs with Quiver's Arrow 1.1 vectorization model. Great for turning logos and icons into editable vectors.
CRAISEE · Sep 4, 2026
Convert complex raster images into high-fidelity SVGs with Quiver's Arrow 1.1 Max vectorization model. Tuned for detailed images where fine structure and alignment matter.

CRAISEE · Sep 4, 2026
Veo 3.1 Fast is a specialized, high-speed variant of Google DeepMind’s Veo 3.1 text-to-video model, optimized for rapid generation of 8-second, high-fidelity videos. It is designed to create cinematic, 1080p, or 720p content with improved prompt adherence and native audio, making it ideal for creating quick, high-quality video clips, social media content, and ad creatives.

CRAISEE · Sep 4, 2026
All-in-one video generation model supporting text-to-video, image-to-video, first/last-frame, and omni-modal reference-based generation (image, video, and audio references) with synchronized audio, at up to 30 seconds per clip.

CRAISEE · Sep 4, 2026
Veo 3.1 Lite Preview is a high-efficiency, developer-first video model providing high-fidelity video generation, editing, and cinematic control. It leverages the state-of-the-art Veo 3.1 model to democratize professional-grade video AI by offering a scalable, programmable interface for creators and enterprises.

CRAISEE · Sep 4, 2026
Veo 3.1 is Google's state-of-the-art model for generating high-fidelity, 8-second 720p, 1080p or 4k videos featuring stunning realism and natively generated audio.

CRAISEE · Sep 4, 2026
ByteDance-Seedream-5.0-lite is the latest image generation model released by BytePlus. For the first time, it introduces web-connected retrieval, enabling the model to fuse real-time online information to significantly improve the timeliness and relevance of generated images. The model’s reasoning and comprehension capabilities are further upgraded, allowing it to accurately interpret complex prompts and visual inputs. In addition, ByteDance-Seedream-5.0-lite delivers notable improvements in global knowledge coverage, reference consistency, and professional-grade scene generation, making it well suited for enterprise-level visual creation workflows.

CRAISEE · Sep 4, 2026
Seedream-5.0-Pro, ByteDance's newest image generation model, delivers comprehensive upgrades for complex, lifelike image creation and editing, ushering in a new phase of controllable visual production. It stands out with precise editing control, robust commercial applicability and natural rendering results.

CRAISEE · Sep 4, 2026
V4.1 is built on the same visual aesthetic and the same eye for what looks right, just pushed further across every dimension. The photorealism is more natural. The gradients are dreamier. Illustration styles are now possible that simply weren't before. And this is a model that reads a short prompts easier and creates something worth stopping for. V4.1 Utility is designed for when restraint is the aesthetic choice, with flat lighting, front-facing composition, and simple, controlled scenes.

CRAISEE · Sep 4, 2026
V4.1 is built on the same visual aesthetic and the same eye for what looks right, just pushed further across every dimension. The photorealism is more natural. The gradients are dreamier. Illustration styles are now possible that simply weren't before. And this is a model that reads a short prompts easier and creates something worth stopping for. V4.1 Utility is designed for when restraint is the aesthetic choice, with flat lighting, front-facing composition, and simple, controlled scenes.

CRAISEE · Sep 4, 2026
ByteDance's flagship image editing model. Edit and compose with up to 10 reference images — product swaps, logo placement, multi-image fusion. Served via fal.ai.

CRAISEE · Sep 4, 2026
V4.1 is built on the same visual aesthetic and the same eye for what looks right, just pushed further across every dimension. The photorealism is more natural. The gradients are dreamier. Illustration styles are now possible that simply weren't before. And this is a model that reads a short prompts easier and creates something worth stopping for. V4.1 Pro generates higher-resolution images for when the idea deserves more room.

CRAISEE · Sep 4, 2026
V3 introduced major advances in photorealism and text rendering. It was the first Recraft model to generate mid-size text accurately and, as of 2025, is the only model capable of placing text at specific positions in an image.

CRAISEE · Sep 4, 2026
V4.1 is built on the same visual aesthetic and the same eye for what looks right, just pushed further across every dimension. The photorealism is more natural. The gradients are dreamier. Illustration styles are now possible that simply weren't before. And this is a model that reads a short prompts easier and creates something worth stopping for.

CRAISEE · Sep 4, 2026
MiniMax H3 Max is a next-generation general-purpose multimodal video model ranking #1 for overall quality, prompt understanding, and aesthetics in all evaluations, while generating a 5-second video in under 3 seconds.

CRAISEE · Sep 4, 2026
Muse Image is the first image generation model from Meta Superintelligence Labs, it uses advanced reasoning to understand complex prompts, seamlessly blending multiple photos into high-quality creations you can download and share anywhere.

CRAISEE · Sep 4, 2026
H3 is a next-generation open-weights, general-purpose multimodal video model. Rather than being limited to specialized tasks such as generating, editing, or referencing, H3 understands multimodal contexts that bring together text, images, video, and audio. This enables it to interpret creative intent in a unified way and deliver more natural, coherent generation and expression.

CRAISEE · Sep 4, 2026
Build upon an All-in-One product framework, the Kling 3.0 model series supports full multimodal input and output spanning text, images, audio, and video, bringing the understanding, generation, and editing of video together in one streamlined AI workflow. The models integrate multiple tasks, including text-to-video, image-to-video, reference-to-video, and in-video editing, into a single, native multimodal architecture, enabling the models to follow complex narrative logic, deliver precise shot control, and maintain strong prompt adherence.

CRAISEE · Sep 4, 2026
Kling 3.0 delivers a major leap in character fidelity for motion-driven generation, with stable facial features across multi-angle and long-duration motion, accurate complex emotions from multi-image face references, identity preservation through partial occlusions (hats, hands, fans), and steady clarity as the camera zooms, pans, or tracks.

CRAISEE · Sep 4, 2026
Build upon an All-in-One product framework, the Kling 3.0 model series supports full multimodal input and output spanning text, images, audio, and video, bringing the understanding, generation, and editing of video together in one streamlined AI workflow. The models integrate multiple tasks, including text-to-video, image-to-video, reference-to-video, and in-video editing, into a single, native multimodal architecture, enabling the models to follow complex narrative logic, deliver precise shot control, and maintain strong prompt adherence.

CRAISEE · Sep 4, 2026
Kling 2.6 introduces a groundbreaking "Native Audio" capability, enabling the generation of complete videos in a single go, including natural voice, action sound effects, and environmental ambient sounds, providing an immersive "what you see if what you hear" experience.

CRAISEE · Sep 4, 2026
Kling 2.6 introduces a groundbreaking "Native Audio" capability, enabling the generation of complete videos in a single go, including natural voice, action sound effects, and environmental ambient sounds, providing an immersive "what you see if what you hear" experience.

CRAISEE · Sep 4, 2026
Kling 2.5 Turbo is a major update to the AI video generation model focused on significantly improving speed, video quality, temporal stability, and creative control for creators, making professional-grade AI-generated video faster, more coherent, and easier to direct from text prompts.

CRAISEE · Sep 4, 2026
Kling 2.6 introduces a groundbreaking "Native Audio" capability, enabling the generation of complete videos in a single go, including natural voice, action sound effects, and environmental ambient sounds, providing an immersive "what you see if what you hear" experience.

CRAISEE · Sep 4, 2026
Kling 2.5 Turbo is a major update to the AI video generation model focused on significantly improving speed, video quality, temporal stability, and creative control for creators, making professional-grade AI-generated video faster, more coherent, and easier to direct from text prompts.

CRAISEE · Sep 4, 2026
Imagen 4 Fast is Google’s speed-optimized variant of the Imagen 4 text-to-image model, designed for rapid, high-volume image generation. It’s ideal for workflows like quick drafts, mockups, and iterative creative exploration. Despite emphasizing speed, it still benefits from the broader Imagen 4 family’s improvements in clarity, text rendering, and stylistic flexibility, and supports high-resolution outputs up to 2K.

CRAISEE · Sep 4, 2026
Imagen 4 Ultra: Highest quality image generation model for detailed and photorealistic outputs.

CRAISEE · Sep 4, 2026
Imagen 4: Google's flagship text-to-image model that serves as the go-to choice for a wide variety of high-quality image generation tasks, featuring significant improvements in text rendering over previous models. It now supports up to 2K resolution generation for creating detailed and crisp visuals, making it suitable for everything from marketing assets to artistic compositions.

CRAISEE · Sep 4, 2026
Complete guide to using Grok Imagine Video 1.5 Preview on CRAISEE.

CRAISEE · Sep 4, 2026
Complete guide to using Grok Imagine Video 1.5 Preview on CRAISEE.

CRAISEE · Sep 4, 2026
Complete guide to using Grok Imagine Video 1.5 on CRAISEE.

CRAISEE · Sep 4, 2026
Complete guide to using Grok Imagine Image 2.0 on CRAISEE.

CRAISEE · Sep 4, 2026
Complete guide to using Grok Imagine Image 2.0 Preview on CRAISEE.

CRAISEE · Sep 4, 2026
Complete guide to using Grok Imagine Image 2.0 on CRAISEE.

CRAISEE · Sep 4, 2026
Generate high-quality images from text prompts with xAI's imagine API.

CRAISEE · Sep 4, 2026
State-of-the-art video generation across quality, cost, and latency. Grok Imagine is x.AI's most powerful video-audio generative model yet. Bring an image to life, start from a simple text prompt, or even refine a complex cinematic sequence.

CRAISEE · Sep 4, 2026
FLUX.2 [dev] with custom LoRA support — apply HuggingFace or URL-hosted style adapters (up to 3, 5GB total). Hosted via fal.ai.

CRAISEE · Sep 4, 2026
Newest frontier OpenAI model for complex professional work

CRAISEE · Sep 4, 2026
Fast Gemini 3.1 model
CRAISEE · Sep 4, 2026
Generate video with synchronized audio from text, images, or video. FLUX 3 is Black Forest Labs' multimodal model (early access preview).

CRAISEE · Jul 7, 2026
Nano Banana 2 is a high-speed image generation and editing model developed by Google, built on Gemini 3.1 Flash Image, and officially announced on February 26, 2026. Its defining strength is delivering professional-grade visual quality alongside Flash-level speed and cost efficiency. The model is optimized for both creators who need advanced capabilities such as conversational image editing, multi-image fusion, and character consistency, as well as developers running high-volume image generation workflows.

CRAISEE · Jul 7, 2026
Nano Banana 2 Lite is the fastest and most affordable image generation model on CRAISEE, built on Google's Gemini 3.1 Flash-Lite Image. Optimized for rapid prototyping, bulk image generation, and cost-efficient creative work, it offers two core capabilities: generating images from text prompts alone, and editing existing images. This model is especially recommended for creators who need quick results, high-volume content producers, and anyone in the early ideation or sketching phase of a project.

CRAISEE · Jul 7, 2026
Nano Banana 2는 Google이 Gemini 3.1 Flash Image 기반으로 개발한 고속 이미지 생성 및 편집 모델로, 2026년 2월 26일 공식 발표되었습니다. 프로급 시각적 품질을 유지하면서도 Flash 수준의 속도와 비용 효율성을 동시에 제공하는 것이 핵심 특징입니다. 대화형 이미지 편집, 다중 이미지 융합, 캐릭터 일관성 유지 등 고급 기능을 필요로 하는 크리에이터와 대용량 이미지 생성 워크플로우를 운영하는 개발자 모두에게 최적화된 모델입니다.

CRAISEE · Jul 7, 2026
Nano Banana 2 Lite는 Google의 Gemini 3.1 Flash-Lite Image 기반으로 구축된 가장 빠르고 저렴한 이미지 생성 모델입니다. 빠른 프로토타이핑, 대량 이미지 생성, 비용 효율적인 창작 작업에 최적화된 이 모델은 텍스트 프롬프트만으로 이미지를 생성하거나, 기존 이미지를 편집하는 두 가지 핵심 기능을 제공합니다. CRAISEE에서 빠른 결과물이 필요한 크리에이터, 대량 콘텐츠 제작자, 아이디어 스케치 단계의 작업자에게 특히 권장되는 모델입니다.

CRAISEE · Jul 3, 2026
HappyHorse-1.1 완전 가이드: 아키텍처·벤치마크·프롬프트 작성법까지 한눈에 정리. 블라인드 평가 1위 AI 비디오 모델의 모든 것.

CRAISEE · Jun 25, 2026
The complete guide to Seedance 2.0 Mini: specs, pricing, and prompt patterns for ByteDance's lightweight AI video model at a glance. Outperforms the Fast tier at 50% of the standard cost.

CRAISEE · Jun 25, 2026
Seedance 2.0 Mini 완전 가이드: ByteDance 경량 AI 비디오 모델의 스펙·가격·프롬프트 패턴을 한눈에 정리. 표준 대비 50% 비용으로 Fast 티어를 능가하는 품질.

CRAISEE · Jun 23, 2026
Anthropic의 빠른 텍스트 모델

CRAISEE · Jun 23, 2026
최신 균형형 MiniMax 모델

CRAISEE · Jun 23, 2026
Anthropic의 이전 세대 플래그십 모델

CRAISEE · Jun 23, 2026
저지연 Grok 변형 모델

CRAISEE · Jun 23, 2026
신뢰할 수 있는 Sonnet 세대 모델

CRAISEE · Jun 23, 2026
최첨단 이미지 생성 모델

CRAISEE · Jun 23, 2026
OpenAI 코딩 특화 모델

CRAISEE · Jun 23, 2026
Eleven v3 (eleven_v3): 기능, Audio Tags 문법, 멀티 스피커 API, 가격, 그리고 ElevenLabs의 가장 표현력 높은 TTS 모델을 위한 검증된 프롬프트.

CRAISEE · Jun 23, 2026
최고 수준의 추론 능력과 다국어 지원

CRAISEE · Jun 23, 2026
최신 DeepSeek 텍스트 모델

CRAISEE · Jun 23, 2026
Z.ai의 최신 범용 모델

CRAISEE · Jun 23, 2026
MiniMax의 최신 플래그십 모델

CRAISEE · Jun 23, 2026
복잡한 작업을 위한 OpenAI의 가장 강력한 모델

CRAISEE · Jun 23, 2026
가장 큰 오픈 웨이트 모델

CRAISEE · Jun 23, 2026
균형 잡힌 오픈 모델

CRAISEE · Jun 23, 2026
GPT-5.4 시리즈 중 가장 경량화된 모델

CRAISEE · Jun 23, 2026
고성능 Gemini 텍스트 모델

CRAISEE · Jun 23, 2026
일상적인 작업을 위한 빠르고 효율적인 모델

CRAISEE · Jun 23, 2026
PixVerse V6의 주요 기능과 실전 활용법, 프롬프트 작성 팁까지 한눈에 정리했습니다. 15초 1080p 생성, 멀티샷, 오디오 기능이 콘텐츠 제작 현장에서 어떤 차이를 만드는지 확인해보세요.

CRAISEE · Jun 23, 2026
코드 생성에 최적화된 모델

CRAISEE · Jun 23, 2026
Anthropic의 최신 균형형 모델

CRAISEE · Jun 23, 2026
추론 기능이 탑재된 Grok 변형 모델

CRAISEE · Jun 23, 2026
에이전틱 코딩의 획기적인 성능 향상, 향상된 비전, 강력한 다단계 추론을 갖춘 Anthropic의 가장 강력한 모델

CRAISEE · Jun 23, 2026
Gemini 3.1 Flash Lite Preview 완벽 분석: 사양, 가격($0.25/백만 토큰), 100만 토큰 컨텍스트, 296 t/s 속도, Google의 가장 빠른 저비용 AI 모델 사용법.

CRAISEE · Jun 23, 2026
전문가급 영상 생성

CRAISEE · Jun 23, 2026
저비용 Gemini 2.5 텍스트 모델

CRAISEE · Jun 23, 2026
OpenAI의 최신 고속 모델

CRAISEE · Jun 23, 2026
lipsync-speed의 기능, 작동 원리, 활용 시점을 알아보세요. 빠른 더빙 워크플로우를 위한 최적 프롬프트, 싱크 모드, 파이프라인 팁을 포함합니다.

CRAISEE · Jun 23, 2026
heygen/lipsync-precision 완전 정복: HeyGen의 최고 품질 AI 립싱크 모델의 아키텍처 심층 분석, 입력 파라미터, Replicate 배포 팁 (2025).

CRAISEE · Jun 23, 2026
지원 종료: 이 모델은 xai/grok-imagine-image로 대체되었습니다. 업스트림 grok-2-image-1212 모델은 2026년 2월 24일 xAI에 의해 지원이 종료되었습니다.
CRAISEE · Jun 23, 2026
인물, 얼굴, 제품에 최적화된 고정밀 영상 업스케일러. Clarity AI가 구동하는 업스케일 모드 중 하나입니다. X: https://x.com/philz1337x

CRAISEE · Jun 23, 2026
Kimi K2 Thinking: Moonshot AI가 개발한 1조 파라미터 오픈 웨이트 추론 모델. 아키텍처, 벤치마크, 가격, 제한 사항, 에이전틱 AI 프롬프트 패턴을 다룹니다.
CRAISEE · Jun 23, 2026
xAI의 Grok Imagine Video 모델로 영상을 확장하세요. 소스 영상을 제공하고 다음에 일어날 내용을 설명하면 됩니다.

CRAISEE · Jun 23, 2026
단안 메트릭 깊이 추정
CRAISEE · Jun 23, 2026
Kling O1은 하나의 엔진 안에 영상 생성, 편집, 정제 기능을 통합합니다. 아키텍처, 기능, 출력 사양, 전문가 프롬프팅 팁을 알아보세요.
CRAISEE · Jun 23, 2026
단일 이미지와 구동 영상만으로 인간, 만화 캐릭터, 동물, 비인간 캐릭터까지 모든 캐릭터를 애니메이션으로 만들어보세요

CRAISEE · Jun 23, 2026
Granite-speech-3.3-8b는 자동 음성 인식(ASR)과 자동 음성 번역(AST)을 위해 특별히 설계된 경량 고효율 음성-언어 모델입니다.

CRAISEE · Jun 23, 2026
kling-avatar-v2는 단 하나의 이미지를 1080p/48fps의 립싱크 영상으로 변환합니다. 스펙, 가격, 프롬프트 팁, 개발자 및 크리에이터를 위한 API 활용법까지 총정리한 완벽 가이드입니다.

CRAISEE · Jun 23, 2026
텍스트 프롬프트로 고품질 2K 해상도 이미지 생성

CRAISEE · Jun 23, 2026
GPT-4o mini를 활용하여 오디오를 텍스트로 변환하는 음성 인식 모델

CRAISEE · Jun 23, 2026
ltx-2.3-fast: Lightricks의 220억 파라미터 증류 DiT 모델로 8단계 만에 최대 20초 분량의 4K/50FPS 클립을 생성합니다. 배포 방법, 프롬프트 작성법, 벤치마크까지 모두 수록.
CRAISEE · Jun 23, 2026
Sync Labs의 lipsync-2 작동 방식을 알아보세요: 제로샷 립싱크, 청크 기반 아키텍처, 초당 $0.04 가격, 최적 프롬프트, 그리고 sync-3로 업그레이드해야 할 시점까지 다룹니다.
CRAISEE · Jun 23, 2026
Sync Labs의 lipsync-2-pro: 4K 디퓨전 기반 AI 립싱크 모델. 아키텍처, API 통합, 프로덕션 팀을 위한 배포 방법까지 2025–2026 완전 가이드.
CRAISEE · Jun 23, 2026
Kling Lip Sync 완전 참조 가이드: 아키텍처, 제로샷 화자 적응, 입력 형식, 비용, 그리고 프로덕션 수준의 얼굴 애니메이션을 위한 프롬프트 공식까지 총망라합니다.

CRAISEE · Jun 23, 2026
Kimi K2.5: Moonshot AI의 1조 파라미터 희소 MoE 모델. 아키텍처, 벤치마크, 14개 제공업체의 가격 정책, 모드 선택 팁을 이 기술 가이드에서 확인하세요.

CRAISEE · Jun 23, 2026
ElevenLabs의 가장 빠른 음성 합성 모델
CRAISEE · Jun 23, 2026
텍스처 충실도와 지오메트리 정밀도를 갖춘 3D 모델
CRAISEE · Jun 23, 2026
xAI의 Grok Imagine Video 모델을 사용하여 참조 이미지 기반의 영상을 생성하세요

CRAISEE · Jun 23, 2026
GPT-4o를 활용해 오디오를 텍스트로 변환하는 음성 인식 모델

CRAISEE · Jun 23, 2026
Bria Background Generation은 텍스트 프롬프트 또는 참조 이미지를 통해 이미지의 배경을 효율적으로 교체할 수 있으며, 사실적이고 완성도 높은 결과물을 제공합니다. 라이선스가 부여된 데이터만으로 학습되어 상업적 사용에 안전하고 위험 부담이 없습니다.

CRAISEE · Jun 23, 2026
이미지 생성을 위한 강력한 네이티브 멀티모달 모델 (PrunaAI 경량화 버전)

CRAISEE · Jun 23, 2026
Flux 2를 1초로 단축한 스텝 증류 버전입니다.
CRAISEE · Jun 23, 2026
사실적인 인체 동작, 시네마틱 VFX, 표현력 있는 캐릭터, 그리고 텍스트-투-비디오 및 이미지-투-비디오 워크플로우 전반에 걸친 강력한 프롬프트·스타일 충실도를 위해 최적화된 고품질 영상 생성 모델

CRAISEE · Jun 23, 2026
다양한 활용 사례를 시작하는 데 도움이 되는 레시피 워크북을 제공하는 Granite 커뮤니티에 참여하세요. https://github.com/ibm-granite-community

CRAISEE · Jun 23, 2026
정밀한 스타일 제어, 향상된 텍스트 렌더링, 상업 디자인 최적화를 갖춘 4MP 텍스트-이미지 생성 모델로, 영화적 품질의 이미지를 생성합니다.

CRAISEE · Jun 23, 2026
OpenAI의 최신 플래그십 모델

CRAISEE · Jun 23, 2026
Anthropic의 최신 플래그십 모델

CRAISEE · Jun 23, 2026
30가지 음성과 70개 이상의 언어를 지원하는 Google의 빠르고 표현력 있는 텍스트 음성 변환 모델

CRAISEE · Jun 23, 2026
Hailuo 2.3의 저지연 이미지-투-비디오 버전으로, 핵심 모션 품질, 시각적 일관성, 스타일화 성능을 유지하면서 더 빠른 반복 작업을 가능하게 합니다.

CRAISEE · Jun 23, 2026
Granite-4.0-H-Small은 오픈소스 명령어 데이터셋과 내부 합성 데이터셋을 결합하여 Granite-4.0-H-Small-Base로부터 파인튜닝된 320억 파라미터 규모의 장문 컨텍스트 명령어 수행 모델입니다.

CRAISEE · Jun 23, 2026
OpenAI의 120b 오픈 웨이트 언어 모델
CRAISEE · Jun 23, 2026
최첨단 영상 모션 품질, 프롬프트 충실도 및 시각적 완성도

CRAISEE · Jun 23, 2026
최고 품질의 이미지 생성 및 편집, 최대 10개의 참조 이미지 지원

CRAISEE · Jun 23, 2026
전문적인 깊이 인식 이미지 생성. 공간적 관계를 유지하면서 이미지를 편집하세요.

CRAISEE · Jun 23, 2026
Bria Expand는 이미지를 고품질로 경계 너머까지 확장합니다. 새로운 픽셀을 생성하여 원하는 종횡비로 이미지를 리사이징합니다. 안전하고 위험 없는 상업적 사용을 위해 라이선스 데이터만으로 학습되었습니다.

CRAISEE · Jun 23, 2026
Kling v3 Motion Control (kling-v3-motion-control): 레퍼런스 영상 기반 모션 전이에서 ELO 1위를 기록한 AI 영상 모델. 설정, 벤치마크, 배포 팁까지 담은 완전 기술 가이드.
CRAISEE · Jun 23, 2026
sync-3, VEED Fabric, MuseTalk 등 주요 AI 립싱크 모델을 정확한 가격, 실패 사례, 입력 팁과 함께 비교합니다. 더빙, 아바타, 실시간 에이전트에 최적화된 선택을 도와드립니다.

CRAISEE · Jun 23, 2026
속도, 품질, 비용의 균형을 맞추세요. Ideogram v3는 놀라운 사실감, 창의적인 디자인, 일관된 스타일의 이미지를 생성합니다

CRAISEE · Jun 23, 2026
최첨단 객체 제거 기술로, 이미지에서 불필요한 객체를 정밀하게 제거하면서도 높은 품질의 결과물을 유지합니다. 안전하고 위험 없는 상업적 사용을 위해 라이선스 데이터만으로 학습되었습니다.

CRAISEE · Jun 23, 2026
Deepseek의 최신 하이브리드 사고 모델

CRAISEE · Jun 23, 2026
단일 참조 이미지로 일관된 캐릭터를 생성합니다. 다양한 스타일로 출력할 수 있으며, 인페인팅을 활용해 기존 이미지에 캐릭터를 추가할 수도 있습니다.

CRAISEE · Jun 23, 2026
kling-v3-omni-video(Kling 3.0 Omni) 완벽 가이드: 아키텍처, 벤치마크, 입력 모드, 2026년 최고 평점 AI 영상 모델의 API 활용법까지 총정리.

CRAISEE · Jun 23, 2026
OpenAI의 고성능 채팅 모델

CRAISEE · Jun 23, 2026
Kling v2.6은 영상, 음성, 오디오를 한 번의 생성으로 만들어냅니다. 아키텍처, 가격, 프롬프트 팁, 경쟁 모델 비교까지 알아보고 CRAISEE에서 직접 사용해보세요.

CRAISEE · Jun 23, 2026
Google의 가장 지능적인 모델로, 향상된 추론 능력과 새로운 중간 사고 수준을 제공합니다
CRAISEE · Jun 23, 2026
xAI의 Grok Imagine Video 모델로 영상을 생성하세요

CRAISEE · Jun 23, 2026
Gemini 2.5의 Google 최신 이미지 생성 모델

CRAISEE · Jun 23, 2026
오픈 웨이트 깊이 인식 이미지 생성. 공간적 관계를 유지하면서 이미지를 편집하세요.

CRAISEE · Jun 23, 2026
Google의 가장 진보된 추론 Gemini 모델

CRAISEE · Jun 23, 2026
Imagen 4 Ultra는 Google의 최상위 텍스트-이미지 생성 모델입니다 (GA: 2025년 8월). 아키텍처, 가격, 프롬프트 전략, 경쟁 모델 비교까지 모두 알아보세요.

CRAISEE · Jun 23, 2026
FLUX.2 [klein]의 4단계 증류 버전. 최대한의 유연성과 제어를 위한 기반 모델

CRAISEE · Jun 23, 2026
강화학습으로 훈련된 추론 모델로, OpenAI o1에 필적하는 성능을 자랑합니다

CRAISEE · Jun 23, 2026
Ideogram v3의 최고 품질 모델. v3는 놀라운 사실감, 창의적인 디자인, 일관된 스타일의 이미지를 생성합니다

CRAISEE · Jun 23, 2026
OpenAI의 플래그십 GPT-5 모델의 고속 버전

CRAISEE · Jun 23, 2026
xAI의 최첨단 이미지 모델

CRAISEE · Jun 23, 2026
AI 이미지 업스케일 모델의 작동 원리를 이해하고, GAN과 디퓨전 아키텍처를 비교하며, 4K 품질의 결과물을 위한 최적 프롬프트를 확인하세요 — CRAISEE와 함께.

CRAISEE · Jun 23, 2026
Black Forest Labs의 최고 품질 이미지 생성 모델

CRAISEE · Jun 23, 2026
레퍼런스 이미지를 지원하는 고품질 이미지 생성 및 편집

CRAISEE · Jun 23, 2026
빠른 Gemini 3 텍스트 모델

CRAISEE · Jun 23, 2026
Llama 4 Maverick: 4000억 파라미터 MoE 모델, 활성 파라미터 170억, 100만 토큰 컨텍스트, 멀티모달 지원. 사양·벤치마크·배포 방법 완벽 정리 — CRAISEE에서 직접 사용해보세요.

CRAISEE · Jun 23, 2026
Google DeepMind의 Imagen 4 Fast: 사양, 벤치마크, 가격($0.02/이미지), 그리고 2.7초 텍스트-이미지 모델을 위한 프롬프트 패턴. CRAISEE에서 직접 사용해보세요.

CRAISEE · Jun 23, 2026
속도와 비용 효율성에 최적화된 Google의 하이브리드 '사고형' AI 모델

CRAISEE · Jun 23, 2026
Imagen 4 가이드: 아키텍처, 3단계 가격 책정($0.02–$0.06), 강점, 한계, 그리고 Google DeepMind의 최고 텍스트-이미지 모델을 위한 검증된 프롬프트 패턴.

CRAISEE · Jun 23, 2026
Turbo는 가장 빠르고 저렴한 Ideogram v3입니다. v3는 놀라운 사실감, 창의적인 디자인, 일관된 스타일의 이미지를 생성합니다

CRAISEE · Jun 23, 2026
자연어 프롬프트로 이미지를 변환하는 최고 성능과 향상된 타이포그래피 생성을 제공하는 프리미엄 텍스트 기반 이미지 편집 모델

CRAISEE · Jun 23, 2026
더 향상된 지시 사항 준수와 프롬프트 충실도를 갖춘 OpenAI의 최신 이미지 생성 모델

CRAISEE · Jun 23, 2026
매우 빠른 이미지 생성 및 편집 모델. 4단계 증류 방식으로 서브초 추론을 지원하며, 프로덕션 및 준실시간 애플리케이션에 적합합니다.

CRAISEE · Jun 17, 2026
P Image는 Pruna AI가 개발한 텍스트-투-이미지 생성 모델로, 이미지 1장을 약 1초 만에 생성하는 속도·비용 최적화를 핵심으로 한다. Quantization과 Pruning이라는 두 가지 경량화 기술을 결합해 최첨단 모델 수준의 시각 품질을 대폭 낮은 연산 비용으로 구현한다는 점에서, 매력적인 선택지로 평가된다

CRAISEE · Jun 15, 2026
MiniMax Speech, 추천 Voice ID부터 한국어·영어·일본어·중국어·스페인어 음성 목록까지 총정리

CRAISEE · Jun 11, 2026
A complete overview of PixVerse V6 — from key features and practical usage to prompt writing tips. See how 15-second 1080p generation, multi-shot, and audio features make a real difference in content production.

CRAISEE · Jun 3, 2026
grok-imagine-video-1.5 by xAI generates synchronized video and audio in one pass, tops leaderboards, and undercuts rivals on price. Full developer guide inside.

CRAISEE · May 6, 2026
HappyHorse-1.0 완전 가이드: Alibaba의 15B 파라미터 T2V·I2V 모델 아키텍처, 벤치마크 1위 성과, 프롬프트 공식까지 한눈에 정리.

CRAISEE · Apr 25, 2026
Generate videos using xAI's Grok Imagine Video model

CRAISEE · Apr 25, 2026
FLUX.2 Pro 완전 가이드: 32B Rectified Flow Transformer 아키텍처, 멀티 레퍼런스, CJK 렌더링, 프롬프트 전략까지 실무 적용법 총정리.

CRAISEE · Apr 25, 2026
Anthropic's most capable model with a step-change improvement in agentic coding, better vision, and stronger multi-step reasoning

CRAISEE · Apr 25, 2026
Top-tier reasoning and multilingual

CRAISEE · Apr 25, 2026
Newest MiniMax flagship model

CRAISEE · Apr 25, 2026
Recent balanced MiniMax model

CRAISEE · Apr 25, 2026
Balanced open model

CRAISEE · Apr 25, 2026
Largest open-weight model

CRAISEE · Apr 25, 2026
State-of-the-art image generation

CRAISEE · Apr 25, 2026
Reasoning-enabled Grok variant

CRAISEE · Apr 25, 2026
Low-latency Grok variant

CRAISEE · Apr 25, 2026
Professional video generation

CRAISEE · Apr 25, 2026
Gemini 3.1 Flash Lite Preview: specs, pricing ($0.25/M tokens), 1M-token context, 296 t/s speed, and how to use Google's fastest budget AI model.

CRAISEE · Apr 25, 2026
Low-cost Gemini 2.5 text model

CRAISEE · Apr 25, 2026
Newest fast OpenAI model

CRAISEE · Apr 25, 2026
OpenAI coding-specialized model

CRAISEE · Apr 25, 2026
Fast and efficient model for everyday tasks

CRAISEE · Apr 25, 2026
Most capable OpenAI model for complex tasks

CRAISEE · Apr 25, 2026
Recent general-purpose Z.ai model

CRAISEE · Apr 25, 2026
Eleven v3 (eleven_v3): features, Audio Tags syntax, multi-speaker API, pricing, and proven prompts for ElevenLabs' most expressive TTS model.

CRAISEE · Apr 25, 2026
Latest DeepSeek text model

CRAISEE · Apr 25, 2026
Optimized for code generation

CRAISEE · Apr 25, 2026
Newest balanced Anthropic model

CRAISEE · Apr 25, 2026
Reliable Sonnet generation

CRAISEE · Apr 25, 2026
Previous-gen Anthropic flagship

CRAISEE · Apr 25, 2026
Fast Anthropic text model