Mjegulla

Modele AI

Çdo model kryesor AI. Një platformë.

AI që ndërton sajtin tënd ekzekutohet mbi Claude Opus 5.5 nga Anthropic. Dhe çdo aplikacion që krijon mund të përdorë 160 modele të tjera — për bisedë, imazhe, video, zë, përkthim dhe kërkim — pa asnjë çelës API.

160

modele

34

ofrues

10

aftësi

OpenAI OpenAI
Google Google
Recraft Recraft
xAI xAI
Anthropic Anthropic
Meta Meta
Pruna AI Pruna AI
MiniMax MiniMax
Alibaba Alibaba
Qwen Qwen
ByteDance ByteDance
BAAI BAAI
DeepSeek DeepSeek
Z.ai (GLM) Z.ai (GLM)
Black Forest Labs Black Forest Labs
Deepgram Deepgram
Inworld Inworld
Krea
Runway Runway
PixVerse PixVerse

Si funksionon AI te Mjegulla

Dy shtresa: AI që ndërton për ty, dhe AI që aplikacionet e tua mund të përdorin.

Ndërtuesi

Kur përshkruan një sajt ose një aplikacion, një agjent AI e planifikon, shkruan kodin dhe tekstin, kontrollon pamjen paraprake për gabime dhe rregullon çfarë gjen. Ti zgjedh sa fort mendon për çdo mesazh.

I shpejtë

Redaktime të shpejta dhe ndryshime të vogla

I balancuar

Modeli i duhur i zgjedhur për çdo detyrë

Më e mira

Modeli më i aftë për ndërtime të mëdha

  • Claude Opus 5.5Planifikon dhe ndërton sajte dhe aplikacione të mëdha
  • Claude Sonnet 5.5Shkruan dhe redakton kod
  • Claude Haiku 4.5Kontrolle të shpejta dhe ndryshime të vogla
  • Kimi K2.7Motor alternativ ndërtimi për ekzekutime të gjata agjentike

Modaliteti Plan i lejon agjentit të propozojë një plan para se të ndërtojë.

AI brenda aplikacioneve të tua

Kërkoji AI Studio një chatbot, një gjenerues imazhesh ose një veçori zëri dhe ai lidh modelin e duhur në aplikacionin tënd. Përdorimi paguhet me kredite dhe ti cakton kufij shpenzimi ditorë dhe mujorë.

  1. 1Zgjidh një model nga skeda e modeleve AI në Studio, ose përmende direkt në bisedë.
  2. 2Agjenti e lidh atë — pa çelësa API, pa llogari ofruesish.
  3. 3Shiko çdo thirrje dhe koston e saj në regjistrin e përdorimit AI.

Katalogu i modeleve

Kërko dhe filtro çdo model të disponueshëm për aplikacionet e tua.

160 modele

Claude Sonnet 5

Anthropic

AnthropicText Generation Claude Sonnet 5 is Anthropic's most agentic Sonnet model yet, built for coding, tool use, reasoning, and long-horizon professional work at lower cost than Opus-class models.

@claude-sonnet-5

Claude Sonnet 4.6

Anthropic

AnthropicText Generation Claude Sonnet 4.6 is Anthropic's latest balanced model offering strong coding, reasoning, and agentic capabilities with improved instruction following.

@claude-sonnet-4.6

Claude Sonnet 4.5

Anthropic

AnthropicText Generation Claude Sonnet 4.5 is the best coding model to date, with significant improvements across the entire development lifecycle.

@claude-sonnet-4.5

Claude Opus 4.8

Anthropic

AnthropicText Generation Claude Opus 4.8 is Anthropic's most capable generally available model, with a step-change improvement in agentic coding over Claude Opus 4.7. It uses adaptive thinking to calibrate reasoning per task and supports a one million token context window at standard pricing.

@claude-opus-4.8

Claude Opus 4.7

Anthropic

AnthropicText Generation Claude Opus 4.7 is Anthropic's most capable generally available model, with a step-change improvement in agentic coding over Claude Opus 4.6. It uses adaptive thinking to calibrate reasoning per task and supports a one million token context window at standard pricing.

@claude-opus-4.7

Claude Opus 4.6

Anthropic

AnthropicText Generation Claude Opus 4.6 is Anthropic's flagship language model built for complex, multi-step work in coding, financial analysis, and legal reasoning. It uses extended thinking to work through complex problems carefully and features a one million token context window.

@claude-opus-4.6

Claude Opus 4.5

Anthropic

AnthropicText Generation Claude Opus 4.5 brings further reasoning, coding, and agentic improvements over Opus 4.1, with stronger tool use and tighter instruction following.

@claude-opus-4.5

Claude Haiku 4.5

Anthropic

AnthropicText Generation Claude Haiku 4.5 delivers similar levels of coding performance at one-third the cost and more than twice the speed of larger models.

@claude-haiku-4.5

Whisper Large v3 Turbo

OpenAI

Whisper is a pre-trained model for automatic speech recognition (ASR) and speech translation.

Zë në tekst @whisper-large-v3-turbo

Whisper

OpenAI

Whisper is a general-purpose speech recognition model. It is trained on a large dataset of diverse audio and is also a multitasking model that can perform multilingual speech recognition, speech translation, and language identification.

Zë në tekst @whisper

TTS 1 Hd

OpenAI

OpenAIText-to-Speech OpenAI's high-definition text-to-speech model producing higher quality audio output.

@tts-1-hd

TTS 1

OpenAI

OpenAIText-to-Speech OpenAI's text-to-speech model optimized for real-time use with low latency.

@tts-1

O4 Mini

OpenAI

OpenAIText Generation OpenAI's fast, lightweight reasoning model optimized for multi-step problem solving at lower cost.

@o4-mini

O3 Mini

OpenAI

OpenAIText Generation o3-mini is the lightweight, low-cost reasoning variant of o3, well suited to quick analytical tasks at scale.

@o3-mini

O3

OpenAI

OpenAIText Generation o3 is OpenAI’s general-purpose reasoning model, balancing strong analytical performance with reasonable latency and cost.

@o3

GPT OSS 120b

OpenAI

OpenAI’s open-weight models designed for powerful reasoning, agentic tasks, and versatile developer use cases – gpt-oss-120b is for production, general purpose, high reasoning use-cases.

Tekst & bisedë 128k kontekst @gpt-oss-120b

GPT OSS 20b

OpenAI

OpenAI’s open-weight models designed for powerful reasoning, agentic tasks, and versatile developer use cases – gpt-oss-20b is for lower latency, and local or specialized use-cases.

Tekst & bisedë 128k kontekst @gpt-oss-20b

GPT Image 2

OpenAI

OpenAIText-to-Image OpenAI's next-generation image model that creates and edits images from text prompts, with support for multiple quality levels, sizes, and output formats. Note: transparent backgrounds are not supported — use openai/gpt-image-1.5 for transparent PNGs.

@gpt-image-2

GPT Image 1.5

OpenAI

OpenAIText-to-Image OpenAI's image generation model that creates and edits images from text prompts, supporting multiple quality levels and output sizes.

@gpt-image-1.5

GPT-5.5 Pro

OpenAI

OpenAIText Generation GPT-5.5 pro uses OpenAI's Responses API with built-in tools, improved reasoning, and stateful context management.

@gpt-5.5-pro

GPT-5.5

OpenAI

OpenAIText Generation GPT-5.5 is OpenAI's flagship model with strong coding, reasoning, and multimodal capabilities.

@gpt-5.5

GPT-5.4 Pro

OpenAI

OpenAIText Generation GPT-5.4 pro uses OpenAI's Responses API with built-in tools, improved reasoning, and stateful context management.

@gpt-5.4-pro

GPT-5.4 Nano

OpenAI

OpenAIText Generation GPT-5.4 nano is OpenAI's smallest and fastest model, optimized for edge and low-latency use cases.

@gpt-5.4-nano

GPT-5.4 Mini

OpenAI

OpenAIText Generation GPT-5.4 mini is a smaller, faster, and more cost-efficient version of GPT-5.4 for lightweight tasks.

@gpt-5.4-mini

GPT-5.4

OpenAI

OpenAIText Generation GPT-5.4 is OpenAI's flagship model with strong coding, reasoning, and multimodal capabilities.

@gpt-5.4

GPT-5.1 Chat

OpenAI

Text GenerationGPT-5.1 Chat is the chat-tuned variant of GPT-5.1, optimised for back-and-forth conversation and instruction following.

@gpt-5.1-chat

GPT-5.1

OpenAI

OpenAIText Generation GPT-5.1 is OpenAI’s incremental improvement over GPT-5, with stronger coding, reasoning, and writing.

@gpt-5.1

GPT-5 Nano

OpenAI

OpenAIText Generation GPT-5 Nano is OpenAI’s smallest GPT-5 variant, optimized for low latency and cheap, high-throughput tasks.

@gpt-5-nano

GPT-5 Mini

OpenAI

OpenAIText Generation GPT-5 Mini is the lightweight, low-cost variant of GPT-5, well suited to high-volume coding and reasoning tasks.

@gpt-5-mini

GPT-5 Chat

OpenAI

Text GenerationGPT-5 Chat is the chat-tuned variant of GPT-5, optimised for back-and-forth conversation and instruction following.

@gpt-5-chat

GPT-5

OpenAI

OpenAIText Generation OpenAI's model excelling at coding, writing, and reasoning.

@gpt-5

GPT-4o Mini

OpenAI

OpenAIText Generation GPT-4o Mini is the lightweight, low-cost variant of GPT-4o, well suited to high-volume tasks with multimodal inputs.

@gpt-4o-mini

GPT-4o

OpenAI

OpenAIText Generation GPT-4o is OpenAI’s multimodal flagship, accepting text and images and responding quickly across a wide range of tasks.

@gpt-4o

GPT-4.1 Nano

OpenAI

OpenAIText Generation GPT-4.1 Nano is OpenAI’s smallest and cheapest GPT-4.1 variant, optimized for high-throughput, low-latency tasks.

@gpt-4.1-nano

GPT-4.1 Mini

OpenAI

OpenAIText Generation Fast, affordable version of GPT-4.1 with a million-token context window.

@gpt-4.1-mini

GPT-4.1

OpenAI

OpenAIText Generation OpenAI's flagship GPT model for complex tasks with a million-token context window.

@gpt-4.1

Veo 3.1 Fast

Google

GoogleText-to-Video A faster version of Veo 3.1 optimized for lower latency while maintaining high-quality video and audio output.

@veo-3.1-fast

Veo 3.1

Google

GoogleText-to-Video Google's latest video generation model with improved quality, motion, and audio generation.

@veo-3.1

Veo 3 Fast

Google

Text-to-VideoA faster version of Veo 3 optimized for lower latency video generation with audio support.

@veo-3-fast

Veo 3

Google

Text-to-VideoGoogle's video generation model capable of producing high-quality videos with optional audio from text prompts.

@veo-3

Nano Banana Pro

Google

GoogleText-to-Image Google's higher-quality image generation model with improved detail and prompt adherence.

@nano-banana-pro

Nano Banana 2

Google

GoogleText-to-Image Google's second-generation image generation model with improved quality and speed.

@nano-banana-2

Nano Banana

Google

GoogleText-to-Image Google's fast image generation model producing high-quality images from text prompts.

@nano-banana

Imagen 4

Google

Text-to-ImageGoogle's latest image generation model producing high-quality, photorealistic images from text prompts with support for multiple aspect ratios.

@imagen-4

Gemma 4 26b A4b IT

Google

Gemma 4 is Google's most intelligent family of open models, built from Gemini 3 research to maximize intelligence-per-parameter.

Tekst & bisedë 256k kontekst @gemma-4-26b-a4b-it

Gemini 3.1 Pro

Google

GoogleText Generation Google's most intelligent Gemini model with improved reasoning, a medium thinking level, and a 1M token context window.

@gemini-3.1-pro

Gemini 3.1 Flash TTS

Google

GoogleText-to-Speech

@gemini-3.1-flash-tts

Gemini 3.1 Flash Lite

Google

GoogleText Generation Google's lightest and most cost-efficient Gemini model for high-throughput tasks.

@gemini-3.1-flash-lite

Gemini 3 Flash

Google

GoogleText Generation Gemini 3 Flash is Google's fast multimodal model with frontier intelligence, superior search, and grounding capabilities.

@gemini-3-flash

Gemini 2.5 Pro

Google

GoogleText Generation Google's most capable Gemini 2.5 model with strong reasoning, thinking support, and a 1M token context window.

@gemini-2.5-pro

Gemini 2.5 Flash Lite

Google

GoogleText Generation Google's lightest and most cost-efficient Gemini 2.5 model for high-throughput tasks.

@gemini-2.5-flash-lite

Gemini 2.5 Flash

Google

GoogleText Generation Google's fast multimodal Gemini 2.5 model with strong reasoning and a 1M token context window.

@gemini-2.5-flash

Grok TTS

xAI

xAIText-to-Speech xAI's Grok text-to-speech model. Generates high-fidelity spoken audio in 5 expressive voices (eve, ara, rex, sal, leo) with 20+ supported languages. Supports inline speech tags for laughter, whispers, and pauses.

@grok-tts

Grok Stt

xAI

xAIAutomatic Speech Recognition xAI's Grok speech-to-text model. Transcribes audio files into text across 25 languages with word-level timestamps, multichannel transcription, speaker diarization, and key-term biasing.

@grok-stt

Grok Imagine Video 1.5 Preview

xAI

xAIImage-to-Video xAI's next-generation video generation model. Generates, edits, and extends videos from text and image inputs. Supports multiple aspect ratios and resolutions with improved quality over the previous generation.

@grok-imagine-video-1.5-preview

Grok Imagine Video

xAI

xAIText-to-Video xAI's video generation model. Generates, edits, and extends videos from text and image inputs with native synchronized audio including dialogue, sound effects, and music. Supports multiple creative modes (normal, fun, custom).

@grok-imagine-video

Grok Imagine Image Quality

xAI

xAIText-to-Image xAI's higher-fidelity text-to-image model optimized for sharper details, more accurate compositions, and stronger text rendering. Supports image editing via reference images and masks. Trades speed for quality compared to grok-imagine-image. Default output at 2k resolution.

@grok-imagine-image-quality

Grok Imagine Image

xAI

xAIText-to-Image xAI's Grok Imagine image model. Generates and edits images from text and reference-image inputs with configurable aspect ratio and resolution.

@grok-imagine-image

Grok 4.20 Multi Agent 0309

xAI

xAIText Generation xAI's Grok 4.20 multi-agent model with a 2M-token context window. Multiple agents collaborate in parallel to perform deep research tasks, with function calling, structured outputs, and reasoning capabilities.

@grok-4.20-multi-agent-0309

Grok 4.20 0309 Reasoning

xAI

xAIText Generation xAI's Grok 4.20 reasoning model. Uses extended thinking to work through complex problems, returning a reasoning trace alongside the final answer.

@grok-4.20-0309-reasoning

Grok 4.20 0309 Non Reasoning

xAI

xAIText Generation xAI's Grok 4.20 non-reasoning model. Skips the thinking trace for fast, single-pass responses while keeping the same training as the reasoning variant.

@grok-4.20-0309-non-reasoning

Grok 4.3

xAI

xAIText Generation xAI's Grok 4.3 model with a 1M-token context window and strong agentic tool calling with minimal hallucinations. Accepts text and image inputs, and supports function calling, structured outputs, and configurable reasoning effort (none, low, medium, high).

@grok-4.3

M2M100 1.2b

Meta

Multilingual encoder-decoder (seq-to-seq) model trained for Many-to-Many multilingual translation

Përkthimi @m2m100-1.2b

Llama Guard 3 8b

Meta

Llama Guard 3 is a Llama-3.1-8B pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM inputs (prompt classification) and in LLM responses (response classification). It acts as an LLM – it generates text in its output that indicates whether a given prompt or response is safe or unsafe, and if unsafe, it also lists the content categories violated.

Tekst & bisedë 131k kontekst @llama-guard-3-8b

Llama 4 Scout 17b 16e Instruct

Meta

Meta's Llama 4 Scout is a 17 billion parameter model with 16 experts that is natively multimodal. These models leverage a mixture-of-experts architecture to offer industry-leading performance in text and image understanding.

Tekst & bisedë 131k kontekst @llama-4-scout-17b-16e-instruct

Llama 3.3 70b Instruct FP8 Fast

Meta

Llama 3.3 70B quantized to fp8 precision, optimized to be faster.

Tekst & bisedë 24k kontekst @llama-3.3-70b-instruct-fp8-fast

Llama 3.2 11b Vision Instruct

Meta

The Llama 3.2-Vision instruction-tuned models are optimized for visual recognition, image reasoning, captioning, and answering general questions about an image.

Tekst & bisedë 128k kontekst @llama-3.2-11b-vision-instruct

Llama 3.2 3b Instruct

Meta

The Llama 3.2 instruction-tuned text only models are optimized for multilingual dialogue use cases, including agentic retrieval and summarization tasks.

Tekst & bisedë 80k kontekst @llama-3.2-3b-instruct

Llama 3.2 1b Instruct

Meta

The Llama 3.2 instruction-tuned text only models are optimized for multilingual dialogue use cases, including agentic retrieval and summarization tasks.

Tekst & bisedë 60k kontekst @llama-3.2-1b-instruct

Llama 3.1 8b Instruct FP8

Meta

Llama 3.1 8B quantized to FP8 precision

Tekst & bisedë 32k kontekst @llama-3.1-8b-instruct-fp8

Deepseek v4 Pro

DeepSeek

d deepseekText Generation DeepSeek V4 Pro is a high-capability reasoning model from DeepSeek, served via Fireworks infrastructure for production-grade inference.

@deepseek-v4-pro

Kimi K2.6

Moonshot AI (Kimi)

Kimi K2.6 is a frontier-scale open-source 1T parameter model with a 262.1k context window, multi-turn tool calling, vision inputs, and structured outputs for agentic workloads.

Tekst & bisedë 262k kontekst @kimi-k2.6

Mistral Small 3.1 24b Instruct

Mistral AI

Building upon Mistral Small 3 (2501), Mistral Small 3.1 (2503) adds state-of-the-art vision understanding and enhances long context capabilities up to 128k tokens without compromising text performance. With 24 billion parameters, this model achieves top-tier capabilities in both text and vision tasks.

Tekst & bisedë 128k kontekst @mistral-small-3.1-24b-instruct

Qwq 32b

Qwen

QwQ is the reasoning model of the Qwen series. Compared with conventional instruction-tuned models, QwQ, which is capable of thinking and reasoning, can achieve significantly enhanced performance in downstream tasks, especially hard problems. QwQ-32B is the medium-sized reasoning model, which is capable of achieving competitive performance against state-of-the-art reasoning models, e.g., DeepSeek-R1, o1-mini.

Tekst & bisedë 24k kontekst @qwq-32b

Qwen3.8 27b

Qwen

Qwen 3.8 27B is a 27-billion-parameter instruction-tuned language model from Alibaba's Qwen family, designed for vision, efficient general-purpose text generation and agentic workloads.

Tekst & bisedë 262k kontekst @qwen3.8-27b

Qwen3 Embedding 0.6b

Qwen

The Qwen3 Embedding model series is the latest proprietary model of the Qwen family, specifically designed for text embedding and ranking tasks.

Embeddings & kërkim 8k kontekst @qwen3-embedding-0.6b

Qwen3 30b A3b FP8

Qwen

Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. Built upon extensive training, Qwen3 delivers groundbreaking advancements in reasoning, instruction-following, agent capabilities, and multilingual support.

Tekst & bisedë 33k kontekst @qwen3-30b-a3b-fp8

Qwen2.5 Coder 32b Instruct

Qwen

Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as CodeQwen). As of now, Qwen2.5-Coder has covered six mainstream model sizes, 0.5, 1.5, 3, 7, 14, 32 billion parameters, to meet the needs of different developers. Qwen2.5-Coder brings the following improvements upon CodeQwen1.5:

Tekst & bisedë 33k kontekst @qwen2.5-coder-32b-instruct

GLM 5.3 Flash

Z.ai (GLM)

The first natively multimodal model in the GLM-5 series. With 320B total parameters and just 18B active parameters, it outperforms GLM-5.2 across benchmarks and real-world workloads at one-tenth the price, while approaching Claude Opus 4.8 on coding and agentic benchmarks.

Tekst & bisedë 1,311k kontekst @glm-5.3-flash

GLM 5.3

Z.ai (GLM)

GLM-5.3 is Z.ai's flagship agentic coding model, pairing a 1M-token context window with reasoning, function calling, and structured outputs to power multi-step, tool-driven development workflows.

Tekst & bisedë 1,311k kontekst @glm-5.3

GLM 5.2

Z.ai (GLM)

Z.ai's flagship agentic coding model

Tekst & bisedë 262k kontekst @glm-5.2

GLM 4.7 Flash

Z.ai (GLM)

GLM-4.7-Flash is a fast and efficient multilingual text generation model with a 131,072 token context window. Optimized for dialogue, instruction-following, and multi-turn tool calling across 100+ languages.

Tekst & bisedë 131k kontekst @glm-4.7-flash

Flux 2 Pro Preview

Black Forest Labs

Black Forest LabsText-to-Image FLUX.2 [pro] Preview is Black Forest Labs' recommended default for production image generation and editing — tracks the latest [pro] weights with strong multi-reference support.

@flux-2-pro-preview

Flux 2 Max

Black Forest Labs

Black Forest LabsText-to-Image FLUX.2 [max] is Black Forest Labs' highest-quality image model — top editing consistency, strongest prompt following, and grounding search for visualizations of real-time information.

@flux-2-max

Flux 2 Flex

Black Forest Labs

Black Forest LabsText-to-Image FLUX.2 [flex] is Black Forest Labs' fine-grained control variant of FLUX.2 — exposes tunable inference steps, guidance, and prompt upsampling for typography-heavy and production workflows.

@flux-2-flex

Flux 1 Schnell

Black Forest Labs

FLUX.1 [schnell] is a 12 billion parameter rectified flow transformer capable of generating images from text descriptions.

Gjenerim imazhesh @flux-1-schnell

Recraftv4 Vector

Recraft

RecraftText-to-Image Generate production-ready SVG vector graphics from text prompts with clean geometry, structured layers, and editable paths.

@recraftv4-vector

Recraftv4 Pro Vector

Recraft

RecraftText-to-Image Generate detailed, production-ready SVG vector graphics from text prompts with fine geometry, scalable to any size for print and design work.

@recraftv4-pro-vector

Recraftv4 Pro

Recraft

RecraftText-to-Image Recraft V4 Pro generates high-resolution, art-directed images at 2048px+ with strong composition, text rendering, and design taste. Built for print and production work.

@recraftv4-pro

Recraftv4 1 Vector

Recraft

RecraftText-to-Image Generate production-ready SVG vector graphics from text prompts with high aesthetic quality, clean geometry, structured layers, and editable paths.

@recraftv4-1-vector

Recraftv4 1 Utility Vector

Recraft

RecraftText-to-Image Generate production-ready SVG vector graphics from text prompts with a general-purpose model suited for a wide range of design and illustration tasks.

@recraftv4-1-utility-vector

Recraftv4 1 Utility Pro Vector

Recraft

RecraftText-to-Image Generate detailed, high-resolution SVG vector graphics from text prompts with a general-purpose model, scalable to any size for print and large-scale design work.

@recraftv4-1-utility-pro-vector

Recraftv4 1 Utility Pro

Recraft

RecraftText-to-Image Recraft V4.1 Utility Pro is a general-purpose text-to-image model producing high-resolution 2048px+ output for a wide range of production and print use cases.

@recraftv4-1-utility-pro

Recraftv4 1 Utility

Recraft

RecraftText-to-Image Recraft V4.1 Utility is a general-purpose text-to-image model balancing quality and flexibility for a wide range of everyday use cases at standard resolution.

@recraftv4-1-utility

Recraftv4 1 Pro Vector

Recraft

RecraftText-to-Image Generate detailed, high-resolution SVG vector graphics from text prompts with high aesthetic quality, fine geometry, scalable to any size for print and design work.

@recraftv4-1-pro-vector

Recraftv4 1 Pro

Recraft

RecraftText-to-Image Recraft V4.1 Pro generates high-resolution, art-directed images at 2048px+ tuned for high aesthetics, with strong composition, text rendering, and refined design taste. Built for print and production work.

@recraftv4-1-pro

Recraftv4 1

Recraft

RecraftText-to-Image Recraft V4.1 generates art-directed images tuned for high aesthetics, with strong composition, accurate text rendering, and refined design taste. Fast and cost-efficient at standard resolution.

@recraftv4-1

Recraftv4

Recraft

RecraftText-to-Image Recraft V4 generates art-directed images with strong composition, accurate text rendering, and design taste built in. Fast and cost-efficient at standard resolution.

@recraftv4

Recraftv3

Recraft

RecraftText-to-Image Recraft V3 is the previous-generation text-to-image model from Recraft, well-suited to design-quality compositions, brand-aware imagery, and accurate text rendering.

@recraftv3

Gen 4.5

Runway

RunwayMLText-to-Video RunwayML's video generation model supporting both text-to-video and image-to-video with customizable duration, aspect ratio, and content moderation controls.

@gen-4.5

Aleph 2

Runway

RunwayMLText-to-Video RunwayML's video editing model. Edit one frame to update your whole video, make changes across multiple shots, and work with up to 30 seconds of video. Supports keyframe-guided editing for precise control over specific moments in the clip.

@aleph-2

Speech 2.8 Turbo

MiniMax

MiniMaxText-to-Speech MiniMax Speech 2.8 Turbo turns text into natural, expressive speech with voice cloning, emotion control, and 40+ language support at faster speeds.

@speech-2.8-turbo

Speech 2.8 Hd

MiniMax

MiniMaxText-to-Speech MiniMax Speech 2.8 HD focuses on studio-grade audio generation with emotion control, multilingual support (40+ languages), and voice cloning.

@speech-2.8-hd

Music 2.6

MiniMax

MiniMaxMusic Generation MiniMax's music generation model that creates full-length songs with vocals from text prompts and lyrics, or instrumental tracks. Supports BPM/key control and auto-generated lyrics.

@music-2.6

M3

MiniMax

MiniMaxText Generation MiniMax's M3 language model with frontier coding and agentic capabilities, a 1M token context window, and multilingual support.

@m3

M2.7

MiniMax

MiniMaxText Generation MiniMax's M2.7 language model with multilingual capabilities.

@m2.7

Hailuo 2.3 Fast

MiniMax

MiniMaxText-to-Video A lower-latency version of Hailuo 2.3 that preserves core motion quality, visual consistency, and stylization while enabling faster iteration.

@hailuo-2.3-fast

Hailuo 2.3

MiniMax

MiniMaxText-to-Video A high-fidelity video generation model optimized for realistic human motion, cinematic VFX, expressive characters, and strong prompt and style adherence across text-to-video and image-to-video workflows.

@hailuo-2.3

Seedream 5 Lite

ByteDance

ByteDanceText-to-Image Seedream 5 Lite is a lighter, faster version of the Seedream 5 family with multi-reference and batch generation support.

@seedream-5-lite

Seedream 4.5

ByteDance

ByteDanceText-to-Image Seedream 4.5 builds on 4.0 with multi-reference image support, batch generation, and sequential image generation.

@seedream-4.5

Seedream 4.0

ByteDance

ByteDanceText-to-Image Seedream 4.0 is ByteDance's image creation model that combines text-to-image generation and image editing into a single architecture, offering fast, high-resolution output up to 4K.

@seedream-4.0

Seedance 2.0 Fast

ByteDance

ByteDanceText-to-Video Faster variant of ByteDance's Seedance 2.0 video model. Trades some quality for speed while sharing the same multimodal architecture. Supports text-to-video, image-to-video, native audio generation, multimodal references (images, videos, audio), video editing, and video extension.

@seedance-2.0-fast

Seedance 2.0

ByteDance

ByteDanceText-to-Video ByteDance's next-generation video model with a unified multimodal architecture. Generates high-quality video with synchronized audio from text, images, video clips, and audio inputs. Supports multimodal references (up to 9 images, 3 videos, 3 audio files), native audio generation, video editing, video extension, intelligent duration, and adaptive aspect ratio.

@seedance-2.0

Aura 2 Es

Deepgram

Aura-2 is a context-aware text-to-speech (TTS) model that applies natural pacing, expressiveness, and fillers based on the context of the provided text. The quality of your text input directly impacts the naturalness of the audio output.

Zë & të folur @aura-2-es

Aura 2 En

Deepgram

Aura-2 is a context-aware text-to-speech (TTS) model that applies natural pacing, expressiveness, and fillers based on the context of the provided text. The quality of your text input directly impacts the naturalness of the audio output.

Zë & të folur @aura-2-en

Aura 1

Deepgram

Aura is a context-aware text-to-speech (TTS) model that applies natural pacing, expressiveness, and fillers based on the context of the provided text. The quality of your text input directly impacts the naturalness of the audio output.

Zë & të folur @aura-1

Wan 2.7 I2v

Alibaba

AlibabaImage-to-Video Alibaba's Wan 2.7 image-to-video model that generates videos from a reference image with optional text prompts. Supports 720P and 1080P output with durations from 2 to 15 seconds.

@wan-2.7-i2v

Wan 2.6 Image

Alibaba

AlibabaText-to-Image Alibaba's Wan 2.6 text-to-image model generating images from text prompts with optional negative prompts and customizable dimensions.

@wan-2.6-image

v6

PixVerse

PixVerseText-to-Video Pixverse v6 is the latest Pixverse video model with support for up to 15-second videos, customizable duration from 1 to 15 seconds, and audio generation.

@v6

v5.6

PixVerse

PixVerseText-to-Video Pixverse v5.6 is a video generation model supporting text-to-video and image-to-video with audio generation, customizable aspect ratios, and up to 1080p output.

@v5.6

Duke shfaqur 120 të parat — përsos kërkimin tënd për të parë më shumë.

Ndalo së planifikuari. Fillo të shesësh.

E vetmja gjë midis teje dhe klientit tënd të parë është një fjali dhe pak minuta.

20 kredite AI falas për të filluar · Anulo kur të duash