AI Model library

LLMs, image, video, and audio models — all on Venice.

Every AI model below runs on Venice, either privately with no prompt logging, or anonymously to protect your identity. Compare pricing and capabilities, and try any of them free.

LLM

106 models
Aion 2.0
AionLabs

AionLabs' affordable text-only model featuring reasoning, web search, and a 128K context window.

$1 in · $2 out / 1M
Aion 3.0 Mini
AionLabs

AionLabs' multi-model collaborative text system built on DeepSeek, tuned for immersive roleplay and storytelling with reasoning and tool support.

$0.88 in · $1.75 out / 1M
Aion 3.0
AionLabs

AionLabs' multi-model collaborative text system for roleplaying and storytelling, built on GLM with tool use and reasoning.

$3.75 in · $7.50 out / 1M
Claude Fable 5
Anthropic

Anthropic's first Mythos-class model for general use — state-of-the-art reasoning, coding, and agentic work with a 1M context window.

$12 in · $60 out / 1M
Claude Opus 4.5
Anthropic

Anthropic's frontier coding and reasoning model with vision, tool use, and a 198K context window.

$6 in · $30 out / 1M
Claude Opus 4.6
Anthropic

Anthropic's flagship multimodal model with a 1M token context window, state-of-the-art coding and reasoning, and dynamic agentic capabilities.

$6 in · $30 out / 1M
Claude Opus 4.7
Anthropic

Anthropic's flagship LLM for agentic coding, long-horizon reasoning, and high-resolution vision with a 1M-token context window.

$6 in · $30 out / 1M
Claude Opus 4.8
Anthropic

Anthropic's premier frontier model, optimized for advanced coding, autonomous agentic loops, and deep reasoning with a massive 1M-token context.

$6 in · $30 out / 1M
Claude Opus 5
Anthropic

Anthropic's high-intelligence model for complex coding, agentic tasks, and professional work — 1M context, reasoning on by default, vision, and web search.

$6 in · $30 out / 1M
Claude Sonnet 4.5
Anthropic

Anthropic's mid-tier frontier model — elite coding, agentic tool use, computer control, and reasoning with vision input.

$3.75 in · $18.75 out / 1M
Claude Sonnet 4.6
Anthropic

Anthropic's hybrid reasoning mid-tier model with a 1M context window, built for coding, agents, and enterprise workflows.

$3.60 in · $18 out / 1M
Claude Sonnet 5
Anthropic

Anthropic's most agentic Sonnet yet — near-Opus coding and reasoning at mid-tier pricing.

$3 in · $15 out / 1M
DeepSeek V3.2OPEN
DeepSeek-AI

DeepSeek's open-weight reasoning model with sparse attention, native tool-use thinking, and GPT-5-level performance at a fraction of frontier pricing.

$0.33 in · $0.48 out / 1M
DeepSeek V4 Flash 0731OPEN
DeepSeek AI

DeepSeek V4 Flash 0731 is a high-performance, agentic-focused model with 284B parameters and 13B active, now officially released with enhanced tool use and reasoning.

$0.17 in · $0.35 out / 1M
DeepSeek V4 FlashOPEN
DeepSeek

DeepSeek's ultra-efficient, open-weights MoE model featuring a 1M context window, hybrid attention, and strong agentic coding capabilities.

$0.17 in · $0.35 out / 1M
DeepSeek V4 ProOPEN
DeepSeek

DeepSeek's flagship 1.6T parameter Mixture-of-Experts model, delivering frontier-class coding, math, and agentic reasoning with an ultra-efficient 1M context window.

$1.73 in · $3.80 out / 1M
DeepSeek V4 FlashOPEN
DeepSeek-AI

DeepSeek's fast 284B-parameter MoE text model with 1M context, 13B active params, and open weights for coding, reasoning, and agentic workflows.

$0.18 in · $0.37 out / 1M
Gemma 3 27BOPEN
Google DeepMind

Google’s 27B-parameter open-weight transformer with multilingual support, long-context architecture, and vision-language design.

$0.14 in · $0.50 out / 1M
Gemma 4 26B A4B Uncensored
Google DeepMind (Base) / Phala (Fine-tune)

An uncensored, privacy-first Mixture-of-Experts (MoE) model based on Google's Gemma 4, optimized for unbiased reasoning, coding, and search.

$0.19 in · $0.88 out / 1M
Gemma 4 31B InstructOPEN
Google DeepMind

Google DeepMind's dense 31B-parameter open-source instruct model with strong reasoning and web search capabilities.

$0.14 in · $0.43 out / 1M
GLM 4.7OPEN
Z.AI

Z.AI's open-weight coding and reasoning model with multi-mode thinking, MIT-licensed weights, and strong agentic performance.

$1.10 in · $4.15 out / 1M
GLM 5.1OPEN
Z.AI

Z.AI's open-weights MoE flagship for agentic coding and long-horizon reasoning, runnable privately on Venice.

$1.10 in · $4.15 out / 1M
GLM 5.2OPEN
Z.ai (Zhipu AI)

Z.ai's open-weights MIT-licensed flagship for long-horizon coding and reasoning, with a 524K context, native tool use, and permissionless self-hosting.

$1.75 in · $5.75 out / 1M
GPT OSS 120BOPEN
OpenAI

OpenAI's largest open-weight reasoning model — 117B MoE parameters, Apache 2.0 licensed, with web search and full chain-of-thought.

$0.13 in · $0.65 out / 1M
GPT OSS 20BOPEN
OpenAI

OpenAI's open-weight, 20B-parameter reasoning model for low-latency, local, and agentic use cases — customizable, uncensored, and privacy-first on Venice.

$0.05 in · $0.19 out / 1M
Qwen 2.5 7BOPEN
Alibaba

Qwen 2.5 7B is an open-weights, instruction-tuned LLM by Alibaba, optimized for coding, math, and multilingual tasks with strong privacy on Venice.

$0.05 in · $0.13 out / 1M
Qwen3 30B A3BOPEN
Alibaba (Qwen Team)

Alibaba's nimble MoE model that switches between deep reasoning and fast chat, with open weights and tool use.

$0.19 in · $0.69 out / 1M
Qwen 3.6 27B FP8OPEN
Alibaba Cloud

Alibaba's 27B open-weight coding specialist with hybrid DeltaNet attention, agentic reasoning, and near-lossless FP8 quantization.

$0.35 in · $3.46 out / 1M
Qwen3.6 35B A3B UncensoredOPEN
Alibaba Cloud (base) / HauhauCS (uncensored)

An uncensored, highly optimized 35B Mixture-of-Experts model from the Qwen 3.6 family, built for agentic coding and unrestricted reasoning.

$0.38 in · $1.88 out / 1M
Qwen 3.6 35B A3B FP8OPEN
Alibaba

Alibaba's open-weights 35B-parameter MoE with 3B active per token, built for agentic coding, reasoning, and tool use.

$0.18 in · $1.18 out / 1M
Qwen3 VL 30B A3BOPEN
Alibaba Cloud (Qwen team)

Alibaba's open-weights vision-language model with tool use, web search, and private TEE inference on Venice.

$0.25 in · $0.90 out / 1M
Venice Uncensored 1.1OPEN
Cognitive Computations & Venice.ai

A highly steerable, 24B-parameter uncensored model co-developed by Venice and Dolphin, running with end-to-end encryption.

$0.25 in · $1.15 out / 1M
Gemini 3.1 Pro Preview
Google DeepMind

Google’s 1M-context multimodal reasoning model with native tool use, vision, and web search for agentic coding and complex analysis.

$2.50 in · $15 out / 1M
Gemini 3.5 Flash-Lite
Google DeepMind

Google's fastest, most cost-efficient 3.5-class model — optimized for high-throughput agentic tasks, document parsing, and low-latency reasoning.

$0.38 in · $3.13 out / 1M
Gemini 3.5 Flash
Google DeepMind

Google's fast, agent-first multimodal model, delivering frontier-level reasoning and coding at Flash speeds.

$1.55 in · $9.45 out / 1M
Gemini 3.6 Flash
Google DeepMind

Google's efficient, multimodal reasoning model optimized for agentic workflows, coding, and real-time tasks at scale.

$1.88 in · $9.38 out / 1M
Gemini 3 Flash Preview
Google DeepMind

Google's fast, multimodal reasoning model built for agentic coding and high-frequency workflows at a fraction of flagship cost.

$0.70 in · $3.75 out / 1M
Gemma 4 UncensoredOPEN
Google DeepMind (community derivative)

A community derivative of Google's Gemma 4 26B MoE with reduced safety alignment, offering 256K context, vision, and tool use at very low cost.

$0.16 in · $0.50 out / 1M
Google Gemma 3 27B InstructOPEN
Google DeepMind

Google's 27B open-weight multimodal model with vision, tool use, and 140+ language support — efficient enough for consumer hardware.

$0.12 in · $0.20 out / 1M
Google Gemma 4 26B A4B InstructOPEN
Google DeepMind

Google's highly efficient 26B Mixture-of-Experts (MoE) model with 4B active parameters, offering multimodal reasoning, vision, and tool use under an Apache 2.0 license.

$0.16 in · $0.50 out / 1M
Google Gemma 4 31B InstructOPEN
Google DeepMind

Google’s open-weights dense multimodal model with reasoning, tool use, and 256K context.

$0.12 in · $0.36 out / 1M
Grok 4.20 Multi-Agent
xAI

xAI's collaborative multi-agent model — four specialized AIs debate in real time to deliver deeply researched, cited answers with real-time X access.

$1.42 in · $2.83 out / 1M
Grok 4.20
xAI

xAI's flagship reasoning model with 2M-token context, low hallucination rate, and agentic tool calling — available on Venice with zero retention.

$1.42 in · $2.83 out / 1M
Grok 4.3
xAI

xAI's frontier LLM with a 1M context, built-in reasoning, vision, web search, and tool use — run privately with zero retention.

$1.42 in · $2.83 out / 1M
Grok 4.5
SpaceXAI

SpaceXAI's frontier mixture-of-experts model for coding, agentic tool use, and long-context knowledge work.

$2.27 in · $6.80 out / 1M
Grok Build 0.1
xAI

xAI's agentic coding model purpose-built for terminal-based software engineering with always-on reasoning, vision, and tool use.

$1 in · $2 out / 1M
Hermes 3 Llama 3.1 405bOPEN
Nous Research

Nous Research's flagship 405B open-weights model, fine-tuned for advanced agentic reasoning, structured JSON, and unmatched steerability.

$1.10 in · $3 out / 1M
InklingOPEN
Thinking Machines Lab

A 975B-parameter open-weights MoE multimodal model from Thinking Machines Lab that processes text, images, and audio through a 1M-token context window.

$2.34 in · $5.85 out / 1M
Kimi K2.5OPEN
Moonshot AI

Open-weight multimodal agentic model with Agent Swarm, vision-to-code, and 256K context — private on Venice with zero retention.

$0.56 in · $3.50 out / 1M
Kimi K2.6OPEN
Moonshot AI

Moonshot AI's 1T-parameter open-weight MoE built for agentic coding, long-horizon execution, and parallel agent swarms.

$0.75 in · $3.50 out / 1M
Kimi K2.7 CodeOPEN
Moonshot AI

Open-weight, coding-focused agentic model with 1T parameters, 256K context, and strong performance on long-horizon software tasks.

$0.75 in · $3.50 out / 1M
Kimi K3OPEN
Moonshot AI

Moonshot AI's 2.8T-parameter open-weight flagship with native vision, 1M context, and frontier coding capabilities.

$4.69 in · $23.44 out / 1M
Llama 3.2 3BOPEN
Meta

Meta's tiny open-weights workhorse — 3B parameters, 128K context, and tool use for edge and budget inference.

$0.15 in · $0.60 out / 1M
Mercury 2
Inception Labs

Mercury 2 is the world's fastest reasoning LLM, built on diffusion architecture for 5x faster generation and real-time agent workflows.

$0.31 in · $0.94 out / 1M
MiniMax M2.5OPEN
MiniMax AI

MiniMax M2.5 is a high-performance, agent-native language model optimized for coding, tool use, and real-world productivity tasks with SOTA scores in agentic benchmarks.

$0.27 in · $0.95 out / 1M
MiniMax M2.7
MiniMax

MiniMax M2.7 is a self-evolving, code-optimized reasoning model with strong agentic capabilities, delivering near-opus-level performance at a fraction of the cost.

$0.38 in · $1.50 out / 1M
MiniMax M3 PreviewOPEN
MiniMax

MiniMax's open-weights 428B MoE with native vision, video, and sparse attention for long-context coding and agentic work.

$0.30 in · $1.20 out / 1M
Mistral Small 4OPEN
Mistral AI

Mistral Small 4 unifies instruct, reasoning, and vision in a single open, efficient MoE model — deployable on-premise or via API with configurable reasoning effort.

$0.19 in · $0.75 out / 1M
Mistral Small 3.2 24B InstructOPEN
Mistral AI

Mistral's open-weight 24B instruction-tuned model with tool use, web search, and structured output — a production-ready upgrade to Small 3.1.

$0.09 in · $0.25 out / 1M
NVIDIA Nemotron 3 Nano 30BOPEN
NVIDIA

NVIDIA's open hybrid MoE model with Mamba-2 layers, configurable reasoning, and agentic tool use.

$0.07 in · $0.30 out / 1M
NVIDIA Nemotron 3 UltraOPEN
NVIDIA

NVIDIA's flagship open-weights frontier model — a 550B-parameter hybrid Mamba-MoE architecture built for agentic reasoning, tool use, and long-context throughput.

$0.63 in · $3.13 out / 1M
Nemotron Cascade 2 30B A3BOPEN
NVIDIA

NVIDIA's open 30B MoE that punches at frontier scale — gold-medal math, coding, and agentic reasoning with only 3B active parameters.

$0.14 in · $0.80 out / 1M
GLM 4.7 Flash HereticOPEN
Olafangensan (community mod; Z.AI base)

A community-abliterated, open-weights variant of GLM-4.7-Flash built for fast inference, reasoning, and tool use with relaxed refusal behavior.

$0.07 in · $0.40 out / 1M
GPT-4o
OpenAI

OpenAI's flagship multimodal model — fast, intelligent, and versatile across text, vision, and voice with real-time responsiveness.

$3.13 in · $12.50 out / 1M
GPT-4o Mini
OpenAI

OpenAI's most cost-efficient small model — fast, multimodal, and ideal for high-volume tasks with vision and tool use.

$0.19 in · $0.75 out / 1M
GPT-5.2 Codex
OpenAI

OpenAI's specialized agentic coding model — long-horizon software engineering, cybersecurity, and tool use with a 256K context window.

$2.19 in · $17.50 out / 1M
GPT-5.2
OpenAI

OpenAI's flagship reasoning model for professional knowledge work, coding, and agentic tasks with tool use and web search.

$2.19 in · $17.50 out / 1M
GPT-5.3 Codex
OpenAI

OpenAI's most capable agentic coding model — autonomous software engineering with vision, reasoning, and tool use.

$2.19 in · $17.50 out / 1M
GPT-5.4 Mini
OpenAI

OpenAI's fastest, most capable small model — 400K context, vision, tool use, and reasoning for coding and subagents.

$0.94 in · $5.63 out / 1M
GPT-5.4 Pro
OpenAI

OpenAI's highest-performance frontier model — native computer-use, adjustable reasoning, and 1M context for complex professional work.

$37.50 in · $225 out / 1M
GPT-5.4
OpenAI

OpenAI's frontier professional model — 1M context, configurable reasoning, and multimodal agentic capabilities with vision and tool use.

$3.13 in · $18.80 out / 1M
GPT-5.5 Pro
OpenAI

OpenAI's highest-intelligence frontier model with extended test-time compute, multimodal reasoning, and a 1M-token context window for deep research and agentic work.

$37.50 in · $225 out / 1M
GPT-5.5
OpenAI

OpenAI's April 2026 frontier model for agentic coding, research, and multi-step tool use with vision and reasoning.

$6.25 in · $37.50 out / 1M
GPT-5.6 Luna Pro
OpenAI

OpenAI's cost-efficient GPT-5.6 tier with a 1M context, vision, reasoning, and tool use for high-volume workloads.

$1.25 in · $7.50 out / 1M
GPT-5.6 Luna
OpenAI

OpenAI's fastest, most cost-efficient GPT-5.6 model — 1M context, vision, tool use, and web search for high-volume agentic work.

$1.25 in · $7.50 out / 1M
GPT-5.6 Sol Pro
OpenAI

OpenAI's flagship GPT-5.6 model with a 1M context window, native multi-agent reasoning, vision, and tool use for frontier coding and knowledge work.

$6.25 in · $37.50 out / 1M
GPT-5.6 Sol
OpenAI

OpenAI's flagship GPT-5.6 model — multimodal reasoning, agentic coding, and a 1M-token context window for frontier knowledge work.

$6.25 in · $37.50 out / 1M
GPT-5.6 Terra Pro
OpenAI

OpenAI's balanced multimodal workhorse with 1,000K context, reasoning, and tool use for production workflows.

$3.13 in · $18.75 out / 1M
GPT-5.6 Terra
OpenAI

OpenAI's balanced mid-tier frontier model — 1M context, multimodal reasoning, and native tool use for production workloads.

$3.13 in · $18.75 out / 1M
OpenAI GPT OSS 120BOPEN
OpenAI

OpenAI's largest open-weight reasoning model — 117B MoE parameters, agentic tool use, and full chain-of-thought under Apache 2.0.

$0.07 in · $0.30 out / 1M
Qwen 3.6 Plus Uncensored
Alibaba

Alibaba's Qwen 3.6 Plus Uncensored — a reasoning-optimized, multimodal model with 1M context, now uncensored and private on Venice.

$0.63 in · $3.75 out / 1M
Qwen 3.7 Max
Alibaba (Qwen Team)

Alibaba's agent-frontier text model with 1M context, tool use, and code-optimized reasoning for long-horizon autonomous workflows.

$2.70 in · $8.05 out / 1M
Qwen 3.7 Plus
Alibaba Cloud

Alibaba's cost-effective multimodal agent model — strong in agentic workflows, vision-language tasks, and GUI automation at low cost.

$0.50 in · $2 out / 1M
Qwen 3.8 Max
Alibaba

Alibaba's 2.4T MoE flagship — vision, reasoning, and code-optimized with 1M context and private execution on Venice.

$2.50 in · $7.50 out / 1M
Qwen 3 235B A22B Instruct 2507OPEN
Alibaba Cloud / Qwen Team

Alibaba's updated 235B-parameter MoE language model with 22B active params, optimized for reasoning, coding, and long-context instruction following.

$0.15 in · $0.75 out / 1M
Qwen 3 235B A22B Thinking 2507OPEN
Alibaba Cloud

Alibaba's flagship open-weight thinking MoE — 235B total, 22B active, reasoning-only mode with tool use and 262K native context.

$0.45 in · $3.50 out / 1M
Qwen 3.5 35B A3BOPEN
Alibaba

Qwen 3.5 35B A3B is an open-weight, multimodal reasoning model from Alibaba with strong vision, code, and agent capabilities at aggressive pricing.

$0.31 in · $1.25 out / 1M
Qwen 3.5 397BOPEN
Alibaba Cloud

Alibaba's flagship open-weight multimodal model with 397B parameters, native vision, and efficient MoE architecture — now on Venice.

$0.75 in · $4.50 out / 1M
Qwen 3.5 9BOPEN
Alibaba

Alibaba's 9B open-weight multimodal model with hybrid attention, 256K context, and native tool use.

$0.10 in · $0.15 out / 1M
Qwen 3.6 27B
Qwen (Alibaba Group)

A 27B dense multimodal model from Alibaba's Qwen team, optimized for agentic coding, reasoning, and long-context tasks.

$0.33 in · $3.25 out / 1M
Qwen 3.6 35B A3BOPEN
Alibaba (Qwen team)

Alibaba's open-weight MoE coding specialist — 35B total, 3B active, built for agentic development and long-context reasoning.

$0.15 in · $1 out / 1M
Qwen 3 Coder 480B TurboOPEN
Alibaba Cloud (Qwen team)

Alibaba's flagship open-weight code model — a 480B-parameter MoE with 35B active, built for agentic coding and 256K-context repo work.

$0.35 in · $1.50 out / 1M
Qwen 3 Next 80bOPEN
Alibaba

Alibaba's 80B/3B sparse MoE with hybrid attention and 256K context, open-weight under Apache 2.0.

$0.35 in · $1.90 out / 1M
Qwen3 VL 235BOPEN
Alibaba Cloud

Alibaba's 235B-parameter open-weights vision-language MoE with tool use, web search, and visual agent capabilities.

$0.21 in · $1.90 out / 1M
Seed 2.1 Turbo
ByteDance

Seed 2.1 Turbo is ByteDance's high-throughput, agent-capable AI model optimized for cost-sensitive production workloads with strong vision and code execution.

$0.63 in · $3.13 out / 1M
Venice Uncensored 1.2OPEN
Venice AI

Venice's flagship 24B uncensored model — built on Mistral architecture with native vision, web search, and zero refusal behavior.

$0.20 in · $0.90 out / 1M
Venice Role Play UncensoredOPEN
Venice AI

Venice's fine-tuned 24B roleplay model, optimized for highly expressive, low-refusal character interactions with zero retention.

$0.50 in · $2 out / 1M
MiMo-V2.5OPEN
Xiaomi

Xiaomi's open-weight, omnimodal AI with 1M context and strong agentic capabilities — built for developers who want sovereignty and uncensored, private inference.

$0.14 in · $0.28 out / 1M
GLM 5 TurboOPEN
Z.ai

Z.ai's speed-optimized agentic model with selectable reasoning modes, native tool calling, and a 200K context window for OpenClaw workflows.

$1.20 in · $4 out / 1M
GLM 5V Turbo
Zhipu AI (Z.ai)

Zhipu AI's native multimodal coding model for vision-to-code and agentic workflows.

$1.50 in · $5 out / 1M
GLM 4.6OPEN
Z.ai (Zhipu AI)

Z.ai's open-weight flagship MoE model for coding, reasoning, and agentic tasks.

$0.43 in · $1.75 out / 1M
GLM 4.7 FlashOPEN
Z.ai

Z.ai's lightweight 30B MoE text model built for fast coding, tool use, and agentic tasks.

$0.13 in · $0.50 out / 1M
GLM 4.7OPEN
Z.AI

Z.AI's open-weight coding and reasoning model that runs privately on Venice with tool use and zero retention.

$0.55 in · $2.65 out / 1M
GLM 5.1OPEN
Z.AI

Z.AI's open-weights flagship LLM for agentic engineering and long-horizon coding tasks.

$1.54 in · $4.84 out / 1M
GLM 5.2OPEN
Z.ai (Zhipu AI)

Z.ai's flagship open-weights MoE with 1M context, reasoning, and tool use for long-horizon coding.

$1.40 in · $4.40 out / 1M
GLM 5OPEN
Zhipu AI

Zhipu AI's 744B-parameter open-weight MoE flagship built for agentic engineering, reasoning, and long-horizon coding tasks.

$1 in · $3.20 out / 1M

Image

34 models
Background Remover
Bria AI

AI-powered background remover for clean alpha mattes — optimized for VFX, compositing, and e-commerce use cases.

$0.03 / image
ChromaOPEN
Lodestones

Open-weights text-to-image model at $0.01 per image with zero-retention privacy.

$0.01 / image
Flux 2 Max
Black Forest Labs

Black Forest Labs' flagship FLUX.2 image model — top-tier photorealism, multi-reference editing, and up to 4MP output.

$0.09 / image
Flux 2 Pro
Black Forest Labs

Black Forest Labs' flagship commercial model — 4MP photorealism, advanced prompt understanding, and precise design control.

$0.04 / image
GPT Image 1.5
OpenAI

OpenAI's high-fidelity image model — precise edits, strong text rendering, and reliable face preservation for professional workflows.

$0.26 / image
GPT Image 2
OpenAI

OpenAI's flagship text-to-image model with built-in reasoning, near-perfect in-image text, and up to 4K output.

From $0.02 / image
Grok Imagine 2.0
xAI

xAI's high-fidelity image generation and editing model, optimized for precision, layout-aware design, and iterative creative workflows.

From $0.05 / image
Grok Imagine High Quality (SOTA)
xAI

xAI's state-of-the-art Quality Mode model delivering photorealistic textures, clean multilingual text rendering, and advanced multi-image composition.

From $0.08 / image
Grok Imagine
xAI

xAI's unified image generation and editing API for product placement, restyling, and precision edits up to 2K.

From $0.04 / image
Hunyuan Image 3.0OPEN
Tencent

Hunyuan Image 3.0 is a powerful open-weight, multimodal image generator with 80B total parameters (13B active), offering high-fidelity text-to-image synthesis and strong multilingual support.

$0.09 / image
Ideogram V4OPEN
Ideogram

Ideogram's premier 9.3B open-weight image model — the gold standard for in-image typography, structured layout control, and 2K design assets.

$0.15 / image
ImagineArt 1.5 Pro
ImagineArt

ImagineArt's flagship image model — native 4K output, enhanced realism, accurate text rendering, and composition intelligence for professional-grade creative work.

$0.06 / image
Krea 2 TurboOPEN
Krea AI

Krea 2 Turbo is a fast, distilled version of Krea 2, optimized for rapid ideation and low-cost iteration in expressive illustration and design exploration.

From $0.04 / image
Krea v2 Large
Krea

Krea v2 Large is a high-fidelity image model optimized for expressive photorealism and advanced style control, with per-image pricing and anonymized privacy on Venice.

$0.07 / image
Krea v2 Medium
Krea AI

Krea v2 Medium is a creatively focused image model built from scratch for expressive aesthetics, advanced style transfer, and full creative control — with 1K resolution cap.

$0.04 / image
Luma Uni-1 Max
Luma AI

Luma's highest-quality image model — unified autoregressive architecture with reasoning-driven generation, strong reference fidelity, and 2K output.

$0.12 / image
Luma Uni-1
Luma AI

Luma Uni-1 is a unified autoregressive model that reasons before generating pixels, enabling precise control, strong spatial logic, and culture-aware visuals.

$0.05 / image
Lustify SDXL
Community (based on Stability AI SDXL)

Community fine-tune of SDXL optimized for photorealistic characters, natural lighting, and NSFW content with strong response to camera and film cues.

$0.01 / image
Lustify v8
Community

The ultimate uncensored SDXL checkpoint for photorealistic character rendering and native 1536px support.

$0.01 / image
Nano Banana 2 Lite
Google DeepMind

Google's fastest and most cost-efficient image model — built for high-speed generation and editing at scale with real-world knowledge grounding.

$0.06 / image
Nano Banana 2
Google DeepMind

Google's fast, knowledgeable image model that blends Pro-grade quality with Flash-tier speed and precise text rendering.

From $0.10 / image
Nano Banana Pro
Google (Gemini 3 Pro Image)

Google's flagship image generation and editing model — enhanced reasoning, real-time knowledge, and native 4K output.

From $0.18 / image
Qwen Image 2 Pro
Alibaba Qwen

Alibaba's high-fidelity image model with professional typography, precise editing, and 2K output — optimized for complex layouts and multilingual text.

$0.10 / image
Qwen Image 2
Alibaba

Alibaba's unified image generation and editing model with professional bilingual typography and native 2K output.

$0.05 / image
Recraft V4 Pro
Recraft

Recraft's premium design-centric model — delivering art-directed, print-ready raster images at 2048px resolution with exceptional visual taste.

$0.29 / image
Recraft V4
Recraft

Recraft’s design-first image model — art-directed raster and editable vector SVG generation with strong typographic accuracy.

$0.05 / image
Seedream V4.5
ByteDance

ByteDance's Seedream 4.5 delivers high-fidelity image generation and precise editing with strong consistency across subjects, text, and lighting.

$0.05 / image
Seedream V5 Lite
ByteDance

ByteDance's intelligent image generator with Chain of Thought reasoning and real-time web search for accurate, intent-aligned visuals.

$0.05 / image
Seedream V5 Pro
ByteDance

ByteDance's production-grade image model for complex layouts, infographics, and native multilingual text rendering.

From $0.06 / image
Venice SD35OPEN
Stability AI

Venice's custom-configured Stable Diffusion 3.5 engine, delivering high-fidelity photorealism and creative freedom with zero prompt retention.

$0.01 / image
Anime (WAI)
WAI Community

Highly-rated anime-specialized model built on Illustrious XL, optimized for authentic Japanese animation aesthetics and character accuracy.

$0.01 / image
Wan 2.7 Pro
Alibaba

Alibaba's high-fidelity image generation model with native 4K output, superior text rendering, and structured reasoning for complex scenes.

$0.09 / image
Wan 2.7OPEN
Alibaba

Alibaba's Wan 2.7 is a unified image and video generation model with open weights, native audio, and instruction-based editing — built for production workflows.

$0.04 / image
Z-Image TurboOPEN
Tongyi Lab (Alibaba Group)

An open-weight, high-speed text-to-image model from Alibaba’s Tongyi Lab, optimized for photorealism and bilingual text rendering with minimal inference steps.

$0.01 / image

Video

41 models
Gemini Omni Flash
Google

Google's fast, multimodal video generation model — creates and edits 10-second clips from text, images, or reference media with conversational control.

from $0.55 / clip
Grok Imagine 1.5
xAI

xAI's flagship video generation model — cinematic motion, synchronized audio, and 15-second 1080p clips with privacy-first processing on Venice.

from $0.09 / clip
Grok Imagine
xAI

xAI's premier video model family — text-, image- and reference-to-video generation at 720p with native synchronized audio, sound effects, and music.

from $0.32 / clip
HappyHorse 1.0
Alibaba Taotian Group (ATH Innovation Unit)

Alibaba's breakthrough AI video model with native audio-video sync generated in a single pass.

from $0.46 / clip
HappyHorse 1.1
Alibaba

Alibaba's multimodal video model that generates 1080p clips with native audio and supports reference-driven subject consistency across text, image, and reference-to-video modes.

from $0.46 / clip
Kling 2.5 Turbo Pro
Kuaishou Technology

Kling 2.5 Turbo Pro delivers cinematic, high-fidelity video generation with industry-leading prompt adherence and motion realism — now more affordable and accessible via Venice.

from $0.39 / clip
Kling 2.6 Pro
Kuaishou Technology

Kuaishou's flagship video model that generates 5–10s cinematic clips with simultaneous audio, voiceovers, and sound effects from text or image prompts.

from $0.77 / clip
Kling O3 4K
Kuaishou Technology

Kuaishou's flagship unified multimodal video model — 4K output, native audio, and visual chain-of-thought reasoning for director-grade clips.

from $1.39 / clip
Kling O3 Pro
Kuaishou

Kuaishou's premium unified multimodal video model — cinematic 3–15s clips with native audio from text, images, or references.

from $0.46 / clip
Kling O3 Standard
Kuaishou

Kuaishou's efficient O3-tier video model — cinematic quality with native audio, character consistency, and reference-driven workflows.

from $0.37 / clip
Kling V3 4K
Kuaishou

Kuaishou's flagship native-4K video generation model, producing up to 15-second clips with synchronized multilingual audio.

from $1.39 / clip
Kling V3 Pro
Kuaishou

Kling V3 Pro delivers cinematic, multi-shot video with native audio and precise director-style control — all in a single model.

from $0.55 / clip
Kling V3 Standard
Kuaishou

Kuaishou's cost-efficient video generation tier, producing cinematic 3–15 second clips with native audio across text, image, and motion-control inputs.

from $0.42 / clip
Kling V3 Turbo Pro
Kuaishou

Kuaishou's speed-optimized video generation model — cinematic text-to-video and image-to-video clips up to 15 seconds with strong human motion.

from $0.46 / clip
Kling V3 Turbo Standard
Kuaishou

Kuaishou's speed-optimized standard-tier video model for 3–15s text-to-video and image-to-video generation.

from $0.37 / clip
Longcat DistilledOPEN
Meituan

Meituan's open-source, uncensored video generation model — unified architecture for text-to-video and image-to-video with efficient long-duration output.

from $0.09 / clip
Longcat Full QualityOPEN
Meituan

Meituan's 13.6B open-source video model — generating coherent, high-quality private clips up to 30 seconds on Venice.

from $0.25 / clip
LTX Video 2.3 FastOPEN
Lightricks

Lightricks' open-source video engine — speed-optimized, native 4K, portrait framing, and synchronized audio.

from $0.40 / clip
LTX Video 2.3 Full QualityOPEN
Lightricks

LTX Video 2.3 Full Quality is an open-weights, audio-visual foundation model delivering high-fidelity 4K video with synchronized sound, native portrait output, and strong prompt adherence — now on Venice with zero retention.

from $0.53 / clip
MiniMax H3OPEN
MiniMax

MiniMax H3 is a general-purpose, omni-modal video generation model that supports text-to-video, image-to-video, and reference-to-video with native stereo audio, up to 2K resolution and 15 seconds duration.

from $0.45 / clip
OviOPEN
Character.AI

Open-source, uncensored image-to-video model with synchronized audio generation — runs privately on Venice with zero retention.

from $0.22 / clip
PixVerse C1
PixVerse

PixVerse C1 is a cinematic AI video model built for film production, delivering physics-accurate motion, fantasy VFX, and multi-shot storyboarding up to 15s at 1080p with synchronized audio.

from $0.13 / clip
PixVerse v5.6
PixVerse

PixVerse v5.6 delivers cinematic, audio-rich AI video generation with strong motion control and multilingual vocal synthesis, available in multiple modes across 1080p resolutions.

from $1.01 / clip
Runway Gen-4.5
Runway

Runway Gen-4.5 is a state-of-the-art AI video model that excels in cinematic quality, motion realism, and prompt adherence for both text-to-video and image-to-video generation.

from $0.32 / clip
Runway Gen-4 Turbo
Runway

Runway's fast, controllable image-to-video model — generates 5–10s clips from an image and prompt in seconds, optimized for rapid creative iteration.

from $0.14 / clip
Seedance 2.0
ByteDance

ByteDance's unified multimodal video model generating cinematic 1080p clips up to 15 seconds with synchronized audio.

from $0.35 / clip
Sora 2 Pro
OpenAI

OpenAI's flagship video generation model — cinematic 20-second clips with synchronized audio, physics-accurate motion, and world-state persistence.

from $1.32 / clip
Sora 2
OpenAI

OpenAI's flagship video generation model — cinematic realism, synchronized audio, and precise physics simulation up to 12 seconds.

from $0.44 / clip
Topaz Video Upscale
Topaz Labs

Topaz Video Upscale delivers cinematic-grade AI video enhancement with 2x and 4x upscaling, artifact reduction, and stabilization — now accessible via Venice without stored prompts.

See pricing
Veo 3.1 Fast
Google DeepMind

Google's high-fidelity video generation model with native audio, cinematic control, and image-to-video capabilities — now optimized for speed.

from $0.66 / clip
Veo 3.1 Full Quality
Google DeepMind

Google DeepMind's flagship video model — native 4K, synchronized audio, and cinematic realism in up to 60-second clips.

from $1.76 / clip
Veo 3 Fast
Google DeepMind

Google's speed-optimized video model — generates 8-second clips with native audio from text or images, starting at $0.44 per clip.

from $0.44 / clip
Veo 3 Full Quality
Google DeepMind

Google's flagship video generation model — cinematic 1080p video with native audio, precise prompt adherence, and real-world physics simulation.

from $0.88 / clip
Vidu Q3
ShengShu Technology

Vidu Q3 is ShengShu's flagship AI video model — the first to generate native audio and video in one pass, supporting up to 16-second cinematic clips with synchronized sound, dialogue, and music.

from $0.27 / clip
Wan 2.1 ProOPEN
Alibaba Tongyi Lab

Wan 2.1 Pro is Alibaba's open-weight, photorealistic image-to-video model that animates still images with cinematic motion and strong subject coherence.

from $0.88 / clip
Wan 2.5 PreviewOPEN
Alibaba Cloud (Tongyi Labs)

Alibaba's open-weight video generation model with native audio synchronization, available in text-to-video and image-to-video variants.

from $0.28 / clip
Wan 2.6 FlashOPEN
Alibaba

Wan 2.6 Flash is Alibaba's speed-optimized image-to-video model, generating up to 15-second clips at 720p or 1080p with optional audio — ideal for fast iteration and high-volume use.

from $0.28 / clip
Wan 2.6OPEN
Alibaba Cloud

Wan 2.6 is Alibaba's open-weights, production-grade AI video model — generating cinematic 15-second clips with multi-shot storytelling, native audio, and character consistency across scenes.

from $0.55 / clip
Wan 2.7 EnhancedOPEN
Alibaba

Alibaba's open-weights video model that generates 1080p, audio-enabled clips up to 15 seconds in text-to-video and image-to-video modes.

from $0.72 / clip
Wan 2.7 UncensoredOPEN
Alibaba Tongyi Lab

Alibaba's flagship 27B open-weight video model — uncensored, featuring native audio synchronization and exceptional character consistency.

from $0.57 / clip
Wan 2.7OPEN
Alibaba

Wan 2.7 is Alibaba's open-weight, uncensored video generation suite — 27B MoE architecture, native audio, 15s clips, and four production-ready modes under Apache 2.0.

from $0.55 / clip

Audio

14 models
ACE-Step 1.5OPEN
ACE Studio & StepFun

A highly efficient open-source music foundation model combining a planning Language Model with a Diffusion Transformer for rapid, high-quality track generation.

from $0.03 / track
ElevenLabs Music
ElevenLabs

ElevenLabs Music generates studio-grade instrumental or vocal tracks from natural language prompts, with commercial rights cleared through industry partnerships.

from $0.69 / track
ElevenLabs Sound Effects
ElevenLabs

ElevenLabs' AI-generated sound effects model — turns text prompts into high-quality, precisely timed audio effects for film, games, and content.

$0 / second
ElevenLabs Multilingual v2
ElevenLabs

ElevenLabs' foundational multilingual text-to-speech model, delivering lifelike, emotionally rich speech synthesis across 29 languages.

$0.12 / 1K chars
ElevenLabs TTS v3
ElevenLabs

ElevenLabs' most expressive TTS model — human-like delivery with emotion, dialogue, and non-verbal cues across 70+ languages.

$0.12 / 1K chars
Lyria 3 Pro
Google DeepMind

Google DeepMind's flagship music generation model that composes full-length songs with structured sections, custom lyrics, and provenance watermarking from text or image prompts.

$0.10 / image
MiniMax Music 2.0
MiniMax

MiniMax Music 2.0 is a next-generation AI music model that generates full songs with expressive vocals and instrumental arrangements from text and lyrics prompts.

$0.04 / image
MiniMax Music 2.5
MiniMax

MiniMax's AI music model with paragraph-level structural control, humanized vocals, and studio-grade mixing.

$0.18 / image
MiniMax Music 2.6
MiniMax

MiniMax Music 2.6 is a high-fidelity AI music model that generates full songs with realistic vocals and instrumentation from text prompts, featuring precise BPM/key control, auto lyrics, and instrumental-only mode.

$0.18 / image
MMAudio V2
Ho Kei Cheng et al. (UIUC, Sony AI, Sony Group Corporation)

MMAudio V2 is a 157M-parameter flow-matching audio model from UIUC and Sony AI researchers that generates synchronized sound from text or video, with a v2 checkpoint tuned for stronger real-world generalization.

$0 / second
Seed Audio 1.0
ByteDance

Seed Audio 1.0 is ByteDance's all-in-one audio scene generator — produces multi-character dialogue, music, SFX, and ambience from one prompt with precise timing control.

$0 / second
Sonilo V1.1 Music
Sonilo

Sonilo V1.1 Music composes original, commercially licensed soundtracks that sync precisely to video pacing, mood, and edits — no prompts needed.

$0 / second
Sonilo V1.1 Sound Effects
Sonilo

Sonilo V1.1 Sound Effects generates royalty-free, commercially licensed audio from text or video input with precise synchronization and high technical fidelity.

$0 / second
Stable Audio 2.5
Stability AI

Stability AI's enterprise text-to-audio model for 3-minute instrumental tracks, sound effects, and audio inpainting with sub-two-second inference.

$0.19 / image

Text-to-Speech

11 models
Chatterbox HD (Resemble AI)
Resemble AI

Resemble AI's high-definition text-to-speech model, delivering highly expressive, natural voice synthesis with zero-shot cloning.

$50 / 1M chars
ElevenLabs Turbo v2.5
ElevenLabs

High-quality, low-latency text-to-speech model supporting 32 languages with fast response times for real-time applications.

$62.50 / 1M chars
Gemini 3.1 Flash TTS
Google

Google's expressive, low-latency TTS model with natural-language control over tone, pace, and emotion — optimized for high-volume, cost-efficient speech generation.

$187.50 / 1M chars
Gradium TTS
Gradium

Gradium's public beta TTS model handles complex text natively—phone numbers, emails, IBANs, time expressions—with ultra-low latency and no preprocessing.

$47.50 / 1M chars
Inworld TTS-1.5 Max
Inworld AI

Inworld TTS-1.5 Max is the highest-ranked text-to-speech model globally, delivering ultra-low latency, expressive speech, and zero data retention for production-grade voice AI.

$12.50 / 1M chars
Kokoro Text to SpeechOPEN
hexgrad

An ultra-lightweight, open-weight text-to-speech model delivering studio-quality synthesis with incredible speed and efficiency.

$3.50 / 1M chars
MiniMax Speech-02 HD
MiniMax

High-definition text-to-speech with emotional control, voice customization, and multilingual support — optimized for audiobooks and voiceovers.

$125 / 1M chars
Orpheus TTSOPEN
Canopy AI

Open-weight, human-sounding TTS with zero-shot voice cloning and low-latency streaming, built on Llama-3b.

$62.50 / 1M chars
Qwen 3 TTS 0.6BOPEN
Alibaba Qwen Team

Open-source, multilingual TTS with 3-second voice cloning, natural-language voice control, and ultra-low-latency streaming — now on Venice.

$87.50 / 1M chars
Qwen 3 TTS 1.7BOPEN
Alibaba Cloud

Open-weight, multilingual TTS with ultra-low-latency streaming and voice cloning, developed by Alibaba's Qwen team.

$112.50 / 1M chars
xAI TTS v1
xAI

xAI's high-fidelity, low-latency text-to-speech API with expressive voices, inline speech tags, and enterprise-grade multilingual support — now on Venice with anonymized processing.

$18.75 / 1M chars

Start creating. Privately.

Every model, no prompt logging, no data used for training. Free to start — no credit card.