Model library

Every model, run privately.

Every model below runs on Venice — privately, with no prompt logging. Compare pricing and capabilities, then try any of them free.

LLM

65 models
Aion 2.0
AionLabs

AionLabs' affordable text-only model featuring reasoning, web search, and a 128K context window.

$1 in · $2 out / 1M
Aion 3.0 Mini
AionLabs

AionLabs' multi-model collaborative text system built on DeepSeek, tuned for immersive roleplay and storytelling with reasoning and tool support.

$0.88 in · $1.75 out / 1M
Aion 3.0
AionLabs

AionLabs' multi-model collaborative text system for roleplaying and storytelling, built on GLM with tool use and reasoning.

$3.75 in · $7.50 out / 1M
Claude Fable 5
Anthropic

Anthropic's first Mythos-class model for general use — state-of-the-art reasoning, coding, and agentic work with a 1M context window.

$12 in · $60 out / 1M
Claude Opus 4.5
Anthropic

Anthropic's frontier coding and reasoning model with vision, tool use, and a 198K context window.

$6 in · $30 out / 1M
Claude Opus 4.6
Anthropic

Anthropic's flagship multimodal model with a 1M token context window, state-of-the-art coding and reasoning, and dynamic agentic capabilities.

$6 in · $30 out / 1M
Claude Opus 4.7
Anthropic

Anthropic's flagship LLM for agentic coding, long-horizon reasoning, and high-resolution vision with a 1M-token context window.

$6 in · $30 out / 1M
Claude Opus 4.8
Anthropic

Anthropic's premier frontier model, optimized for advanced coding, autonomous agentic loops, and deep reasoning with a massive 1M-token context.

$6 in · $30 out / 1M
Claude Sonnet 4.5
Anthropic

Anthropic's mid-tier frontier model — elite coding, agentic tool use, computer control, and reasoning with vision input.

$3.75 in · $18.75 out / 1M
Claude Sonnet 4.6
Anthropic

Anthropic's hybrid reasoning mid-tier model with a 1M context window, built for coding, agents, and enterprise workflows.

$3.60 in · $18 out / 1M
Claude Sonnet 5
Anthropic

Anthropic's most agentic Sonnet yet — near-Opus coding and reasoning at mid-tier pricing.

$3 in · $15 out / 1M
DeepSeek V3.2OPEN
DeepSeek-AI

DeepSeek's open-weight reasoning model with sparse attention, native tool-use thinking, and GPT-5-level performance at a fraction of frontier pricing.

$0.33 in · $0.48 out / 1M
DeepSeek V4 FlashOPEN
DeepSeek

DeepSeek's ultra-efficient, open-weights MoE model featuring a 1M context window, hybrid attention, and strong agentic coding capabilities.

$0.17 in · $0.35 out / 1M
DeepSeek V4 ProOPEN
DeepSeek

DeepSeek's flagship 1.6T parameter Mixture-of-Experts model, delivering frontier-class coding, math, and agentic reasoning with an ultra-efficient 1M context window.

$1.73 in · $3.80 out / 1M
Gemma 4 26B A4B Uncensored
Google DeepMind (Base) / Phala (Fine-tune)

An uncensored, privacy-first Mixture-of-Experts (MoE) model based on Google's Gemma 4, optimized for unbiased reasoning, coding, and search.

$0.19 in · $0.88 out / 1M
Qwen3.6 35B A3B UncensoredOPEN
Alibaba Cloud (base) / HauhauCS (uncensored)

An uncensored, highly optimized 35B Mixture-of-Experts model from the Qwen 3.6 family, built for agentic coding and unrestricted reasoning.

$0.38 in · $1.88 out / 1M
Venice Uncensored 1.1OPEN
Cognitive Computations & Venice.ai

A highly steerable, 24B-parameter uncensored model co-developed by Venice and Dolphin, running with end-to-end encryption.

$0.25 in · $1.15 out / 1M
Gemini 3.1 Pro Preview
Google DeepMind

Google’s 1M-context multimodal reasoning model with native tool use, vision, and web search for agentic coding and complex analysis.

$2.50 in · $15 out / 1M
Gemini 3.5 Flash
Google DeepMind

Google's fast, agent-first multimodal model, delivering frontier-level reasoning and coding at Flash speeds.

$1.55 in · $9.45 out / 1M
Gemini 3 Flash Preview
Google DeepMind

Google's fast, multimodal reasoning model built for agentic coding and high-frequency workflows at a fraction of flagship cost.

$0.70 in · $3.75 out / 1M
Google Gemma 3 27B InstructOPEN
Google DeepMind

Google's 27B open-weight multimodal model with vision, tool use, and 140+ language support — efficient enough for consumer hardware.

$0.12 in · $0.20 out / 1M
Google Gemma 4 26B A4B InstructOPEN
Google DeepMind

Google's highly efficient 26B Mixture-of-Experts (MoE) model with 4B active parameters, offering multimodal reasoning, vision, and tool use under an Apache 2.0 license.

$0.16 in · $0.50 out / 1M
Grok 4.3
xAI

xAI's frontier LLM with a 1M context, built-in reasoning, vision, web search, and tool use — run privately with zero retention.

$1.42 in · $2.83 out / 1M
Grok 4.5
SpaceXAI

SpaceXAI's frontier mixture-of-experts model for coding, agentic tool use, and long-context knowledge work.

$2.27 in · $6.80 out / 1M
Hermes 3 Llama 3.1 405bOPEN
Nous Research

Nous Research's flagship 405B open-weights model, fine-tuned for advanced agentic reasoning, structured JSON, and unmatched steerability.

$1.10 in · $3 out / 1M
InklingOPEN
Thinking Machines Lab

A 975B-parameter open-weights MoE multimodal model from Thinking Machines Lab that processes text, images, and audio through a 1M-token context window.

$2.34 in · $5.85 out / 1M
Llama 3.2 3BOPEN
Meta

Meta's tiny open-weights workhorse — 3B parameters, 128K context, and tool use for edge and budget inference.

$0.15 in · $0.60 out / 1M
MiniMax M3 PreviewOPEN
MiniMax

MiniMax's open-weights 428B MoE with native vision, video, and sparse attention for long-context coding and agentic work.

$0.30 in · $1.20 out / 1M
Mistral Small 3.2 24B InstructOPEN
Mistral AI

Mistral's open-weight 24B instruction-tuned model with tool use, web search, and structured output — a production-ready upgrade to Small 3.1.

$0.09 in · $0.25 out / 1M
NVIDIA Nemotron 3 Nano 30BOPEN
NVIDIA

NVIDIA's open hybrid MoE model with Mamba-2 layers, configurable reasoning, and agentic tool use.

$0.07 in · $0.30 out / 1M
NVIDIA Nemotron 3 UltraOPEN
NVIDIA

NVIDIA's flagship open-weights frontier model — a 550B-parameter hybrid Mamba-MoE architecture built for agentic reasoning, tool use, and long-context throughput.

$0.63 in · $3.13 out / 1M
Nemotron Cascade 2 30B A3BOPEN
NVIDIA

NVIDIA's open 30B MoE that punches at frontier scale — gold-medal math, coding, and agentic reasoning with only 3B active parameters.

$0.14 in · $0.80 out / 1M
GLM 4.7 Flash HereticOPEN
Olafangensan (community mod; Z.AI base)

A community-abliterated, open-weights variant of GLM-4.7-Flash built for fast inference, reasoning, and tool use with relaxed refusal behavior.

$0.07 in · $0.40 out / 1M
GPT-5.2 Codex
OpenAI

OpenAI's specialized agentic coding model — long-horizon software engineering, cybersecurity, and tool use with a 256K context window.

$2.19 in · $17.50 out / 1M
GPT-5.2
OpenAI

OpenAI's flagship reasoning model for professional knowledge work, coding, and agentic tasks with tool use and web search.

$2.19 in · $17.50 out / 1M
GPT-5.3 Codex
OpenAI

OpenAI's most capable agentic coding model — autonomous software engineering with vision, reasoning, and tool use.

$2.19 in · $17.50 out / 1M
GPT-5.4 Pro
OpenAI

OpenAI's highest-performance frontier model — native computer-use, adjustable reasoning, and 1M context for complex professional work.

$37.50 in · $225 out / 1M
GPT-5.4
OpenAI

OpenAI's frontier professional model — 1M context, configurable reasoning, and multimodal agentic capabilities with vision and tool use.

$3.13 in · $18.80 out / 1M
GPT-5.5 Pro
OpenAI

OpenAI's highest-intelligence frontier model with extended test-time compute, multimodal reasoning, and a 1M-token context window for deep research and agentic work.

$37.50 in · $225 out / 1M
GPT-5.5
OpenAI

OpenAI's April 2026 frontier model for agentic coding, research, and multi-step tool use with vision and reasoning.

$6.25 in · $37.50 out / 1M
GPT-5.6 Luna Pro
OpenAI

OpenAI's cost-efficient GPT-5.6 tier with a 1M context, vision, reasoning, and tool use for high-volume workloads.

$1.25 in · $7.50 out / 1M
GPT-5.6 Luna
OpenAI

OpenAI's fastest, most cost-efficient GPT-5.6 model — 1M context, vision, tool use, and web search for high-volume agentic work.

$1.25 in · $7.50 out / 1M
GPT-5.6 Sol Pro
OpenAI

OpenAI's flagship GPT-5.6 model with a 1M context window, native multi-agent reasoning, vision, and tool use for frontier coding and knowledge work.

$6.25 in · $37.50 out / 1M
GPT-5.6 Sol
OpenAI

OpenAI's flagship GPT-5.6 model — multimodal reasoning, agentic coding, and a 1M-token context window for frontier knowledge work.

$6.25 in · $37.50 out / 1M
GPT-5.6 Terra Pro
OpenAI

OpenAI's balanced multimodal workhorse with 1,000K context, reasoning, and tool use for production workflows.

$3.13 in · $18.75 out / 1M
GPT-5.6 Terra
OpenAI

OpenAI's balanced mid-tier frontier model — 1M context, multimodal reasoning, and native tool use for production workloads.

$3.13 in · $18.75 out / 1M
OpenAI GPT OSS 120BOPEN
OpenAI

OpenAI's largest open-weight reasoning model — 117B MoE parameters, agentic tool use, and full chain-of-thought under Apache 2.0.

$0.07 in · $0.30 out / 1M
Qwen 3.7 Max
Alibaba (Qwen Team)

Alibaba's agent-frontier text model with 1M context, tool use, and code-optimized reasoning for long-horizon autonomous workflows.

$2.70 in · $8.05 out / 1M
Qwen 3 235B A22B Instruct 2507OPEN
Alibaba Cloud / Qwen Team

Alibaba's updated 235B-parameter MoE language model with 22B active params, optimized for reasoning, coding, and long-context instruction following.

$0.15 in · $0.75 out / 1M
Qwen 3 235B A22B Thinking 2507OPEN
Alibaba Cloud

Alibaba's flagship open-weight thinking MoE — 235B total, 22B active, reasoning-only mode with tool use and 262K native context.

$0.45 in · $3.50 out / 1M
Qwen 3.5 9BOPEN
Alibaba

Alibaba's 9B open-weight multimodal model with hybrid attention, 256K context, and native tool use.

$0.10 in · $0.15 out / 1M
Qwen 3.6 27B
Qwen (Alibaba Group)

A 27B dense multimodal model from Alibaba's Qwen team, optimized for agentic coding, reasoning, and long-context tasks.

$0.33 in · $3.25 out / 1M
Qwen 3 Coder 480B TurboOPEN
Alibaba Cloud (Qwen team)

Alibaba's flagship open-weight code model — a 480B-parameter MoE with 35B active, built for agentic coding and 256K-context repo work.

$0.35 in · $1.50 out / 1M
Qwen 3 Next 80bOPEN
Alibaba

Alibaba's 80B/3B sparse MoE with hybrid attention and 256K context, open-weight under Apache 2.0.

$0.35 in · $1.90 out / 1M
Qwen3 VL 235BOPEN
Alibaba Cloud

Alibaba's 235B-parameter open-weights vision-language MoE with tool use, web search, and visual agent capabilities.

$0.21 in · $1.90 out / 1M
Venice Uncensored 1.2OPEN
Venice AI

Venice's flagship 24B uncensored model — built on Mistral architecture with native vision, web search, and zero refusal behavior.

$0.20 in · $0.90 out / 1M
Venice Role Play UncensoredOPEN
Venice AI

Venice's fine-tuned 24B roleplay model, optimized for highly expressive, low-refusal character interactions with zero retention.

$0.50 in · $2 out / 1M
GLM 5 TurboOPEN
Z.ai

Z.ai's speed-optimized agentic model with selectable reasoning modes, native tool calling, and a 200K context window for OpenClaw workflows.

$1.20 in · $4 out / 1M
GLM 5V Turbo
Zhipu AI (Z.ai)

Zhipu AI's native multimodal coding model for vision-to-code and agentic workflows.

$1.50 in · $5 out / 1M
GLM 4.6OPEN
Z.ai (Zhipu AI)

Z.ai's open-weight flagship MoE model for coding, reasoning, and agentic tasks.

$0.43 in · $1.75 out / 1M
GLM 4.7 FlashOPEN
Z.ai

Z.ai's lightweight 30B MoE text model built for fast coding, tool use, and agentic tasks.

$0.13 in · $0.50 out / 1M
GLM 4.7OPEN
Z.AI

Z.AI's open-weight coding and reasoning model that runs privately on Venice with tool use and zero retention.

$0.55 in · $2.65 out / 1M
GLM 5.1OPEN
Z.AI

Z.AI's open-weights flagship LLM for agentic engineering and long-horizon coding tasks.

$1.54 in · $4.84 out / 1M
GLM 5.2OPEN
Z.ai (Zhipu AI)

Z.ai's flagship open-weights MoE with 1M context, reasoning, and tool use for long-horizon coding.

$1.40 in · $4.40 out / 1M
GLM 5OPEN
Zhipu AI

Zhipu AI's 744B-parameter open-weight MoE flagship built for agentic engineering, reasoning, and long-horizon coding tasks.

$1 in · $3.20 out / 1M

Image

14 models
ChromaOPEN
Lodestones

Open-weights text-to-image model at $0.01 per image with zero-retention privacy.

$0.01 / image
Flux 2 Max
Black Forest Labs

Black Forest Labs' flagship FLUX.2 image model — top-tier photorealism, multi-reference editing, and up to 4MP output.

$0.09 / image
Flux 2 Pro
Black Forest Labs

Black Forest Labs' flagship commercial model — 4MP photorealism, advanced prompt understanding, and precise design control.

$0.04 / image
GPT Image 2
OpenAI

OpenAI's flagship text-to-image model with built-in reasoning, near-perfect in-image text, and up to 4K output.

From $0.02 / image
Grok Imagine High Quality (SOTA)
xAI

xAI's state-of-the-art Quality Mode model delivering photorealistic textures, clean multilingual text rendering, and advanced multi-image composition.

From $0.08 / image
Grok Imagine
xAI

xAI's unified image generation and editing API for product placement, restyling, and precision edits up to 2K.

From $0.04 / image
Ideogram V4OPEN
Ideogram

Ideogram's premier 9.3B open-weight image model — the gold standard for in-image typography, structured layout control, and 2K design assets.

$0.15 / image
Lustify v8
Community

The ultimate uncensored SDXL checkpoint for photorealistic character rendering and native 1536px support.

$0.01 / image
Nano Banana 2
Google DeepMind

Google's fast, knowledgeable image model that blends Pro-grade quality with Flash-tier speed and precise text rendering.

From $0.10 / image
Qwen Image 2
Alibaba

Alibaba's unified image generation and editing model with professional bilingual typography and native 2K output.

$0.05 / image
Recraft V4 Pro
Recraft

Recraft's premium design-centric model — delivering art-directed, print-ready raster images at 2048px resolution with exceptional visual taste.

$0.29 / image
Recraft V4
Recraft

Recraft’s design-first image model — art-directed raster and editable vector SVG generation with strong typographic accuracy.

$0.05 / image
Seedream V5 Pro
ByteDance

ByteDance's production-grade image model for complex layouts, infographics, and native multilingual text rendering.

From $0.06 / image
Venice SD35OPEN
Stability AI

Venice's custom-configured Stable Diffusion 3.5 engine, delivering high-fidelity photorealism and creative freedom with zero prompt retention.

$0.01 / image

Video

6 models

Audio

1 model

Text-to-Speech

2 models

Start creating. Privately.

Every model, no prompt logging, no data used for training. Free to start — no credit card.

Room