Skip to content
InnerZero logoInnerZero

Supported AI models

InnerZero runs open-source language models directly on your PC via Ollama, LM Studio, or your own llama.cpp server. The main families are Qwen 3, Gemma 3, and gpt-oss, covering everything from a 1B voice model on a basic laptop to a 235B reasoning model on a datacenter-class workstation. Local AI is free, no subscription, no account required.

Cloud support is optional. If you want frontier-quality models for hard tasks, subscribe to InnerZero's managed cloud plans, or bring your own API keys for any of seven providers with zero markup. Keys stay encrypted on your machine.

Local models

Installed automatically based on your hardware tier. These run entirely on your machine, offline if you want.

Qwen 3

4B8B14B30B32B235B
Minimum hardware
Basic tier
Backend
Ollama, LM Studio, or llama.cpp

Apache 2.0

Official page

Gemma 3

1B4B12B
Minimum hardware
Entry tier
Backend
Ollama, LM Studio, or llama.cpp

Gemma Terms of Use

Official page

gpt-oss

120B
Minimum hardware
Pro tier
Backend
Ollama, LM Studio, or llama.cpp

Apache 2.0

Official page

How we choose supported models

Supported local families run reliably on consumer hardware, ship under permissive licences, are actively maintained by their upstream teams, and work cleanly through Ollama, LM Studio, or llama.cpp without custom glue. We prefer families with a range of sizes so the same recipe scales from a budget laptop up to a frontier workstation. Models that satisfy those constraints land in the default hardware tiers and get tested as new versions ship.

Cloud models

This section is optional, and it is the bring-your-own-key catalogue: add your own provider API key and your prompts go directly to that provider, at the provider's own prices with no markup. Managed cloud plans are separate: they run on InnerZero's own cloud service with monthly credits and per-tier routing rather than a per-model choice. See pricing for the plans.

DeepSeek

DeepSeek V4 Flash (chat)deepseek-v4-flash
DeepSeek V4 Flash (reasoning)deepseek-v4-flash-thinking
DeepSeek V4 Prodeepseek-v4-pro

OpenAI

GPT-5.6 Solgpt-5.6-sol
GPT-5.6 Terragpt-5.6-terra
GPT-5.6 Lunagpt-5.6-luna
GPT-5.5gpt-5.5
GPT-5.5 Progpt-5.5-pro
GPT-5.4gpt-5.4
GPT-5.4 Minigpt-5.4-mini
GPT-5.4 Nanogpt-5.4-nano
GPT-4.1gpt-4.1
GPT-4.1 Minigpt-4.1-mini
GPT-4.1 Nanogpt-4.1-nano
GPT-4ogpt-4o
o3o3
o4 Minio4-mini
GPT-5.3 Codexgpt-5.3-codex
GPT-5.3 Chatgpt-5.3-chat

Anthropic

Claude Opus 4.8claude-opus-4-8
Claude Sonnet 5claude-sonnet-5
Claude Fable 5claude-fable-5
Claude Opus 4.7claude-opus-4-7
Claude Opus 4.6claude-opus-4-6
Claude Sonnet 4.6claude-sonnet-4-6
Claude Haiku 4.5claude-haiku-4-5-20251001

Google

Gemini 3.5 Flashgemini-3.5-flash
Gemini 3.1 Flash Litegemini-3.1-flash-lite
Gemini 3.1 Pro (Preview)gemini-3.1-pro-preview
Gemini 2.5 Progemini-2.5-pro
Gemini 2.5 Flashgemini-2.5-flash

Qwen

Qwen 3.7 Maxqwen3.7-max
Qwen 3.7 Plusqwen3.7-plus
Qwen Maxqwen-max
Qwen Plusqwen-plus
Qwen Turboqwen-turbo
Qwen3 Coder Plusqwen3-coder-plus

xAI

Grok 4.5grok-4.5
Grok 4.3grok-4.3

Kimi

Kimi K3kimi-k3
Kimi K2.6kimi-k2.6
Kimi K2.7 Codekimi-k2.7-code
Kimi K2.5kimi-k2.5

Model support last reviewed 16 August 2026. Pricing is set by each provider and changes over time; follow the pricing links above for current rates. Local models are free to download and run.