Qwen 3
- Minimum hardware
- Basic tier
- Backend
- Ollama, LM Studio, or llama.cpp
Apache 2.0
Official pageInnerZero runs open-source language models directly on your PC via Ollama, LM Studio, or your own llama.cpp server. The main families are Qwen 3, Gemma 3, and gpt-oss, covering everything from a 1B voice model on a basic laptop to a 235B reasoning model on a datacenter-class workstation. Local AI is free, no subscription, no account required.
Cloud support is optional. If you want frontier-quality models for hard tasks, subscribe to InnerZero's managed cloud plans, or bring your own API keys for any of seven providers with zero markup. Keys stay encrypted on your machine.
Installed automatically based on your hardware tier. These run entirely on your machine, offline if you want.
Apache 2.0
Official pageGemma Terms of Use
Official pageApache 2.0
Official pageSupported local families run reliably on consumer hardware, ship under permissive licences, are actively maintained by their upstream teams, and work cleanly through Ollama, LM Studio, or llama.cpp without custom glue. We prefer families with a range of sizes so the same recipe scales from a budget laptop up to a frontier workstation. Models that satisfy those constraints land in the default hardware tiers and get tested as new versions ship.
This section is optional, and it is the bring-your-own-key catalogue: add your own provider API key and your prompts go directly to that provider, at the provider's own prices with no markup. Managed cloud plans are separate: they run on InnerZero's own cloud service with monthly credits and per-tier routing rather than a per-model choice. See pricing for the plans.
gpt-5.6-solgpt-5.6-terragpt-5.6-lunagpt-5.5gpt-5.5-progpt-5.4gpt-5.4-minigpt-5.4-nanogpt-4.1gpt-4.1-minigpt-4.1-nanogpt-4oo3o4-minigpt-5.3-codexgpt-5.3-chatModel support last reviewed 16 August 2026. Pricing is set by each provider and changes over time; follow the pricing links above for current rates. Local models are free to download and run.