AI Chat & Models

GizAIAuto
Auto routing
Modalities
text / image / file → text
Capabilities
Multimodal routing · Tools · Reasoning
Routing
Best fit per request
Billing
Selected model rate

Lets GizAI route each request to the most suitable chat model automatically.

OpenAIGPT-5.6 Luna
Free 5/1 dayafter $0.2 in · $1.2 out / 1M
Modalities
file / image / text / pdf → text
Released
Jul 9, 2026
In / out price
$0.2 in · $1.2 out / 1M
Context
1.05M
Max output
128K
Capabilities
Tools · Reasoning · Structured output

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...

xAIGrok 4.6
Free 2/2 hrafter $2 in · $6 out / 1M
Modalities
text / image / file → text
Released
Aug 12, 2026
In / out price
$2 in · $6 out / 1M
Context
500K
Max output
450K
Capabilities
Tools · Reasoning · Structured output

Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.

DeepSeekDeepSeek V4 Flash 0731
Free 6/1 hrafter $0.0798 in · $0.161 out / 1M
Modalities
text / image → text
Released
Jul 31, 2026
In / out price
$0.0798 in · $0.161 out / 1M
Context
1M
Max output
384K
Capabilities
Tools · Reasoning · Structured output

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....

DeepSeekDeepSeek V4 Pro
Free 6/1 hrafter $0.87 in · $1.74 out / 1M
Modalities
text / image / file → text
Released
Apr 24, 2026
In / out price
$0.87 in · $1.74 out / 1M
Context
1M
Max output
384K
Capabilities
Tools · Reasoning · Structured output

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...

GoogleGemini 3.5 Flash-Lite
Free 2/1 dayafter $0.15 in · $1.25 out / 1M
Modalities
text / image / video / file / audio / pdf → text
Released
Jul 21, 2026
In / out price
$0.15 in · $1.25 out / 1M
Context
1.05M
Max output
65.5K
Capabilities
Tools · Reasoning · Structured output

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

GoogleGemma 4 31B
Free 2/1 hrafter $0.09 in · $0.34 out / 1M
Modalities
image / text / video / pdf → text
Released
Apr 2, 2026
In / out price
$0.09 in · $0.34 out / 1M
Context
262K
Max output
16.4K
Capabilities
Tools · Reasoning · Structured output

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

Z.aiGLM 4.7 Flash
Free 1/1 hrafter $0.06 in · $0.4 out / 1M
Modalities
text → text
Released
Jan 19, 2026
In / out price
$0.06 in · $0.4 out / 1M
Context
203K
Max output
16.4K
Capabilities
Tools · Reasoning · Structured output

As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...

MiniMaxMiniMax M3
Free 1/1 hrafter $0.23 in · $0.96 out / 1M
Modalities
text / image / video / pdf → text
Released
May 31, 2026
In / out price
$0.23 in · $0.96 out / 1M
Context
1.05M
Max output
512K
Capabilities
Tools · Reasoning · Structured output

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

Moonshot AIKimi K2.6
Free 1/1 hrafter $0.537 in · $2.26 out / 1M
Modalities
text / image / video → text
Released
Apr 20, 2026
In / out price
$0.537 in · $2.26 out / 1M
Context
262K
Max output
8.19K
Capabilities
Tools · Reasoning · Structured output

Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...

AlibabaQwen3.7 Plus
Free 1/2 hrafter $0.32 in · $1.28 out / 1M
Modalities
text / image / pdf / file → text
Released
Jun 3, 2026
In / out price
$0.32 in · $1.28 out / 1M
Context
1M
Max output
131K
Capabilities
Tools · Reasoning · Structured output

Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series. It supports text and image input with text output, building on the series' text capabilities with a comprehensive upgrade to its...

OpenAIGPT-5.6 Terra
Upgrade required
Modalities
file / image / text / pdf → text
Released
Jul 9, 2026
In / out price
$2 in · $12 out / 1M
Context
1.05M
Max output
128K
Capabilities
Tools · Reasoning · Structured output

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...

OpenAIGPT-5.6 Sol
Upgrade required
Modalities
file / image / text / pdf → text
Released
Jul 9, 2026
In / out price
$2 in · $10 out / 1M
Context
1.05M
Max output
128K
Capabilities
Tools · Reasoning · Structured output

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...

AnthropicClaude Sonnet 5
Upgrade required
Modalities
text / image / file / pdf → text
Released
Jun 30, 2026
In / out price
$0.0235 in · $0.118 out / 1M
Context
1M
Max output
128K
Capabilities
Tools · Reasoning · Structured output

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

AnthropicClaude Opus 5
Upgrade required
Modalities
text / image / file / pdf → text
Released
Jul 24, 2026
In / out price
$0.0588 in · $0.294 out / 1M
Context
1M
Max output
128K
Capabilities
Tools · Reasoning · Structured output

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

GoogleGemini 3.7 Flash
Upgrade required
Modalities
text / image / video / file / audio / pdf → text
Released
Aug 13, 2026
In / out price
$0.188 in · $0.938 out / 1M
Context
1.05M
Max output
65.5K
Capabilities
Tools · Reasoning · Structured output

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Z.aiGLM 5.2
Upgrade required
Modalities
text / image / file → text
Released
Jun 16, 2026
In / out price
$0.487 in · $1.56 out / 1M
Context
1.05M
Max output
262K
Capabilities
Tools · Reasoning · Structured output

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

InclusionAILing 3.0 Flash Fin
$0 in · $0 out / 1M
Modalities
text → text
Released
Aug 27, 2026
In / out price
$0 in · $0 out / 1M
Context
256K
Max output
32K
Capabilities
Tools · Reasoning

Ling 3.0 Flash Fin is InclusionAI’s finance-enhanced MoE language model, combining 124 billion total parameters with approximately 5.1 billion active parameters for efficient financial reasoning. Its 256K context window, function calling, and support for complex, multi-step investment workflows make it ideal for financial research, analysis, long-horizon planning, and execution, while retaining strong capabilities in coding and mathematics.

InclusionAILing 3.0 Flash Fin (Free)
$0 in · $0 out / 1M
Modalities
text → text
Released
Aug 27, 2026
In / out price
$0 in · $0 out / 1M
Context
256K
Max output
32K
Capabilities
Tools · Reasoning

Ling 3.0 Flash Fin is InclusionAI’s finance-enhanced MoE language model, combining 124 billion total parameters with approximately 5.1 billion active parameters for efficient financial reasoning. Its 256K context window, function calling, and support for complex, multi-step investment workflows make it ideal for financial research, analysis, long-horizon planning, and execution, while retaining strong capabilities in coding and mathematics.

AlibabaQwen3.8 Flash
$0.15 in · $0.47 out / 1M
Modalities
text / image / video / pdf → text
Released
Aug 26, 2026
In / out price
$0.15 in · $0.47 out / 1M
Context
1M
Max output
131K
Capabilities
Tools · Reasoning · Structured output

Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.

Z.aiGLM 5.3 Flash
$0.075 in · $0.25 out / 1M
Modalities
text / image / video → text
Released
Aug 26, 2026
In / out price
$0.075 in · $0.25 out / 1M
Context
1.31M
Max output
131K
Capabilities
Tools · Reasoning · Structured output

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

IBMgranite-4.2-30b
$0.168 in · $0.682 out / 1M
Modalities
text → text
Released
Aug 24, 2026
In / out price
$0.168 in · $0.682 out / 1M
Context
131K
Max output
131K
Capabilities
Tools · Structured output

Granite-4.2-30B is the flagship reasoning model in the Granite 4.2 family. It delivers the strongest performance across reasoning-intensive tasks by leveraging built-in <think>...</think> chain-of-thought. It supports flexible thinking modes — full thinking (default), non-thinking, and low-effort — allowing users to balance depth vs. latency on a per-query basis.

IBMgranite-4.2-8b
$0.063 in · $0.263 out / 1M
Modalities
text → text
Released
Aug 24, 2026
In / out price
$0.063 in · $0.263 out / 1M
Context
131K
Max output
131K
Capabilities
Tools · Structured output

Granite-4.2-8B is the mid-size reasoning model in the Granite 4.2 family. It delivers strong performance on reasoning-intensive tasks by leveraging built-in <think>...</think> chain-of-thought. It supports flexible thinking modes — full thinking (default), non-thinking, and low-effort — allowing users to balance depth vs. latency on a per-query basis.

IBMgranite-4.2-3b
$0.0315 in · $0.126 out / 1M
Modalities
text → text
Released
Aug 24, 2026
In / out price
$0.0315 in · $0.126 out / 1M
Context
131K
Max output
131K
Capabilities
Tools · Structured output

Granite-4.2-3B is the compact reasoning model in the Granite 4.2 family. Despite its small parameter count, it delivers strong performance on reasoning-intensive tasks by leveraging built-in <think>...</think> chain-of-thought. It supports flexible thinking modes — full thinking (default), non-thinking, and low-effort — allowing users to balance depth vs. latency on a per-query basis.