GizAI
Auto
Modalities
text / image / file → text
Capabilities
Multimodal routing · Tools · Reasoning
Routing
Best fit per request
Billing
Selected model rate
Lets GizAI route each request to the most suitable chat model automatically.
Google
Nano Banana 2
text → text / image
Gemini image model for visual reasoning, prompt-guided edits, and creative image output.
Gemini 3.5 Flash-Lite
text →
Cost-efficient Gemini model for high-volume agentic tasks, multimodal analysis, and document extraction.
Gemini 3.7 Flash
Fast Gemini reasoning model for agentic work, coding, multimodal analysis, and quick synthesis.
Gemma 4 31B
Google DeepMind open model for efficient multilingual chat, coding, and concise reasoning.
text / image / video / file / audio → text
Released
Aug 14, 2026
Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that
Gemini 3.7 Flash (batch)
gemma-4-31B-it-Ultra
text / image → text
Jul 27, 2026
Ultra speed version of gemma-4-31B-it
Gemini 3.6 Flash
text / image / video / file / audio / pdf → text
Jul 21, 2026
Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished
Gemini 3.6 Flash (batch)
Gemini 3.5 Flash Lite
Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks
Gemini 3.5 Flash Lite (batch)
gemma-4-E4B-it
text → text
Jul 15, 2026
DeepInfra-hosted google/gemma-4-E4B-it text-generation model.
Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image)
text / image → text / image
Jun 30, 2026
Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) is Google's fastest, most cost-efficient Gemini image model, built for high-velocity developer pipeli
Gemini Omni Flash
text / image / pdf / video → text / video
Gemini Omni Flash (Preview) is a multimodal model designed for video, image, and text tasks. It is optimized for video generation, offering video outp
Nano Banana 2 (Gemini 3.1 Flash Image)
Jun 18, 2026
Gemini 3.1 Flash Image, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual qu
Nano Banana Pro (Gemini 3 Pro Image)
Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro. It extends the original Nano Banana with signific
Gemini 3.5 Flash
May 19, 2026
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly
Gemini 3.5 Flash (batch)
Gemini 3.1 Flash Lite
May 7, 2026
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video
Gemini 3.1 Flash Lite (batch)
Gemma 4 26B A4B
image / text / video / pdf → text
Apr 3, 2026
Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per
Apr 2, 2026
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context
Gemini 3.1 Flash Lite Preview
Mar 3, 2026
Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall q
Gemini 3.1 Pro Preview Custom Tools
text / audio / image / video / file → text
Feb 26, 2026
Gemini 3.1 Pro Preview Custom Tools is a variant of Gemini 3.1 Pro that improves tool selection behavior by preventing overuse of a general bash tool
Gemini 3.1 Pro Preview
audio / file / image / text / video / pdf → text
Feb 19, 2026
Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and m
Gemini 3.1 Pro Preview (batch)
audio / file / image / text / video → text
Gemini 3 Flash Preview
text / image / file / audio / video / pdf → text
Dec 17, 2025
Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers
Gemini 3 Flash Preview (batch)
text / image / file / audio / video → text
Nano Banana (Gemini 2.5 Flash Image)
Oct 8, 2025
Gemini 2.5 Flash Image, a.k.a. "Nano Banana," is now generally available. It is a state of the art image generation model with contextual understandin
Gemini 2.5 Flash Lite
Jul 22, 2025
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improv
Gemini 2.5 Flash Lite (batch)
Gemini 2.5 Flash
file / image / text / audio / video / pdf → text
Jun 17, 2025
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks
Gemini 2.5 Flash (batch)
file / image / text / audio / video → text
Gemini 2.5 Pro
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking”
Gemini 2.5 Pro (batch)
Gemini 2.5 Pro Preview 06-05
file / image / text / audio → text
Jun 5, 2025
Gemma 3n 4B
May 21, 2025
Gemma 3n E4B-it is optimized for efficient execution on mobile and low-resource devices, such as phones, laptops, and tablets. It supports multimodal
Gemini 2.5 Pro Preview 05-06
May 7, 2025
Gemma 3 4B
Mar 14, 2025
Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 14
Gemma 3 12B
Gemma 3 27B
Mar 12, 2025
Gemma 2 27B
Jul 13, 2024
Gemma 2 27B by Google is an open model built from the same research and technology used to create the Gemini models. Gemma models are well-suited for
Gemini 2.0 Flash
Gemini 2.0 Flash Lite
Deepseek
DeepSeek V4 Flash 0731
Official DeepSeek V4 Flash 0731 release with enhanced agent and coding capabilities.
DeepSeek V4 Pro
DeepSeek V4 Pro model for agentic tool use, long-context reasoning, coding, and math.
Zai
GLM 5.2
Stronger GLM model for coding, math, multilingual reasoning, and multimodal analysis.
MiniMax
MiniMax M3
MiniMax model for efficient chat, writing, roleplay, and high-volume assistant workloads.
OpenAI
GPT-5.6 Luna
Efficient GPT-5.6 model for cost-sensitive, high-volume workloads.
GPT-5.6 Terra
Balanced GPT-5.6 model for agentic coding, analysis, and production knowledge work.
GPT-5.6 Sol
Frontier GPT-5.6 model for complex coding, research, and high-depth reasoning.
GPT-4o
Omni GPT model for natural conversation, image understanding, writing, and creative work.
Anthropic
Claude Sonnet 5
Advanced Claude model for agentic coding, reasoning, document work, and production workflows.
Claude Opus 5
Highest-capability Claude model for complex agents, coding, research, and sustained reasoning.
xAI
Grok 4.6
xAI's frontier model for agentic coding, analysis, knowledge work, and multimodal prompts.
GLM 4.7 Flash
Fast GLM model for inexpensive multilingual chat and lightweight reasoning.
Moonshotai
Kimi K2.6
Kimi long-context model for document-heavy chat, coding, structured reasoning, and synthesis.
Alibaba
Qwen3.7 Plus
Alibaba MoE reasoning model for math, coding, multilingual analysis, and structured answers.
ByteDance
Seed 2.1 Turbo
text / image / video → text
Aug 12, 2026
Seed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows. It is suited for end-to-end software delivery, m
Using Auto · free