MODELSOne Humiris key. Every production route.
humiris
HUMIRIS MODEL CLOUD

Leading AI models,
one production API.

Build with text, coding, voice, image and video models from OpenAI, Anthropic, Google, Meta, Mistral, xAI, NVIDIA, Qwen, Kimi and GLM. Humiris handles routing, governance and observability.

No provider key required OpenAI-compatible API No model markup
LIVE MODEL ROUTES 107 ready
OpenAIGPT-5.6 Sol
1.05M
OpenAIGPT-5.6 Terra
1.05M
AnthropicClaude Sonnet 5
1M
Google GeminiGemini 3.5 Flash
1M
Meta · LlamaLlama 4 Maverick
1M
Mistral AIMistral Large 3
256K
HUMIRIS AUTOPolicy-aware model selection
ONLINE
Adaptive routingCost · speed · quality
Governed by defaultIdentity · data · budgets
One API contractModels · agents · tools
PROVIDERS

Every model family.
A consistent way to build.

Browse model families, compare pricing and open the exact route you need.

OFFICIAL BENCHMARKS

Compare evidence,
not a synthetic rank.

20sourced profilesacross 107 routes

Scores come from provider model cards and launch reports. Humiris keeps benchmark versions separate, links the source and shows an explicit gap when an exact route has no comparable public result.

Coding

SWE-Bench Pro

Repository-scale software engineering tasks across multiple languages.

050100
Scores depend on the agent scaffold, tools, compute budget and benchmark revision.
01Same benchmark only

No blended score across unrelated evaluations.

02Exact version retained

2.0 and 2.1 results never share one chart.

03Provider methodology

Every published score opens its primary source.

04Missing means missing

No inferred or fabricated result for uncovered routes.

MODEL CATALOG

Find the right model for the route.

107 models · pricing per 1M tokens

OpenAI

GPT-5.6 Sol

NEW

Frontier model for complex reasoning, coding and agent execution.

ReasoningToolsVisionCoding
INPUT / 1M$5.00
OUTPUT / 1M$30.00
CONTEXT1.05M
OpenAI

GPT-5.6 Terra

NEW

Balanced frontier intelligence with faster production throughput.

ReasoningToolsVisionCoding
INPUT / 1M$2.50
OUTPUT / 1M$15.00
CONTEXT1.05M
OpenAI

GPT-5.6 Luna

NEW

Efficient reasoning for high-volume agentic applications.

ReasoningToolsVisionCoding
INPUT / 1M$1.00
OUTPUT / 1M$6.00
CONTEXT1.05M
OpenAI

GPT-5.4

High-accuracy general intelligence for demanding production work.

ReasoningToolsVisionCoding
INPUT / 1M$2.50
OUTPUT / 1M$15.00
CONTEXT256K
OpenAI

GPT-5.4 Mini

Fast multimodal reasoning with a lower cost profile.

ReasoningToolsVision
INPUT / 1M$0.75
OUTPUT / 1M$4.50
CONTEXT256K
OpenAI

GPT-5.4 Nano

Compact model for classification, extraction and routing.

ReasoningTools
INPUT / 1M$0.20
OUTPUT / 1M$1.25
CONTEXT128K
OpenAI

GPT-5 Mini

Reliable everyday intelligence at high throughput.

ReasoningTools
INPUT / 1M$0.25
OUTPUT / 1M$2.00
CONTEXT128K
OpenAI

GPT-5 Nano

The lowest-cost OpenAI route for focused tasks.

Tools
INPUT / 1M$0.050
OUTPUT / 1M$0.40
CONTEXT128K
OpenAI

GPT Image 2

NEW

State-of-the-art image generation and editing through the Images API.

ImageVision
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTImage
OpenAI

GPT-Realtime 2.1

NEW

Realtime speech model with reasoning and native tool use.

VoiceAudioReasoningTools
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTRealtime
OpenAI

GPT-Realtime 2.1 Mini

NEW

Lower-cost realtime voice route with reasoning and tool support.

VoiceAudioReasoningTools
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTRealtime
OpenAI

GPT-Realtime 2

Production speech-to-speech model for responsive voice agents.

VoiceAudioReasoningTools
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTRealtime
OpenAI

GPT-Realtime Translate

NEW

Streaming speech-to-speech translation for multilingual experiences.

VoiceAudio
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTRealtime
OpenAI

GPT-Realtime Whisper

NEW

Streaming speech recognition for live transcription workflows.

VoiceAudio
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTStreaming
OpenAI

GPT-4o Transcribe

High-accuracy speech-to-text model for recorded audio.

VoiceAudio
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTAudio
OpenAI

GPT-4o Mini Transcribe

Fast, economical transcription for high-volume audio.

VoiceAudio
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTAudio
OpenAI

Text Embedding 3 Large

High-quality vector representations for retrieval, clustering and RAG.

Embeddings
INPUT / 1M$0.13
OUTPUT / 1MRoute pricing
CONTEXT8K
OpenAI

Text Embedding 3 Small

Cost-efficient embeddings for semantic search and classification.

Embeddings
INPUT / 1M$0.020
OUTPUT / 1MRoute pricing
CONTEXT8K
Anthropic

Claude Sonnet 5

NEW

Advanced reasoning and coding with strong agent reliability.

ReasoningToolsVisionCoding
INPUT / 1M$2.00
OUTPUT / 1M$10.00
CONTEXT1M
Anthropic

Claude Fable 5

NEW

Anthropic's highest-capability route for complex knowledge work.

ReasoningToolsVisionLong context
INPUT / 1M$10.00
OUTPUT / 1M$50.00
CONTEXT1M
Anthropic

Claude Opus 4.8

Deep reasoning for long-running engineering and research tasks.

ReasoningToolsVisionCoding
INPUT / 1M$5.00
OUTPUT / 1M$25.00
CONTEXT1M
Anthropic

Claude Opus 4.7

Reliable high-depth analysis and agentic coding.

ReasoningToolsVisionCoding
INPUT / 1M$5.00
OUTPUT / 1M$25.00
CONTEXT200K
Anthropic

Claude Opus 4.6

High-capability Claude route for long-running coding and enterprise tasks.

ReasoningToolsVisionCoding
INPUT / 1M$5.00
OUTPUT / 1M$25.00
CONTEXT200K
Anthropic

Claude Opus 4.5

Deep reasoning model for complex engineering and research workflows.

ReasoningToolsVisionCoding
INPUT / 1M$5.00
OUTPUT / 1M$25.00
CONTEXT200K
Anthropic

Claude Sonnet 4.6

Balanced intelligence for coding and business workflows.

ReasoningToolsVisionCoding
INPUT / 1M$3.00
OUTPUT / 1M$15.00
CONTEXT200K
Anthropic

Claude Sonnet 4.5

Reliable coding and analysis model with broad cloud availability.

ReasoningToolsVisionCoding
INPUT / 1M$3.00
OUTPUT / 1M$15.00
CONTEXT200K
Anthropic

Claude Haiku 4.5

Fast Claude model for responsive user-facing experiences.

ToolsVision
INPUT / 1M$1.00
OUTPUT / 1M$5.00
CONTEXT200K
Google Gemini

Gemini 3.5 Flash

NEW

Native multimodal reasoning with a one-million-token context.

ReasoningToolsVisionLong context
INPUT / 1M$1.50
OUTPUT / 1M$9.00
CONTEXT1M
Google Gemini

Gemini 3.1 Pro Preview

NEW

High-depth preview model for complex multimodal work.

ReasoningToolsVisionCoding
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT1M
Google Gemini

Gemini 3 Flash Preview

Fast preview route for high-volume multimodal applications.

ReasoningToolsVisionLong context
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT1M
Google Gemini

Gemini 3.1 Flash-Lite

Cost-efficient Gemini route for extraction and classification.

ReasoningToolsVisionLong context
INPUT / 1M$0.25
OUTPUT / 1M$1.50
CONTEXT1M
Google Gemini

Gemini 2.5 Flash

Stable price-performance model for low-latency multimodal reasoning.

ReasoningToolsVisionLong context
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT1M
Google Gemini

Gemini 2.5 Flash-Lite

Fastest budget Gemini route for high-throughput multimodal work.

ToolsVisionLong context
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT1M
Google Gemini

Nano Banana 2

NEW

High-efficiency native image generation and editing.

ImageVision
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTImage
Google Gemini

Nano Banana 2 Lite

NEW

Ultra-low-latency image generation for interactive applications.

ImageVision
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTImage
Google Gemini

Nano Banana Pro

High-context native image generation and editing for premium creative work.

ImageVision
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTImage
Google Gemini

Gemini 3.1 Flash Live

NEW

Low-latency Live API model for voice-first multimodal applications.

VoiceAudioVisionTools
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTRealtime
Google Gemini

Gemini Embedding 2

NEW

Unified embeddings for text, image, video, audio and document retrieval.

EmbeddingsVisionAudio
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTMultimodal
Meta · Llama

Llama 4 Maverick

Large multimodal open-weight model for general workloads.

ToolsVisionOpen weightsLong context
INPUT / 1M$0.27
OUTPUT / 1M$0.85
CONTEXT1M
Meta · Llama

Llama 4 Scout

Efficient multimodal Llama route with broad context.

ToolsVisionOpen weightsLong context
INPUT / 1M$0.18
OUTPUT / 1M$0.59
CONTEXT512K
Meta · Llama

Llama 3.3 70B Turbo

Production-tuned open model for chat and agent workflows.

ToolsOpen weights
INPUT / 1M$0.88
OUTPUT / 1M$0.88
CONTEXT128K
Meta · Llama

Llama 3.3 8B

Fine-tunable compact Llama model for specialized private workloads.

ToolsOpen weights
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT128K
Meta · Llama

Llama 3.2 90B Vision

Large open multimodal model for image and document understanding.

VisionToolsOpen weights
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT128K
Meta · Llama

Llama 3.2 11B Vision

Efficient open vision model for private multimodal applications.

VisionToolsOpen weights
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT128K
Meta · Llama

Llama 3.1 405B

Large open-weight foundation model for demanding inference workloads.

ToolsOpen weightsCoding
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT128K
Meta · Llama

Llama 3.1 8B Instant

Ultra-fast open route for simple high-volume tasks.

ToolsOpen weights
INPUT / 1M$0.050
OUTPUT / 1M$0.080
CONTEXT128K
Mistral AI

Mistral Large 3

Mistral's production generalist for multilingual work.

ToolsCodingLong context
INPUT / 1M$0.50
OUTPUT / 1M$1.50
CONTEXT256K
Mistral AI

Mistral Medium 3.5

NEW

Deep reasoning with efficient European deployment options.

ReasoningToolsCoding
INPUT / 1M$1.50
OUTPUT / 1M$7.50
CONTEXT256K
Mistral AI

Mistral Small 4

NEW

Compact production model for agents and structured output.

ReasoningToolsCoding
INPUT / 1M$0.15
OUTPUT / 1M$0.60
CONTEXT128K
Mistral AI

Ministral 3 14B

Open text-and-vision model for capable private deployments.

VisionToolsOpen weights
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT256K
Mistral AI

Ministral 3 8B

Efficient open model for edge and private deployments.

ToolsOpen weights
INPUT / 1M$0.15
OUTPUT / 1M$0.15
CONTEXT128K
Mistral AI

Ministral 3 3B

Small open-weight route for focused automations.

ToolsOpen weights
INPUT / 1M$0.10
OUTPUT / 1M$0.10
CONTEXT64K
Mistral AI

Devstral 2

Frontier code-agent model for repository-scale software engineering.

ReasoningToolsCoding
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT256K
Mistral AI

Codestral

Code completion model optimized for low-latency developer tooling.

Coding
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT256K
Mistral AI

Voxtral TTS

Multilingual text-to-speech with zero-shot voice cloning.

VoiceAudio
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTAudio
Mistral AI

Voxtral Mini Transcribe Realtime

NEW

Open realtime transcription model for streaming audio.

VoiceAudioOpen weights
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTRealtime
Mistral AI

Voxtral Mini Transcribe 2

Efficient transcription model for recorded speech.

VoiceAudio
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTAudio
Mistral AI

Mistral OCR 4

NEW

Document parsing with structural blocks and paragraph-level coordinates.

VisionTools
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTDocuments
Mistral AI

Mistral Moderation 2

Safety classification with jailbreak and policy-risk detection.

Safety
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT128K
Mistral AI

Mistral Embed

Semantic vectors for multilingual retrieval and RAG.

Embeddings
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT8K
Qwen

Qwen 3.7 Max

NEW

Qwen's flagship reasoning and coding model.

ReasoningToolsVisionCoding
INPUT / 1M$2.50
OUTPUT / 1M$7.50
CONTEXT1M
Qwen

Qwen 3.6 Plus

Long-context multilingual reasoning at efficient cost.

ReasoningToolsCodingLong context
INPUT / 1M$0.50
OUTPUT / 1M$3.00
CONTEXT1M
Qwen

Qwen 3.6 Flash

Fast multimodal Qwen route for interactive products.

ToolsVisionLong context
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT1M
Qwen

Qwen 3.5 397B A17B

Open mixture-of-experts model for reasoning and code.

ReasoningToolsOpen weightsCoding
INPUT / 1M$0.60
OUTPUT / 1M$3.60
CONTEXT256K
Qwen

Qwen 3 235B A22B

Large open mixture-of-experts model for math, code and agents.

ReasoningToolsOpen weightsCoding
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT128K
Qwen

Qwen 3 30B A3B

Efficient open MoE route for local reasoning and tool use.

ReasoningToolsOpen weightsCoding
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT128K
Qwen

Qwen 2.5 VL 32B

Open vision-language model for documents, images and visual reasoning.

VisionReasoningOpen weights
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT128K
Kimi

Kimi K2.5

Multimodal Moonshot model for reasoning, coding and agent workflows.

ReasoningToolsVisionCoding
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT256K
Kimi

Kimi K3

NEW

Moonshot's flagship model for long-horizon agent work.

ReasoningToolsCodingLong context
INPUT / 1M$3.00
OUTPUT / 1M$15.00
CONTEXT1M
Kimi

Kimi K2.7 Code

NEW

Specialized coding model for repository-scale tasks.

ReasoningToolsCodingLong context
INPUT / 1M$0.95
OUTPUT / 1M$4.00
CONTEXT512K
Kimi

Kimi K2.6

Balanced agentic reasoning with reliable tool use.

ReasoningToolsCoding
INPUT / 1M$0.95
OUTPUT / 1M$4.00
CONTEXT256K
GLM · Z.ai

GLM 5.2

NEW

Long-horizon reasoning and software engineering model.

ReasoningToolsCodingLong context
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT1M
GLM · Z.ai

GLM 5.1 HighSpeed

Low-latency GLM route for interactive agents.

ReasoningToolsCoding
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT256K
GLM · Z.ai

GLM 5 Turbo

Fast general reasoning and multilingual tool use.

ReasoningToolsCoding
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT256K
GLM · Z.ai

GLM 5.1

NEW

Agentic engineering model for sustained autonomous software work.

ReasoningToolsCoding
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT200K
GLM · Z.ai

GLM 5

Open-weight foundation model for long-horizon engineering agents.

ReasoningToolsCodingOpen weights
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT200K
GLM · Z.ai

GLM 5V Turbo

NEW

Multimodal coding model specialized for visual programming.

ReasoningToolsVisionCoding
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT200K
GLM · Z.ai

GLM 4.7

Stable multi-step reasoning and agentic coding model.

ReasoningToolsCoding
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT200K
GLM · Z.ai

GLM 4.7 Flash

Lightweight high-speed GLM route for interactive agent applications.

ReasoningToolsCoding
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT200K
GLM · Z.ai

GLM 4.6V

Vision-language model with native function calling and thinking control.

ReasoningToolsVision
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT128K
GLM · Z.ai

GLM Image

Open image generation model for complex prompts and text rendering.

ImageVisionOpen weights
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTImage
GLM · Z.ai

GLM OCR

Document parsing and structured information extraction model.

VisionTools
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTDocuments
xAI · Grok

Grok 4.5

NEW

xAI's flagship reasoning model for agents, code and multimodal work.

ReasoningToolsVisionCoding
INPUT / 1M$2.00
OUTPUT / 1M$6.00
CONTEXT500K
xAI · Grok

Grok Build 0.1

NEW

Coding-focused Grok route for repository work and software agents.

ReasoningToolsCoding
INPUT / 1M$1.00
OUTPUT / 1M$2.00
CONTEXT256K
xAI · Grok

Grok 4.3

High-throughput general reasoning with a one-million-token context.

ReasoningToolsVisionLong context
INPUT / 1M$1.25
OUTPUT / 1M$2.50
CONTEXT1M
xAI · Grok

Grok 4.20 Multi-Agent

NEW

Native multi-agent orchestration for parallel research and execution.

ReasoningToolsCodingLong context
INPUT / 1M$1.25
OUTPUT / 1M$2.50
CONTEXT1M
xAI · Grok

Grok Imagine Image Quality

NEW

High-quality image generation and editing up to 2K resolution.

ImageVision
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT2K image
xAI · Grok

Grok Imagine Image

Fast image generation and editing for interactive products.

ImageVision
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT2K image
xAI · Grok

Grok Imagine Video 1.5

NEW

Image-to-video generation with 480p, 720p and 1080p outputs.

VideoVision
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT1080p
xAI · Grok

Grok Imagine Video

Text, image and video conditioned generation and video extension.

VideoVision
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT720p
xAI · Grok

Grok Voice Think Fast 1.0

NEW

Realtime voice agent model with fast reasoning and tool use.

VoiceAudioReasoningTools
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTRealtime
xAI · Grok

Grok Text to Speech

Natural speech synthesis through xAI's Text to Speech API.

VoiceAudio
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTAudio
xAI · Grok

Grok Speech to Text

Batch and streaming speech recognition across 25 languages.

VoiceAudio
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTStreaming
NVIDIA

Nemotron 3 Ultra 550B A55B

NEW

Frontier-scale open reasoning model for complex agents and long-context code.

ReasoningToolsCodingOpen weights
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT1M
NVIDIA

Nemotron 3 Super 120B A12B

Efficient open MoE for reasoning, planning, coding and tool calling.

ReasoningToolsCodingOpen weights
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXT1M
NVIDIA

Nemotron 3 Nano Omni 30B

NEW

Omnimodal model that understands video, speech, images and text.

ReasoningVisionVideoAudio
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTMultimodal
NVIDIA

Nemotron 3 VoiceChat

NEW

Full-duplex speech-to-speech model that listens and speaks simultaneously.

VoiceAudioReasoningOpen weights
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTRealtime
NVIDIA

Nemotron ASR Streaming

Low-latency streaming speech recognition for enterprise voice systems.

VoiceAudioOpen weights
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTStreaming
NVIDIA

Magpie TTS Zeroshot

Expressive zero-shot text-to-speech from a short voice sample.

VoiceAudioOpen weights
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTAudio
NVIDIA

Cosmos 3 Nano

NEW

Physics-aware video world generation from text or image prompts.

VideoVisionOpen weights
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTVideo
NVIDIA

Cosmos 3 Nano Reasoner

Structured physical-world reasoning over video and image inputs.

VideoVisionReasoningOpen weights
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTVideo
NVIDIA

Nemotron OCR v1

Fast OCR with layout, table and document structure extraction.

VisionOpen weights
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTDocuments
Google Gemini

Gemini 3.1 Flash TTS

NEW

Low-latency, steerable speech generation with expressive audio tags.

VoiceAudio
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTAudio
Google Gemini

Gemini 3.5 Live Translate

NEW

Realtime speech-to-speech translation across more than 70 languages.

VoiceAudio
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTRealtime
Google Gemini

Gemini Omni Flash Preview

NEW

Conversational video generation and multi-turn editing from text and images.

VideoVisionAudioReasoning
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTVideo
Google Gemini

Veo 3.1 Preview

Cinematic video generation with synchronized audio and reference controls.

VideoVisionAudio
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTVideo + audio
Google Gemini

Veo 3.1 Lite Preview

Lower-cost Veo route for production video generation and editing.

VideoVisionAudio
INPUT / 1MRoute pricing
OUTPUT / 1MRoute pricing
CONTEXTVideo + audio
ONE KEY · EVERY ROUTE

Stop integrating models one by one.

Use Humiris managed access or bring provider keys when your governance policy requires it.

Get started Manage providers