PRICING — SOVEREIGN AI

Sovereign AI.
Transparent Pricing.

Two ways to put frontier AI to work without giving up control of your data: consume models through the IG1 AI API, or run your own FullStack AI platform.

IG1 AI API Sovereign API

IG1 AI API

Frontier models through a sovereign API hosted in France. No infrastructure to set up — applications connect instantly.

Monthly subscription

100 €/ month
  • Included usage100 €
  • Beyond thatPer-token rates
  • Monthly ceilingScales with your subscription1 000 €
  • Minimum commitmentNone
FullStack AI Dedicated platform

FullStack AI

Your own private AI platform: sovereign infrastructure, intelligent software, and expert support, deployed securely at scale.

Starting at

$3,500/month
  • On your infrastructurePublic cloud or your hardware$3,500 /month
  • On our infrastructureNvidia GPU$4,900 /month
  • CustomTailored deploymentOn quote

IG1 AI API

Predictable base.
Pay for what you use.

A flat monthly subscription with usage included, per-token rates beyond it, and a ceiling you control. Light months stay light, heavy months stay capped.

01

100 € / month

Flat subscription that includes 100 € of usage across every model in the catalog.

02

Per-token rates

Beyond the included allowance, usage is metered at the published rates below.

03

1 000 € ceiling

Applies to the 100 € subscription and scales with higher tiers. When you reach it, access pauses until the next billing cycle. You are never charged beyond it.

04

Billed on the 1st

Previous month’s overage plus the coming month’s subscription, by SEPA direct debit or saved card.

Need more headroom?

Higher monthly tiers, with a larger included allowance and ceiling, are available for teams that need more.

Talk to Sales
Model Pricing

Our Models

Transparent per-token pricing. All models are sovereign, hosted in France, with no hidden fees and no minimum commitment.

Language Models

Model Description 1M Input Tokens 1M Output Tokens
GLM-5.3-FlashSeptember 2026 Natively multimodal efficient frontier model (video, image, text, file) for reasoning, code and agentic workflows. Max context: 1M tokens, 128K output. 0,40 €Cached input: 0,09 € 1,30 €
Qwen3.5-122B-A10B Efficient high-performance model for production chat, code, and reasoning. Default instruct profile. 0,60 € 2,50 €
Qwen3.5-122B-A10B Creative Same as Qwen3.5-122B-A10B with hyperparameters tuned for creative ideation, synthesis, and high-quality content generation. 0,60 € 2,50 €
Qwen3.5-122B-A10B Thinking Same as Qwen3.5-122B-A10B with thinking mode enforced for structured reasoning and analysis. Max context: 256K tokens. 0,60 € 2,50 €
Qwen3.5-122B-A10B Thinking Coder Qwen3.5-122B-A10B configured for advanced coding, architecture, refactoring, debugging, and thinking-enforced development workflows. 0,60 € 2,50 €
Qwen3.8-Flash-NextSeptember 2026 Next mid-size tier, set to replace Qwen3.5-122B-A10B for production chat, code and multi-turn workflows. 0,60 € 2,50 €
Qwen3.8-27BSeptember 2026 Newer-generation compact model for production chat, code, extraction, and multi-turn workflows. Default instruct profile. 0,60 € 2,50 €
Qwen3.8-27B CreativeSeptember 2026 Same as Qwen3.8-27B with hyperparameters tuned for creative ideation, synthesis, and high-quality content generation. 0,60 € 2,50 €
Qwen3.8-27B ThinkingSeptember 2026 Same as Qwen3.8-27B with thinking mode enforced for structured reasoning and analysis. 0,60 € 2,50 €
Qwen3.8-27B Thinking CoderSeptember 2026 Qwen3.8-27B configured for advanced coding, refactoring, debugging, and thinking-enforced development workflows. 0,60 € 2,50 €
BGE-m3 Text embeddings for RAG and semantic search. 0,05 € N/A
BGE-reranker-v2-m3 Reranking for search result optimization. 0,30 € N/A
Document Convert any document in Markdown to be used by LLM. Useful for CAG – Cached RAG including vision RAG use cases. Document images uses IG1 Standard to be described as text. N/A 2,00 €

Image Models

Model Description 1M Input Tokens 1M Output Tokens
Qwen Image Image generation. Token count based on image resolution and quality level. 0,60 € 2,50 €
Qwen Image PE Image generation with Prompt Enhancing by LLM. Prompt enhancement uses Qwen3.5-122B-A10B (charged separately). Token count based on resolution and quality level. 0,60 € 2,50 €
Qwen Image Edit Image editing. Token count based on source and generated image resolution and quality level. 0,60 € 2,50 €
Qwen Image Edit PE Image editing with Prompt Enhancing by LLM. Prompt enhancement uses Qwen3.5-122B-A10B (charged separately). Token count based on source and generated image resolution and quality. 0,60 € 2,50 €

FullStack AI

Your own
sovereign AI platform.

Choose the deployment model that fits your needs. Scale up anytime.

On Your Infrastructure

Starting at

$3,500 /month

On Public Cloud or Your Hardware

Get Started
Recommended
On Our Infrastructure

Starting at

$4,900 /month

On Nvidia GPU

Get Started
Custom
Custom

Get in touch for a custom quote

Contact Us

FAQ

Pricing & Billing

Everything about the IG1 AI API subscription and billing.

How does pricing work?
Two parts. A flat 100 € / month subscription that includes 100 € of usage across every model in the catalog. Beyond that, usage is metered at the published per-token rates (GLM 5.3 Flash at 0,40 € input, 0,09 € cached input and 1,30 € output; Qwen3.5-122B-A10B and Qwen3.8-27B at 0,60 € / 2,50 €), up to a monthly ceiling of 1 000 € on the 100 € subscription (the ceiling scales with your tier). No hidden fees.
What happens if I hit my cap?
Your monthly ceiling is 1 000 € on the 100 € subscription, and scales with higher tiers. When you reach it, access pauses until the next billing cycle — or until you settle that amount. You are never charged beyond the ceiling you control. If you consistently hit it, we’ll discuss raising your limit.
How is overage billed?
On the 1st of each month we bill the previous month’s overage — usage above your included allowance — together with the coming month’s subscription. Payment is by SEPA direct debit or saved card.
Why a subscription plus metered usage?
You get the predictability of a fixed monthly base and the flexibility to scale when a project demands it — no renegotiation. Light months stay light, heavy months stay capped. You only pay for what you use, up to a ceiling you set.
Can I commit to a higher plan?
Yes. Higher monthly subscription tiers — with a larger included allowance and ceiling — are available for teams that need more. Contact sales to arrange a commitment that fits your volume.
Is there a free trial?
Contact us for pilot program options.

Not sure which option
fits your needs?

Tell us about your workloads. Our team will help you size the right setup: API, dedicated platform, or both.

Talk to Sales sales@ig1.com