Skip to main content
Connic
Build

Connic Model Catalog

Exact model IDs, supported inputs, context limits, and EUR token prices for Connic-managed inference.

Last updated
Enable Auto-refill before production

On prepaid Projects, enable Auto-refill under Project → Billing before sending production traffic to connic/* models. If available credit runs out, new Project usage pauses until credit is added.

Configuration

Use an exact catalog ID as the primary or fallback model. A Project can mix Connic-managed and BYOK models.

agents/assistant.yaml
model: connic/glm-5.2
fallback_model: connic/gpt-5.6-luna # or anthropic/claude-sonnet-4-6 with Anthropic BYOK configured
temperature: 0.3
reasoning_effort: high

Available models

Prices are net EUR per one million tokens. A dash means no cached-input discount is available. Listed reasoning effort values are the model-specific YAML overrides. Omit reasoning_effort or use auto for model-managed behavior. Standard and Fast IDs share one row and the same prices. Open the info icon beside a Fast ID for its supported reasoning settings, context limit, and input modalities.

Non-EU-native provider
Claude Haiku 4.5
connic/claude-haiku-4-5

Fast model for high-volume coding and agent tasks.

Reasoning effort: low, medium, high
TextVisionReasoning
Context
200k
Input / 1M
€1.10
Cached input / 1M
€0.15
Output / 1M
€5.10
Non-EU-native provider
Claude Opus 5
connic/claude-opus-5

Advanced model for long-running agents, complex coding, and professional work.

Reasoning effort: low, medium, high
TextVisionReasoning
Context
1M
Input / 1M
€5.20
Cached input / 1M
€0.55
Output / 1M
€26.00
Non-EU-native provider
Claude Sonnet 5
connic/claude-sonnet-5

High-performance model for coding, agents, and everyday professional tasks.

Reasoning effort: low, medium, high
TextVisionReasoning
Context
1M
Input / 1M
€2.10
Cached input / 1M
€0.25
Output / 1M
€10.40
EU-native provider
Codestral 2508
connic/codestral-2508

Low-latency coding model for completion, correction, and test generation.

Reasoning effort: Not supported
Text
Context
256k
Input / 1M
€0.35
Cached input / 1M
€0.05
Output / 1M
€0.95
EU-native provider
DeepSeek V4 Flash 0731
connic/deepseek-v4-flash-0731
connic/deepseek-v4-flash-0731-fast

Updated open-weight reasoning model for long-context coding and agentic work.

Reasoning effort: none, low, high, max
TextReasoning
Context
256k
Input / 1M
€0.45
Cached input / 1M
€0.10
Output / 1M
€0.85
EU-native provider
DeepSeek V4 Pro
connic/deepseek-v4-pro-fast

Frontier open-weight reasoning model for complex long-context coding and agentic work.

Reasoning effort: high, max
TextReasoning
Context
1M
Input / 1M
€1.70
Cached input / 1M
€0.45
Output / 1M
€3.30
Non-EU-native provider
Gemini 3.1 Flash Lite
connic/gemini-3.1-flash-lite

Low-latency multimodal model for cost-sensitive, high-volume workloads.

Reasoning effort: low, medium, high
TextVisionReasoning
Context
1M
Input / 1M
€0.30
Cached input / 1M
€0.05
Output / 1M
€1.60
Non-EU-native provider
Gemini 3.5 Flash
connic/gemini-3.5-flash

Fast frontier multimodal model for agents, coding, and large-scale tasks.

Reasoning effort: low, medium, high
TextVisionReasoning
Context
1M
Input / 1M
€1.60
Cached input / 1M
€0.20
Output / 1M
€9.40
Non-EU-native provider
Gemini 3.5 Flash Lite
connic/gemini-3.5-flash-lite

Low-latency multimodal model for high-volume subagent workflows and document parsing.

Reasoning effort: low, medium, high
TextVisionReasoning
Context
1M
Input / 1M
€0.35
Cached input / 1M
€0.05
Output / 1M
€2.60
Non-EU-native provider
Gemini 3.6 Flash
connic/gemini-3.6-flash

Fast multimodal model for agentic coding, long-context reasoning, and multi-step workflows.

Reasoning effort: low, medium, high
TextVisionReasoning
Context
1M
Input / 1M
€0.75
Cached input / 1M
€0.10
Output / 1M
€3.60
Non-EU-native provider
Gemini 3.7 Flash
connic/gemini-3.7-flash

Fast multimodal model for agentic coding, long-context reasoning, and multi-step workflows.

Reasoning effort: low, medium, high
TextVisionReasoning
Context
1M
Input / 1M
€0.75
Cached input / 1M
€0.10
Output / 1M
€3.60
Non-EU-native provider
Gemini 3.8 Flash
connic/gemini-3.8-flash

Fast multimodal model for coding, agents, and long-context reasoning.

Reasoning effort: low, medium, high
TextVisionReasoning
Context
1M
Input / 1M
€0.80
Cached input / 1M
€0.10
Output / 1M
€3.90
EU-native provider
Gemma 4 26B A4B
connic/gemma-4-26b-a4b-it

Efficient open-weight agentic and reasoning model with image understanding.

Reasoning effort: none, low, medium, high
TextVisionReasoning
Context
256k
Input / 1M
€0.30
Cached input / 1M
Output / 1M
€0.55
EU-native provider
Gemma 4 31B IT
connic/gemma-4-31b-it

Multimodal instruction model for reasoning, coding, and vision-language tasks.

Reasoning effort: low, medium, high
TextVisionReasoning
Context
128k
Input / 1M
€0.30
Cached input / 1M
Output / 1M
€0.55
EU-native provider
GLM 5.2
connic/glm-5.2
connic/glm-5.2-fast

Open-weight reasoning model for long-horizon agentic work.

Reasoning effort: none, high, max
TextReasoning
Context
256k
Input / 1M
€1.90
Cached input / 1M
€0.40
Output / 1M
€5.80
EU-native provider
GPT OSS 120B
connic/gpt-oss-120b
connic/gpt-oss-120b-fast

Large open-weight mixture-of-experts model for reasoning, agents, and developer workloads.

Reasoning effort: low, medium, high
TextReasoning
Context
128k
Input / 1M
€0.25
Cached input / 1M
Output / 1M
€0.70
Non-EU-native provider30-day provider retention
GPT-5.4
connic/gpt-5.4

Frontier multimodal model for professional workflows, coding, and agents.

Reasoning effort: low, medium, high
TextVisionReasoning
Context
1M
Input / 1M
€2.80
Cached input / 1M
€0.25
Output / 1M
€14.60
Non-EU-native provider30-day provider retention
GPT-5.6 Luna
connic/gpt-5.6-luna

Fast multimodal model for high-volume work and lightweight reasoning.

Reasoning effort: low, medium, high
TextVisionReasoning
Context
1M
Input / 1M
€0.25
Cached input / 1M
€0.05
Output / 1M
€1.30
Non-EU-native provider30-day provider retention
GPT-5.6 Sol
connic/gpt-5.6-sol

Flagship multimodal model for advanced reasoning, deep research, and complex multi-step execution.

Reasoning effort: low, medium, high
TextVisionReasoning
Context
1M
Input / 1M
€5.20
Cached input / 1M
€0.55
Output / 1M
€31.20
Non-EU-native provider30-day provider retention
GPT-5.6 Terra
connic/gpt-5.6-terra

Balanced multimodal model for professional work, coding, and reasoning.

Reasoning effort: low, medium, high
TextVisionReasoning
Context
1M
Input / 1M
€2.10
Cached input / 1M
€0.25
Output / 1M
€12.50
EU-native provider
Kimi K2.7 Code
connic/kimi-k2.7-code-fast

Open-weight agentic coding model for long-horizon software engineering.

Reasoning effort: Automatic (not configurable)
TextVisionReasoning
Context
256k
Input / 1M
€1.20
Cached input / 1M
€0.30
Output / 1M
€4.20
EU-native provider
Kimi K3
connic/kimi-k3-fast

Frontier open-weight multimodal model for long-horizon coding, reasoning, and agentic work.

Reasoning effort: Automatic (not configurable)
TextVisionReasoning
Context
1M
Input / 1M
€2.90
Cached input / 1M
Output / 1M
€14.20
EU-native provider
Llama 3.3 70B Instruct
connic/llama-3.3-70b-instruct
connic/llama-3.3-70b-instruct-fast

Established open-weight multilingual instruction model for text workloads.

Reasoning effort: Not supported
Text
Context
100k
Input / 1M
€0.95
Cached input / 1M
Output / 1M
€0.95
EU-native provider
MiniMax M2.7
connic/minimax-m2.7

Autonomous agent model for multi-agent coordination and enterprise-scale work.

Reasoning effort: low, medium, high
TextReasoning
Context
197k
Input / 1M
€0.65
Cached input / 1M
Output / 1M
€2.60
EU-native provider
Ministral 3 14B
connic/ministral-14b-2512

Compact frontier-level multimodal model for text, vision, agents, and local deployment.

Reasoning effort: Not supported
TextVision
Context
256k
Input / 1M
€0.25
Cached input / 1M
€0.05
Output / 1M
€0.25
EU-native provider
Mistral Large 2512
connic/mistral-large-2512

Open-weight multimodal mixture-of-experts model for text and vision tasks.

Reasoning effort: Not supported
TextVision
Context
256k
Input / 1M
€0.55
Cached input / 1M
€0.10
Output / 1M
€1.60
EU-native provider
Mistral Medium 3.5 128B
connic/mistral-medium-3.5-128b

Unified open-weight model for instruction following, reasoning, and image understanding.

Reasoning effort: none, high
TextVisionReasoning
Context
180k
Input / 1M
€1.60
Cached input / 1M
€0.15
Output / 1M
€7.90
EU-native provider
Mistral Small 3.2 24B
connic/mistral-small-3.2-24b-instruct-2506

Efficient open-weight text and vision model for responsive agent workloads.

Reasoning effort: Not supported
TextVision
Context
128k
Input / 1M
€0.15
Cached input / 1M
Output / 1M
€0.35
EU-native provider
Mistral Small 4 2603
connic/mistral-small-2603
connic/mistral-small-2603-fast

Efficient multimodal model for automatic reasoning, agents, and coding.

Reasoning effort: Automatic (not configurable)
TextVisionReasoning
Context
256k
Input / 1M
€0.25
Cached input / 1M
€0.05
Output / 1M
€0.65
EU-native provider
Qwen 3 235B A22B Instruct
connic/qwen3-235b-a22b-instruct-2507

Large open-weight multilingual instruction model for demanding text workloads.

Reasoning effort: Not supported
Text
Context
250k
Input / 1M
€0.75
Cached input / 1M
Output / 1M
€2.30
EU-native provider
Qwen 3.5 397B A17B
connic/qwen3.5-397b-a17b
connic/qwen3.5-397b-a17b-fast

Frontier open-weight reasoning model with image understanding.

Reasoning effort: none, low, medium, high
TextVisionReasoning
Context
250k
Input / 1M
€0.65
Cached input / 1M
Output / 1M
€3.80
EU-native provider
Qwen 3.6 35B A3B
connic/qwen3.6-35b-a3b
connic/qwen3.6-35b-a3b-fast

Fast open-weight model for agentic tasks, reasoning, and image understanding.

Reasoning effort: none, low, medium, high
TextVisionReasoning
Context
256k
Input / 1M
€0.30
Cached input / 1M
Output / 1M
€1.60
EU-native provider
Voxtral Small 2507
connic/voxtral-small-2507

Efficient Mistral model for multilingual text understanding and generation.

Reasoning effort: Not supported
Text
Context
32k
Input / 1M
€0.15
Cached input / 1M
€0.05
Output / 1M
€0.35

Provider labels beside a model name report whether the provider is EU-native. EU-native provider marks an EU-native provider, while Non-EU-native provider marks a model whose provider is not EU-native. The additional 30-day provider retention marks one whose provider retains request data for 30 days under its own policy. Inference still runs on EU capacity in both cases.

Fast IDs

An ID ending in -fast selects a latency-optimized model variant. Most models list standard and Fast IDs at the same price; some list only a Fast ID.

agents/classifier.yaml
model: connic/glm-5.2-fast
reasoning_effort: high   # -fast accepts a narrower set than connic/glm-5.2
  • Fast and standard IDs can produce different output for the same prompt.
  • A Fast ID can be narrower than its standard ID in three ways: fewer reasoning_effort values, a smaller context window, or fewer input modalities. Setting a value the Fast ID does not accept is rejected.
  • The info icon beside a Fast ID lists its differences from the standard ID. No listed differences means the supported settings match.

Fallbacks and retries

Model request failures use retry_options. Without a fallback, the primary uses the configured attempt budget. With a configured fallback_model, the primary is tried once and the fallback uses that budget. A switch marks the run with context.fallback_model_used.

Retries apply to the individual model request, not the whole run, so completed tool calls keep their results. Configure attempts and backoff under execution limits and retries.

Cancelled requests

A managed request that ends early, because the run hit its timeout or the run was stopped, is settled at zero tokens and the credit reserved for it is released back to the Project balance. This applies to streaming and non-streaming requests alike.

EU inference and data handling

  • Every connic/* request runs in the EU. A request fails when EU execution is unavailable.
  • Managed inference is not used to train models. Connic records model and token counts for billing rather than prompt or response content; narrow security and error-retention exceptions are described in the Privacy Policy.

EU Project boundary

The connic/* guarantee covers the model call. A complete Project remains EU-only when every customer-configured component, including BYOK providers, tools, guardrails, judges, and external data destinations, also stays in the EU.