AI Platform

AI Inference served in the EU with zero data retention.

Hosted models behind one OpenAI-compatible API. Keys, usage and privacy controls live in the same project as your app.

# any OpenAI client works, only the base URL changes
curl https://api.nodion.ai/v1/chat/completions \
  -H "Authorization: Bearer $NODION_AI_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "deepseek/deepseek-v4.1-flash",
    "messages": [{"role": "user", "content": "Hello"}]
  }'

Models

The catalog.

Text, images, audio and documents behind one API. Pay only for what you call.

Text and reasoning

DeepSeek V4.1 Flashdeepseek/deepseek-v4.1-flash
Context1M Input0,30 € 0,010 € cached Output1,05 € EU (Germany)
Gemma 4 31Bgoogle/gemma4-31b
Context256K Input0,25 € 0,15 € cached Output0,50 € EU (Germany)
Qwen3.8 27Bqwen/qwen3.8-27b
Context256K Input0,35 € 0,030 € cached Output2,60 € EU (Germany)

per 1M tokens

Document OCR

PDF Inspectorfirecrawl/pdf-inspector
Max pages500 Price0,20 € EU (Germany)
Mistral OCR 4mistral/mistral-ocr-4-0
Max pages100 Price3,50 € EU (France)

per 1,000 pages

Endpoints

Web Fetchweb-fetch
Price1,80 € -
Web Searchweb-search
Price4,50 € -

per 1,000 calls

Net prices in EUR, excl. VAT, from the live catalog. USD prices and request logs are in the dashboard.

Built for production

Inference you can put in front of customers.

Drop-in

OpenAI-compatible.

Point any OpenAI SDK at api.nodion.ai/v1. Chat, tools, JSON mode and streaming work as you expect.

Keys

Limits per key.

Optional expiry and a daily spend limit on every key. Only project admins create or revoke them.

Privacy

Your data, your rules.

Restrict which regions may serve a project, and turn on zero data retention in the project settings.

Billing

Prepaid wallet.

AI spend comes from a wallet you top up by card or PayPal, separate from the platform invoice.

Give your agent a cloud.

Earn up to 15 € in credit when signing up.

Continue with GitHub

or with email