LLM HOSTING

Tokens on tap. Hosted in Sweden.

A growing catalogue of open models, billed per token, with no monthly commitment. Or models we run on our own hardware, when your requirements reach the compute as well as the data.

Two Ways to Buy

Both run in Sweden, on the same OpenAI-compatible API. Pick the one that matches your constraints.

Monthly subscription

In-house models

Models we run ourselves, on hardware Aixia owns and administers in Aixia datacenters. The answer when the requirements cover where the compute runs, not only where the data rests.

  • Hardware Aixia owns and administers
  • No sub-processors with access to your data
  • Included token allowance, priced per month
  • Dedicated infrastructure on Enterprise
SEE SUBSCRIPTIONS

Core Values

What holds true whichever way you buy.

Hosted in Sweden

Models and hardware run in Sweden, under GDPR and ISO 27001, with no dependency on a non-European cloud. Nothing leaves the country on either track.

A Catalogue That Keeps Up

New models reach the catalogue as they prove themselves. Moving to one is a change of model ID, not a change of architecture.

Expert-Backed Operations

Support for model testing, validation, optimization, and deployment from Aixia's AI experts.

Buy What You Use

Start on tokens with nothing committed. Move to a subscription when the workload settles. Neither direction is a migration.

Cost Transparency

Published per-model rates and per-key metering. You can see what a workload costs before you run it, and what it cost after.

Open Models, No Lock-In

Everything we serve is open weight, behind an OpenAI-compatible API. Nothing here is a proprietary endpoint you cannot leave.

The Catalogue

Rates in SEK per million tokens, excl. VAT. You are billed on the tokens you actually send and receive, per API key.

Model Input / 1M (SEK) Output / 1M (SEK) Good at
Kimi K2.6kimi-k2.6-swe 13.86 55.45 Reasoning, coding, long context. The one to reach for first.
Gemma 4 26Bgemma-4-26b-swe 1.44 5.55 Fast and cheap. Built for high volume and short turnarounds.
Qwen3.8 27Bqwen3.8-27b-swe 8.32 33.27 General purpose. A capable default for mixed workloads.
Qwen3 Embedding 8BQwen3-Embedding-8B-swe 1.11 1.11 Multilingual semantic search and RAG retrieval.
Qwen3 Reranker 4BQwen3-Reranker-4B-swe 0.55 0.00 Reranking retrieved passages before they reach the model.

Catalogue current as of August 2026. It is extended as new models prove themselves. Reranking is charged on input only. Models and hardware are hosted in Sweden. Catalogue compute is operated by a Swedish partner under a data processing agreement. If your requirements rule out a processor entirely, that is what the in-house subscription is for.

In-House Subscriptions

Models we run on hardware Aixia owns and administers. A dedicated VPS for agents, retrieval and preprocessing is included in every tier. All prices excl. VAT.

Free Trial
0 SEK
7 days
Try It Out
  • No credit card required
  • Instant API key
  • Test with any of our models
  • Catalogue and in-house alike
GET FREE KEY
Pilot
1 500 SEK/mo
1M tokens included (max 10M)
PoC, Innovation & Testing
  • Extra: $5/M input, $15/M output
  • Standard model library
  • Agent framework VPS included
  • Shared support
  • Scheduled optimization checkpoints
CONTACT SALES
Enterprise
7 500 SEK/mo
50M tokens included (unlimited)
Mission-Critical Operations
  • Extra: $3/M input, $10/M output
  • Extended model library
  • Agent framework VPS included
  • Dedicated support
  • Dedicated infrastructure
  • Advanced security & compliance
  • Direct access to Aixia experts
CONTACT SALES

Try It Free

Try our chat portal for free at chat-demo.aiqu.ai

or get an API key

Valid for 7 days

In Detail

What you get on each track, and who each one is for.

Models on tap

Per-model rates

No monthly fee, no included allowance

Purpose: Scaling usage without committing capacity

  • Billed on tokens sent and received, per API key
  • Full catalogue, extended as models prove themselves
  • Models and hardware hosted in Sweden
  • Catalogue compute run by a Swedish partner under a DPA

Best for: Teams whose volume moves week to week, and anyone who wants the newest models without a migration

Pilot

1 500 SEK / month

1M tokens included (max 10M)

Purpose: Proof-of-Concept, innovation, and testing on in-house models

  • Extra tokens: $5/M input, $15/M output
  • Standard model library
  • Agent framework VPS included
  • Shared support
  • Scheduled optimization checkpoints

Best for: Organizations exploring AI capabilities or validating use-cases

Standard

3 500 SEK / month

10M tokens included (max 50M)

Purpose: Production workloads and multiple parallel use-cases

  • Extra tokens: $4/M input, $12/M output
  • Extended model library
  • Agent framework VPS included
  • Priority support
  • Frequent optimization sessions

Best for: Companies running LLMs in production with higher availability needs

Enterprise

7 500 SEK / month

50M tokens included (unlimited)

Purpose: Business-critical, high-scale operations

  • Extra tokens: $3/M input, $10/M output
  • Extended model library
  • Agent framework VPS included
  • Dedicated support
  • Dedicated infrastructure
  • Advanced security & compliance
  • Direct access to Aixia's AI experts

Best for: Organizations where AI is mission-critical and performance cannot be compromised

Swedish models, whichever way you buy.

Take a key and start on the catalogue today. Talk to us when the workload needs to sit on hardware we own.