Language models on Swedish GPUs
Some data is too sensitive to leave Sweden. For those workloads we run the strongest open models in our own data center. The API is OpenAI compatible, so your existing code works as soon as you point it at our endpoint. Your data never leaves Sweden and is never used for training.
tokens every month
Our own team and the solutions we operate for customers consume hundreds of billions of tokens every month, so we know what it takes to run language models in production.
What you get
API inference
OpenAI-compatible API to the strongest open models.
Your data is yours
Nothing is stored longer than necessary. Nothing is used for training.
GDPR and the AI Act
All processing happens in Sweden. We help you document compliance.
Dedicated capacity
Your own NVIDIA RTX PRO 6000 with guaranteed throughput and your choice of model.
From pilot to production
Our AI consultants take you from pilot to production, with agent workflows and tailored search that make the models answer accurately from your data.
A data lake for agents
Your agents get permission-controlled access to your data through our platform LimiLake: structured data, files and SQL in one place.
Pricing per model
You pay per token. The table below is the whole price list. All prices excl. VAT.
| Model | Input | Output |
|---|---|---|
| GLM 5.2 | 17 kr | 52 kr |
| Kimi K2.6 | 9 kr | 42 kr |
| Qwen 3.5 397B | 8 kr | 43 kr |
| GPT-OSS 120B | 3 kr | 9 kr |
| Qwen 3.6 35B | 2 kr | 8 kr |
| Mistral Small 3.2 24B | 4 kr | 4 kr |
| Gemma 4 27B | 3 kr | 6 kr |
| E5 Large (embeddings) | 0.4 kr |
Prices in SEK per million tokens.
Example prices. The model lineup is updated continuously.
Dedicated capacity
GPU capacity on NVIDIA RTX PRO 6000 with 96 GB VRAM. Rent a full GPU or a MIG instance. All prices excl. VAT.
from 3,995 SEK/month
- Isolated slice of an RTX PRO 6000
- Guaranteed VRAM and compute
- Private endpoint
- Fits smaller models and test workloads
from 12,995 SEK/month
- NVIDIA RTX PRO 6000, 96 GB VRAM
- Any model, including fine-tuned
- Splittable into up to four MIG instances
- Monitoring included
Custom quote
- Multiple GPUs and models
- GPU nodes in your Kubernetes cluster
- Custom SLA
- Data processing agreement
Example prices. Contact us for a review of your use case.
Frequently asked questions
Make your first API call the same day
Book a demo and we will show you the API and discuss how language models can create value in your business.