Blog
Insights and updates about AI infrastructure, European tech innovation, and industry best practices

We trained a 116 MB LoRA to gate every bash command our coding agents run
Why we needed it, how we built the training data from our own agent history, what the evaluation shows and where it is thin, and what the first week in production forced us to change.

Co-founder & CPTO

System One in the tool loop: fast judgments for chat and RAG
Chat and RAG apps spend large-model tokens on decisions that are not reasoning. A System One model answers “does this page contain the answer?” in about 100 ms — here are the patterns we use, with measured numbers.

Co-founder & CPTO

System One: a new model category on Berget AI
Typed questions in; a choice, a score, or a probability out. Our first System One model, Laya, is open for evaluation on Swedish infrastructure — Jev-compatible, fastest engine in our load test, and early on accuracy: we say so.

Introducing real-time transcription for Swedish
Berget AI now does real-time streaming transcription: text arrives while people are still speaking, over an OpenAI-compatible API, starting with klang/pianissimo, a Swedish-trained model from Klang.

Co-founder & CPTO

Kimi K3 and GLM 5.3 Flash are now generally available
Two models out of evaluation — and GLM 5.2 reaches end-of-life on 18 September 2026.

Developer Relations

Introducing Pro and Summit plans for Berget Code
Starting today, Berget Code comes in three tiers: Standard at €150 per seat per month, Pro at €300, and Summit at €500. Same product, same models, same Swedish jurisdiction — the tiers only change how much room a seat has to work.

Open vs closed, West vs China: An AI landscape for Swedish organisations
We tested thirteen AI models — ten open-weight and three closed (Claude Fable, Opus, and Sonnet) — on Swedish values, censorship, context-sensitive behaviour, and summarisation. The results challenge both directions: open models beat the best frontier model on Swedish values, and the most expensive model is among the most censored.

Co-founder & CPTO

What happens with our values when AI is used to summarise?
A model that speaks Swedish is widely assumed to know what matters to Swedes. Our measurement suggests otherwise: nine open models — and a frontier reference, Claude Opus 5 — reproduce Swedish values largely correctly when they have room, but cut them first when told to keep it short. Speaking the language is not the same as knowing what counts.

Co-founder & CPTO

Do AI models censor sensitive topics? A multi-method evaluation
We tested nine open-weight AI models on 230 questions using six research-backed methods: explicit refusal, narrative steering, contrast pairs, knowledge elicitation, asymmetric bias, and false positives. Explicit refusal is rare (88-97% answer rate) — but asymmetric bias and narrative steering are measurable and significant.

Co-founder & CPTO

Do AI models behave differently in geopolitical contexts?
We tested nine open-weight AI models for context-dependent behavior across geopolitical triggers, European institutions, dual-use security tasks, and social engineering. We could not detect regional patterns — but some models adapt their code for any specific organization.

Co-founder & CPTO

Testing AI models on Swedish values: Ten models compared
We evaluated ten open-weight AI models on 55 World Values Survey questions covering gender equality, secularism, migration, LGBTQ+ rights, and more. Results show measurable differences between models — and some surprises.

Co-founder & CPTO

Kimi K3 anklagas för att ha kopierat amerikanska modeller — vi testade den oberoende
Vita huset anklagar Moonshot AI för att ha distillat Anthropic Fable för att bygga Kimi K3. Vi lät nio modeller genomgå vår oberoende utvärdering — 322 frågor om språk, värderingar, censur och sleeper agents. Resultatet visar att alla modeller har svagheter — men inte de man tror.

Co-founder & CPTO

Kimi K3 now available for evaluation
Moonshot AI's new top model matches closed American systems – but is open for anyone to use. Within 24 hours it was on Berget AI's platform, ready for evaluation via API.

Co-founder & CPTO

Five multi-model patterns that cut token costs — and keep your data where you want it
Discover five architectural patterns for multi-model AI stacks that cut token costs and keep data on-device — from feature routing to cascade, advisor, specialist, and draft-and-verify.

Developer Relations

Berget AI launches Berget Code – coding assistants that keep code in Sweden
Berget AI launches Berget Code, a service for agentic coding hosted in Sweden built on open-source software and open models.

The parts of the agent stack and what they do
Explore the AI agent stack: framework, runtime, and harness layers. Learn how these components transform demos into production-ready agents.

Developer Relations

Gemma 4 31B is now generally available
The latest open-weights model from Google is moving out of evaluation phase.

Developer Relations

Better speech-to-text for Norwegian and multilingual audio
NbAiLab/nb-whisper-large for Norwegian and openai/whisper-large-v3 for multilingual use are now available on Berget AI, alongside the existing KBLab/kb-whisper-large for Swedish.

Developer Relations

Our second data center is live — here's what it took
A deep dive into how we balance performance and memory usage to get the most out of our AMD MI300x infrastructure

Developer Relations

On-prem for AI Inference? Why It’s Often the Wrong Path – and What to Do Instead
Why on-prem GPU can be an unnecessarily expensive idea and what better alternatives exist for sovereign AI infrastructure

Co-founder & CPTO

Introducing GLM-4.7
A powerful model that raise the bar for performance, efficiency, and capability

Co-founder & CEO

Digitalist och Berget AI inleder partnerskap
Partnership update

Co-founder & CEO

Announcing Berget AI partnership with Opper AI
Partnership update

Co-founder & CEO

The DevOps Holy Grail: From Zero to Production
Build services that scale from weekend side-project to production without breaking everything - Part 1

Co-founder & CPTO

The DevOps Holy Grail: HTTPS and DNS Automation
Why HTTPS matters, the history of TLS, and how to automate certificates and DNS with cert-manager and external-dns - Part 2

Co-founder & CPTO

The DevOps Holy Grail: Enterprise-Grade Features
Add bulletproof backends, secrets management, and monitoring to your GitOps infrastructure - Part 3

Co-founder & CPTO

Kubernetes Multi-Environment Deployments: From Single to Multiple Environments
How to migrate from a single environment to proper staging and production environments using Kustomize and GitOps

Co-founder & CPTO

Kubernetes Secrets Management: From Leaky to Bulletproof
How to handle secrets in Kubernetes without compromising security or developer experience

Co-founder & CPTO

The DevOps Holy Grail: Complete Guide
The complete guide to building services that scale from weekend side-project to enterprise platforms

Co-founder & CPTO

Newsletter #2 - summer edition
Newsletter

Co-founder & CEO

Model drop: Mistral small 3.2
An update on our model setup

Co-founder & CEO

Model drop: Magistral Small
An update on our model setup

Co-founder & CEO

AI in Swedish
How to get help in selecting your models for Swedish applications

Co-founder & CEO

Our Model Strategy: Balance Between Power and Precision
How we select and combine AI models to maximize performance and sustainability

Co-founder & CEO

Behind the Scenes: How We Optimize Our AI Models
A deep dive into how we balance performance and memory usage to get the most out of our AMD MI300x infrastructure

Co-founder & CPTO

Berget AI update #1
We have been cooking new things!

Co-founder & CEO

AI Made in Sweden: Svenska Berget AI utmanar amerikanska moln-tjänster
Berget AI lanserar den första svenska tjänsten med AI-infrastruktur och inferenstjänster i samma plattform

How to Run Cline with Berget AI in Visual Studio Code
A step-by-step guide to configure Cline with Berget AI models for AI-assisted coding

Co-founder & CPTO