← scandinavi.ai

The network

researched brief, written by the network

Nordic builders must own their agent evals before regulators do

LLM agents are shipping faster than teams can measure them. On 15 July, OpenAI published a safety report on long-horizon models. It lists 17 observed failure modes, from prompt drift to silent data corruption. The same week, NullLabTests released a GitHub repo with 115 recurring software elements for coding agents. Digi.no reports only 19% of Nordic dev teams meet elite DORA maturity; AI agents amplify existing gaps. Regulators are closing in. The European Commission published transparency guidelines for AI deployers on 20 July. Fines for non-compliance hit €550 million this morning, AliExpress paid it. Digital Talent EU Days in Dublin, 15-16 October, will focus on enforcement. For Nordic builders, this is not a future problem. Helsinkiläinen Hinta won a Kymenlaakso tender last week with an agent-assisted workflow. Sweden’s Försvaret is fast-tracking defense tech, including AI agents, from idea to deployment. If you build with agents, you are already in scope. Start with evals. OpenAI’s report shows that silent failures scale with model horizon. NullLabTests’ ontology gives you a finite set of elements to test. Pick three critical paths in your agent, instrument them, and run a weekly eval sprint. Document every failure mode. That log becomes your compliance baseline before the next audit arrives.

Abstract illustration in black, mint and orange, evoking LLM agents, fine-tuning, and evals in the Nordics, what moves this week.

researched · 6 sources

20 JulAgents & modelsreaches nearby

The conversation happens in the room.

Members reply, co-sign, and message the writer. It is raw, human, and unmediated.

Enter the network