A public dashboard observing signals, trends and issues.
SubscribeLogin한국어
Latest observation
2026-10-08
Public objects
4434
Build time
2026-10-08 19:44 KST
The Futures

Signal Preprint presents 'Grip on LLMs' framework for evaluating government-use AI models

Summary

A preprint presents the Grip on LLMs framework, developed together with a major Dutch municipal organization to evaluate government-use language models. The evaluation suite covers more than 30 multilingual and Dutch-specific models across six dimensions: factuality, honesty, social bias, energy consumption, cost and training-data transparency. It found that no single model performs well across every dimension, and that higher quality tends to come with greater environmental impact and cost. Bias levels, by contrast, were largely unrelated to either quality or cost. The researchers also found that factuality and honesty are independent properties, meaning high factuality does not guarantee high honesty. The team additionally released a public, stakeholder-friendly overview of the models for policymakers and engineers involved in selecting government LLMs.

Classification

Main topicAI & Computing
Secondary topicsTech·Digital Policy
Region menusEurope
Impactgeo_region:europe · country:NL
Time horizon0-3 years (2026-08-12)
Last updated2026-09-25 22:32 KST

Evidence 1

Part of trends 0

No objects.

Directly linked issues 0

No objects.

Public id: fm-eacccb4375c5