A public dashboard observing signals, trends and issues.
SubscribeLogin한국어
Latest observation
2026-10-08
Public objects
4434
Build time
2026-10-08 19:44 KST
The Futures

Signal Do Large Language Models Know Colombian Law? A Reliability Benchmark for the Colombian Legal System

Summary

The paper introduces an expert-validated benchmark for evaluating LLM reliability on the Colombian legal system. It contains 1,042 items across ten areas of law and three question formats, built through a human-in-the-loop pipeline. Across 15 models, closed-question accuracy ranged from 0.577 to 0.905. No model exceeded 0.45 factual correctness on free-text legal answers. Models often sounded responsive while being wrong, and only about half of the norms they cited were correct. The authors conclude that current LLMs need expert supervision for Colombian legal tasks, and that grounding answers in authoritative sources is promising.

Classification

Main topicAI & Computing
Secondary topicsTech·Digital Policy
Region menusGlobal
Impactscope:global
Time horizon0-3 years (2026-10-05)
Last updated2026-10-07 08:51 KST

Evidence 1

Part of trends 0

No objects.

Directly linked issues 0

No objects.

Public id: fm-e32e4dabead9