The GPT-NL project from TNO is easy to file under national-pride LLM and ignore. The part worth your attention is the data governance, not the model.

That last bullet is the real tension. Most enterprises I’ve seen don’t actually want a sovereign model — they want a frontier model with GPT-NL’s data lineage, and those two things are still in conflict. The HN discussion splits on exactly this: is provable provenance worth a capability gap?

If your compliance team had to choose between a frontier model with murky training data and a weaker one with a documented, licensed corpus, which one actually ships to production — and which one does legal sign off on?