And more details from @pieterdelobelle
Pieter Delobelle (@pieterdelobelle)
Today we release Nemotron-Personas-Belgium in collaboration with @NVIDIAAI , a synthetic dataset of 4x300k Belgian personas in Dutch, French, German and English.
The dataset is modeled on the @Statbel_en census, modeling age, occupations, household, income, names etc.
— https://nitter.net/pieterdelobelle/status/2067275709978915059#m