Elysium X 500 FR: public-data upgrade

Same 500-emotion schema, weights continued-trained on the team dataset plus 8,250 mapped rows from 11 public emotion datasets.

Headline: 23.3% exact on rows without the emotion name (n=403), up from 19.4%

Test slice (n=1,195)OriginalThis releaseChange
No emotion name (n=403), headline19.35%23.33%+3.98
All rows69.54%70.71%+1.17
Emotion name in text (n=792)95.08%94.82%-0.26 (2 of 792)

We set a bar to beat all three numbers. This release beats two and misses the with-name number by 2 answers. We say so. A repeat run (with a training-loop bug) scored lower and is not released.

No-name exact by language

DE50.0%
PT45.5%
EN39.5%
FR35.7%
KO27.8%
Hinglish25.0%
ZH23.1%
AR21.1%
JA17.2%
ES16.4%
BN7.1%
HI3.8%

Caveats: test covers part of the 500 coordinates; team data is synthetic and templated; labels are human-monitored, not gold; small cells per language; one run, no confidence intervals. Licence CC BY-NC-SA 4.0, non-commercial; some source datasets have unknown licences (listed on the model card). Not for medical or profiling use.

Model · GGUF · Dataset · Paper