New table: 200k rows x 300d, full Persian Wikipedia dump, 5 epochs,
vs md's 50k rows / 400k documents.
- configs/fa_ner_lg.cfg, project.yml ent-lg workflow (vectors-lg
through smoke-ent-lg)
- scripts/compare_tiers.py: generalized sm/md pair to N tiers; ent
NER test now includes lg; fixed sm baseline to the file that's
actually scored (perdt-ner-test.json, not the missing ent-test.json)
- scripts/finalize_pipeline.py: FLORET_LG source and vectors_note_lg
corrected to full Wikipedia, 5 epochs (were a generic Wikipedia +
OSCAR placeholder)
- docs/MODELS.md §7: PerDT NER test ENTS_F 75.94 (sm 71.87, md
74.71), full per-label table, cost (217 MB wheel)
Not built: fa_dep_news_lg / fa_core_news_lg.