2 posts on DeBERTa.
The raw 27-cell Kaggle notebook for the Naamapadam PII fine-tune. What eleven Indian languages of NER data actually gave me.
The multilingual model was the obvious pick and it lost by six F1 points on Bundesdatenschutzgesetz entities. Tokenizer coverage beat parameter breadth on German.