Hinglish Named Entity Recognition Benchmark
Prepared COMI-LINGUA for entity-level evaluation, fine-tuned mBERT and XLM-RoBERTa, and compared them with zero-shot GPT-4o and Claude 3.5 Sonnet under a common protocol. The repository includes data preparation, evaluation, and error-slice analysis code.
- 78%
- entity-level F1 for fine-tuned XLM-RoBERTa
- 76%
- F1 for the zero-shot GPT-4o baseline
- 2 scripts
- Roman and Devanagari Hinglish evaluation