返回
Telugu named entity recognition using bert
DOI:10.1007/s41060-021-00305-w.png)
摘要
En 中文
Named entity recognition (NER) is a fundamental step for many Natural Language Processing tasks that aim to classify words into a predefined set of named entities (NE). For high-resource languages like English, many deep learning architectures have produced good results. However, the NER task has not yet achieved much progress for Telugu, a low resource Language. This paper performs the NER task on Telugu Language using Word2Vec, Glove, FastText, Contextual String embedding, and bidirectional encoder representations from transformers (BERT) embeddings generated using Telugu Wikipedia articles. These embeddings have been used as input to build deep learning models. We also investigated the effect of concatenating handcrafted features with the word embeddings on the deep learning model's performance. Our experimental results demonstrate that embeddings generated from BERT added with handcrafted features have outperformed other word embedding models with an F1-Score 96.32%.
Keyword:
Named entity recognition
Telugu
Word2vec
Glove
FastText
Contextual string embeddding
BERT
期刊
I
IF:
2.8
论文数:
1.1K
被引数:
1.3K
机构
引用论文
Cobalt and Copper Composite Oxides as Efficient Catalysts for Preferential Oxidation of CO in H2-Rich Stream钴和铜复合氧化物作为H2-Rich流中CO优先氧化的有效催化剂
Transposase-Derived Transcription Factors Regulate Light Signaling in
Arabidopsis转座酶衍生的转录因子调节光信号
拟南芥
Science
IF0

