Improving large language models for clinical named entity recognition via prompt engineering
Yan Hu, Qingyu Chen, Jingcheng Du, et al.
This paper shows that task-specific prompt engineering, incorporating medical knowledge and few-shot examples, significantly improves GPT-3.5 and GPT-4 performance on clinical NER tasks, though still below BioClinicalBERT.