arrow
Return

AI-Driven Dental Procedure Coding: A Multi-Model Framework for CDT Extraction from Clinical Text

delete2026-06-02
delete0
delete
OA
AI
P
Pranav Annareddy
A
Ali Noori
D
Deepthi Kollipara
P
Prashanti Manda *
DOI:10.3390/dj14060339delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
Background and Objectives: Dental procedure coding is essential for accurate billing, reimbursement, and clinical documentation, yet it remains largely manual, time-consuming, and error-prone. While natural language processing (NLP) has enabled significant advances in automated medical coding, limited work has focused on the dental domain, particularly the assignment of Code on Dental Procedures and Nomenclature (CDT) codes from free-text clinical notes. This study aims to develop and evaluate an artificial intelligence framework that integrates large language models (LLMs) and traditional deep learning methods to automate CDT code extraction from narrative dental documentation. Methods: We evaluated three LLM-based strategies—zero-shot prompting, QLoRA fine-tuning, and parameter-efficient fine-tuning (PEFT) using LoRA—alongside a supervised Bidirectional GRU (Bi-GRU) classifier. Experiments were conducted using a synthetic dataset designed to emulate real-world dental encounters. Structured JSON output schemas, few-shot prompting, and scalable batch inference pipelines were employed to ensure consistent and interpretable predictions. Model performance was assessed using micro- and macro-averaged F1 scores, precision, recall, exact-match accuracy, and Hamming loss. Results: The zero-shot LLM achieved the highest micro-F1 score (0.9614) and perfect recall for frequent CDT codes, demonstrating strong baseline reasoning without task-specific training; however, performance declined for rare procedures and diagnostic code hallucinations were common. Fine-tuning improved domain alignment, with the non-quantized PEFT LoRA model outperforming QLoRA across all metrics, though both fine-tuned LLMs showed tendencies to over-generate plausible but incorrect codes. The Bi-GRU model achieved balanced performance (micro-F1 = 0.9362, macro-F1 = 0.9377) with minimal hallucinations but occasionally missed context-dependent procedures. Conclusions: These findings highlight complementary strengths between LLM-based and supervised approaches. LLMs provide strong contextual understanding and rapid deployment, while traditional models offer stable and precise multi-label classification. This work supports the development of hybrid, schema-constrained systems for scalable dental procedure coding.
Keywords:
dental informatics
automated procedure coding
Large Language Models (LLMs)
Clinical Natural Language Processing (NLP)
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Dentistry Journal cover
Dentistry Journal
IF:
3.1
Papers:
1.4K
Citations:
3.7K

Organization

S
summit dental
Scholars:
2
Papers: 1
Citations: 0
U
university of nebraska omaha
Scholars:
118
Papers: 71
Citations: 0
U
University of North Carolina Greensboro
Scholars:
118
Papers: 90
Citations: 0
researcher View more organizations