Automated Medical Labels Detection and Text Extraction Using Tesseract
摘要
Medical labels are widely employed in the field of pharmacology to convey critical information about medications, including their names, prices, and other relevant details. In Algeria, the absence of an approved standardized model by the Minister of Health has resulted in the lack of automated identification systems for medical labels. Tesseract has proven to be highly effective in OCR and text extraction tasks, making it a suitable choice for this application. This study introduces an artificial intelligence-based method designed to automatically recognize and extract information from medical labels using Tesseract. Experimental results demonstrate the overall effectiveness of the system, with Tesseract proving particularly effective in text extraction tasks from medical labels images.