A Specialized Dataset of Persian Pharmaceutical Questions and Expert Responses

Published: 13 December 2024| Version 4 | DOI: 10.17632/hzzf2zkzcz.4
Contributors:
, pedram porbaha

Description

The dataset contains 12,399 Persian comments categorized by drug name, with 4,719 of these (38.1%) receiving expert responses, including details about the experts' specialties. Additionally, the dataset includes drug names, Martindale categories for each drug, and information about the experts' categories and specialties. This dataset can be used to fine-tune large language models (LLMs) for question and answering in Persian about pharmaceutical topics. It also allows for the analysis of common questions and answers about drugs. Furthermore, the English translations of the Persian data are included in the dataset.

Files

Institutions

  • Shiraz University of Medical Sciences

Categories

Pharmacy, Artificial Intelligence, Natural Language Processing, Machine Learning, Pharmacist, Large Language Model

Licence