نوع مقاله : مقاله پژوهشی
نویسندگان
1 گروه مدیریت فناوری اطلاعات، واحد بین المللی کیش، دانشگاه آزاد اسلامی، کیش، ایران
2 گروه مدیریت صنعتی، واحد تهران مرکزی، دانشگاه آزاد اسلامی، تهران، ایران نویسنده مسئول: moh.afsharkazemi@iauctb.ac.ir
3 گروه مدیریت صنعتی و فناوری اطلاعات، دانشکده مدیریت و حسابداری، دانشگاه شهید بهشتی، تهران، ایران
4 گروه مدیریت عملیات و فناوری اطلاعات، دانشکده مدیریت و حسابداری، دانشگاه علامه طباطبائی، تهران، ایران
کلیدواژهها
عنوان مقاله English
نویسندگان English
Given the importance of the tourism industry in the country’s economy and culture, and the shortage of domain-specific datasets for training Persian language models, this research aims to fill the existing gap. The objective of this study is to improve the performance of Persian tourism-related question answering systems and to provide accurate responses to queries concerning Iranian attractions and destinations. To achieve this, a collection of reliable tourism texts was gathered and processed, and using advanced natural language processing approaches, including Prompt Engineering, approximately 20,000 question–answer records were generated. In addition, the application of the innovative Answer Window strategy enhanced the quality of the produced data. Subsequently, the ToKA-BERT model was fine-tuned with this dataset, and evaluation results revealed its significant superiority in specialized tourism question answering compared to general-purpose Persian QA models. Both the dataset and the final model have been published on the Hugging Face platform and can serve as valuable resources for developing intelligent systems in the tourism industry and enhancing user experience.
کلیدواژهها English