<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE ArticleSet PUBLIC "-//NLM//DTD PubMed 2.7//EN" "https://dtd.nlm.nih.gov/ncbi/pubmed/in/PubMed.dtd">
<ArticleSet>
<Article>
<Journal>
				<PublisherName>پژوهشگاه علوم و فناوری اطلاعات ایران (ایرانداک)</PublisherName>
				<JournalTitle>پژوهشنامه پردازش و مدیریت اطلاعات</JournalTitle>
				<Issn>2251-8223</Issn>
				<Volume>40</Volume>
				<Issue>4</Issue>
				<PubDate PubStatus="epublish">
					<Year>2025</Year>
					<Month>06</Month>
					<Day>22</Day>
				</PubDate>
			</Journal>
<ArticleTitle>Text Recognition in Printed Persian Documents Based on Recurrent Neural Networks</ArticleTitle>
<VernacularTitle>تشخیص متن در اسناد فارسی چاپی بر اساس شبکه‌های عصبی بازگشتی</VernacularTitle>
			<FirstPage>1283</FirstPage>
			<LastPage>1305</LastPage>
			<ELocationID EIdType="pii">725228</ELocationID>
			
<ELocationID EIdType="doi">10.22034/jipm.2025.2052358.1926</ELocationID>
			
			<Language>FA</Language>
<AuthorList>
<Author>
					<FirstName>آزاده</FirstName>
					<LastName>فخرزاده</LastName>
<Affiliation>دکتری تخصصی پردازش تصویر کامپیوتری       استادیار، پژوهشگاه علوم و فناوری اطلاعات (ایرانداک)، تهران، ایران.</Affiliation>

</Author>
<Author>
					<FirstName>امیرحسین</FirstName>
					<LastName>صدیقی</LastName>
<Affiliation>دکتری تخصصی مهندسی صنایع 
   استادیار، پژوهشگاه علوم و فناوری اطلاعات (ایرانداک) )، تهران، ایران.</Affiliation>

</Author>
<Author>
					<FirstName>محمد</FirstName>
					<LastName>عشرت آبادی</LastName>
<Affiliation>دستیار پژوهشی، پژوهشگاه علوم و فناوری اطلاعات (ایرانداک) )، تهران، ایران</Affiliation>

</Author>
<Author>
					<FirstName>البرز</FirstName>
					<LastName>اسفندیاری</LastName>
<Affiliation>دستیار پژوهشی، پژوهشگاه علوم و فناوری اطلاعات (ایرانداک)، تهران، ایران.</Affiliation>

</Author>
</AuthorList>
				<PublicationType>Journal Article</PublicationType>
			<History>
				<PubDate PubStatus="received">
					<Year>2025</Year>
					<Month>02</Month>
					<Day>02</Day>
				</PubDate>
			</History>
		<Abstract>&lt;span style=&quot;font-size: 12.0pt; mso-bidi-font-size: 11.0pt; line-height: 115%; font-family: &#039;Times New Roman&#039;,serif; mso-ascii-theme-font: major-bidi; mso-hansi-theme-font: major-bidi; mso-bidi-theme-font: major-bidi;&quot;&gt;Automatic Persian text recognition has always been challenging due to the unique characteristics of the Persian script, including its connected structure, the high visual similarity between letters, and the significant variation in the shape of letters depending on their position within a word. The aim of this research is to develop an optical character recognition (OCR) model capable of converting Persian printed and scientific documents, including theses, articles, and books, into editable texts. Such a model is essential for tasks like labeling, indexing, and information retrieval in databases. This paper proposes a hybrid approach based on deep learning architectures for Persian text recognition. In this method, convolutional neural networks (CNNs) are used for feature extraction and recurrent neural networks (RNNs) for word recognition. The main advantage of this model is its ability to directly recognize Persian printed text without relying on complex preprocessing steps, such as letter segmentation. The proposed model is trained on a large and dedicated dataset, comprising over two million samples generated in five common Persian fonts. The model achieves an accuracy of 81 per cent in recognizing Persian letters and 60 per cent in recognizing words. The most common errors occur in words related to semi-spaces and signs.&lt;/span&gt;</Abstract>
			<OtherAbstract Language="FA">&lt;span lang=&quot;FA&quot; style=&quot;font-size: 10.0pt; mso-ansi-font-size: 9.0pt; line-height: 90%; font-family: &#039;B Zar&#039;; mso-bidi-language: FA;&quot;&gt;تشخیص خودکار متن فارسی&lt;/span&gt;&lt;span dir=&quot;LTR&quot; lang=&quot;FA&quot; style=&quot;font-size: 9.0pt; mso-bidi-font-size: 10.0pt; line-height: 90%; mso-bidi-language: FA;&quot;&gt; &lt;/span&gt;&lt;span lang=&quot;FA&quot; style=&quot;font-size: 10.0pt; mso-ansi-font-size: 9.0pt; line-height: 90%; font-family: &#039;B Zar&#039;; mso-bidi-language: FA;&quot;&gt;به‌دلیل ویژگی‌های یکتای خط فارسی از جمله ساختار پیوسته، اشتراک بالای ویژگی‌های بصری بین حروف، و تنوع بالای نوشتاری حروف با توجه به موقعیت آنان در کلمه‌&lt;/span&gt;&lt;span dir=&quot;LTR&quot; lang=&quot;FA&quot; style=&quot;font-size: 10.0pt; mso-ansi-font-size: 9.0pt; line-height: 90%; font-family: &#039;B Zar&#039;; mso-bidi-language: FA;&quot;&gt;‎&lt;/span&gt;&lt;span lang=&quot;FA&quot; style=&quot;font-size: 10.0pt; mso-ansi-font-size: 9.0pt; line-height: 90%; font-family: &#039;B Zar&#039;; mso-bidi-language: FA;&quot;&gt; همواره چالش‌برانگیز بوده است. هدف این پژوهش ارائه یک مدل نویسه‌خوانی نوری است که بتواند اسناد چاپی و علمی فارسی را که شامل پایان‌نامه‌ها، مقالات و کتب فارسی است، به متن قابل ویرایش تبدیل کند. این امر برای برچسب‌گذاری، فهرست‌بندی و بازیابی اطلاعات در پایگاه داده‌ها یک ضرورت محسوب می‌شود. این مقاله رویکردی ترکیبی مبتنی ‌بر معماری‌های یادگیری عمیق برای تشخیص متن فارسی ارائه می‌دهد. در این روش از شبکه‌های عصبی پیچشی برای استخراج ویژگی‌ها و از شبکه‌های عصبی بازگشتی برای تشخیص کلمات استفاده ‌می‌شود. &lt;/span&gt;&lt;span lang=&quot;AR-SA&quot; style=&quot;font-size: 10.0pt; mso-ansi-font-size: 9.0pt; line-height: 90%; font-family: &#039;B Zar&#039;;&quot;&gt;مزیت اصلی این مدل، توانایی آن در تشخیص مستقیم متن چاپی فارسی بدون نیاز به پیش‌پردازش‌های پیچیده مانند ناحیه‌بندی حروف است&lt;/span&gt;&lt;span dir=&quot;LTR&quot; style=&quot;font-size: 9.0pt; mso-bidi-font-size: 10.0pt; line-height: 90%; mso-bidi-language: FA;&quot;&gt;.&lt;/span&gt;&lt;span lang=&quot;FA&quot; style=&quot;font-size: 10.0pt; mso-ansi-font-size: 9.0pt; line-height: 90%; font-family: &#039;B Zar&#039;; mso-bidi-language: FA;&quot;&gt; مدل پیشنهادی با استفاده از یک مجموعه داده اختصاصی و بزرگ، شامل بیش از دو میلیون نمونه که با پنج فونت متداول فارسی تولید شده‌، آموزش داده شده است. مدل معرفی‌شده دقت 81 درصد در تشخیص حروف فارسی و 60 درصد در تشخیص کلمات دارد. عمده‌ترین خطاها در کلمات مرتبط با نیم‌فاصله و علائم بود.&lt;/span&gt;</OtherAbstract>
		<ObjectList>
			<Object Type="keyword">
			<Param Name="value">تشخیص کاراکتر نوری، حافظه طولانی کوتاه‌مدت، شبکه‌ عصبی بازگشتی، شبکه عصبی پیچشی</Param>
			</Object>
		</ObjectList>
<ArchiveCopySource DocType="pdf">https://jipm.irandoc.ac.ir/article_725228_b2bcadbc255e198c694db98fbc0f5e0a.pdf</ArchiveCopySource>
</Article>
</ArticleSet>
